Seshat AIDocumentation

Documentation / Reference

Providers and models

Every provider Seshat supports, with the models it knows, their context window and their capabilities.

This page is generated from the provider registry of the runtime. How to set keys and endpoints is in Configuration, and the code side in Providers and authentication.

Providers

Provider IDServiceAuthenticationVisionPrompt cache
anthropicAnthropicapi_keyYesYes
openaiOpenAIapi_keyYesYes
mistralMistral AIapi_keyYesNo
geminiGoogle Geminiapi_keyYesNo
deepseekDeepSeekapi_keyYesNo
ollamaOllamanoneNoNo
openrouterOpenRouterapi_keyYesNo
codexCodexoauthNoNo
z-aiZ.aiapi_keyYesYes
minimaxMiniMaxapi_keyYesYes
bedrockAWS BedrockawsYesYes
vertexGoogle Cloud Vertex AIgcpYesYes
foundryAzure AI Foundryapi_keyYesYes
workers-aiCloudflare Workers AIapi_keyNoNo
opencodeOpenCode Zenapi_keyYesNo
kimiKimi (Moonshot AI)api_keyYesYes

Models

Anthropic (anthropic)

Model IDContextMax outputPrompt cache
claude-sonnet-4-20250514200K64KYes
claude-3-5-sonnet-20241022200K8KYes
claude-3-5-haiku-20241022200K8KYes

OpenAI (openai)

Model IDContextMax outputPrompt cache
gpt-5.5272K32KYes
gpt-5.4-mini272K16KYes
gpt-4o128K16KYes

Mistral AI (mistral)

Model IDContextMax outputPrompt cache
mistral-large-latest131K16KNo
mistral-small-latest131K8KNo
open-mistral-7b32K4KNo

Google Gemini (gemini)

Model IDContextMax outputPrompt cache
gemini-2.0-pro2M8KNo
gemini-2.0-flash1M8KNo
gemini-1.5-flash1M8KNo

DeepSeek (deepseek)

Model IDContextMax outputPrompt cache
deepseek-v4-flash1M393KNo
deepseek-v4-pro1M393KNo
deepseek-v4-flash-vision-exp1M393KNo

OpenRouter (openrouter)

Model IDContextMax outputPrompt cache
anthropic/claude-3.5-sonnet200K8KNo
openai/gpt-4o128K16KNo
deepseek/deepseek-r164K32KNo

Codex (codex)

Model IDContextMax outputPrompt cache
gpt-5.6-sol200K100KNo
gpt-5.5200K100KNo
gpt-5.4-mini272K16KNo

Z.ai (z-ai)

Model IDContextMax outputPrompt cache
glm-5.1200K128KYes
glm-4.71M8KNo
glm-4.5200K8KNo

MiniMax (minimax)

Model IDContextMax outputPrompt cache
MiniMax-M2.7204K128KYes
MiniMax-M2.5204K128KYes
MiniMax-M2.1204K8KNo

AWS Bedrock (bedrock)

Model IDContextMax outputPrompt cache
anthropic.claude-3-5-sonnet-20241022200K8KYes
anthropic.claude-3-5-haiku-20241022200K8KYes
anthropic.claude-3-opus-20240229200K4KYes

Google Cloud Vertex AI (vertex)

Model IDContextMax outputPrompt cache
claude-3-5-sonnet@20241022200K8KYes
claude-3-5-haiku@20241022200K8KYes
claude-3-opus@20240229200K4KYes

Azure AI Foundry (foundry)

Model IDContextMax outputPrompt cache
claude-3-5-sonnet-20241022200K8KYes
claude-3-5-haiku-20241022200K8KYes
claude-3-opus-20240229200K4KYes

Cloudflare Workers AI (workers-ai)

Model IDContextMax outputPrompt cache
@cf/meta/llama-3.1-70b-instruct128K4KNo
@cf/deepseek-ai/deepseek-r146K4KNo
@cf/qwen/qwen2.5-coder-7b32K4KNo

OpenCode Zen (opencode)

Model IDContextMax outputPrompt cache
claude-sonnet-4200K64KNo
gpt-5.3-codex200K100KNo
glm-5.1200K128KNo

Kimi (Moonshot AI) (kimi)

Model IDContextMax outputPrompt cache
kimi-k31M1MYes
kimi-k2.7-code262K262KYes
kimi-k2.7-code-highspeed262K262KYes
kimi-k2.6262K262KYes
kimi-k2.5262K262KYes

Updated on 2026-10-07