Data source: models.dev. Providers may change prices and availability; check the provider's information before use. models.dev
Mistral Small 4 · Opper
Model ID: opper/mistral/mistral-small-2603. Context window: 256000 tokens. Input / output per million tokens: $0.15 / $0.6.
Mistral Small 4 · EmpirioLabs AI
Model ID: empiriolabs/mistral-small-4. Context window: 256000 tokens. Input / output per million tokens: $0.15 / $0.6.
Mistral Small 4 · Kilo Gateway
Model ID: kilo/mistralai/mistral-small-2603. Context window: 262144 tokens. Input / output per million tokens: $0.15 / $0.6.
GLM-5-Turbo · Kilo Gateway
Model ID: kilo/z-ai/glm-5-turbo. Context window: 202752 tokens. Input / output per million tokens: $1.2 / $4.
GLM-5 Turbo · Merge Gateway
Model ID: merge-gateway/zai/glm-5-turbo. Context window: 200000 tokens. Input / output per million tokens: $1.2 / $4.
Mistral Small (latest) · Merge Gateway
Model ID: merge-gateway/mistral/mistral-small-latest. Context window: 256000 tokens. Input / output per million tokens: $0.15 / $0.6.
Mistral Small 4 · OpenRouter
Model ID: openrouter/mistralai/mistral-small-2603. Context window: 262144 tokens. Input / output per million tokens: $0.15 / $0.6.
GLM-5-Turbo · OpenRouter
Model ID: openrouter/z-ai/glm-5-turbo. Context window: 202752 tokens. Input / output per million tokens: $1.2 / $4.
GLM-5-Turbo · Ofox
Model ID: ofox/z-ai/glm-5-turbo. Context window: 200000 tokens. Input / output per million tokens: $1.2 / $4.
glm-5-turbo · 302.AI
Model ID: 302ai/glm-5-turbo. Context window: 200000 tokens. Input / output per million tokens: $0.72 / $3.2.
grok-4.20-beta-0309-reasoning · 302.AI
Model ID: 302ai/grok-4.20-beta-0309-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
GLM-5-Turbo · Z.AI
Model ID: zai/glm-5-turbo. Context window: 200000 tokens. Input / output per million tokens: $1.2 / $4.
Mistral Small (latest) · Mistral
Model ID: mistral/mistral-small-latest. Context window: 256000 tokens. Input / output per million tokens: $0.15 / $0.6.
Mistral Small 4 · Mistral
Model ID: mistral/mistral-small-2603. Context window: 256000 tokens. Input / output per million tokens: $0.15 / $0.6.
GLM 5 Turbo · ZenMux
Model ID: zenmux/z-ai/glm-5-turbo. Context window: 200000 tokens. Input / output per million tokens: $0.73 / $3.19.
Mistral Small 4 · Cortecs
Model ID: cortecs/mistral-small-2603. Context window: 262144 tokens. Input / output per million tokens: $0.156 / $0.625.
GLM-5-Turbo · Cortecs
Model ID: cortecs/glm-5-turbo. Context window: 202752 tokens. Input / output per million tokens: $1.186 / $3.955.
GLM-5-Turbo · Eden AI
Model ID: edenai/zai/glm-5-turbo. Context window: 202752 tokens. Input / output per million tokens: $1.2 / $4.
Mistral Small (latest) · Eden AI
Model ID: edenai/mistral/mistral-small-latest. Context window: 262144 tokens. Input / output per million tokens: $0.15 / $0.6.
Mistral Small 4 · Eden AI
Model ID: edenai/mistral/mistral-small-2603. Context window: 262144 tokens. Input / output per million tokens: $0.15 / $0.6.
Mistral Small 4 119B · Regolo AI
Model ID: regolo-ai/mistral-small-4-119b. Context window: 256000 tokens. Input / output per million tokens: $0.75 / $3.
GLM 5 Turbo · Venice AI
Model ID: venice/z-ai-glm-5-turbo. Context window: 200000 tokens. Input / output per million tokens: $1.2 / $4.
Body Builder (beta) · Kilo Gateway
Model ID: kilo/openrouter/bodybuilder. Context window: 128000 tokens. Input / output per million tokens: $0 / $0.
Auto Router · Kilo Gateway
Model ID: kilo/openrouter/auto. Context window: 2000000 tokens. Input / output per million tokens: $0 / $0.
Grok-4.20-Multi-Agent · Poe
Model ID: poe/xai/grok-4.20-multi-agent. Context window: 128000 tokens. Input / output per million tokens: $2 / $6.
Qwen3.5 9B · CrofAI
Model ID: crof/qwen3.5-9b. Context window: 262144 tokens. Input / output per million tokens: $0.04 / $0.15.
GPT-5.4-Mini · Poe
Model ID: poe/openai/gpt-5.4-mini. Context window: 400000 tokens. Input / output per million tokens: $0.68 / $4.
Grok 4.20 · Venice AI
Model ID: venice/grok-4-20. Context window: 2000000 tokens. Input / output per million tokens: $1.42 / $2.83.
Grok 4.20 Multi-Agent · Venice AI
Model ID: venice/grok-4-20-multi-agent. Context window: 2000000 tokens. Input / output per million tokens: $1.42 / $2.83.
NVIDIA Nemotron 3 Super 120B (Public Preview) · DigitalOcean
Model ID: digitalocean/nvidia-nemotron-3-super-120b. Context window: 1000000 tokens. Input / output per million tokens: $0.3 / $0.65.
nemotron-3-super · Ollama Cloud
Model ID: ollama-cloud/nemotron-3-super. Context window: 262144 tokens. Input / output per million tokens: $0.015 / $0.6.
Nemotron 3 Super 120B A12B · Kenari
Model ID: kenari/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0 / $0.
Nemotron 3 Super 120B A12B (Free) · Kenari
Model ID: kenari/nemotron-3-super-120b-a12b:free. Context window: 262144 tokens. Input / output per million tokens: $0 / $0.
GPT-5.4-Nano · Poe
Model ID: poe/openai/gpt-5.4-nano. Context window: 400000 tokens. Input / output per million tokens: $0.18 / $1.1.
Nemotron-3-Super-120B-A12B · Nebius Token Factory
Model ID: nebius/nvidia/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.3 / $0.9.
Nemotron 3 Super 120B A12B · Pioneer
Model ID: pioneer/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-FP8. Context window: 256000 tokens. Input / output per million tokens: $0.09 / $0.45.
nvidia-nemotron-3-super-120b-a12b · Requesty
Model ID: requesty/nvidia-nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.1 / $0.5.
Nemotron 3 Super 120B A12B · Requesty
Model ID: requesty/nemotron-3-super-120b-a12b. Context window: 1048576 tokens. Input / output per million tokens: $0 / $0.
Nemotron 3 Super · Nvidia
Model ID: nvidia/nvidia/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.2 / $0.8.
Nemotron Super · Baseten
Model ID: baseten/nvidia/Nemotron-120B-A12B. Context window: 202800 tokens. Input / output per million tokens: $0.3 / $0.75.
NVIDIA Nemotron 3 Super 120B A12B · Vercel AI Gateway
Model ID: vercel/nvidia/nemotron-3-super-120b-a12b. Context window: 256000 tokens. Input / output per million tokens: $0.15 / $0.65.
Grok 4.20 Beta Reasoning · Vercel AI Gateway
Model ID: vercel/spacexai/grok-4.20-reasoning-beta. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Multi Agent Beta · Vercel AI Gateway
Model ID: vercel/spacexai/grok-4.20-multi-agent-beta. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Beta Non-Reasoning · Vercel AI Gateway
Model ID: vercel/spacexai/grok-4.20-non-reasoning-beta. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Nemotron 3 Super 120B · Perplexity Agent
Model ID: perplexity-agent/nvidia/nemotron-3-super-120b-a12b. Context window: 1000000 tokens. Input / output per million tokens: $0.25 / $2.5.
Nemotron 3 Super 120B A12B · Crusoe
Model ID: crusoe/nvidia/NVIDIA-Nemotron-3-Super-120B-A12B. Context window: 262144 tokens. Input / output per million tokens: $0.3 / $2.4.
Nemotron 3 Super 120B A12B · Kilo Gateway
Model ID: kilo/nvidia/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.08 / $0.45.
NVIDIA: Nemotron 3 Super (free) · Kilo Gateway
Model ID: kilo/nvidia/nemotron-3-super-120b-a12b:free. Context window: 262144 tokens. Input / output per million tokens: $0 / $0.
NVIDIA Nemotron 3 Super 120B A12B · Amazon Bedrock
Model ID: amazon-bedrock/nvidia.nemotron-super-3-120b. Context window: 262144 tokens. Input / output per million tokens: $0.15 / $0.65.
Nemotron 3 Super 120B A12B · OpenRouter
Model ID: openrouter/nvidia/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.08 / $0.45.
Nemotron 3 Super (free) · OpenRouter
Model ID: openrouter/nvidia/nemotron-3-super-120b-a12b:free. Context window: 262144 tokens. Input / output per million tokens: $0 / $0.
Nemotron 3 Super 120B A12B · Synthetic
Model ID: synthetic/hf:nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4. Context window: 262144 tokens. Input / output per million tokens: $0.3 / $1.
Nemotron 3 Super 120B · Cloudflare Workers AI
Model ID: cloudflare-workers-ai/@cf/nvidia/nemotron-3-120b-a12b. Context window: 256000 tokens. Input / output per million tokens: $0.5 / $1.5.
Nemotron 3 Super 120B A12B (Nebius) · Eden AI
Model ID: edenai/nebius/nvidia/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.3 / $0.9.
qwen3.5-2b · Requesty
Model ID: requesty/qwen3.5-2b. Context window: 262144 tokens. Input / output per million tokens: $0.02 / $0.1.
Grok 4.20 Reasoning · Vercel AI Gateway
Model ID: vercel/spacexai/grok-4.20-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Multi-Agent · Vercel AI Gateway
Model ID: vercel/spacexai/grok-4.20-multi-agent. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Non-Reasoning · Vercel AI Gateway
Model ID: vercel/spacexai/grok-4.20-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Reasoning) · Impossibl
Model ID: impossibl/xai/grok-4.20-0309-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Non-Reasoning) · Impossibl
Model ID: impossibl/xai/grok-4.20-0309-non-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Non-Reasoning (Vertex AI (OpenAI-compatible)) · LLM Gateway
Model ID: llmgateway-providers/vertex-openai/grok-4-20-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Reasoning (Vertex AI (OpenAI-compatible)) · LLM Gateway
Model ID: llmgateway-providers/vertex-openai/grok-4-20-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Beta Reasoning (0309) (xAI) · LLM Gateway
Model ID: llmgateway-providers/xai/grok-4-20-beta-0309-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 Beta Non-Reasoning (0309) (xAI) · LLM Gateway
Model ID: llmgateway-providers/xai/grok-4-20-beta-0309-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 (Reasoning) · Vertex
Model ID: google-vertex/xai/grok-4.20-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Non-Reasoning) · Vertex
Model ID: google-vertex/xai/grok-4.20-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
S2 Pro · Vercel AI Gateway
Model ID: vercel/fish-audio/s2-pro. Context window: 0 tokens. Input / output per million tokens: — / —.
Grok 4.20 (Reasoning) · DevPass (LLM Gateway)
Model ID: llmgateway/grok-4-20-beta-0309-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 (Non-Reasoning) · DevPass (LLM Gateway)
Model ID: llmgateway/grok-4-20-beta-0309-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 (Non-Reasoning) · DevPass (LLM Gateway)
Model ID: llmgateway/grok-4-20-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Reasoning) · DevPass (LLM Gateway)
Model ID: llmgateway/grok-4-20-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Reasoning) · Cloudflare AI Gateway
Model ID: cloudflare-ai-gateway/xai/grok-4.20-0309-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 (Non-Reasoning) · Cloudflare AI Gateway
Model ID: cloudflare-ai-gateway/xai/grok-4.20-0309-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 · Merge Gateway
Model ID: merge-gateway/xai/grok-4.20-0309-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Non-Reasoning · Merge Gateway
Model ID: merge-gateway/xai/grok-4.20-0309-non-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Reasoning) · Ofox
Model ID: ofox/x-ai/grok-4.20. Context window: 2000000 tokens. Input / output per million tokens: $4 / $12.
Grok 4.20 (Reasoning) · xAI
Model ID: xai/grok-4.20-0309-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 Multi-Agent · xAI
Model ID: xai/grok-4.20-multi-agent-0309. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Non-Reasoning) · xAI
Model ID: xai/grok-4.20-0309-non-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.2 Fast Non Reasoning · ZenMux
Model ID: zenmux/x-ai/grok-4.2-fast-non-reasoning. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.2 Fast · ZenMux
Model ID: zenmux/x-ai/grok-4.2-fast. Context window: 2000000 tokens. Input / output per million tokens: $2 / $6.
Grok 4.20 (Reasoning) · OCI Generative AI
Model ID: oci/xai.grok-4.20-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Non-Reasoning) · OCI Generative AI
Model ID: oci/xai.grok-4.20-non-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Reasoning) · Eden AI
Model ID: edenai/xai/grok-4.20-0309-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Grok 4.20 (Non-Reasoning) · Eden AI
Model ID: edenai/xai/grok-4.20-0309-non-reasoning. Context window: 1000000 tokens. Input / output per million tokens: $1.25 / $2.5.
Tencent HY 2.0 Think · Tencent Coding Plan (China)
Model ID: tencent-coding-plan/hunyuan-2.0-thinking. Context window: 131072 tokens. Input / output per million tokens: $0 / $0.
Hunyuan-T1 · Tencent Coding Plan (China)
Model ID: tencent-coding-plan/hunyuan-t1. Context window: 131072 tokens. Input / output per million tokens: $0 / $0.
Hunyuan-TurboS · Tencent Coding Plan (China)
Model ID: tencent-coding-plan/hunyuan-turbos. Context window: 131072 tokens. Input / output per million tokens: $0 / $0.
Auto · Tencent Coding Plan (China)
Model ID: tencent-coding-plan/tc-code-latest. Context window: 131072 tokens. Input / output per million tokens: $0 / $0.
Tencent HY 2.0 Instruct · Tencent Coding Plan (China)
Model ID: tencent-coding-plan/hunyuan-2.0-instruct. Context window: 131072 tokens. Input / output per million tokens: $0 / $0.
GPT 5.4 · NanoGPT
Model ID: nano-gpt/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
OpenAI GPT-5.4 Pro · DigitalOcean
Model ID: digitalocean/openai-gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
OpenAI GPT-5.4 · DigitalOcean
Model ID: digitalocean/openai-gpt-5.4. Context window: 400000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Vivgrid
Model ID: vivgrid/gpt-5.4. Context window: 400000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · FreeModel
Model ID: freemodel/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Impossibl
Model ID: impossibl/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Impossibl
Model ID: impossibl/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 (Azure) · LLM Gateway
Model ID: llmgateway-providers/azure/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro (Azure) · LLM Gateway
Model ID: llmgateway-providers/azure/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 (OpenAI) · LLM Gateway
Model ID: llmgateway-providers/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro (OpenAI) · LLM Gateway
Model ID: llmgateway-providers/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · AIHubMix
Model ID: aihubmix/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · OrcaRouter
Model ID: orcarouter/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · OrcaRouter
Model ID: orcarouter/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
gpt-5.4 · SAP AI Core
Model ID: sap-ai-core/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Azure Cognitive Services
Model ID: azure-cognitive-services/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Azure Cognitive Services
Model ID: azure-cognitive-services/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4-Pro · Poe
Model ID: poe/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $27 / $160.
GPT-5.4 · Pioneer
Model ID: pioneer/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
MiniMax-M2.5 · CloudFerro Sherlock
Model ID: cloudferro-sherlock/MiniMaxAI/MiniMax-M2.5. Context window: 196000 tokens. Input / output per million tokens: $0.3 / $1.2.
GPT-5.4 · UnoRouter
Model ID: unorouter/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $1.8 / $10.8.
GPT-5.4 · UnoRouter
Model ID: unorouter/gpt-5.4:free. Context window: 1050000 tokens. Input / output per million tokens: $0 / $0.
GPT-5.4 (EU) · Requesty
Model ID: requesty/gpt-5.4@eu. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Requesty
Model ID: requesty/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.75 / $16.5.
GPT-5.4 Pro · Requesty
Model ID: requesty/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · DaoXE
Model ID: daoxe/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · CrossModel
Model ID: crossmodel/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.5 · FrogBot
Model ID: frogbot/gpt-5-5. Context window: 272000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Xpersona
Model ID: xpersona/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $0.75 / $6.
GPT 5.4 · Vercel AI Gateway
Model ID: vercel/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT 5.4 (Fast) · Vercel AI Gateway
Model ID: vercel/openai/gpt-5.4-fast. Context window: 1050000 tokens. Input / output per million tokens: $5 / $30.
GPT 5.4 Pro · Vercel AI Gateway
Model ID: vercel/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
Qwen 3.5 9B · Venice AI
Model ID: venice/qwen3-5-9b. Context window: 256000 tokens. Input / output per million tokens: $0.1 / $0.15.
GPT-5.4 · Venice AI
Model ID: venice/openai-gpt-54. Context window: 1000000 tokens. Input / output per million tokens: $3.13 / $18.8.
GPT-5.4 Pro · Venice AI
Model ID: venice/openai-gpt-54-pro. Context window: 1000000 tokens. Input / output per million tokens: $37.5 / $225.
GPT-5.4 · DevPass (LLM Gateway)
Model ID: llmgateway/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · DevPass (LLM Gateway)
Model ID: llmgateway/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · Cloudflare AI Gateway
Model ID: cloudflare-ai-gateway/openai/gpt-5.4. Context window: 1000000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Cloudflare AI Gateway
Model ID: cloudflare-ai-gateway/openai/gpt-5.4-pro. Context window: 1000000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · GitHub Copilot
Model ID: github-copilot/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Perplexity Agent
Model ID: perplexity-agent/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Opper
Model ID: opper/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Opper
Model ID: opper/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · Databricks
Model ID: databricks/databricks-gpt-5-4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
Qwen 3.5 9B (Q4_K_M) · Atomic Chat
Model ID: atomic-chat/Qwen3_5-9B-Q4_K_M. Context window: 32768 tokens. Input / output per million tokens: $0 / $0.
Qwen 3.5 9B (MLX 4-bit) · Atomic Chat
Model ID: atomic-chat/Qwen3_5-9B-MLX-4bit. Context window: 32768 tokens. Input / output per million tokens: $0 / $0.
GPT-5.4 · Kilo Gateway
Model ID: kilo/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Kilo Gateway
Model ID: kilo/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · Azure
Model ID: azure/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Azure
Model ID: azure/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · Amazon Bedrock
Model ID: amazon-bedrock/openai.gpt-5.4. Context window: 272000 tokens. Input / output per million tokens: $2.75 / $16.5.
GPT-5.4 · Merge Gateway
Model ID: merge-gateway/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Abacus
Model ID: abacus/gpt-5.4. Context window: 400000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · OpenCode Zen
Model ID: opencode/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · OpenCode Zen
Model ID: opencode/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · NEAR AI Cloud
Model ID: nearai/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · OpenRouter
Model ID: openrouter/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · OpenRouter
Model ID: openrouter/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · Model Oracle AI
Model ID: model-oracle-ai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: — / —.
GPT-5.4 · Ofox
Model ID: ofox/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2 / $12.
GPT-5.4 Pro · Ofox
Model ID: ofox/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $24 / $144.
GPT-5.4 · Neon
Model ID: neon/gpt-5-4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · OpenAI
Model ID: openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · OpenAI
Model ID: openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.4 · Snowflake Cortex
Model ID: snowflake-cortex/openai-gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: — / —.
gpt-5.4 · 302.AI
Model ID: 302ai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · AI-ROUTER
Model ID: ai-router/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 · Cortecs
Model ID: cortecs/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.898 / $15.453.
GPT-5.4 · AnyAPI
Model ID: anyapi/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: — / —.
Agentic Chat (GPT-5.4) · GitLab Duo
Model ID: gitlab/duo-chat-gpt-5-4. Context window: 1050000 tokens. Input / output per million tokens: $0 / $0.
GPT-5.4 · Eden AI
Model ID: edenai/openai/gpt-5.4. Context window: 1050000 tokens. Input / output per million tokens: $2.5 / $15.
GPT-5.4 Pro · Eden AI
Model ID: edenai/openai/gpt-5.4-pro. Context window: 1050000 tokens. Input / output per million tokens: $30 / $180.
GPT-5.3-Codex-Spark · Poe
Model ID: poe/openai/gpt-5.3-codex-spark. Context window: 128000 tokens. Input / output per million tokens: $0 / $0.
Kling v3.0 Motion Control · Vercel AI Gateway
Model ID: vercel/klingai/kling-v3.0-motion-control. Context window: 0 tokens. Input / output per million tokens: — / —.
Inception: Mercury 2 · Kilo Gateway
Model ID: kilo/inception/mercury-2. Context window: 128000 tokens. Input / output per million tokens: $0.25 / $0.75.
Mercury 2 · OpenRouter
Model ID: openrouter/inception/mercury-2. Context window: 128000 tokens. Input / output per million tokens: $0.25 / $0.75.
Solar Pro 3 · NanoGPT
Model ID: nano-gpt/upstage/solar-pro-3. Context window: 131072 tokens. Input / output per million tokens: $0.15 / $0.6.
Gemini 3.1 Flash Lite Preview · Vivgrid
Model ID: vivgrid/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
Qwen/Qwen3.5-4B · SiliconFlow (China)
Model ID: siliconflow-cn/Qwen/Qwen3.5-4B. Context window: 262144 tokens. Input / output per million tokens: $0 / $0.
Qwen/Qwen3.5-9B · SiliconFlow (China)
Model ID: siliconflow-cn/Qwen/Qwen3.5-9B. Context window: 262144 tokens. Input / output per million tokens: $0.22 / $1.74.
Qwen Image 2.0 Pro · Alibaba Token Plan
Model ID: alibaba-token-plan/qwen-image-2.0-pro. Context window: 8192 tokens. Input / output per million tokens: $0 / $0.
Qwen Image 2.0 · Alibaba Token Plan
Model ID: alibaba-token-plan/qwen-image-2.0. Context window: 8192 tokens. Input / output per million tokens: $0 / $0.
Gemini 3.1 Flash Lite Preview · OrcaRouter
Model ID: orcarouter/google/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
GPT-5.3-Instant · Poe
Model ID: poe/openai/gpt-5.3-instant. Context window: 128000 tokens. Input / output per million tokens: $1.6 / $13.
gliner-pii · Nvidia
Model ID: nvidia/nvidia/gliner-pii. Context window: 128000 tokens. Input / output per million tokens: $0 / $0.
GPT-5.3 Chat (latest) · Opper
Model ID: opper/openai/gpt-5.3-chat-latest. Context window: 128000 tokens. Input / output per million tokens: $1.75 / $14.
Gemini 3.1 Flash Lite Preview · Databricks
Model ID: databricks/databricks-gemini-3-1-flash-lite. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
Gemini 3.1 Flash Lite Preview · Kilo Gateway
Model ID: kilo/google/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.125 / $0.75.
Gemini 3.1 Flash Lite Preview · Merge Gateway
Model ID: merge-gateway/google/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
GPT-5.3 Chat Latest · Merge Gateway
Model ID: merge-gateway/openai/gpt-5.3-chat-latest. Context window: 128000 tokens. Input / output per million tokens: $1.75 / $14.
Gemini 3.1 Flash Lite Preview · OpenRouter
Model ID: openrouter/google/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
Gemini 3.1 Flash Lite Preview · Neon
Model ID: neon/gemini-3-1-flash-lite. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
Qwen/Qwen3.5-9B · SiliconFlow
Model ID: siliconflow/Qwen/Qwen3.5-9B. Context window: 262144 tokens. Input / output per million tokens: $0.1 / $0.15.
Gemini 3.1 Flash Lite Preview · 302.AI
Model ID: 302ai/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
GPT-5.3 Chat (latest) · 302.AI
Model ID: 302ai/gpt-5.3-chat-latest. Context window: 128000 tokens. Input / output per million tokens: $1.75 / $14.
Qwen3.5 9B · Together AI
Model ID: togetherai/Qwen/Qwen3.5-9B. Context window: 262144 tokens. Input / output per million tokens: $0.17 / $0.25.
Qwen Image 2.0 Pro · Alibaba Token Plan (China)
Model ID: alibaba-token-plan-cn/qwen-image-2.0-pro. Context window: 8192 tokens. Input / output per million tokens: $0 / $0.
Qwen Image 2.0 · Alibaba Token Plan (China)
Model ID: alibaba-token-plan-cn/qwen-image-2.0. Context window: 8192 tokens. Input / output per million tokens: $0 / $0.
Gemini 3.1 Flash Lite Preview · Eden AI
Model ID: edenai/google/gemini-3.1-flash-lite-preview. Context window: 1048576 tokens. Input / output per million tokens: $0.25 / $1.5.
Qwen3.5 4B · EmpirioLabs AI
Model ID: empiriolabs/qwen3-5-4b. Context window: 262144 tokens. Input / output per million tokens: $0.04 / $0.07.
Nvidia Nemotron 3 Super 120B Thinking · NanoGPT
Model ID: nano-gpt/nvidia/nemotron-3-super-120b-a12b:thinking. Context window: 262144 tokens. Input / output per million tokens: $0.05 / $0.25.
Nvidia Nemotron 3 Super 120B · NanoGPT
Model ID: nano-gpt/nvidia/nemotron-3-super-120b-a12b. Context window: 262144 tokens. Input / output per million tokens: $0.05 / $0.25.
Claude Sonnet Latest · NanoGPT
Model ID: nano-gpt/anthropic/claude-sonnet-latest. Context window: 1000000 tokens. Input / output per million tokens: $2 / $10.
GPT-OSS-20B · Regolo AI
Model ID: regolo-ai/gpt-oss-20b. Context window: 128000 tokens. Input / output per million tokens: $0.4 / $1.8.
Qwen3-Coder-Next · Regolo AI
Model ID: regolo-ai/qwen3-coder-next. Context window: 262144 tokens. Input / output per million tokens: $0.3 / $1.2.
Qwen-Image · Regolo AI
Model ID: regolo-ai/qwen-image. Context window: 8192 tokens. Input / output per million tokens: $0.5 / $2.
Step TTS 2 · StepFun (Global)
Model ID: stepfun-ai/step-tts-2. Context window: 0 tokens. Input / output per million tokens: — / —.
Step TTS 2 · StepFun (China)
Model ID: stepfun/step-tts-2. Context window: 0 tokens. Input / output per million tokens: — / —.
Transcribe-1 · Vercel AI Gateway
Model ID: vercel/fish-audio/transcribe-1. Context window: 0 tokens. Input / output per million tokens: — / —.
Standard Compute · Standard Compute
Model ID: standardcompute/standardcompute. Context window: 1000000 tokens. Input / output per million tokens: $0 / $0.