Appearance
Fresh 2026
Gateway Models
Merge Gateway routes across 225 models from many providers. Any model works, including new releases; Gateway infers cost tiers from provider pricing.
Routing policies map requests to these models by provider, priority, performance, or ML complexity score. See Routing Policies and Gateway Routing Flow.
Available models
- Amazon Nova 2 Lite
- Amazon Nova Lite (US Cross-Region)
- Amazon Nova Micro (US Cross-Region)
- Amazon Nova Premier
- Amazon Nova Pro
- Amazon.Nova 2 Sonic V1:0
- ByteDance-Seed/UI-TARS-1.5-7B
- Claude Haiku 4.5 20251001
- Claude Opus 4
- Claude Opus 4 1 20250805
- Claude Opus 4 5 20251101
- Claude Opus 4 6
- Claude Opus 4.7
- Claude Sonnet 4 20250514
- Claude Sonnet 4 5 20250929
- Claude Sonnet 4 6
- Codestral
- Codestral 2508
- Command A 03 2025
- Command R 08 2024
- Command R Plus 08 2024
- Command R7B 12 2024
- Computer Use Preview
- DeepSeek R1 (0528)
- DeepSeek V3
- DeepSeek V3.2
- DeepSeek V4 Flash
- DeepSeek V4 Pro
- Deepseek V32
- Devstral 2512
- Devstral Latest
- Devstral Medium 2507
- Devstral Medium Latest
- Devstral Small 2507
- Dola Seed 2.0 Code (preview)
- Dola Seed 2.0 Lite
- Dola Seed 2.0 Mini
- Dola Seed 2.0 Pro
- GLM 4.7
- GLM 4.7 Flash
- GLM-4-32B-0414-128K
- GLM-4.5
- GLM-4.5-Air
- GLM-4.5-AirX
- GLM-4.5-X
- GLM-4.6
- GLM-4.7-FlashX
- GLM-5
- GLM-5-Turbo
- GPT-3.5 Turbo
- GPT-3.5 Turbo (0125)
- GPT-4
- GPT-4 Turbo
- GPT-4o
- GPT-4o (2024-08-06)
- GPT-4o (2024-11-20)
- GPT-4o Mini
- GPT-4o Mini (2024-07-18)
- GPT-5.5
- Gemini 2.5 Computer Use Preview 10 2025
- Gemini 2.5 Flash
- Gemini 2.5 Flash Lite
- Gemini 2.5 Pro
- Gemini 3 Flash Preview
- Gemini 3 Pro Preview
- Gemini 3.1 Flash Lite
- Gemini 3.1 Flash Lite Preview
- Gemini 3.1 Pro Preview
- Gemini 3.1 Pro Preview Customtools
- Gemini 3.5 Flash
- Gemini Flash Latest
- Gemini Flash Lite Latest
- Gemini Pro Latest
- Gemma 3 12B
- Gemma 3 27B
- Gemma 3 4B
- Glm 5
- Glm47
- Gpt 3.5 Turbo 1106
- Gpt 3.5 Turbo 16K
- Gpt 4 0613
- Gpt 4 Turbo 2024 04 09
- Gpt 4.1
- Gpt 4.1 2025 04 14
- Gpt 4.1 Mini
- Gpt 4.1 Mini 2025 04 14
- Gpt 4.1 Nano
- Gpt 4.1 Nano 2025 04 14
- Gpt 4O 2024 05 13
- Gpt 4O Mini Search Preview
- Gpt 4O Mini Search Preview 2025 03 11
- Gpt 4O Search Preview
- Gpt 4O Search Preview 2025 03 11
- Gpt 5
- Gpt 5 2025 08 07
- Gpt 5 Chat Latest
- Gpt 5 Mini
- Gpt 5 Mini 2025 08 07
- Gpt 5 Nano
- Gpt 5 Nano 2025 08 07
- Gpt 5 Search Api
- Gpt 5 Search Api 2025 10 14
- Gpt 5.1
- Gpt 5.1 2025 11 13
- Gpt 5.1 Chat Latest
- Gpt 5.2
- Gpt 5.2 2025 12 11
- Gpt 5.2 Chat Latest
- Gpt 5.3 Chat Latest
- Gpt 5.4
- Gpt 5.4 2026 03 05
- Gpt 5.4 Mini
- Gpt 5.4 Mini 2026 03 17
- Gpt 5.4 Nano
- Gpt 5.4 Nano 2026 03 17
- Gpt 5.5 2026 04 23
- Grok 3
- Grok 3 Mini
- Grok 4 0709
- Grok 4 1 Fast Non Reasoning
- Grok 4 1 Fast Reasoning
- Grok 4 Fast Non Reasoning
- Grok 4 Fast Reasoning
- Grok 4.20
- Grok 4.3
- Grok Code Fast 1
- Jamba 1.5 Large
- Jamba 1.5 Mini
- Kimi K2 0711 Preview
- Kimi K2 0905 Preview
- Kimi K2 Thinking
- Kimi K2 Thinking Turbo
- Kimi K2 Turbo Preview
- Kimi K2.5
- Kimi K2.6
- Llama 3 8B
- Llama 3.1 70B
- Llama 3.1 8B
- Llama 3.2 11B
- Llama 3.2 1B
- Llama 3.2 90B
- Llama 33 70B Fp8
- Llama 4 Maverick 17B
- Llama 4 Maverick Instruct Fp8
- Llama 4 Scout 17B
- Magistral Medium 2509
- Magistral Medium Latest
- Magistral Small Latest
- Meta.Llama3 70B Instruct V1:0
- MiniMax M2
- MiniMax M2.1
- MiniMax M2.5 Highspeed
- MiniMax M2.7
- MiniMax M2.7 Highspeed
- MiniMaxAI/MiniMax-M2.5
- Minimax M25
- Ministral 8B 2512
- Mistral Large
- Mistral Large 2411
- Mistral Large 2512
- Mistral Medium
- Mistral Medium 2505
- Mistral Nemo
- Mistral Small
- Model catalog
- Nemotron Nano 12B VL
- Nemotron Nano 3 30B
- Nemotron Nano 9B
- Nvidia.Nemotron Super 3 120B
- O1
- O1 2024 12 17
- O3
- O3 2025 04 16
- O3 Mini
- O3 Mini 2025 01 31
- O4 Mini
- O4 Mini 2025 04 16
- Open Mistral Nemo 2407
- Openai.Gpt Oss 120B 1:0
- Openai.Gpt Oss 20B 1:0
- Openai.Gpt Oss Safeguard 120B
- Openai.Gpt Oss Safeguard 20B
- Pixtral Large 2411
- Pixtral Large Latest
- Qwen 3 Next 80B Instruct
- Qwen.Qwen3 Next 80B A3B
- Qwen.Qwen3 Vl 235B A22B
- Qwen/Qwen2.5-VL-72B-Instruct
- Qwen/Qwen3-235B-A22B-Instruct-2507-tput
- Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8
- Qwen/Qwen3-Coder-Next
- Qwen/Qwen3-Next-80B-A3B-Instruct
- Qwen/Qwen3-VL-235B-A22B-Instruct
- Qwen/Qwen3-VL-8B-Instruct
- Qwen/Qwen3.5-35B-A3B
- Qwen/Qwen3.5-397B-A17B
- Qwen25 Vl 72B Instruct
- Qwen3 235B
- Qwen3 32B
- Qwen3 Coder 30B
- Qwen3.6 Plus
- Qwen35 397B A17B
- Qwen3P5 35B A3B
- Qwen3Vl 8B Instruct
- Titan Embed Text V2
- Titan Text Large
- Twelvelabs.Pegasus 1 2 V1:0
- Ui Tars 1P5 7B
- Writer.Palmyra Vision 7B
- Zai.Glm 5
- arcee-ai/Trinity-Large-Thinking
- deepseek-ai/DeepSeek-R1
- deepseek-ai/DeepSeek-V3.1
- google/gemma-4-26B-A4B-it
- google/gemma-4-31B-it
- meta-llama/Llama-3.3-70B-Instruct
- meta-llama/Llama-3.3-70B-Instruct-Turbo
- meta-llama/Llama-4-Maverick-17B-128E-Instruct
- nvidia/Nemotron-120B-A12B
- openai/gpt-oss-120b
- openai/gpt-oss-20b
- palmyra-x4
- palmyra-x5
- zai-org/GLM-4.7
- zai-org/GLM-5.1