Model catalog

22 models. 9 providers. One key.

These are the live specs Gate's Smart Router reads on every request — list price per million tokens, context window, capabilities and routing tier. Call any model by its native name; if a provider isn't connected yet, Gate transparently serves the nearest equivalent and stamps the substitution on your receipt.

routable today cataloged — served via equivalent until connected

Showing 22 of 22 models

Alibaba

Qwen3 235B

qwen/qwen3-235b

via equivalent

$0.22

$ / 1M in

$0.88

$ / 1M out

262K

context

balancedquality 81~1686 msTool callingStreamingJSON modeVisionOpen weights

Also answers to: qwen3-235b, qwen-max, qwen2.5-72b, qwen-plus

Until Alibaba is connected, Gate serves this as google/gemini-3.7-flash and marks the receipt.

Anthropic

Claude Haiku 4.5

anthropic/claude-haiku-4.5

via equivalent

$1.00

$ / 1M in

$5.00

$ / 1M out

200K

context

fastquality 72~707 msTool callingStreamingJSON modeVision

Also answers to: claude-haiku-4.5, claude-3-5-haiku, claude-3-5-haiku-20241022, claude-3-haiku, +1 more

Until Anthropic is connected, Gate serves this as google/gemini-2.5-flash and marks the receipt.

Claude Sonnet 4.5

anthropic/claude-sonnet-4.5

via equivalent

$3.00

$ / 1M in

$15.00

$ / 1M out

200K

context

deepquality 93~4454 msTool callingStreamingJSON modeVisionExtended reasoning

Also answers to: claude-sonnet-4.5, claude-sonnet-4, claude-3-7-sonnet, claude-3-5-sonnet, +2 more

Until Anthropic is connected, Gate serves this as openai/gpt-5.6-sol and marks the receipt.

Claude Opus 4.1

anthropic/claude-opus-4.1

via equivalent

$15.00

$ / 1M in

$75.00

$ / 1M out

200K

context

deepquality 96~4499 msTool callingStreamingJSON modeVisionExtended reasoning

Also answers to: claude-opus-4.1, claude-opus-4, claude-3-opus, claude-opus

Until Anthropic is connected, Gate serves this as openai/gpt-5.6-sol and marks the receipt.

Cohere

Command A

cohere/command-a

via equivalent

$2.50

$ / 1M in

$10.00

$ / 1M out

256K

context

deepquality 79~4241 msTool callingStreamingJSON modeExtended reasoning

Also answers to: command-a, command-r-plus, command-r, command

Until Cohere is connected, Gate serves this as google/gemini-2.5-pro and marks the receipt.

DeepSeek

DeepSeek V3.2

deepseek/deepseek-v3.2

via equivalent

$0.28

$ / 1M in

$0.42

$ / 1M out

128K

context

balancedquality 78~1668 msTool callingStreamingJSON modeOpen weights

Also answers to: deepseek-chat, deepseek-v3, deepseek-v3.2, deepseek/deepseek-chat

Until DeepSeek is connected, Gate serves this as google/gemini-2.5-flash and marks the receipt.

DeepSeek R1

deepseek/deepseek-r1

via equivalent

$0.55

$ / 1M in

$2.19

$ / 1M out

128K

context

deepquality 86~4347 msTool callingStreamingJSON modeExtended reasoningOpen weights

Also answers to: deepseek-reasoner, deepseek-r1, deepseek/deepseek-r1

Until DeepSeek is connected, Gate serves this as google/gemini-2.5-pro and marks the receipt.

Google

Gemini 2.5 Flash Lite

google/gemini-2.5-flash-lite

routable

$0.10

$ / 1M in

$0.40

$ / 1M out

1M

context

fastquality 60~676 msTool callingStreamingJSON modeVision1M+ context

Also answers to: gemini-2.5-flash-lite, gemini-flash-lite, gemini-1.5-flash-8b

Gemini 2.5 Flash

google/gemini-2.5-flash

routable

$0.30

$ / 1M in

$2.50

$ / 1M out

1M

context

balancedquality 76~1656 msTool callingStreamingJSON modeVision1M+ context

Also answers to: gemini-2.5-flash, gemini-1.5-flash, gemini-flash, gemini-2.0-flash

Gemini 3.7 Flash

google/gemini-3.7-flash

routable

$0.30

$ / 1M in

$2.50

$ / 1M out

1M

context

balancedquality 84~1704 msTool callingStreamingJSON modeVision1M+ context

Also answers to: gemini-3.7-flash, gemini-3.5-flash, gemini-3-flash, gemini-3.6-flash

Gemini 2.5 Pro

google/gemini-2.5-pro

routable

$1.25

$ / 1M in

$10.00

$ / 1M out

2M

context

deepquality 90~4408 msTool callingStreamingJSON modeVisionExtended reasoning1M+ context

Also answers to: gemini-2.5-pro, gemini-1.5-pro, gemini-pro, gemini-3.1-pro

Meta

Llama 3.3 70B

meta/llama-3.3-70b

via equivalent

$0.23

$ / 1M in

$0.40

$ / 1M out

128K

context

balancedquality 70~1620 msTool callingStreamingJSON modeOpen weights

Also answers to: llama-3.3-70b, llama-3.3-70b-instruct, meta-llama/llama-3.3-70b-instruct, llama-3.1-70b, +1 more

Until Meta is connected, Gate serves this as google/gemini-2.5-flash and marks the receipt.

Llama 3.1 8B

meta/llama-3.1-8b

via equivalent

$0.02

$ / 1M in

$0.03

$ / 1M out

128K

context

fastquality 46~640 msTool callingStreamingJSON modeOpen weights

Also answers to: llama-3.1-8b, llama-3.1-8b-instruct, llama3-8b

Until Meta is connected, Gate serves this as google/gemini-2.5-flash-lite and marks the receipt.

Llama 4 Maverick

meta/llama-4-maverick

via equivalent

$0.27

$ / 1M in

$0.85

$ / 1M out

1M

context

deepquality 80~4256 msTool callingStreamingJSON modeVisionExtended reasoning1M+ contextOpen weights

Also answers to: llama-4-maverick, llama-4, llama-4-scout

Until Meta is connected, Gate serves this as google/gemini-3.7-flash and marks the receipt.

Mistral

Mistral Small 3.2

mistral/mistral-small-3.2

via equivalent

$0.10

$ / 1M in

$0.30

$ / 1M out

128K

context

fastquality 55~663 msTool callingStreamingJSON modeOpen weights

Also answers to: mistral-small, mistral-small-latest, ministral-8b, open-mistral-nemo

Until Mistral is connected, Gate serves this as google/gemini-2.5-flash-lite and marks the receipt.

Mistral Large 2

mistral/mistral-large-2

via equivalent

$2.00

$ / 1M in

$6.00

$ / 1M out

128K

context

deepquality 82~4286 msTool callingStreamingJSON modeExtended reasoningOpen weights

Also answers to: mistral-large, mistral-large-latest, mistral-medium, codestral

Until Mistral is connected, Gate serves this as google/gemini-2.5-pro and marks the receipt.

OpenAI

GPT-5 Nano

openai/gpt-5-nano

routable

$0.05

$ / 1M in

$0.40

$ / 1M out

400K

context

fastquality 62~681 msTool callingStreamingJSON modeVision

Also answers to: gpt-5-nano, gpt-4o-mini, gpt-4.1-nano, gpt-3.5-turbo, +1 more

GPT-5 Mini

openai/gpt-5-mini

routable

$0.25

$ / 1M in

$2.00

$ / 1M out

400K

context

balancedquality 74~1644 msTool callingStreamingJSON modeVision

Also answers to: gpt-5-mini, gpt-4.1-mini, gpt-4o, chatgpt-4o-latest, +1 more

GPT-5.5

openai/gpt-5.5

routable

$1.25

$ / 1M in

$10.00

$ / 1M out

400K

context

deepquality 92~4438 msTool callingStreamingJSON modeVisionExtended reasoning

Also answers to: gpt-5.5, gpt-5, gpt-4.1, gpt-4-turbo, +2 more

GPT-5.6 Sol

openai/gpt-5.6-sol

routable

$1.25

$ / 1M in

$10.00

$ / 1M out

400K

context

deepquality 95~4484 msTool callingStreamingJSON modeVisionExtended reasoning

Also answers to: gpt-5.6-sol, gpt-5.6, gpt-5.4, gpt-5.2

xAI

Grok 4

xai/grok-4

via equivalent

$3.00

$ / 1M in

$15.00

$ / 1M out

256K

context

deepquality 91~4423 msTool callingStreamingJSON modeVisionExtended reasoning

Also answers to: grok-4, grok-3, grok-beta, grok

Until xAI is connected, Gate serves this as openai/gpt-5.5 and marks the receipt.

Grok 4 Fast

xai/grok-4-fast

via equivalent

$0.20

$ / 1M in

$0.50

$ / 1M out

2M

context

fastquality 68~697 msTool callingStreamingJSON modeVision1M+ context

Also answers to: grok-4-fast, grok-3-mini, grok-2-mini

Until xAI is connected, Gate serves this as google/gemini-2.5-flash and marks the receipt.

Route across all of them today

Set cost, latency and capability rules once — Gate picks the cheapest model that fits every prompt and shows the savings on your dashboard.

Access Now (beta)