Get started · Models

The model catalog.

Every model in the catalog, with its tier, context window, capabilities, and per-million-token list pricing. Request one by name with the model field. Every paid plan includes the whole catalog; the Free plan includes the open-model lineup. This is a snapshot; GET /v1/models is always the live source of truth.

sdn-opus-5
Frontier

The frontier ceiling — deepest reasoning and vision for the hardest problems.

Input /Mtok
$5
Output /Mtok
$25
1M ctxToolsReasoningVision
sdn-opus-4.7
Frontier

Prior Opus flagship — frontier reasoning, pinned in place.

Input /Mtok
$5
Output /Mtok
$25
1M ctxToolsReasoningVision
sdn-gpt-5.5
Frontier

Frontier generalist — broad, sharp, and steady on long tool chains.

Input /Mtok
$2.5
Output /Mtok
$15
400K ctxToolsReasoningVision
sdn-gpt-5.4
Frontier

The prior GPT-5 flagship — most of 5.5's strength for less.

Input /Mtok
$1.25
Output /Mtok
$10
400K ctxToolsReasoningVision
sdn-gemini-3-pro
Frontier

Frontier scale — a 1M-token window and strong multimodal reasoning.

Input /Mtok
$2
Output /Mtok
$12
1M ctxToolsReasoningVision
sdn-kimi-k3
Frontier

Moonshot's flagship — deep reasoning, vision, and a 1M-token window.

Input /Mtok
$3
Output /Mtok
$15
1M ctxToolsReasoningVision
sdn-deepseek-v4-pro
Frontier

DeepSeek's flagship — frontier-grade reasoning at a startling price.

Input /Mtok
$0.435
Output /Mtok
$0.87
1M ctxToolsReasoningVision
sdn-qwen3.7-max
Frontier

The current Qwen flagship — frontier scale with a 1M-token window.

Input /Mtok
$2.5
Output /Mtok
$7.5
1M ctxToolsReasoningVision
sdn-inkling
Frontier

Thinking Machines' frontier debut — text, images, and audio in one model.

Input /Mtok
$1
Output /Mtok
$4.05
1M ctxToolsReasoningVision
sdn-gpt-oss-120b
General purpose

Default. Strong general-purpose agent model with fast responses.

Input /Mtok
$0.35
Output /Mtok
$0.75
128K ctxToolsReasoningVision
sdn-sonnet-4.6
Premium

The premium ceiling — top reasoning and vision for the hardest work.

Input /Mtok
$3
Output /Mtok
$15
200K ctxToolsReasoningVision
sdn-sonnet-5
Premium

The newest Sonnet — near-frontier quality for everyday premium work.

Input /Mtok
$3
Output /Mtok
$15
200K ctxToolsReasoningVision
sdn-haiku-4.5
Premium

The fast Claude — premium-family quality at a fraction of the latency.

Input /Mtok
$1
Output /Mtok
$5
200K ctxToolsReasoningVision
sdn-grok-4.3
Premium

Deep reasoning over a very large window, at the low end of premium.

Input /Mtok
$1.25
Output /Mtok
$2.5
1M ctxToolsReasoningVision
sdn-glm-5.2
Coding

Highest-quality GLM. Flagship tuned for quality over raw speed — higher, variable latency.

Input /Mtok
$1.4
Output /Mtok
$4.4
131K ctxToolsReasoningVision
sdn-glm-5
Coding

Fast, reliable coding flagship for everyday heavy work.

Input /Mtok
$1
Output /Mtok
$3.2
200K ctxToolsReasoningVision
sdn-qwen3-coder-480b
Coding

Heavy coding flagship — large MoE built for complex code.

Input /Mtok
$0.45
Output /Mtok
$1.8
131K ctxToolsReasoningVision
sdn-kimi-k2.7-code
Coding

Coding-specialist flagship. Quality-first; higher, variable latency.

Input /Mtok
$0.95
Output /Mtok
$4
131K ctxToolsReasoningVision
sdn-qwen3-max
Coding

The prior Qwen flagship — a heavyweight that codes exceptionally well.

Input /Mtok
$1.2
Output /Mtok
$6
262K ctxToolsReasoningVision
sdn-qwen3-235b
Reasoning

Big reasoning model at a cheap-tier price — standout value.

Input /Mtok
$0.22
Output /Mtok
$0.88
262K ctxToolsReasoningVision
sdn-deepseek-v3.2
Reasoning

Strong reasoning. Best on open-ended analysis, not strict tool loops.

Input /Mtok
$0.57
Output /Mtok
$1.71
131K ctxToolsReasoningVision
sdn-kimi-k2-thinking
Reasoning

Extended reasoning; budget output tokens for its hidden chain-of-thought.

Input /Mtok
$0.6
Output /Mtok
$2.5
262K ctxToolsReasoningVision
sdn-minimax-m2.5
Reasoning

Cheap reasoning with a large context window.

Input /Mtok
$0.3
Output /Mtok
$1.2
197K ctxToolsReasoningVision
sdn-ernie-x1
Reasoning

Baidu's reasoning specialist — deep deliberation at a low price.

Input /Mtok
$0.28
Output /Mtok
$1.1
128K ctxToolsReasoningVision
sdn-hunyuan-t1
Reasoning

Tencent's reasoning model — strong long-form thinking, very cheap.

Input /Mtok
$0.14
Output /Mtok
$0.55
200K ctxToolsReasoningVision
sdn-step-3
Reasoning

StepFun's multimodal reasoner — thinks, sees, and calls tools.

Input /Mtok
$0.57
Output /Mtok
$1.42
256K ctxToolsReasoningVision
sdn-qwen3.5-397b
Reasoning

Qwen's biggest open-weight reasoner — flagship thinking, mid-tier price.

Input /Mtok
$0.6
Output /Mtok
$3.6
262K ctxToolsReasoningVision
sdn-nemotron-3-ultra
Reasoning

The top Nemotron — 550B of deliberate reasoning at a mid-tier price.

Input /Mtok
$0.6
Output /Mtok
$2.4
202K ctxToolsReasoningVision
sdn-nemotron-super-3-120b
General purpose

Strong all-round workhorse with a very large context.

Input /Mtok
$0.15
Output /Mtok
$0.65
262K ctxToolsReasoningVision
sdn-nemotron-3-120b
General purpose

Hybrid MoE, strong on multi-agent, 256K context.

Input /Mtok
$0.5
Output /Mtok
$1.5
256K ctxToolsReasoningVision
sdn-qwen3-next-80b
General purpose

Efficient workhorse — big context at a low price.

Input /Mtok
$0.14
Output /Mtok
$1.2
262K ctxToolsReasoningVision
sdn-gemma-4-31b
General purpose

Google's open workhorse — reasoning and a 256K window near the price floor.

Input /Mtok
$0.14
Output /Mtok
$0.4
256K ctxToolsReasoningVision
sdn-mistral-large-3-675b
General purpose

Large general-purpose model, fast and capable.

Input /Mtok
$0.5
Output /Mtok
$1.5
262K ctxToolsReasoningVision
sdn-ernie-5.1
General purpose

Baidu's flagship generalist — broad knowledge with vision.

Input /Mtok
$0.59
Output /Mtok
$2.65
128K ctxToolsReasoningVision
sdn-hunyuan-turbos
General purpose

Tencent's fast generalist — quick answers at a rock-bottom price.

Input /Mtok
$0.11
Output /Mtok
$0.28
200K ctxToolsReasoningVision
sdn-gemini-3-flash
General purpose

Google's fast frontier model — 1M context, multimodal, cheap.

Input /Mtok
$0.5
Output /Mtok
$3
1M ctxToolsReasoningVision
sdn-minimax-m3
General purpose

MiniMax's frontier agent model — 1M context and vision at a workhorse price.

Input /Mtok
$0.3
Output /Mtok
$1.2
1M ctxToolsReasoningVision
sdn-kimi-k2.6
General purpose

Kimi's vision generalist — thinks when asked, sees what you show it.

Input /Mtok
$0.95
Output /Mtok
$4
262K ctxToolsReasoningVision
sdn-glm-4.7
General purpose

The GLM workhorse — flagship instincts at an everyday price.

Input /Mtok
$0.6
Output /Mtok
$2.2
200K ctxToolsReasoningVision
sdn-qwen3.7-plus
General purpose

Qwen's balanced mid-tier — a 1M-token window at an everyday price.

Input /Mtok
$0.4
Output /Mtok
$1.6
1M ctxToolsReasoningVision
sdn-qwen3-vl-235b
General purpose

Affordable eyes — a 235B vision model at workhorse money.

Input /Mtok
$0.4
Output /Mtok
$1.6
262K ctxToolsReasoningVision
sdn-hunyuan-hy3
General purpose

Tencent's newest generalist — reasoning and tools near the price floor.

Input /Mtok
$0.14
Output /Mtok
$0.58
262K ctxToolsReasoningVision
sdn-llama-3.3-70b
General purpose

Meta's dependable open workhorse — a known quantity everywhere.

Input /Mtok
$0.9
Output /Mtok
$0.9
131K ctxToolsReasoningVision
sdn-qwen3-coder-30b
Everyday coding

Cheap coding offload for routine changes.

Input /Mtok
$0.15
Output /Mtok
$0.6
262K ctxToolsReasoningVision
sdn-qwen3-coder-next
Everyday coding

Balanced coding model with a large context.

Input /Mtok
$0.5
Output /Mtok
$1.2
262K ctxToolsReasoningVision
sdn-kimi-k2.5
Everyday coding

Fast coding-general model.

Input /Mtok
$0.6
Output /Mtok
$3
262K ctxToolsReasoningVision
sdn-devstral-2-123b
Everyday coding

Coding-specialist tuned for software tasks.

Input /Mtok
$0.4
Output /Mtok
$2
262K ctxToolsReasoningVision
sdn-minimax-m2.7
Everyday coding

MiniMax's coding workhorse — agentic edits at a budget rate.

Input /Mtok
$0.3
Output /Mtok
$1.2
205K ctxToolsReasoningVision
sdn-kat-coder-pro
Everyday coding

Kuaishou's coding specialist — direct edits, no reasoning overhead.

Input /Mtok
$0.3
Output /Mtok
$1.2
256K ctxToolsReasoningVision
sdn-gemma-4-26b
Long context

256K context for the price of a small model.

Input /Mtok
$0.1
Output /Mtok
$0.3
256K ctxToolsReasoningVision
sdn-longcat-2
Coding

Meituan's flagship MoE — a 1M-token window with an unusually deep cached-input discount.

Input /Mtok
$0.75
Output /Mtok
$2.95
1M ctxToolsReasoningVision
sdn-deepseek-v4-flash
Long context

The V4 line's fast sibling — a 1M-token window at cheap-tier prices.

Input /Mtok
$0.14
Output /Mtok
$0.28
1M ctxToolsReasoningVision
sdn-qwen3.5-flash
Long context

A million tokens of context at one flat, tiny price.

Input /Mtok
$0.1
Output /Mtok
$0.4
1M ctxToolsReasoningVision
sdn-mimo-v2.5
Long context

Xiaomi's efficiency play — 1M context and vision for pocket change.

Input /Mtok
$0.168
Output /Mtok
$0.336
1M ctxToolsReasoningVision
sdn-qwen-turbo
Fast

The catalog's cheapest chat tokens — quick, direct answers at volume.

Input /Mtok
$0.05
Output /Mtok
$0.2
98K ctxToolsReasoningVision
sdn-qwen3.8-max
Reasoning

Flagship reasoning with images and a near-million-token window.

Input /Mtok
$2
Output /Mtok
$6
992K ctxToolsReasoningVision
sdn-gpt-oss-20b
Fast

Smallest and fastest tier.

Input /Mtok
$0.2
Output /Mtok
$0.3
128K ctxToolsReasoningVision
sdn-glm-4.7-flash
Fast

Near the price floor — GLM-family quality in the budget tier.

Input /Mtok
$0.0605
Output /Mtok
$0.4
131K ctxToolsReasoningVision
sdn-nemotron-nano-3-30b
Fast

Cheapest tool + reasoning capable model.

Input /Mtok
$0.06
Output /Mtok
$0.24
262K ctxToolsReasoningVision
sdn-step-3.7-flash
Fast

StepFun's quick multimodal — sees, thinks, and answers fast for very little.

Input /Mtok
$0.2
Output /Mtok
$1.15
262K ctxToolsReasoningVision
sdn-bge-m3
Embeddings

BAAI's versatile embedding — multilingual, multi-granularity.

Input /Mtok
$0.02
Dimensions
1,024
8K max inEmbeddings
sdn-bge-base-en
Embeddings

The long-standing default English embedding.

Input /Mtok
$0.07
Dimensions
768
1K max inEmbeddings
sdn-bge-small-en
Embeddings

384 dimensions — the smallest index and the fastest search.

Input /Mtok
$0.03
Dimensions
384
1K max inEmbeddings
sdn-bge-large-en
Embeddings

The most accurate English embedding in the catalog.

Input /Mtok
$0.21
Dimensions
1,024
1K max inEmbeddings
sdn-embed-gemma-300m
Embeddings

Compact Gemma-family embedding with a 2K input window.

Input /Mtok
$0.02
Dimensions
768
2K max inEmbeddings
sdn-qwen3-embed-0.6b
Embeddings

Multilingual retrieval with an 8K input window.

Input /Mtok
$0.02
Dimensions
1,024
8K max inEmbeddings
sdn-plamo-embed-1b
Embeddings

Japanese-specialist embedding — the widest vector in the catalog.

Input /Mtok
$0.02
Dimensions
2,048
4K max inEmbeddings

Prices are USD per one million tokens — the public list rates your plan's included usage is measured at. Click any card for the model's full page: description, strengths, and capabilities.

Note

Call GET /v1/models to read the catalog programmatically — each entry is self-describing, including capabilities and pricing. The gateway is always authoritative over this snapshot.