Get started · Models

The model catalog.

Every model in the catalog, with its tier, context window, capabilities, and per-million-token list pricing. Request one by name with the model field. Every plan includes the whole catalog. This is a snapshot; GET /v1/models is always the live source of truth.

lx1-opus-4.8
Frontier

The frontier ceiling — deepest reasoning and vision for the hardest problems.

Input /Mtok
$15
Output /Mtok
$75
200K ctxToolsReasoningVision
lx1-opus-4.7
Frontier

Prior Opus flagship — frontier reasoning, pinned in place.

Input /Mtok
$15
Output /Mtok
$75
200K ctxToolsReasoningVision
lx1-gpt-5.5
Frontier

Frontier generalist — broad, sharp, and steady on long tool chains.

Input /Mtok
$2.5
Output /Mtok
$15
400K ctxToolsReasoningVision
lx1-gpt-5.4
Frontier

The prior GPT-5 flagship — most of 5.5's strength for less.

Input /Mtok
$1.25
Output /Mtok
$10
400K ctxToolsReasoningVision
lx1-gemini-3-pro
Frontier

Frontier scale — a 1M-token window and strong multimodal reasoning.

Input /Mtok
$2
Output /Mtok
$12
1M ctxToolsReasoningVision
lx1-grok-4.3
Frontier

Fast frontier reasoning with a very large window.

Input /Mtok
$1.25
Output /Mtok
$2.5
1M ctxToolsReasoningVision
lx1-gpt-oss-120b
Workhorse

Default. Strong general-purpose agent model with fast responses.

Input /Mtok
$0.35
Output /Mtok
$0.75
128K ctxToolsReasoningVision
lx1-sonnet-4.6
Premium

The premium ceiling — top reasoning and vision for the hardest work.

Input /Mtok
$3
Output /Mtok
$15
200K ctxToolsReasoningVision
lx1-sonnet-5
Premium

The newest Sonnet — near-frontier quality for everyday premium work.

Input /Mtok
$3
Output /Mtok
$15
200K ctxToolsReasoningVision
lx1-haiku-4.5
Premium

The fast Claude — premium-family quality at a fraction of the latency.

Input /Mtok
$1
Output /Mtok
$5
200K ctxToolsReasoningVision
lx1-glm-5.2
Coding flagship

Highest-quality GLM. Flagship tuned for quality over raw speed — higher, variable latency.

Input /Mtok
$1
Output /Mtok
$3.2
131K ctxToolsReasoningVision
lx1-glm-5
Coding flagship

Fast, reliable coding flagship for everyday heavy work.

Input /Mtok
$1
Output /Mtok
$3.2
200K ctxToolsReasoningVision
lx1-qwen3-coder-480b
Coding flagship

Heavy coding flagship — large MoE built for complex code.

Input /Mtok
$0.45
Output /Mtok
$1.8
131K ctxToolsReasoningVision
lx1-kimi-k2.7-code
Coding flagship

Coding-specialist flagship. Quality-first; higher, variable latency.

Input /Mtok
$0.6
Output /Mtok
$3
131K ctxToolsReasoningVision
lx1-qwen3-max
Coding flagship

Qwen's largest flagship — a heavyweight that codes exceptionally well.

Input /Mtok
$1.2
Output /Mtok
$6
262K ctxToolsReasoningVision
lx1-qwen3-235b
Reasoning

Big reasoning model at a cheap-tier price — standout value.

Input /Mtok
$0.22
Output /Mtok
$0.88
262K ctxToolsReasoningVision
lx1-deepseek-v3.2
Reasoning

Strong reasoning. Best on open-ended analysis, not strict tool loops.

Input /Mtok
$0.62
Output /Mtok
$1.85
164K ctxToolsReasoningVision
lx1-kimi-k2-thinking
Reasoning

Extended reasoning; budget output tokens for its hidden chain-of-thought.

Input /Mtok
$0.6
Output /Mtok
$2.5
262K ctxToolsReasoningVision
lx1-minimax-m2.5
Reasoning

Cheap reasoning with a large context window.

Input /Mtok
$0.3
Output /Mtok
$1.2
197K ctxToolsReasoningVision
lx1-ernie-x1
Reasoning

Baidu's reasoning specialist — deep deliberation at a low price.

Input /Mtok
$0.28
Output /Mtok
$1.1
128K ctxToolsReasoningVision
lx1-hunyuan-t1
Reasoning

Tencent's reasoning model — strong long-form thinking, very cheap.

Input /Mtok
$0.14
Output /Mtok
$0.55
200K ctxToolsReasoningVision
lx1-step-3
Reasoning

StepFun's multimodal reasoner — thinks, sees, and calls tools.

Input /Mtok
$0.57
Output /Mtok
$1.42
256K ctxToolsReasoningVision
lx1-nemotron-super-3-120b
Workhorse

Strong all-round workhorse with a very large context.

Input /Mtok
$0.15
Output /Mtok
$0.65
262K ctxToolsReasoningVision
lx1-nemotron-3-120b
Workhorse

Hybrid MoE, strong on multi-agent, 256K context.

Input /Mtok
$0.5
Output /Mtok
$1.5
256K ctxToolsReasoningVision
lx1-qwen3-next-80b
Workhorse

Efficient workhorse — big context at a low price.

Input /Mtok
$0.14
Output /Mtok
$1.2
262K ctxToolsReasoningVision
lx1-mistral-large-3-675b
Workhorse

Large general-purpose model, fast and capable.

Input /Mtok
$0.5
Output /Mtok
$1.5
262K ctxToolsReasoningVision
lx1-ernie-5.1
Workhorse

Baidu's flagship generalist — broad knowledge with vision.

Input /Mtok
$0.59
Output /Mtok
$2.65
128K ctxToolsReasoningVision
lx1-hunyuan-turbos
Workhorse

Tencent's fast generalist — quick answers at a rock-bottom price.

Input /Mtok
$0.11
Output /Mtok
$0.28
200K ctxToolsReasoningVision
lx1-gemini-3-flash
Workhorse

Google's fast frontier model — 1M context, multimodal, cheap.

Input /Mtok
$0.5
Output /Mtok
$3
1M ctxToolsReasoningVision
lx1-qwen3-coder-30b
Coding

Cheap coding offload for routine changes.

Input /Mtok
$0.15
Output /Mtok
$0.6
262K ctxToolsReasoningVision
lx1-qwen3-coder-next
Coding

Balanced coding model with a large context.

Input /Mtok
$0.5
Output /Mtok
$1.2
262K ctxToolsReasoningVision
lx1-kimi-k2.5
Coding

Fast coding-general model.

Input /Mtok
$0.6
Output /Mtok
$3
262K ctxToolsReasoningVision
lx1-devstral-2-123b
Coding

Coding-specialist tuned for software tasks.

Input /Mtok
$0.4
Output /Mtok
$2
262K ctxToolsReasoningVision
lx1-gemma-4-26b
Cheap · big ctx

256K context for the price of a small model.

Input /Mtok
$0.1
Output /Mtok
$0.3
256K ctxToolsReasoningVision
lx1-longcat-2
Cheap · big ctx

Meituan's efficient MoE — a 1M-token window at a small-model price.

Input /Mtok
$0.3
Output /Mtok
$1.2
1M ctxToolsReasoningVision
lx1-gpt-oss-20b
Cheap · fast

Smallest and fastest tier.

Input /Mtok
$0.2
Output /Mtok
$0.3
128K ctxToolsReasoningVision
lx1-glm-4.7-flash
Cheap · fast

Lowest price per token.

Input /Mtok
$0.0605
Output /Mtok
$0.4
131K ctxToolsReasoningVision
lx1-nemotron-nano-3-30b
Cheap · fast

Cheapest tool + reasoning capable model.

Input /Mtok
$0.06
Output /Mtok
$0.24
262K ctxToolsReasoningVision
lx1-embed-3-large
Embeddings

OpenAI's strongest general embedding — high quality, 3072-dim.

Input /Mtok
$0.13
Dimensions
3,072
8K max inEmbeddings
lx1-embed-3-small
Embeddings

OpenAI's cheap embedding — most of the quality at a fraction of the price.

Input /Mtok
$0.02
Dimensions
1,536
8K max inEmbeddings
lx1-voyage-3.5
Embeddings

Voyage's balanced retrieval embedding — strong quality, cheap.

Input /Mtok
$0.06
Dimensions
1,024
32K max inEmbeddings
lx1-voyage-code-3
Embeddings

Voyage's code-specialized embedding — built for code search.

Input /Mtok
$0.18
Dimensions
1,024
32K max inEmbeddings
lx1-cohere-embed-4
Embeddings

Cohere's multilingual embedding — long inputs, 100+ languages.

Input /Mtok
$0.12
Dimensions
1,536
128K max inEmbeddings
lx1-gemini-embed
Embeddings

Google's embedding — top-ranked general retrieval quality.

Input /Mtok
$0.15
Dimensions
3,072
2K max inEmbeddings
lx1-qwen3-embed-8b
Embeddings

Qwen's open-weight embedding — multilingual, high-dimensional.

Input /Mtok
$0.05
Dimensions
4,096
32K max inEmbeddings
lx1-bge-m3
Embeddings

BAAI's versatile embedding — multilingual, multi-granularity.

Input /Mtok
$0.02
Dimensions
1,024
8K max inEmbeddings
lx1-jina-v3
Embeddings

Jina's multilingual embedding — task-tuned, long inputs.

Input /Mtok
$0.02
Dimensions
1,024
8K max inEmbeddings
lx1-mistral-embed
Embeddings

Mistral's embedding — solid retrieval quality, simple to adopt.

Input /Mtok
$0.1
Dimensions
1,024
8K max inEmbeddings
lx1-nomic-embed
Embeddings

Nomic's fully-open embedding — long inputs, very cheap.

Input /Mtok
$0.01
Dimensions
768
8K max inEmbeddings

Prices are USD per one million tokens — the public list rates your plan's included usage is measured at. Click any card for the model's full page: description, strengths, and capabilities.

Note

Call GET /v1/models to read the catalog programmatically — each entry is self-describing, including capabilities and pricing. The gateway is always authoritative over this snapshot.