The model catalog.
Every model in the catalog, with its tier, context window, capabilities, and per-million-token list pricing. Request one by name with the model field. Every plan includes the whole catalog. This is a snapshot; GET /v1/models is always the live source of truth.
lx1-opus-4.8The frontier ceiling — deepest reasoning and vision for the hardest problems.
lx1-opus-4.7Prior Opus flagship — frontier reasoning, pinned in place.
lx1-gpt-5.5Frontier generalist — broad, sharp, and steady on long tool chains.
lx1-gpt-5.4The prior GPT-5 flagship — most of 5.5's strength for less.
lx1-gemini-3-proFrontier scale — a 1M-token window and strong multimodal reasoning.
lx1-grok-4.3Fast frontier reasoning with a very large window.
lx1-gpt-oss-120bDefault. Strong general-purpose agent model with fast responses.
lx1-sonnet-4.6The premium ceiling — top reasoning and vision for the hardest work.
lx1-sonnet-5The newest Sonnet — near-frontier quality for everyday premium work.
lx1-haiku-4.5The fast Claude — premium-family quality at a fraction of the latency.
lx1-glm-5.2Highest-quality GLM. Flagship tuned for quality over raw speed — higher, variable latency.
lx1-glm-5Fast, reliable coding flagship for everyday heavy work.
lx1-qwen3-coder-480bHeavy coding flagship — large MoE built for complex code.
lx1-kimi-k2.7-codeCoding-specialist flagship. Quality-first; higher, variable latency.
lx1-qwen3-maxQwen's largest flagship — a heavyweight that codes exceptionally well.
lx1-qwen3-235bBig reasoning model at a cheap-tier price — standout value.
lx1-deepseek-v3.2Strong reasoning. Best on open-ended analysis, not strict tool loops.
lx1-kimi-k2-thinkingExtended reasoning; budget output tokens for its hidden chain-of-thought.
lx1-minimax-m2.5Cheap reasoning with a large context window.
lx1-ernie-x1Baidu's reasoning specialist — deep deliberation at a low price.
lx1-hunyuan-t1Tencent's reasoning model — strong long-form thinking, very cheap.
lx1-step-3StepFun's multimodal reasoner — thinks, sees, and calls tools.
lx1-nemotron-super-3-120bStrong all-round workhorse with a very large context.
lx1-nemotron-3-120bHybrid MoE, strong on multi-agent, 256K context.
lx1-qwen3-next-80bEfficient workhorse — big context at a low price.
lx1-mistral-large-3-675bLarge general-purpose model, fast and capable.
lx1-ernie-5.1Baidu's flagship generalist — broad knowledge with vision.
lx1-hunyuan-turbosTencent's fast generalist — quick answers at a rock-bottom price.
lx1-gemini-3-flashGoogle's fast frontier model — 1M context, multimodal, cheap.
lx1-qwen3-coder-30bCheap coding offload for routine changes.
lx1-qwen3-coder-nextBalanced coding model with a large context.
lx1-kimi-k2.5Fast coding-general model.
lx1-devstral-2-123bCoding-specialist tuned for software tasks.
lx1-gemma-4-26b256K context for the price of a small model.
lx1-longcat-2Meituan's efficient MoE — a 1M-token window at a small-model price.
lx1-gpt-oss-20bSmallest and fastest tier.
lx1-glm-4.7-flashLowest price per token.
lx1-nemotron-nano-3-30bCheapest tool + reasoning capable model.
lx1-embed-3-largeOpenAI's strongest general embedding — high quality, 3072-dim.
lx1-embed-3-smallOpenAI's cheap embedding — most of the quality at a fraction of the price.
lx1-voyage-3.5Voyage's balanced retrieval embedding — strong quality, cheap.
lx1-voyage-code-3Voyage's code-specialized embedding — built for code search.
lx1-cohere-embed-4Cohere's multilingual embedding — long inputs, 100+ languages.
lx1-gemini-embedGoogle's embedding — top-ranked general retrieval quality.
lx1-qwen3-embed-8bQwen's open-weight embedding — multilingual, high-dimensional.
lx1-bge-m3BAAI's versatile embedding — multilingual, multi-granularity.
lx1-jina-v3Jina's multilingual embedding — task-tuned, long inputs.
lx1-mistral-embedMistral's embedding — solid retrieval quality, simple to adopt.
lx1-nomic-embedNomic's fully-open embedding — long inputs, very cheap.
Prices are USD per one million tokens — the public list rates your plan's included usage is measured at. Click any card for the model's full page: description, strengths, and capabilities.
Call GET /v1/models to read the catalog programmatically — each entry is self-describing, including capabilities and pricing. The gateway is always authoritative over this snapshot.