Model catalog
30 frontier models. 11 providers. One API.
Every id pinned and price-verified against the provider — never floating -latest aliases. The rate you see is the rate you pay, all-in. Retired models remap to successors — they never 404.
Routing aliases
Or skip model names entirely
Request a capability instead of a model. Each alias is a curated cross-provider chain with automatic failover, and your org's routing policy — data residency, provider and model blocks, PHI allowlists — is enforced per candidate. The response tells you what served via x-conduix-alias and x-conduix-model-served.
General
frontier-bestHighest-quality frontier model, cross-provider failover.GPT-5.5 → Claude Opus 4.8 → Gemini 3.1 Pro
frontier-fastFrontier-quality with production latency.Claude Sonnet 5 → GPT-5.4 → Gemini 3.5 Flash → Grok 4.3
frontier-budgetBest quality per dollar for high-volume workloads.GPT-5.4 mini → Claude Haiku 4.5 → Gemini 3.1 Flash-Lite → DeepSeek V4 Flash
enterpriseConservative chain: US frontier providers with enterprise SLAs.Claude Opus 4.8 → GPT-5.5 → GPT-5.4
Coding
code-premiumStrongest coding models available.GPT-5.3 Codex → Claude Opus 4.8 → Claude Sonnet 5
code-balancedProduction coding default — quality/cost balance.Claude Sonnet 5 → GPT-5.3 Codex → Codestral
code-fastLow-latency completions and quick edits.GPT-5.4 mini → Codestral → Claude Haiku 4.5
code-openOpen-weights coding models only.Kimi K2.7 Code (Together) → Qwen 3.7 Plus (Together)
agentic-codeLong-horizon agentic coding (tool loops, large diffs).GPT-5.3 Codex → Claude Sonnet 5 → Kimi K2.7 Code (Together)
Reasoning
reasoning-bestDeepest deliberate reasoning available.Claude Fable 5 → GPT-5.5 → DeepSeek V4 Pro
reasoning-fastThinking-capable models at interactive latency.Gemini 3.5 Flash → Grok 4.3 → GPT-5.4
reasoning-budgetChain-of-thought quality without frontier pricing.DeepSeek V4 Pro → DeepSeek V4 Flash → GPT-5.4 mini
Vision & documents
vision-bestStrongest image + document understanding.Gemini 3.1 Pro → GPT-5.5 → Claude Opus 4.8
vision-fastFast multimodal understanding.Gemini 3.5 Flash → GPT-5.4 mini → Gemma 4 31B (Together)
document-aiLong-document analysis (contracts, filings, PDFs).Gemini 3.1 Pro → Gemini 3.5 Flash → Claude Sonnet 5
Coming soon: ocr · speech-best · speech-fast · stt · tts · embedding-best · embedding-fast · image-best · image-fast · video-best · rerank-best — these aliases ship with their modality endpoints.
Browse the catalog
30 of 30 models
GPT-5.5
Frontiergpt-5.5OpenAI flagship. 1M context, adjustable reasoning effort, top-tier agentic work.
Claude Fable 5
Frontierclaude-fable-5Anthropic’s most capable model. Thinking always on — the gateway strips sampling parameters this model rejects.
Claude Opus 4.8
Frontierclaude-opus-4-8Anthropic flagship for complex writing, analysis, and agentic coding. 1M context.
Gemini 3.1 Pro
Frontiergemini-3.1-pro-previewGoogle’s top model — strongest multimodal + long-document reasoning. Google designates this tier "Preview".
Grok 4.3
Frontiergrok-4.3xAI flagship. 1M context; reasoning effort replaces the retired fast/mini tiers (set effort low for speed).
GPT-5.4
Premiumgpt-5.4OpenAI premium workhorse — 1M context at workhorse pricing.
GPT-5.3 Codex
Premiumgpt-5.3-codexOpenAI’s most capable agentic coding model.
Claude Sonnet 5
Premiumclaude-sonnet-5Anthropic workhorse — excellent coding + long-context. Introductory provider pricing until 2026-08-31 (operator repricing due 2026-09-01).
DeepSeek V4 Pro
PremiumOpen weightsapacdeepseek-v4-proFrontier-class open-weights reasoning at budget pricing. 1M context.
Mistral Medium 3.5
PremiumOpen weightseumistral-medium-2604Mistral’s most capable model — EU residency with per-request reasoning effort.
Kimi K2.6 (Together)
PremiumOpen weightsmoonshotai/Kimi-K2.6Open-weights frontier chat — top-tier agentic tool use.
Kimi K2.7 Code (Together)
PremiumOpen weightsmoonshotai/Kimi-K2.7-CodeCurrent open-weights coding flagship — agentic repo-scale work.
GLM 5.2 (Together)
PremiumOpen weightszai-org/GLM-5.2Open-weights frontier chat — strong tool calling + multilingual.
GLM 5.2 (Fireworks)
PremiumOpen weightsaccounts/fireworks/models/glm-5p2Same GLM 5.2 via Fireworks — 1M-context alternate route.
DeepSeek V4 Pro (Azure)
PremiumOpen weightsdeepseek-v4-pro-azureDeepSeek V4 Pro hosted on Azure AI Foundry — Global Standard tier, US region. Distinct from the direct `deepseek-v4-pro` route (APAC region, direct DeepSeek): same model, hosted in Microsoft-managed US infrastructure.
GPT-5.4 mini
Midgpt-5.4-miniCheap and capable — default for high-volume routine tasks.
Claude Haiku 4.5
Midclaude-haiku-4-5-20251001Fast, cheap Claude with 200k context. Good fit for classification + summarisation.
Gemini 3.5 Flash
Midgemini-3.5-flashFast thinking-capable Gemini. Great for long-context, high-throughput multimodal pipelines.
Mistral Large 3
MidOpen weightseumistral-large-2512Open-weights (Apache 2.0) EU flagship — strong general chat at mid-tier pricing.
Codestral
Mideucodestral-2508EU-resident coding specialist — completion + fill-in-the-middle.
Qwen 3.7 Plus (Together)
MidOpen weightsQwen/Qwen3.7-PlusOpen-weights 1M-context workhorse — strong multilingual + tool use.
Amazon Nova Pro
Midamazon.nova-pro-v1:0Amazon Nova Pro on AWS Bedrock (Converse API), served from AWS us-east-1 at standard on-demand pricing. Text in/out with streaming and tool calling.
GPT-5.4 nano
Budgetgpt-5.4-nanoLowest-cost OpenAI tier — classification, routing, bulk processing.
Gemini 3.1 Flash-Lite
Budgetgemini-3.1-flash-liteCheapest Gemini tier — bulk multimodal processing with 1M context.
DeepSeek V4 Flash
BudgetOpen weightsapacdeepseek-v4-flashCheapest strong general-purpose model — 1M context, dual thinking/non-thinking modes.
Mistral Small 4
BudgetOpen weightseumistral-small-2603EU-resident budget tier with built-in vision. Classification + light generation.
GPT-OSS 120B (Groq)
BudgetOpen weightsopenai/gpt-oss-120bOpen-weights workhorse at ~500 tokens/sec — Groq’s mainline model.
GPT-OSS 20B (Groq)
BudgetOpen weightsopenai/gpt-oss-20bFastest cheap inference (~1,000 tokens/sec) — drafts, classification, agent loops.
Llama 4 Scout (Groq)
BudgetOpen weightsmeta-llama/llama-4-scout-17b-16e-instructMultimodal open-weights Llama at ~600 tokens/sec — vision + tools at budget price.
Gemma 4 31B (Together)
BudgetOpen weightsgoogle/gemma-4-31B-itLight open-weights multimodal — cost-sensitive vision tasks.
Coming soon — new modalities
Embeddings, image, speech, video, OCR, and rerank models are staged in the catalog but excluded from the API until their endpoints ship — we don't sell dead SKUs. They become routable the day the endpoint goes live.
Embeddings
Coming soon- text-embedding-3-large
OpenAI
- text-embedding-3-small
OpenAI
- Gemini Embedding 2Google Gemini
- Mistral EmbedMistral
- Voyage 4 LargeVoyage AI
- Jina Embeddings v4Jina AI
- Cohere Embed 4Cohere
Rerank
Coming soon- Cohere Rerank 3.5Cohere
- Voyage Rerank 2.5Voyage AI
- Jina Reranker v3Jina AI
Image
Coming soon- GPT Image 2
OpenAI
- Gemini 3 Pro ImageGoogle Gemini
- FLUX.2 [pro]Black Forest Labs
- Ideogram 4.0Ideogram
- Grok Imagine (Quality)
xAI
Speech
Coming soon- Whisper Large v3 (Groq)Groq
- GPT-4o Transcribe
OpenAI
- ElevenLabs Scribe v2ElevenLabs
- ElevenLabs Eleven v3ElevenLabs
- Voxtral SmallMistral
OCR
Coming soon- Mistral OCR 4Mistral
Video
Coming soon- Veo 3.1Google Gemini
- Grok Imagine Video 1.5
xAI
Realtime
Coming soon- GPT Realtime 2
OpenAI
- Gemini 3.1 Flash LiveGoogle Gemini
Chat
Coming soon- Kimi K2.6 (first-party)Moonshot AI
- Cohere Command ACohere
Coding
Coming soon- Grok Build 0.1
xAI
Route everything through one API.
Point your existing OpenAI or Anthropic SDK at Conduix and get this whole catalog — with governance, failover, and spend controls built in.

