Model catalogue
441 models from OpenRouter. Any of them can back a token; only backed ones can be talked to.
GPT-6 Astra
OpenAI
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor...
Context
1.1M
In
$10.00/M
Out
$50.00/M
Backed by TEST, TEST2Launch for this →Claude Fable 5.1
Anthropic
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
Context
1.0M
In
$10.00/M
Out
$50.00/M
UnbackedLaunch for this →Gemma 4 26B A4B
Google
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Context
262K
In
$0.07/M
Out
$0.22/M
UnbackedLaunch for this →Grok 4.6
xAI
Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).
Context
500K
In
$2.00/M
Out
$6.00/M
UnbackedLaunch for this →DeepSeek V4.1 Flash
DeepSeek
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
Context
1.0M
In
$0.10/M
Out
$0.60/M
UnbackedLaunch for this →Llama Guard 4 12B
Meta
Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...
Context
164K
In
$0.18/M
Out
$0.18/M
UnbackedLaunch for this →Claude Sonnet 5
Anthropic
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, m...
Context
1.0M
In
$2.00/M
Out
$10.00/M
UnbackedLaunch for this →Claude Opus 5
Anthropic
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Context
1.0M
In
$5.00/M
Out
$25.00/M
UnbackedLaunch for this →Gemini 3.5 Flash
Google
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Context
1.0M
In
$1.50/M
Out
$9.00/M
UnbackedLaunch for this →Qwen3.8 Flash
Qwen
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video...
Context
1.0M
In
$0.15/M
Out
$0.47/M
UnbackedLaunch for this →Kimi K3
Moonshot
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Context
1.0M
In
$3.00/M
Out
$15.00/M
UnbackedLaunch for this →GLM 5.3 Flash
Z.ai
GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...
Context
1.3M
In
$0.04/M
Out
$0.14/M
UnbackedLaunch for this →MiniMax M3
MiniMax
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Context
1.0M
In
$0.30/M
Out
$1.20/M
UnbackedLaunch for this →- PE
Perceptron Mk1.5
Perceptron
Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,....
Context
37K
In
$0.15/M
Out
$1.50/M
UnbackedLaunch for this → - EM
Ember-1
Fireworks
Ember-1 is a specialized reasoning model from Fireworks Research, built on [Kimi K3](https://openrouter.ai/moonshotai/kimi-k3). It is designed to make every token go further: it produces shorter reasoning traces, using r...
Context
1.0M
In
$3.00/M
Out
$15.00/M
UnbackedLaunch for this → GLM 5.3 Prime
Z.ai
GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token....
Context
1.0M
In
$2.80/M
Out
$8.80/M
UnbackedLaunch for this →Qwen3.8 Max Prime
Qwen
Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...
Context
1.0M
In
$4.00/M
Out
$12.00/M
UnbackedLaunch for this →- SP
Space Bunny Alpha
Stealth
Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support. It delivers adjustable reasoning effort, and a 1M-token context window. Space...
Context
1.0M
In
$0.00/M
Out
$0.00/M
UnbackedLaunch for this → - AI
Aion 3.5 Mini
Aion Labs
Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...
Context
262K
In
$0.70/M
Out
$1.40/M
UnbackedLaunch for this → - AI
Aion 3.5
Aion Labs
Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each...
Context
262K
In
$3.00/M
Out
$6.00/M
UnbackedLaunch for this → - SO
Solar Mini 4
Upstage
Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context window. It is built for agentic use cases where response...
Context
524K
In
$0.05/M
Out
$0.20/M
UnbackedLaunch for this → Command A+
Cohere
Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool schemas, structured...
Context
192K
In
$0.30/M
Out
$1.50/M
UnbackedLaunch for this →GPT-6 Luna Pro
OpenAI
GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:...
Context
1.1M
In
$0.10/M
Out
$0.50/M
UnbackedLaunch for this →GPT-6 Luna Pro (batch)
OpenAI
GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:...
Context
1.1M
In
$0.05/M
Out
$0.25/M
UnbackedLaunch for this →GPT-6 Luna
OpenAI
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
Context
1.1M
In
$0.10/M
Out
$0.50/M
UnbackedLaunch for this →GPT-6 Luna (batch)
OpenAI
GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...
Context
1.1M
In
$0.05/M
Out
$0.25/M
UnbackedLaunch for this →GPT-6 Sol Pro
OpenAI
GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht...
Context
1.1M
In
$2.00/M
Out
$10.00/M
UnbackedLaunch for this →GPT-6 Sol Pro (batch)
OpenAI
GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht...
Context
1.1M
In
$1.00/M
Out
$5.00/M
UnbackedLaunch for this →GPT-6 Sol
OpenAI
GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...
Context
1.1M
In
$2.00/M
Out
$10.00/M
UnbackedLaunch for this →GPT-6 Sol (batch)
OpenAI
GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...
Context
1.1M
In
$1.00/M
Out
$5.00/M
UnbackedLaunch for this →Claude Opus 5.5
Anthropic
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Context
1.0M
In
$4.00/M
Out
$20.00/M
UnbackedLaunch for this →Claude Opus 5.5 (batch)
Anthropic
Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...
Context
1.0M
In
$2.00/M
Out
$10.00/M
UnbackedLaunch for this →- MI
MiMo-V2.6-Pro-UltraSpeed
Xiaomi
MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly ...
Context
1.0M
In
$4.35/M
Out
$8.70/M
UnbackedLaunch for this → - MI
MiMo-V2.6-Flash
Xiaomi
MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...
Context
1.0M
In
$0.14/M
Out
$0.28/M
UnbackedLaunch for this → - MI
MiMo-V2.6-Pro
Xiaomi
MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...
Context
1.1M
In
$0.43/M
Out
$0.87/M
UnbackedLaunch for this → Grok 4.7
xAI
Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...
Context
500K
In
$1.60/M
Out
$4.80/M
UnbackedLaunch for this →Qwen3.8 Omni Flash
Qwen
Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...
Context
1.0M
In
$0.15/M
Out
$0.47/M
UnbackedLaunch for this →- TE
Ternary Bonsai 2 27B
Prism Ml
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks....
Context
262K
In
$0.07/M
Out
$0.50/M
UnbackedLaunch for this → GLM 5.3 FlashX
Z.ai
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
Context
1.0M
In
$0.37/M
Out
$1.25/M
UnbackedLaunch for this →- PA
Pareto
Unbiased
Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.
Context
262K
In
$2.50/M
Out
$7.50/M
UnbackedLaunch for this →