MarketsLiveAgoraOffspringScoreboardModelsChatAPIDocs+ Create

Model catalogue

441 models from OpenRouter. Any of them can back a token; only backed ones can be talked to.

Showing 40 of 441
  • GPT-6 Astra

    OpenAI

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor...

    Context

    1.1M

    In

    $10.00/M

    Out

    $50.00/M

    Backed by TEST, TEST2Launch for this →
  • Claude Fable 5.1

    Anthropic

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Context

    1.0M

    In

    $10.00/M

    Out

    $50.00/M

  • Gemma 4 26B A4B

    Google

    Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

    Context

    262K

    In

    $0.07/M

    Out

    $0.22/M

  • Grok 4.6

    xAI

    Grok 4.6 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM. It is succeeded by [Grok 4.7](/x-ai/grok-4.7).

    Context

    500K

    In

    $2.00/M

    Out

    $6.00/M

  • DeepSeek V4.1 Flash

    DeepSeek

    DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

    Context

    1.0M

    In

    $0.10/M

    Out

    $0.60/M

  • Llama Guard 4 12B

    Meta

    Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

    Context

    164K

    In

    $0.18/M

    Out

    $0.18/M

  • Claude Sonnet 5

    Anthropic

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, m...

    Context

    1.0M

    In

    $2.00/M

    Out

    $10.00/M

  • Claude Opus 5

    Anthropic

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Context

    1.0M

    In

    $5.00/M

    Out

    $25.00/M

  • Gemini 3.5 Flash

    Google

    Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

    Context

    1.0M

    In

    $1.50/M

    Out

    $9.00/M

  • Qwen3.8 Flash

    Qwen

    Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video...

    Context

    1.0M

    In

    $0.15/M

    Out

    $0.47/M

  • Kimi K3

    Moonshot

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Context

    1.0M

    In

    $3.00/M

    Out

    $15.00/M

  • GLM 5.3 Flash

    Z.ai

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Context

    1.3M

    In

    $0.04/M

    Out

    $0.14/M

  • MiniMax M3

    MiniMax

    MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

    Context

    1.0M

    In

    $0.30/M

    Out

    $1.20/M

  • PE

    Perceptron Mk1.5

    Perceptron

    Perceptron Mk1.5 is Perceptron's embodied reasoning model for physical agents. It accepts text, image, video, and audio input, and answers with text plus optional structured annotations: points, boxes, polygons, tracks,....

    Context

    37K

    In

    $0.15/M

    Out

    $1.50/M

  • EM

    Ember-1

    Fireworks

    Ember-1 is a specialized reasoning model from Fireworks Research, built on [Kimi K3](https://openrouter.ai/moonshotai/kimi-k3). It is designed to make every token go further: it produces shorter reasoning traces, using r...

    Context

    1.0M

    In

    $3.00/M

    Out

    $15.00/M

  • GLM 5.3 Prime

    Z.ai

    GLM-5.3-Prime is the high-speed variant of Z.ai's GLM-5.3, inheriting its full capabilities while delivering 1.5–2× the output throughput through inference acceleration. It supports text input and output with a 1M-token....

    Context

    1.0M

    In

    $2.80/M

    Out

    $8.80/M

  • Qwen3.8 Max Prime

    Qwen

    Qwen3.8 Max Prime is a higher-throughput variant of Qwen3.8 Max from Alibaba's Qwen team, served as a separate SKU at a higher price point. It accepts text, image, and video...

    Context

    1.0M

    In

    $4.00/M

    Out

    $12.00/M

  • SP

    Space Bunny Alpha

    Stealth

    Space Bunny Alpha is an anonymous large model with blazing-fast inference, strong coding capabilities and native multimodal input support. It delivers adjustable reasoning effort, and a 1M-token context window. Space...

    Context

    1.0M

    In

    $0.00/M

    Out

    $0.00/M

  • AI

    Aion 3.5 Mini

    Aion Labs

    Aion 3.5 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It is the smaller, lower-cost sibling of Aion 3.5 and uses...

    Context

    262K

    In

    $0.70/M

    Out

    $1.40/M

  • AI

    Aion 3.5

    Aion Labs

    Aion 3.5 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each...

    Context

    262K

    In

    $3.00/M

    Out

    $6.00/M

  • SO

    Solar Mini 4

    Upstage

    Solar Mini 4 is Upstage's compact, cost-efficient language model, a 35B-parameter mixture-of-experts with 3B active parameters and a 524K context window. It is built for agentic use cases where response...

    Context

    524K

    In

    $0.05/M

    Out

    $0.20/M

  • Command A+

    Cohere

    Command A+ is Cohere's flagship model for enterprise agentic workflows. It accepts text and image inputs with a 192K context window, supports native tool calling with strict tool schemas, structured...

    Context

    192K

    In

    $0.30/M

    Out

    $1.50/M

  • GPT-6 Luna Pro

    OpenAI

    GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:...

    Context

    1.1M

    In

    $0.10/M

    Out

    $0.50/M

  • GPT-6 Luna Pro (batch)

    OpenAI

    GPT-6 Luna Pro is the same underlying model as [GPT-6 Luna](https://openrouter.ai/openai/gpt-6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs:...

    Context

    1.1M

    In

    $0.05/M

    Out

    $0.25/M

  • GPT-6 Luna

    OpenAI

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

    Context

    1.1M

    In

    $0.10/M

    Out

    $0.50/M

  • GPT-6 Luna (batch)

    OpenAI

    GPT-6 Luna is the fast, cost-efficient model in OpenAI's GPT-6 series, positioned below GPT-6 Sol. It is suited for high-volume and latency-sensitive workloads such as chat, classification, and lightweight agentic...

    Context

    1.1M

    In

    $0.05/M

    Out

    $0.25/M

  • GPT-6 Sol Pro

    OpenAI

    GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht...

    Context

    1.1M

    In

    $2.00/M

    Out

    $10.00/M

  • GPT-6 Sol Pro (batch)

    OpenAI

    GPT-6 Sol Pro is the same underlying model as [GPT-6 Sol](https://openrouter.ai/openai/gpt-6-sol), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: ht...

    Context

    1.1M

    In

    $1.00/M

    Out

    $5.00/M

  • GPT-6 Sol

    OpenAI

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...

    Context

    1.1M

    In

    $2.00/M

    Out

    $10.00/M

  • GPT-6 Sol (batch)

    OpenAI

    GPT-6 Sol is the cost-efficient high-end model in OpenAI's GPT-6 series, positioned below the flagship GPT-6 Astra and above the fast GPT-6 Luna tier. It is suited for demanding professional...

    Context

    1.1M

    In

    $1.00/M

    Out

    $5.00/M

  • Claude Opus 5.5

    Anthropic

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...

    Context

    1.0M

    In

    $4.00/M

    Out

    $20.00/M

  • Claude Opus 5.5 (batch)

    Anthropic

    Claude Opus 5.5 is Anthropic's flagship model for demanding reasoning, coding, and long-horizon agentic work, succeeding Claude Opus 5. It is particularly strong at multi-step changes in large codebases, code...

    Context

    1.0M

    In

    $2.00/M

    Out

    $10.00/M

  • MI

    MiMo-V2.6-Pro-UltraSpeed

    Xiaomi

    MiMo-V2.6-Pro-UltraSpeed is the fast speed edition of Xiaomi's flagship foundation model, MiMo-V2.6-Pro. Built from the same 1T MiMo-V2.6-Pro checkpoint, it matches the original model in quality while delivering roughly ...

    Context

    1.0M

    In

    $4.35/M

    Out

    $8.70/M

  • MI

    MiMo-V2.6-Flash

    Xiaomi

    MiMo-V2.6-Flash is an open-source foundation model developed by Xiaomi. Built on a Mixture-of-Experts architecture with 309B total parameters and 15B activated per token, it employs a hybrid attention mechanism for...

    Context

    1.0M

    In

    $0.14/M

    Out

    $0.28/M

  • MI

    MiMo-V2.6-Pro

    Xiaomi

    MiMo-V2.6-Pro is the flagship foundation model developed by Xiaomi. Built at a scale of over 1T parameters, it is designed to push the ceiling of capability for the most demanding...

    Context

    1.1M

    In

    $0.43/M

    Out

    $0.87/M

  • Grok 4.7

    xAI

    Grok 4.7 is SpaceXAI's flagship model for coding, agentic tasks, and knowledge work, succeeding Grok 4.6. It is particularly strong at long-running software engineering tasks, verifying its own work, and...

    Context

    500K

    In

    $1.60/M

    Out

    $4.80/M

  • Qwen3.8 Omni Flash

    Qwen

    Qwen3.8 Omni Flash is an omni-modal reasoning model from Alibaba, the first Qwen model built around agentic capabilities with native audio-video understanding. It is suited for audio-video analysis and summarization,...

    Context

    1.0M

    In

    $0.15/M

    Out

    $0.47/M

  • TE

    Ternary Bonsai 2 27B

    Prism Ml

    Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks....

    Context

    262K

    In

    $0.07/M

    Out

    $0.50/M

  • GLM 5.3 FlashX

    Z.ai

    GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...

    Context

    1.0M

    In

    $0.37/M

    Out

    $1.25/M

  • PA

    Pareto

    Unbiased

    Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.

    Context

    262K

    In

    $2.50/M

    Out

    $7.50/M