MarketsModelsChatAPIDocs+ Create

Model catalogue

426 models from OpenRouter. Any of them can back a token; only backed ones can be talked to.

Showing 40 of 426
  • GPT-6 Astra

    OpenAI

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor...

    Context

    1.1M

    In

    $10.00/M

    Out

    $50.00/M

  • Claude Fable 5.1

    Anthropic

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Context

    1.0M

    In

    $10.00/M

    Out

    $50.00/M

  • Gemma 4 26B A4B

    Google

    Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

    Context

    262K

    In

    $0.09/M

    Out

    $0.30/M

  • Grok 4.6

    xAI

    Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

    Context

    500K

    In

    $2.00/M

    Out

    $6.00/M

  • DeepSeek V4.1 Flash

    DeepSeek

    DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

    Context

    1.0M

    In

    $0.30/M

    Out

    $1.20/M

    Backed by TEST2Launch for this →
  • Llama Guard 4 12B

    Meta

    Llama Guard 4 is a Llama 4 Scout-derived multimodal pretrained model, fine-tuned for content safety classification. Similar to previous versions, it can be used to classify content in both LLM...

    Context

    164K

    In

    $0.18/M

    Out

    $0.18/M

  • Claude Sonnet 5

    Anthropic

    Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, m...

    Context

    1.0M

    In

    $2.00/M

    Out

    $10.00/M

    Backed by TESTLaunch for this →
  • Claude Opus 5

    Anthropic

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

    Context

    1.0M

    In

    $5.00/M

    Out

    $25.00/M

  • Gemini 3.5 Flash

    Google

    Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

    Context

    1.0M

    In

    $1.50/M

    Out

    $9.00/M

  • Qwen3.8 Flash

    Qwen

    Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video...

    Context

    1.0M

    In

    $0.15/M

    Out

    $0.47/M

  • Kimi K3

    Moonshot

    Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

    Context

    1.0M

    In

    $2.65/M

    Out

    $13.28/M

  • GLM 5.3 Flash

    Z.ai

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Context

    1.3M

    In

    $0.15/M

    Out

    $0.50/M

  • MiniMax M3

    MiniMax

    MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

    Context

    1.0M

    In

    $0.30/M

    Out

    $1.20/M

  • SC

    Schematron V2 Turbo

    Inference Net

    Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in re...

    Context

    128K

    In

    $0.03/M

    Out

    $0.15/M

  • SC

    Schematron V2 Small

    Inference Net

    Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

    Context

    128K

    In

    $0.05/M

    Out

    $0.23/M

  • GP

    GPT Astra Latest

    ~Openai

    This model always redirects to the latest model in the GPT Astra family.

    Context

    1.1M

    In

    $10.00/M

    Out

    $50.00/M

  • GP

    GPT Sol Latest

    ~Openai

    This model always redirects to the latest model in the GPT Sol family.

    Context

    1.1M

    In

    $2.00/M

    Out

    $10.00/M

  • GP

    GPT Terra Latest

    ~Openai

    This model always redirects to the latest model in the GPT Terra family.

    Context

    1.1M

    In

    $2.00/M

    Out

    $12.00/M

  • GP

    GPT Luna Latest

    ~Openai

    This model always redirects to the latest model in the GPT Luna family.

    Context

    1.1M

    In

    $0.20/M

    Out

    $1.20/M

  • FU

    Fugu Ultra v2

    Sakana

    Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...

    Context

    1.0M

    In

    $5.00/M

    Out

    $30.00/M

  • FU

    Fugu Max

    Sakana

    Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

    Context

    1.0M

    In

    $2.00/M

    Out

    $6.00/M

  • LI

    Ling 3.0 Flash VL

    Inclusionai

    Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

    Context

    131K

    In

    $0.06/M

    Out

    $0.18/M

  • ME

    Mercury 2.5

    Inception

    Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...

    Context

    260K

    In

    $0.04/M

    Out

    $0.15/M

  • GPT-6 Astra (batch)

    OpenAI

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-hor...

    Context

    1.1M

    In

    $5.00/M

    Out

    $25.00/M

  • GPT-6 Astra Pro

    OpenAI

    GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's do...

    Context

    1.1M

    In

    $10.00/M

    Out

    $50.00/M

  • GPT-6 Astra Pro (batch)

    OpenAI

    GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's do...

    Context

    1.1M

    In

    $5.00/M

    Out

    $25.00/M

  • Qwen3.8 Max (0902)

    Qwen

    Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...

    Context

    1.0M

    In

    $2.00/M

    Out

    $6.00/M

  • Muse Spark 1.3 Contributor

    Meta

    Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track in...

    Context

    1.0M

    In

    $0.10/M

    Out

    $0.20/M

  • Muse Spark 1.3

    Meta

    Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It is designed to keep track of information across extended tasks, work through...

    Context

    1.0M

    In

    $1.25/M

    Out

    $4.25/M

  • Gemini 3.8 Flash

    Google

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

    Context

    1.0M

    In

    $0.75/M

    Out

    $3.75/M

  • Gemini 3.8 Flash (batch)

    Google

    Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

    Context

    1.0M

    In

    $0.38/M

    Out

    $1.88/M

  • Claude Fable 5.1 (batch)

    Anthropic

    Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...

    Context

    1.0M

    In

    $5.00/M

    Out

    $25.00/M

  • GR

    Granite 4.2 8B

    Ibm Granite

    Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...

    Context

    131K

    In

    $0.06/M

    Out

    $0.25/M

  • HY

    Hy4 preview

    Tencent

    Tencent: Hy4 preview is a mixture-of-experts model from Tencent, with 49B active parameters out of 770B total. It is designed for coding agents, complex tool-use workflows, and productivity tasks that...

    Context

    1.0M

    In

    $0.83/M

    Out

    $2.50/M

  • LI

    Ling 3.0 Flash Fin

    Inclusionai

    Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

    Context

    262K

    In

    $0.06/M

    Out

    $0.18/M

  • GL

    GLM Flash Latest

    ~Z Ai

    This model always redirects to the latest model in the GLM Flash family.

    Context

    1.3M

    In

    $0.07/M

    Out

    $0.25/M

  • GLM 5.3 Flash (batch)

    Z.ai

    GLM-5.3-Flash is a native multimodal model from Z.ai. It is suited for efficient coding and long-horizon agent tasks. Its hybrid sparse and linear attention architecture maintains accurate long-context behavior while...

    Context

    1.0M

    In

    $0.07/M

    Out

    $0.25/M

  • Muse Spark 1.2 Contributor

    Meta

    Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...

    Context

    1.0M

    In

    $0.10/M

    Out

    $0.20/M

  • DeepSeek V4 Flash Vision Exp

    DeepSeek

    DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base ...

    Context

    1.0M

    In

    $0.22/M

    Out

    $0.66/M

  • DeepSeek V4 Flash Vision Exp (batch)

    DeepSeek

    DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base ...

    Context

    1.0M

    In

    $0.11/M

    Out

    $0.33/M