Model Directory

Search the modelsyour product can ship.

Browse live routes for Seedance 2.5, Seedance 2.0 mini, Seedance 2.0, Happy Horse 1.1, Seedream 5.0, Claude Fable 5, Claude 4.8, Claude 4.7, Claude 4.6, Claude 4.5, ElevenLabs, Suno, and the rest of your image, video, audio, TTS, and text stack.

Seedance 2.5Seedance 2.0 miniSeedance 2.0Happy Horse 1.1Seedream 5.0Claude Fable 5Claude 4.8Claude 4.7Claude 4.6Claude 4.5ElevenLabsSuno
Live catalogOpen the current model pricing dashboardJump to live model prices, availability, API keys, and route setup.

Search by provider, compare context and pricing signals, then route the right model through one API.

Search and filter

Filter by modality, provider, reasoning, or sort order.

Reset

Results

Anthropic

Anthropic: Claude Opus 4.7

TextReasoningLive

anthropic/claude-opus-4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on complex, multi-step tasks and more reliable agentic execution across extended workflows. It is especially effective for asynchronous agent pipelines where tasks unfold over time - large codebases, multi-stage debugging, and end-to-end project orchestration. Beyond coding, Opus 4.7 brings improved knowledge work capabilities - from drafting documents and building presentations to analyzing data. It maintains coherence across very long outputs and extended sessions, making it a strong default for tasks that require persistence, judgment, and follow-through. For users upgrading from earlier Opus versions, see our [official migration guide here](https://openrouter.ai/docs/guides/evaluate-and-optimize/model-migrations/claude-4-7)

TextImage

Context

1M

Group

Claude

Pricing preview

Input Price: $5 /M tokens

Output Price: $25 /M tokens

Need volume pricing or launch support?

Contact sales

Anthropic

Anthropic: Claude Opus 4.6 (Fast)

TextReasoningLive

anthropic/claude-opus-4.6-fast

Fast-mode variant of [Opus 4.6](/anthropic/claude-opus-4.6) - identical capabilities with higher output speed at premium 6x pricing. Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

TextImage

Context

1M

Group

Claude

Pricing preview

Input Price: $30 /M tokens

Output Price: $15 /M tokens

Need volume pricing or launch support?

Contact sales

Amazon Bedrock

Anthropic: Claude Opus 4.6

TextReasoningLive

anthropic/claude-opus-4.6

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective for large codebases, complex refactors, and multi-step debugging that unfolds over time. The model shows deeper contextual understanding, stronger problem decomposition, and greater reliability on hard engineering tasks than prior generations. Beyond coding, Opus 4.6 excels at sustained knowledge work. It produces near-production-ready documents, plans, and analyses in a single pass, and maintains coherence across very long outputs and extended sessions. This makes it a strong default for tasks that require persistence, judgment, and follow-through, such as technical design, migration planning, and end-to-end project execution. For users upgrading from earlier Opus versions, see our [official migration guide here](https://openrouter.ai/docs/guides/guides/model-migrations/claude-4-6-opus)

TextImage

Context

1M

Group

Claude

Pricing preview

Input Price: $5 /M tokens

Output Price: $25 /M tokens

Need volume pricing or launch support?

Contact sales

Anthropic

Anthropic: Claude Sonnet 4.6

TextReasoningLive

anthropic/claude-sonnet-4.6

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with memory, polished document creation, and confident computer use for web QA and workflow automation.

TextImage

Context

1M

Group

Claude

Pricing preview

Input Price: $3 /M tokens

Output Price: $15 /M tokens

Need volume pricing or launch support?

Contact sales

Anthropic

Anthropic: Claude Opus 4.5

TextReasoningLive

anthropic/claude-opus-4.5

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and reasoning benchmarks, and improved robustness to prompt injection. The model is designed to operate efficiently across varied effort levels, enabling developers to trade off speed, depth, and token usage depending on task requirements. It comes with a new parameter to control token efficiency, which can be accessed using the OpenRouter Verbosity parameter with low, medium, or high. Opus 4.5 supports advanced tool use, extended context management, and coordinated multi-agent setups, making it well-suited for autonomous research, debugging, multi-step planning, and spreadsheet/browser manipulation. It delivers substantial gains in structured reasoning, execution reliability, and alignment compared to prior Opus generations, while reducing token overhead and improving performance on long-running tasks.

TextFileImage

Context

200K

Group

Claude

Pricing preview

Input Price: $5 /M tokens

Output Price: $25 /M tokens

Need volume pricing or launch support?

Contact sales

Google Vertex (Europe)

Anthropic: Claude Opus 4

TextReasoningLive

anthropic/claude-opus-4

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in software engineering, achieving leading results on SWE-bench (72.5%) and Terminal-bench (43.2%). Opus 4 supports extended, agentic workflows, handling thousands of task steps continuously for hours without degradation. Read more at the [blog post here](https://www.anthropic.com/news/claude-4)

TextImageFile

Context

200K

Group

Claude

Pricing preview

Input Price: $15 /M tokens

Output Price: $75 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude 3.5 Sonnet

TextLive

anthropic/claude-3.5-sonnet

New Claude 3.5 Sonnet delivers better-than-Opus capabilities, faster-than-Sonnet speeds, at the same Sonnet prices. Sonnet is particularly good at: - Coding: Scores ~49% on SWE-Bench Verified, higher than the last best score, and without any fancy prompt scaffolding - Data science: Augments human data science expertise; navigates unstructured data while using multiple tools for insights - Visual processing: excelling at interpreting charts, graphs, and images, accurately transcribing text to derive insights beyond just the text alone - Agentic tasks: exceptional tool use, making it great at agentic tasks (i.e. complex, multi-step problem solving tasks that require engaging with other systems) #multimodal

TextImageFile

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Google Vertex

Anthropic: Claude Opus 4.1

TextReasoningLive

anthropic/claude-opus-4.1

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains in multi-file code refactoring, debugging precision, and detail-oriented reasoning. The model supports extended thinking up to 64K tokens and is optimized for tasks involving research, data analysis, and tool-assisted reasoning.

TextImageFile

Context

200K

Group

Claude

Pricing preview

Input Price: $15 /M tokens

Output Price: $75 /M tokens

Need volume pricing or launch support?

Contact sales

Google Vertex (Global)

Anthropic: Claude Sonnet 4

TextReasoningLive

anthropic/claude-sonnet-4

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%), Sonnet 4 balances capability and computational efficiency, making it suitable for a broad range of applications from routine coding tasks to complex software development projects. Key enhancements include improved autonomous codebase navigation, reduced error rates in agent-driven workflows, and increased reliability in following intricate instructions. Sonnet 4 is optimized for practical everyday use, providing advanced reasoning capabilities while maintaining efficiency and responsiveness in diverse internal and external scenarios. Read more at the [blog post here](https://www.anthropic.com/news/claude-4)

TextImageFile

Context

1M

Group

Claude

Pricing preview

Input Price: $3 /M tokens

Output Price: $15 /M tokens

Need volume pricing or launch support?

Contact sales

Amazon Bedrock

Anthropic: Claude 3.7 Sonnet

TextReasoningLive

anthropic/claude-3.7-sonnet

Claude 3.7 Sonnet is an advanced large language model with improved reasoning, coding, and problem-solving capabilities. It introduces a hybrid reasoning approach, allowing users to choose between rapid responses and extended, step-by-step processing for complex tasks. The model demonstrates notable improvements in coding, particularly in front-end development and full-stack updates, and excels in agentic workflows, where it can autonomously navigate multi-step processes. Claude 3.7 Sonnet maintains performance parity with its predecessor in standard mode while offering an extended reasoning mode for enhanced accuracy in math, coding, and instruction-following tasks. Read more at the [blog post here](https://www.anthropic.com/news/claude-3-7-sonnet)

TextImageFile

Context

200K

Group

Claude

Pricing preview

Input Price: $3 /M tokens

Output Price: $15 /M tokens

Need volume pricing or launch support?

Contact sales

Google Vertex

Anthropic: Claude 3.7 Sonnet (thinking)

TextReasoningLive

anthropic/claude-3.7-sonnet

Claude 3.7 Sonnet is an advanced large language model with improved reasoning, coding, and problem-solving capabilities. It introduces a hybrid reasoning approach, allowing users to choose between rapid responses and extended, step-by-step processing for complex tasks. The model demonstrates notable improvements in coding, particularly in front-end development and full-stack updates, and excels in agentic workflows, where it can autonomously navigate multi-step processes. Claude 3.7 Sonnet maintains performance parity with its predecessor in standard mode while offering an extended reasoning mode for enhanced accuracy in math, coding, and instruction-following tasks. Read more at the [blog post here](https://www.anthropic.com/news/claude-3-7-sonnet)

TextImageFile

Context

200K

Group

Claude

Pricing preview

Input Price: $3 /M tokens

Output Price: $15 /M tokens

Need volume pricing or launch support?

Contact sales

Google Vertex

Anthropic: Claude Sonnet 4.5

TextReasoningLive

anthropic/claude-sonnet-4.5

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with improvements across system design, code security, and specification adherence. The model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking. Sonnet 4.5 also introduces stronger agentic capabilities, including improved tool orchestration, speculative parallel execution, and more efficient context and memory management. With enhanced context tracking and awareness of token usage across tool calls, it is particularly well-suited for multi-context and long-running workflows. Use cases span software engineering, cybersecurity, financial analysis, research agents, and other domains requiring sustained reasoning and tool use.

TextImageFile

Context

1M

Group

Claude

Pricing preview

Input Price: $3 /M tokens

Output Price: $15 /M tokens

Need volume pricing or launch support?

Contact sales

SiliconFlow

Z.ai: GLM 4.6

TextReasoningLive

z-ai/glm-4.6

Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex agentic tasks. Superior coding performance: The model achieves higher scores on code benchmarks and demonstrates better real-world performance in applications such as Claude Code、Cline、Roo Code and Kilo Code, including improvements in generating visually polished front-end pages. Advanced reasoning: GLM-4.6 shows a clear improvement in reasoning performance and supports tool use during inference, leading to stronger overall capability. More capable agents: GLM-4.6 exhibits stronger performance in tool using and search-based agents, and integrates more effectively within agent frameworks. Refined writing: Better aligns with human preferences in style and readability, and performs more naturally in role-playing scenarios.

Text

Context

204.8K

Group

Other

Pricing preview

Input Price: $0.39 /M tokens

Output Price: $1.9 /M tokens

Need volume pricing or launch support?

Contact sales

Mistral

Mistral: Mistral Medium 3.1

TextLive

mistralai/mistral-medium-3.1

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost compared to traditional large models, making it suitable for scalable deployments across professional and industrial use cases. The model excels in domains such as coding, STEM reasoning, and enterprise adaptation. It supports hybrid, on-prem, and in-VPC deployments and is optimized for integration into custom workflows. Mistral Medium 3.1 offers competitive accuracy relative to larger models like Claude Sonnet 3.5/3.7, Llama 4 Maverick, and Command R+, while maintaining broad compatibility across cloud environments.

TextImage

Context

131.1K

Group

Mistral

Pricing preview

Input Price: $0.4 /M tokens

Output Price: $2 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Inception: Mercury

TextLive

inception/mercury

Mercury is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like GPT-4.1 Nano and Claude 3.5 Haiku while matching their performance. Mercury's speed enables developers to provide responsive user experiences, including with voice agents, search interfaces, and chatbots. Read more in the [blog post] (https://www.inceptionlabs.ai/blog/introducing-mercury) here.

Text

Context

128K

Group

Other

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Mistral

Mistral: Mistral Medium 3

TextLive

mistralai/mistral-medium-3

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost compared to traditional large models, making it suitable for scalable deployments across professional and industrial use cases. The model excels in domains such as coding, STEM reasoning, and enterprise adaptation. It supports hybrid, on-prem, and in-VPC deployments and is optimized for integration into custom workflows. Mistral Medium 3 offers competitive accuracy relative to larger models like Claude Sonnet 3.5/3.7, Llama 4 Maverick, and Command R+, while maintaining broad compatibility across cloud environments.

TextImage

Context

131.1K

Group

Mistral

Pricing preview

Input Price: $0.4 /M tokens

Output Price: $2 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Inception: Mercury Coder

TextLive

inception/mercury-coder

Mercury Coder is the first diffusion large language model (dLLM). Applying a breakthrough discrete diffusion approach, the model runs 5-10x faster than even speed optimized models like Claude 3.5 Haiku and GPT-4o Mini while matching their performance. Mercury Coder's speed means that developers can stay in the flow while coding, enjoying rapid chat-based iteration and responsive code completion suggestions. On Copilot Arena, Mercury Coder ranks 1st in speed and ties for 2nd in quality. Read more in the [blog post here](https://www.inceptionlabs.ai/blog/introducing-mercury).

Text

Context

128K

Group

Other

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Dolphin3.0 R1 Mistral 24B

TextReasoningLive

cognitivecomputations/dolphin3.0-r1-mistral-24b

Dolphin 3.0 R1 is the next generation of the Dolphin series of instruct-tuned models. Designed to be the ultimate general purpose local model, enabling coding, math, agentic, function calling, and general use cases. The R1 version has been trained for 3 epochs to reason using 800k reasoning traces from the Dolphin-R1 dataset. Dolphin aims to be a general purpose reasoning instruct model, similar to the models behind ChatGPT, Claude, Gemini. Part of the [Dolphin 3.0 Collection](https://huggingface.co/collections/QuixiAI/dolphin-30) Curated and trained by [Eric Hartford](https://huggingface.co/ehartford), [Ben Gitter](https://huggingface.co/bigstorm), [BlouseJury](https://huggingface.co/BlouseJury) and [DphnAI](https://huggingface.co/dphn)

Text

Context

32.8K

Group

Mistral

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Dolphin3.0 Mistral 24B

TextLive

cognitivecomputations/dolphin3.0-mistral-24b

Dolphin 3.0 is the next generation of the Dolphin series of instruct-tuned models. Designed to be the ultimate general purpose local model, enabling coding, math, agentic, function calling, and general use cases. Dolphin aims to be a general purpose instruct model, similar to the models behind ChatGPT, Claude, Gemini. Part of the [Dolphin 3.0 Collection](https://huggingface.co/collections/QuixiAI/dolphin-30) Curated and trained by [Eric Hartford](https://huggingface.co/ehartford), [Ben Gitter](https://huggingface.co/bigstorm), [BlouseJury](https://huggingface.co/BlouseJury) and [DphnAI](https://huggingface.co/dphn)

Text

Context

32.8K

Group

Mistral

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Amazon Bedrock (US-WEST)

Anthropic: Claude 3.5 Haiku

TextLive

anthropic/claude-3.5-haiku

Claude 3.5 Haiku features offers enhanced capabilities in speed, coding accuracy, and tool use. Engineered to excel in real-time applications, it delivers quick response times that are essential for dynamic tasks such as chat interactions and immediate coding suggestions. This makes it highly suitable for environments that demand both speed and precision, such as software development, customer service bots, and data management systems. This model is currently pointing to [Claude 3.5 Haiku (2024-10-22)](/anthropic/claude-3-5-haiku-20241022).

TextImage

Context

200K

Group

Claude

Pricing preview

Input Price: $0.8 /M tokens

Output Price: $4 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude 3.5 Haiku (2024-10-22)

TextLive

anthropic/claude-3.5-haiku-20241022

Claude 3.5 Haiku features enhancements across all skill sets including coding, tool use, and reasoning. As the fastest model in the Anthropic lineup, it offers rapid response times suitable for applications that require high interactivity and low latency, such as user-facing chatbots and on-the-fly code completions. It also excels in specialized tasks like data extraction and real-time content moderation, making it a versatile tool for a broad range of industries. It does not support image inputs. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/3-5-models-and-computer-use)

TextImageFile

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Mancer

Magnum v4 72B

TextLive

anthracite-org/magnum-v4-72b

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

Text

Context

16.4K

Group

Qwen

Pricing preview

Input Price: $3 /M tokens

Output Price: $5 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Magnum v2 72B

TextLive

anthracite-org/magnum-v2-72b

From the maker of [Goliath](https://openrouter.ai/models/alpindale/goliath-120b), Magnum 72B is the seventh in a family of models designed to achieve the prose quality of the Claude 3 models, notably Opus & Sonnet. The model is based on [Qwen2 72B](https://openrouter.ai/models/qwen/qwen-2-72b-instruct) and trained with 55 million tokens of highly curated roleplay (RP) data.

Text

Context

32.8K

Group

Qwen

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Meta: Llama 3.1 405B Instruct

TextLive

meta-llama/llama-3.1-405b-instruct

The highly anticipated 400B class of Llama3 is here! Clocking in at 128k context with impressive eval scores, the Meta AI team continues to push the frontier of open-source LLMs. Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 405B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong performance compared to leading closed-source models including GPT-4o and Claude 3.5 Sonnet in evaluations. To read more about the model release, [click here](https://ai.meta.com/blog/meta-llama-3-1/). Usage of this model is subject to [Meta's Acceptable Use Policy](https://llama.meta.com/llama3/use-policy/).

Text

Context

131.1K

Group

Llama3

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Magnum 72B

TextLive

alpindale/magnum-72b

From the maker of [Goliath](https://openrouter.ai/models/alpindale/goliath-120b), Magnum 72B is the first in a new family of models designed to achieve the prose quality of the Claude 3 models, notably Opus & Sonnet. The model is based on [Qwen2 72B](https://openrouter.ai/models/qwen/qwen-2-72b-instruct) and trained with 55 million tokens of highly curated roleplay (RP) data.

Text

Context

16.4K

Group

Qwen

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude 3.5 Sonnet (2024-06-20)

TextLive

anthropic/claude-3.5-sonnet-20240620

Claude 3.5 Sonnet delivers better-than-Opus capabilities, faster-than-Sonnet speeds, at the same Sonnet prices. Sonnet is particularly good at: - Coding: Autonomously writes, edits, and runs code with reasoning and troubleshooting - Data science: Augments human data science expertise; navigates unstructured data while using multiple tools for insights - Visual processing: excelling at interpreting charts, graphs, and images, accurately transcribing text to derive insights beyond just the text alone - Agentic tasks: exceptional tool use, making it great at agentic tasks (i.e. complex, multi-step problem solving tasks that require engaging with other systems) For the latest version (2024-10-23), check out [Claude 3.5 Sonnet](/anthropic/claude-3.5-sonnet). #multimodal

TextImageFile

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Amazon Bedrock

Anthropic: Claude 3 Haiku

TextLive

anthropic/claude-3-haiku

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-haiku) #multimodal

TextImage

Context

200K

Group

Claude

Pricing preview

Input Price: $0.25 /M tokens

Output Price: $1.25 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude 3 Sonnet

TextLive

anthropic/claude-3-sonnet

Claude 3 Sonnet is an ideal balance of intelligence and speed for enterprise workloads. Maximum utility at a lower price, dependable, balanced for scaled deployments. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-family) #multimodal

TextImage

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude 3 Opus

TextLive

anthropic/claude-3-opus

Claude 3 Opus is Anthropic's most powerful model for highly complex tasks. It boasts top-level performance, intelligence, fluency, and understanding. See the launch announcement and benchmark results [here](https://www.anthropic.com/news/claude-3-family) #multimodal

TextImage

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude v2

TextLive

anthropic/claude-2

Claude 2 delivers advancements in key capabilities for enterprises—including an industry-leading 200K token context window, significant reductions in rates of model hallucination, system prompts and a new beta feature: tool use.

Text

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude Instant v1.1

TextLive

anthropic/claude-instant-1.1

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text

Context

100K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude v2.1

TextLive

anthropic/claude-2.1

Claude 2 delivers advancements in key capabilities for enterprises—including an industry-leading 200K token context window, significant reductions in rates of model hallucination, system prompts and a new beta feature: tool use.

Text

Context

200K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Mancer

Mancer: Weaver (alpha)

TextLive

mancer/weaver

An attempt to recreate Claude-style verbosity, but don't expect the same level of coherence or memory. Meant for use in roleplay/narrative situations.

Text

Context

8K

Group

Llama2

Pricing preview

Input Price: $0.75 /M tokens

Output Price: $1 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude Instant v1.0

TextLive

anthropic/claude-instant-1.0

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text

Context

100K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude Instant v1

TextLive

anthropic/claude-instant-1

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text

Context

100K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude v1

TextLive

anthropic/claude-1

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text

Context

100K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude v2.0

TextLive

anthropic/claude-2.0

Anthropic's flagship model. Superior performance on tasks that require complex reasoning. Supports hundreds of pages of text.

Text

Context

100K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Unknown provider

Anthropic: Claude v1.2

TextLive

anthropic/claude-1.2

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text

Context

100K

Group

Claude

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Inception

Inception: Mercury 2

TextReasoningLive

inception/mercury-2

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving >1,000 tokens/sec on standard GPUs. Mercury 2 is 5x+ faster than leading speed-optimized LLMs like Claude 4.5 Haiku and GPT 5 Mini, at a fraction of the cost. Mercury 2 supports tunable reasoning levels, 128K context, native tool use, and schema-aligned JSON output. Built for coding workflows where latency compounds, real-time voice/search, and agent loops. OpenAI API compatible. Read more in the [blog post](https://www.inceptionlabs.ai/blog/introducing-mercury-2).

Text

Context

128K

Group

Other

Pricing preview

Input Price: $0.25 /M tokens

Output Price: $0.75 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Auto Router

TextLive

openrouter/auto

Your prompt will be processed by a meta-model and routed to one of dozens of models (see below), optimizing for the best possible output. To see which model was used, visit [Activity](/activity), or read the `model` attribute of the response. Your response will be priced at the same rate as the routed model. Learn more, including how to customize the models for routing, in our [docs](/docs/guides/routing/routers/auto-router). Requests will be routed to the following models: - [anthropic/claude-haiku-4.5](/anthropic/claude-haiku-4.5) - [anthropic/claude-opus-4.6](/anthropic/claude-opus-4.6) - [anthropic/claude-sonnet-4.5](/anthropic/claude-sonnet-4.5) - [anthropic/claude-sonnet-4.6](/anthropic/claude-sonnet-4.6) - [deepseek/deepseek-r1](/deepseek/deepseek-r1) - [google/gemini-2.5-flash-lite](/google/gemini-2.5-flash-lite) - [google/gemini-3-flash-preview](/google/gemini-3-flash-preview) - [google/gemini-3-pro-preview](/google/gemini-3-pro-preview) - [google/gemini-3.1-pro-preview](/google/gemini-3.1-pro-preview) - [meta-llama/llama-3.3-70b-instruct](/meta-llama/llama-3.3-70b-instruct) - [minimax/minimax-m2.5](/minimax/minimax-m2.5) - [mistralai/codestral-2508](/mistralai/codestral-2508) - [mistralai/mistral-7b-instruct-v0.1](/mistralai/mistral-7b-instruct-v0.1) - [mistralai/mistral-large](/mistralai/mistral-large) - [mistralai/mistral-medium-3.1](/mistralai/mistral-medium-3.1) - [mistralai/mistral-small-3.2-24b-instruct-2506](/mistralai/mistral-small-3.2-24b-instruct-2506) - [moonshotai/kimi-k2-thinking](/moonshotai/kimi-k2-thinking) - [openai/gpt-5](/openai/gpt-5) - [openai/gpt-5-mini](/openai/gpt-5-mini) - [openai/gpt-5-nano](/openai/gpt-5-nano) - [openai/gpt-5.1](/openai/gpt-5.1) - [openai/gpt-5.2](/openai/gpt-5.2) - [openai/gpt-5.2-pro](/openai/gpt-5.2-pro) - [openai/gpt-5.3-chat](/openai/gpt-5.3-chat) - [openai/gpt-oss-120b](/openai/gpt-oss-120b) - [perplexity/sonar](/perplexity/sonar) - [qwen/qwen3-235b-a22b](/qwen/qwen3-235b-a22b) - [x-ai/grok-3](/x-ai/grok-3) - [x-ai/grok-3-mini](/x-ai/grok-3-mini) - [x-ai/grok-4](/x-ai/grok-4) - [x-ai/grok-4.1-fast](/x-ai/grok-4.1-fast) - [z-ai/glm-5](/z-ai/glm-5)

TextImageAudioFileVideo

Context

2M

Group

Router

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Chutes

Xiaomi: MiMo-V2-Flash

TextReasoningLive

xiaomi/mimo-v2-flash

MiMo-V2-Flash is an open-source foundation language model developed by Xiaomi. It is a Mixture-of-Experts model with 309B total parameters and 15B active parameters, adopting hybrid attention architecture. MiMo-V2-Flash supports a hybrid-thinking toggle and a 256K context window, and excels at reasoning, coding, and agent scenarios. On SWE-bench Verified and SWE-bench Multilingual, MiMo-V2-Flash ranks as the top #1 open-source model globally, delivering performance comparable to Claude Sonnet 4.5 while costing only about 3.5% as much. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. [Learn more in our docs](https://openrouter.ai/docs/use-cases/reasoning-tokens#enable-reasoning-with-default-config).

Text

Context

262.1K

Group

Other

Pricing preview

Input Price: $0.09 /M tokens

Output Price: $0.29 /M tokens

Need volume pricing or launch support?

Contact sales

Amazon Bedrock

Anthropic: Claude Haiku 4.5

TextReasoningLive

anthropic/claude-haiku-4.5

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications. It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.

TextImage

Context

200K

Group

Claude

Pricing preview

Input Price: $1 /M tokens

Output Price: $5 /M tokens

Need volume pricing or launch support?

Contact sales

Relace

Relace: Relace Apply 3

TextLive

relace/relace-apply-3

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at 10,000 tokens/sec on average. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code> <update>{edit_snippet}</update> Zero Data Retention is enabled for Relace. Learn more about this model in their [documentation](https://docs.relace.ai/api-reference/instant-apply/apply)

Text

Context

256K

Group

Other

Pricing preview

Input Price: $0.85 /M tokens

Output Price: $1.25 /M tokens

Need volume pricing or launch support?

Contact sales

Unknown provider

Morph: Fast Apply

TextLive

morph/morph-v2

Morph Apply is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at 4000+ tokens per second. The model requires the prompt to be in the following format: <code>${originalCode}</code>\n<update>${updateSnippet}</update> Learn more about this model in their [documentation](https://docs.morphllm.com/)

Text

Context

32K

Group

Other

Pricing preview

No pricing information published for this model.

Need volume pricing or launch support?

Contact sales

Missing a route?

Need a model that is not in the directory yet?

Send the model name, expected traffic, and launch timing. We will confirm availability, onboarding priority, or an equivalent route already live on ImaRouter.

Support

support@imarouter.com

For model requests, onboarding priority, routing strategy, or rollout planning, reach out directly.

AI Model Directory | Seedance, Claude, ElevenLabs, Suno | ImaRouter