Aliyun Bailian (Qwen + Wan) is now live — 13 models across two regions, routed to whichever side is cheaper for each model.Read more
Model Catalog

Every major model behind one OpenAI-compatible API

Router AI lists 250+ AI models with live pricing, context windows, and capabilities — route all via one OpenAI-compatible API.

275 of 275 models

Amazon Nova 2 Lite

amazon/nova-2-lite
Amazon

Nova 2 Lite is an advanced multimodal reasoning model with 1M context. Dynamically adjusts reasoning depth. Extended thinking on complex problems.

Context1M
Max output65K
Input$0.3564
Output$2.97
Params8B
textimage

Amazon Nova Lite

amazon/nova-lite
Amazon

Nova Lite is a multimodal understanding model. Multilingual with reasoning over text, images, and videos. Cost-effective for everyday tasks.

Context300K
Max output5K
Input$0.0648
Output$0.2592
Params7B
textimage

Amazon Nova Micro

amazon/nova-micro
Amazon

Amazon's fastest and most cost-effective text-only model. Ideal for high-throughput, low-latency tasks.

Context128K
Max output5K
Input$0.0378
Output$0.1512
Params7B
text

Amazon Nova Premier

amazon/nova-premier
Amazon

Amazon's most capable multimodal model for complex reasoning tasks. Best teacher for distilling custom models. Supports text, images, and videos.

Context1M
Max output5K
Input$2.70
Output$13.50
Params7B
textimage

Amazon Nova Pro

amazon/nova-pro
Amazon

Amazon Nova Pro is a multimodal understanding model. Multilingual with reasoning over text, images, and videos.

Context300K
Max output5K
Input$0.864
Output$3.46
Params7B
textimage

Amazon Titan Text Embeddings V2

amazon/titan-embed-v2
Amazon

Lightweight, efficient embedding model for high accuracy retrieval tasks. Supports flexible embedding sizes (1024, 512, 256) and 100+ languages.

Context8K
Max output--
Input$0.0216
Output--
Params2B
text

Baidu ERNIE 4.5 Turbo (128K)

baidu/ernie-4.5-turbo-128k
Baidu

Baidu ERNIE 4.5 Turbo — flagship reasoning model with 128K context.

Context131K
Max output8K
Input$0.1296
Output$0.5184
Params7B
text

Baidu ERNIE 4.5 Turbo VL (32K)

baidu/ernie-4.5-turbo-vl-32k
Baidu

Baidu ERNIE 4.5 Turbo vision-language with 32K context.

Context33K
Max output8K
Input$0.486
Output$1.46
Params7B
textimage

Claude Fable 5

anthropic/claude-fable-5
Anthropic

Anthropic's 2026-06 frontier model — top-tier agentic coding and reasoning, 1M context, 128K output.

Context1M
Max output128K
Input$10.80
Output$54.00
Params8B
textimage

Claude Fable 5.1

anthropic/claude-fable-5-1
Anthropic

Anthropic's 2026-09 frontier model — successor to Claude Fable 5, with gains in agentic coding, long-running agentic workflows, and knowledge work; 1M context, 128K output.

Context1M
Max output128K
Input$10.80
Output$54.00
Params8B
textimage

Claude Haiku 4.5

anthropic/claude-haiku-4.5
Anthropic

Claude Haiku 4.5 delivers near-frontier performance for a wide range of use cases, and stands out as one of the best coding and agent models–with the right speed and cost to power free products and high-volume user experiences. Use cases: Powering free tier user experiences: Claude Haiku 4.5 delivers near-frontier performance at a cost and speed that makes powering free agent products and agentic use cases economically viable at scale. Real-time experiences: Claude Haiku 4.5's speed is ideal for real-time applications like customer service agents and chatbots where response time is critical. Coding sub-agents: Use Claude Haiku 4.5 to power sub-agents, enabling multi-agent systems that tackle complex refactors, migrations, and large feature builds with quality and speed. Financial sub-agents: Use Claude Haiku 4.5 to monitor thousands of data streams—tracking regulatory changes, market signals, and portfolio risks to preemptively adapt compliance and trading systems at previously impossible scales. Research sub-agents: Perform parallel analyses across multiple data sources while maintaining fast response times. Ideal for rapid business intelligence, competitive analysis, and real-time decision support. Business tasks: Claude Haiku 4.5 is capable of producing and editing office files like slides, documents, and spreadsheets. It also better supports strategy and campaign planning, business analysis and brainstorming.

Context200K
Max output8K
Input$1.08
Output$5.40
Params7B
textimagepdf

Claude Opus 4

anthropic/claude-opus-4
Anthropic

Claude Opus 4 is Anthropic's most intelligent model and is state-of-the-art for coding and agent capabilities, especially agentic search. It excels for customers needing frontier intelligence: Advanced coding: Independently plan and execute complex development tasks end-to-end. It adapts to your style and maintains high code quality throughout. AI agents: Enable agents to tackle complex, multi-step tasks that require peak accuracy. Agentic search and research: Connect to multiple data sources to synthesize comprehensive insights across repositories. Long-horizon tasks and complex problem solving (virtual collaborator): Unlock new use cases involving long-horizon tasks that require memory, sustained reasoning, and long chains of actions. Content creation: Create human-quality content with natural prose. Produce long-form creative content, technical documentation, marketing copy, and frontend design mockups.

Context200K
Max output32K
Input$16.20
Output$81.00
Params8B
textimagepdf

Claude Opus 4.1

anthropic/claude-opus-4.1
Anthropic

Claude Opus 4.1 is Anthropic's most intelligent model and an industry leader for coding and agent capabilities, especially agentic search. It excels for customers needing frontier intelligence: Advanced coding: Independently plan and execute complex development tasks end-to-end. It adapts to your style, thoughtfully plans and pivots, and maintains high code quality throughout. Long-horizon tasks and complex problem solving (virtual collaborator): Unlock new use cases involving long-horizon tasks that require memory, sustained reasoning, and long chains of actions. AI agents: Enable agents to tackle complex, multi-step tasks that require peak accuracy. Agentic search and research: Connect to multiple data sources to synthesize comprehensive insights across repositories. Content creation: Create human-quality content with natural prose. Produce long-form creative content, technical documentation, marketing copy, and frontend design mockups. Memory and context management: Incorporates memory capabilities that allow it to effectively summarize and reference previous interactions.

Context200K
Max output32K
Input$16.20
Output$81.00
Params8B
textimagepdf

Claude Opus 4.5

anthropic/claude-opus-4.5
Anthropic

The next generation of Anthropic's most intelligent model, Claude Opus 4.5 is an industry leader across coding, agents, computer use, and enterprise workflows. Use cases: Coding: Opus 4.5 can confidently deliver multi-day software development projects in hours, working independently with the technical depth and taste to create efficient and straightforward solutions. It has improved performance across coding languages, with better planning and architecture choices - making it the ideal model for enterprise developers. Agents: Claude Opus 4.5, paired with our advanced tool use capabilities, enables more capable agents with new behaviors. Computer use: Our best computer-using model yet, Claude Opus 4.5 navigates new experiences with confident, consistent approaches that deliver more human-like browsing, enabling better web QA, workflow automation, and advanced user experiences. Enterprise workflows: Opus 4.5 can power agents that manage sprawling professional projects from start to finish. It better leverages memory to maintain context and consistency across files, alongside a step-change improvement in creating spreadsheets, slides, and docs. Financial analysis: Opus 4.5 connects the dots across complex information systems - regulatory filings, market reports, internal data - making sophisticated predictive modeling and proactive compliance possible. Cybersecurity: Opus 4.5 brings professional-grade analysis to security workflows, correlating logs, vulnerability databases, and threat intelligence for proactive threat detection and automated incident response.

Context200K
Max output64K
Input$5.40
Output$27.00
Params8B
textimagepdf

Claude Opus 4.5 (2025-11-01)

anthropic/claude-opus-4.5-20251101
Anthropic

Claude Opus 4.5 dated snapshot — pinned to 2025-11-01 release.

Context200K
Max output32K
Input$5.40
Output$27.00
Params8B
textimage

Claude Opus 4.6

anthropic/claude-opus-4.6
Anthropic

Claude Opus 4.6 is the next generation of our most intelligent model, and the world's best model for coding, enterprise agents, and professional work. Use cases include: Agents: Opus 4.6 is the world's best model for agentic workflows, orchestrating complex tasks across dozens of tools with industry-leading reliability. It proactively spins up subagents, parallelizes work, and drives tasks forward with minimal oversight. Coding: Opus 4.6 is the world's best coding model, excelling at long-horizon projects, complex implementations, and large-scale codebases. It handles the full lifecycle from architecture to deployment—so senior engineers can delegate their most complex work with confidence. Enterprise workflows: Opus 4.6 sets the standard for enterprise workflows, powering agents that manage sprawling projects end-to-end with professional polish, domain awareness, and industry-leading performance on spreadsheets, slides, and docs. Financial analysis: Opus 4.6 is Anthropic's most capable model for financial workflows, surfacing insights that would take analysts days to compile. It handles the nuance and precision that compliance-sensitive work demands. Cybersecurity: Opus 4.6 delivers the deepest reasoning for security workflows, catching subtle patterns and complex attack vectors with unmatched accuracy. Computer use: Opus 4.6 is our most capable computer-use model for complex workflows, bringing deep reasoning to multi-step tasks that span multiple applications and require planning and judgment.

Context1M
Max output128K
Input$5.40
Output$27.00
Params8B
textimagepdf

Claude Opus 4.6 Thinking

anthropic/claude-opus-4.6-thinking
Anthropic

Claude Opus 4.6 with extended-thinking mode pre-enabled.

Context200K
Max output32K
Input$5.40
Output$27.00
Params8B
textimage

Claude Opus 4.7

anthropic/claude-opus-4.7
Anthropic

Claude Opus 4.7 is Anthropic's most capable model — 13% coding lift over Opus 4.6, tripled image resolution (2576px / 3.75 MP), new xhigh effort level, and task budgets for autonomous agent loops. Best for coding, agents, enterprise workflows, cybersecurity, and financial analysis.

Context1M
Max output128K
Input$5.40
Output$27.00
Params8B
textimagepdf

Claude Opus 4.8

anthropic/claude-opus-4.8
Anthropic

Latest Anthropic flagship Opus model — frontier reasoning and tool use, available via partner network.

Context200K
Max output32K
Input$5.40
Output$27.00
Params8B
textimage

Claude Opus 5

anthropic/claude-opus-5
Anthropic

Anthropic's frontier Opus-tier model, released 2026-07-24. Best for coding, agents, and enterprise workflows that already run on the Opus family.

Context1M
Max output128K
Input$5.40
Output$27.00
Params8B
textimage

Claude Sonnet 4

anthropic/claude-sonnet-4
Anthropic

Claude Sonnet 4 balances impressive performance for coding with the right speed and cost for high-volume use cases: Coding: Handle everyday development tasks with enhanced performance-power code reviews, bug fixes, API integrations, and feature development with immediate feedback loops. AI Assistants: Power production-ready assistants for real-time applications—from customer support automation to operational workflows that require both intelligence and speed. Efficient research: Perform focused analysis across multiple data sources while maintaining fast response times. Ideal for rapid business intelligence, competitive analysis, and real-time decision support. Large-scale content: Generate and analyze content at scale with improved quality. Create customer communications, analyze user feedback, and produce marketing materials with the right balance of quality and throughput.

Context1M
Max output64K
Input$3.24
Output$16.20
Params8B
textimagepdf

Claude Sonnet 4.5

anthropic/claude-sonnet-4.5
Anthropic

Claude Sonnet 4.5 is our most capable model to date for building real-world agents and handling complex, long-horizon tasks–balancing the right speed and cost for high-volume use cases: Long-running agents: Power production-ready assistants for multi-step, real-time applications—from customer support automation to complex operational workflows that require peak accuracy, intelligence, and speed. Coding: Handle everyday development tasks with enhanced performance––or plan and execute complex software projects spanning hours or days––with the ability to save, maintain, and reference information across multiple sessions. Cybersecurity: Deploy agents that autonomously patch vulnerabilities before exploitation––shifting from reactive detection to proactive defense. Financial analysis: Conduct entry-level financial analysis, deliver advanced predictive analysis, or preemptively develop intelligent risk management strategies that leverage best-in-class domain knowledge. Computer use: Claude Sonnet 4.5 is our most accurate model for computer use, enabling developers to direct Claude to use computers the way people do. Research: Perform focused analysis across multiple data sources, turning expert analysis into final deliverables. Ideal for complex problem solving, rapid business intelligence, and real-time decision support.

Context1M
Max output64K
Input$3.24
Output$16.20
Params8B
textimagepdf

Claude Sonnet 4.6

anthropic/claude-sonnet-4.6
Anthropic

Claude Sonnet 4.6 delivers frontier intelligence at scale—built for coding, agents, and enterprise workflows.

Context1M
Max output64K
Input$3.24
Output$16.20
Params8B
textimagepdf

Claude Sonnet 5

anthropic/claude-sonnet-5
Anthropic

Anthropic's frontier Sonnet-tier model — top-tier agentic coding and reasoning, 1M context, 128K output.

Context1M
Max output128K
Input$2.16
Output$10.80
Params8B
textimage

Codestral

mistral/codestral
Mistral

Mistral's specialized coding model. Optimized for code generation, completion, and analysis.

Context256K
Max output16K
Input$0.324
Output$0.972
Params7B
text

CogVideoX 3

zhipu/cogvideox-3
Zhipu

Zhipu AI CogVideoX 3 — flagship text/image-to-video generation. Up to 5s or 10s, up to 4K resolution.

Context--
Max output--
Input--
Output--
Params7B
textimagevideo

CogView 4

zhipu/cogview-4
Zhipu

Zhipu AI CogView 4 — text-to-image generation with strong bilingual prompt understanding.

Context--
Max output--
Input--
Output--
Params5B
textimage

Cohere Embed V4

cohere/embed-v4
Cohere

Multilingual multimodal embedding model capable of transforming images, texts, and interleaved content into vector representations. State-of-the-art performance with byte/binary quantization and matryoshka embeddings for compression.

Context8K
Max output--
Input$0.1296
Output--
Params3B
textimage

DeepSeek R1

deepseek/deepseek-r1
Deepseek

DeepSeek R1 (671B total, 37B active MoE) is a reasoning model that uses chain-of-thought with <think> tags to solve complex problems. Excels at math, coding, and scientific reasoning tasks with transparent step-by-step thinking.

Context128K
Max output33K
Input$1.46
Output$5.83
Params6B
text

DeepSeek V3.1

deepseek/deepseek-v3.1
Deepseek

DeepSeek V3.1 is a hybrid model supporting both thinking and non-thinking modes. Features enhanced tool calling capabilities for agent-based tasks. Thinking mode maintains answer quality comparable to DeepSeek-R1 with improved response times.

Context128K
Max output33K
Input$0.6264
Output$1.81
Params8B
text

DeepSeek V3.1 Terminus

deepseek/deepseek-v3.1-terminus
Deepseek

DeepSeek V3.1 Terminus — refined variant of V3.1 optimized for tool calling and structured generation tasks.

Context131K
Max output33K
Input$0.2916
Output$1.08
Params7B
text

DeepSeek V3.2

deepseek/deepseek-v3.2
Deepseek

DeepSeek V3.2 (685B total, 37B active MoE) harmonizes high computational efficiency with superior reasoning and agent performance. Features DeepSeek Sparse Attention for long-context efficiency and a scalable reinforcement learning framework. Excels at long-context reasoning, tool-using agents, function calling, JSON output, and FIM.

Context128K
Max output33K
Input$0.6696
Output$2.00
Params7B
text

DeepSeek V3.2 Exp

deepseek/deepseek-v3.2-exp
Deepseek

DeepSeek V3.2 Exp — experimental variant of V3.2 with enhanced general-purpose capabilities. Strong at tool use, structured output, and multi-turn conversation.

Context131K
Max output16K
Input$0.2916
Output$0.4428
Params7B
text

DeepSeek V4 Flash

deepseek/deepseek-v4-flash
Deepseek

DeepSeek V4 Flash — fast, cost-efficient model with 1M context window. Supports reasoning, tool calling, and structured output.

Context1M
Max output384K
Input$0.162
Output$0.648
Params12B
text

DeepSeek V4 Pro

deepseek/deepseek-v4-pro
Deepseek

DeepSeek V4 Pro — high-capability model with 1M context window. Superior reasoning, coding, and agent performance with tool calling and structured output.

Context1M
Max output384K
Input$0.7128
Output$2.14
Params12B
text

Doubao 1.5 Lite 32k

doubao/doubao-1-5-lite-32k
Doubao

Doubao 1.5 Lite (32k context) — cost-efficient ByteDance chat model for high-volume routine tasks.

Context33K
Max output16K
Input$0.0486
Output$0.0972
Params7B
text

Doubao 1.5 Pro (32K, 250115)

bytedance/doubao-1.5-pro-32k-250115
Bytedance

ByteDance Doubao 1.5 Pro snapshot from 250115, 32K context.

Context33K
Max output8K
Input$0.1296
Output$0.324
Params7B
text

Doubao 1.5 Pro 32k

doubao/doubao-1-5-pro-32k
Doubao

Doubao 1.5 Pro — ByteDance flagship general-purpose chat model with tools and JSON mode.

Context33K
Max output16K
Input$0.1296
Output$0.324
Params7B
text

Doubao 1.5 Vision Pro 32k

doubao/doubao-1-5-vision-pro-32k
Doubao

Doubao 1.5 Vision Pro (32k context) — extended-context vision-language variant.

Context33K
Max output16K
Input$0.486
Output$1.46
Params7B
textimage

Doubao Seed 1.6

bytedance/doubao-seed-1.6
Bytedance

ByteDance Doubao Seed 1.6 — general-purpose chat with strong code and tool-use performance.

Context256K
Max output8K
Input$0.270
Output$2.16
Params7B
text

Doubao Seed 1.6

doubao/doubao-seed-1-6
Doubao

Doubao Seed 1.6 — ByteDance Seed-series next-gen general model with tools and structured output.

Context131K
Max output33K
Input$0.270
Output$2.16
Params7B
text

Doubao Seed 1.6 Flash

doubao/doubao-seed-1-6-flash
Doubao

Doubao Seed 1.6 Flash — ultra-low-latency variant of Seed 1.6, ideal for chat and agent loops.

Context131K
Max output33K
Input$0.081
Output$0.324
Params7B
text

Doubao Seed 1.6 Vision

doubao/doubao-seed-1-6-vision
Doubao

Doubao Seed 1.6 Vision — vision-language Seed 1.6 variant for multimodal understanding.

Context131K
Max output33K
Input$0.1296
Output$1.30
Params7B
textimage

Doubao Seed 1.8

doubao/doubao-seed-1-8
Doubao

Doubao Seed 1.8 — incremental upgrade of Seed 1.6 with improved tool-call reliability.

Context131K
Max output33K
Input$0.270
Output$2.16
Params7B
text

Doubao Seed 2.0 Code

doubao/doubao-seed-2-0-code
Doubao

Doubao Seed 2.0 Code — coding-specialized Seed 2.0 model for code generation, refactor, and review.

Context131K
Max output33K
Input$0.540
Output$3.24
Params7B
text

Doubao Seed 2.0 Code Preview (260215)

bytedance/doubao-seed-2.0-code-preview-260215
Bytedance

ByteDance Doubao Seed 2.0 Code Preview — code-specialist variant with tiered pricing.

Context262K
Max output8K
Input$0.540
Output$3.24
Params7B
text

Doubao Seed 2.0 Lite

doubao/doubao-seed-2-0-lite
Doubao

Doubao Seed 2.0 Lite — cost-efficient Seed 2.0 variant.

Context131K
Max output33K
Input$0.270
Output$2.16
Params7B
text

Doubao Seed 2.0 Lite (260215)

bytedance/doubao-seed-2.0-lite-260215
Bytedance

ByteDance Doubao Seed 2.0 Lite — cost-efficient variant with tiered pricing.

Context262K
Max output8K
Input$0.1836
Output$1.09
Params7B
text

Doubao Seed 2.0 Mini

doubao/doubao-seed-2-0-mini
Doubao

Doubao Seed 2.0 Mini — smallest Seed 2.0 variant for high-QPS edge use cases.

Context131K
Max output33K
Input$0.108
Output$0.432
Params7B
text

Doubao Seed 2.0 Mini (260215)

bytedance/doubao-seed-2.0-mini-260215
Bytedance

ByteDance Doubao Seed 2.0 Mini — ultra-low-cost variant with tiered pricing.

Context262K
Max output8K
Input$0.108
Output$0.432
Params7B
text

Doubao Seed 2.0 Pro

doubao/doubao-seed-2-0-pro
Doubao

Doubao Seed 2.0 Pro — flagship Seed 2.0 model with strongest reasoning and tool use.

Context131K
Max output33K
Input$0.540
Output$3.24
Params7B
text

Doubao Seed 2.0 Pro (260215)

bytedance/doubao-seed-2.0-pro-260215
Bytedance

ByteDance Doubao Seed 2.0 Pro snapshot 260215. Frontier model with tiered pricing.

Context262K
Max output8K
Input$0.540
Output$3.24
Params7B
text

Doubao Seed 2.1 Pro

bytedance/doubao-seed-2.1-pro
Bytedance

ByteDance Doubao Seed 2.1 Pro — frontier model with text + image input. 256K context. Strong coding and agent performance. Cost ¥6/$0.833 / ¥30/$4.167 per MTok (I/O).

Context256K
Max output33K
Input$0.9612
Output$4.81
Params7B
textimage

Doubao Seed 2.1 Turbo

bytedance/doubao-seed-2.1-turbo
Bytedance

ByteDance Doubao Seed 2.1 Turbo — cost-efficient tier. 256K context. Text input only. Cost ¥3/$0.417 / ¥15/$2.083 per MTok (I/O).

Context256K
Max output33K
Input$0.540
Output$2.70
Params7B
text

Doubao Seed Character

doubao/doubao-seed-character
Doubao

Doubao Seed Character — roleplay / persona-driven chat model.

Context131K
Max output33K
Input$0.1296
Output$0.324
Params7B
text

Doubao Seed Code

doubao/doubao-seed-code
Doubao

Doubao Seed Code — code generation and code understanding model.

Context131K
Max output33K
Input$0.1944
Output$1.30
Params7B
text

Doubao Seed Evolving

bytedance/doubao-seed-evolving
Bytedance

ByteDance Doubao Seed Evolving — continuously-updated Seed-series model. Latest improvements auto-deployed. Text only.

Context256K
Max output33K
Input$0.9612
Output$4.81
Params7B
text

Doubao Seedance 2.0

doubao/doubao-seedance-2-0
Doubao

Doubao SeedDance 2.0 — text/image-to-video generation, flagship quality tier.

Context--
Max output--
Input--
Output--
Params5B
textimagevideo

Doubao Seedance 2.0 Fast

doubao/doubao-seedance-2-0-fast
Doubao

Doubao SeedDance 2.0 Fast — faster, lower-cost variant of SeedDance 2.0 for iterative video drafting.

Context--
Max output--
Input--
Output--
Params5B
textimagevideo

Doubao Seedream 4.5

doubao/doubao-seedream-4-5
Doubao

Doubao SeedDream 4.5 — text/image-to-image generation, Chinese-bilingual prompt support.

Context--
Max output--
Input--
Output--
Params5B
textimage
Model Catalog — 275+ AI Models, Pricing & Context Windows | Router AI