Models and providers
Every model has several unofficial providers, and each has its own price for input, output, and cache. DarkRouter unifies them behind one API and always picks the cheapest working route. Expand any model to see all providers and their prices.
Anthropic
10Claude Fable 5
Claude Fable 5 is a Mythos-class model for autonomous knowledge work and coding. It handles text, image, and file inputs to produce text outputs with reasoning support.
anthropic/claude-fable-5
Claude Opus 5
Claude Opus 5 is a flagship model excelling at demanding reasoning, coding, and long-horizon agentic tasks, including software development and visual analysis.
anthropic/claude-opus-5
Claude Sonnet 5
Sonnet 5 is Anthropic's most capable Sonnet model, excelling in coding, agents, and professional tasks with selectable reasoning effort levels.
anthropic/claude-sonnet-5
Claude Haiku 4.5
Fastest, most affordable Claude for everyday tasks.
anthropic/claude-haiku-4.5
Claude Sonnet 4.6
Balanced Claude with strong quality at a moderate price.
anthropic/claude-sonnet-4.6
Claude Opus 4.8
Newest Opus — top-tier coding, analysis, and long-horizon tasks.
anthropic/claude-opus-4.8
Claude Opus 4.7
Most capable Claude for deep reasoning and agentic work.
anthropic/claude-opus-4.7
Claude Opus 4.6
anthropic/claude-opus-4.6
Claude Opus 4.5
anthropic/claude-opus-4.5
Claude Sonnet 4.5
Claude Sonnet 4.5 is an advanced model optimized for agentic workflows and coding. It excels at complex coding tasks and benchmarks.
anthropic/claude-sonnet-4.5
OpenAI
15GPT-5.6 Sol
This model excels at complex reasoning, coding, and agentic workflows, particularly with command-line and multi-step coding tasks.
openai/gpt-5.6-sol
GPT-5.6 Terra
A balanced model for everyday coding, reasoning, and agentic tasks.
openai/gpt-5.6-terra
GPT-5.6 Luna
A fast, cost-efficient model for high-volume, latency-sensitive tasks like chat and classification. It offers capable reasoning for lightweight agentic workflows.
openai/gpt-5.6-luna
GPT-5.5
Latest frontier model with the strongest general capabilities.
openai/gpt-5.5
GPT-5.4 Pro
GPT-5.4 Pro is OpenAI's most advanced model, excelling at complex, high-stakes reasoning tasks with its enhanced capabilities.
openai/gpt-5.4-pro
GPT-5.4
Higher-quality reasoning for complex, multi-step tasks.
openai/gpt-5.4
GPT-5.4 Mini
Fast and cheap for high-volume, latency-sensitive workloads.
openai/gpt-5.4-mini
GPT-5.4 Nano
A fast, cost-efficient model for high-volume tasks, supporting text and image inputs with low latency.
openai/gpt-5.4-nano
GPT-5.4 Image 2
This model combines advanced reasoning with image generation for multimodal workflows. It excels at tasks involving text, code, and visual content.
openai/gpt-5.4-image-2
GPT-5.2
Balanced flagship for everyday reasoning, chat, and tool use.
openai/gpt-5.2
GPT-5 Mini
GPT-5 Mini is a smaller, faster, and cheaper version of GPT-5, ideal for lighter reasoning tasks while retaining instruction-following and safety features.
openai/gpt-5-mini
GPT-4.1
This advanced LLM excels at complex instruction following, software engineering tasks, and reasoning over very long contexts.
openai/gpt-4.1
GPT-4.1 Mini
A mid-sized model offering GPT-4o-level performance with lower latency and cost. It excels with a large context window and strong performance on challenging tasks.
openai/gpt-4.1-mini
GPT-4o
This is OpenAI's latest multimodal model, accepting text and image inputs to generate text outputs. It offers GPT-4 Turbo intelligence with enhanced speed.
openai/gpt-4o
GPT-4o-mini
This advanced small model handles text and image inputs, generating text outputs. It offers powerful capabilities at a significantly reduced cost.
openai/gpt-4o-mini
Google Gemini
9Gemini 3.6 Flash
Gemini 3.6 Flash is a fast, efficient model for coding and agentic workflows, producing polished outputs with minimal edits.
google/gemini-3.6-flash
Gemini 3.5 Flash
Latest Flash generation, optimised for speed and scale.
google/gemini-3.5-flash
Gemini 3.1 Pro
Refined Gemini 3 Pro with an even larger 2M-token context.
google/gemini-3.1-pro-preview
Gemini 3.1 Flash Lite
Gemini 3.1 Flash Lite is a high-efficiency multimodal model for low-latency, high-volume tasks. It handles text, image, video, audio, and PDF inputs.
google/gemini-3.1-flash-lite
Gemini 3.1 Flash Lite Preview
A high-efficiency model for high-volume use cases, offering improved quality and performance.
google/gemini-3.1-flash-lite-preview
Gemini 3 Flash Preview
A fast, cost-effective model for agentic workflows, chat, and coding, offering near Pro-level reasoning.
google/gemini-3-flash-preview
Gemini 2.5 Pro
High-quality reasoning across text, images, and code.
google/gemini-2.5-pro
Gemini 2.5 Flash
Fast, low-cost multimodal model with a huge context window.
google/gemini-2.5-flash
Gemini 2.5 Flash Lite
A fast, cost-efficient reasoning model for low-latency applications. It excels at quick text generation and high throughput.
google/gemini-2.5-flash-lite
DeepSeek
3DeepSeek V4 Pro
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model excelling at advanced reasoning and coding tasks.
deepseek/deepseek-v4-pro
DeepSeek V4 Flash
An efficiency-optimized Mixture-of-Experts model designed for fast inference. It excels at handling long contexts and complex reasoning tasks.
deepseek/deepseek-v4-flash
DeepSeek V3.2
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance.
deepseek/deepseek-v3.2
Kimi
4Kimi K3
Kimi K3 is an open-weight multimodal reasoning model adept at complex coding, knowledge work, and long-horizon agentic workflows.
moonshotai/kimi-k3
Kimi K2.7 Code
Kimi K2.7 Code is a coding-focused model designed for reliable, end-to-end programming tasks over long contexts. It utilizes a multimodal mixture-of-experts architecture.
moonshotai/kimi-k2.7-code
Kimi K2.6
Kimi K2.6 is a multimodal model excelling at long-horizon coding, UI/UX generation, and multi-agent orchestration across Python, Rust, and Go.
moonshotai/kimi-k2.6
Kimi K2.5
Kimi K2.5 is a multimodal model excelling at visual coding and agent swarm coordination. It leverages extensive pretraining for advanced capabilities.
moonshotai/kimi-k2.5
GLM
2GLM 5.2
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering, and complex multi-step automation.
z-ai/glm-5.2
GLM 5.1
GLM-5.1 excels at long-horizon coding tasks, offering independent and continuous operation for complex development.
z-ai/glm-5.1
MiniMax
2MiniMax M3
This multimodal foundation model accepts text, image, and video inputs to generate text outputs, excelling at long-horizon agentic tasks and coding.
minimax/minimax-m3
MiniMax M2.7
MiniMax-M2.7 is a large language model designed for autonomous productivity and continuous improvement. It excels at active participation in its own evolution through advanced agentic capabilities.
minimax/minimax-m2.7
xAI
1Grok 4.5
Grok 4.5 is SpaceXAI's most advanced model, excelling in coding, knowledge work, and STEM tasks.
x-ai/grok-4.5
Qwen
6Qwen3.7 Max
Qwen3.7-Max is a flagship model for agent-centric workloads, excelling at coding and office productivity tasks.
qwen/qwen3.7-max
Qwen3.7 Plus
A multimodal model that excels at text generation and understanding, with enhanced capabilities for image processing.
qwen/qwen3.7-plus
Qwen3.6 Flash
Qwen3.6 Flash is a fast, efficient multimodal model supporting text, image, and video input. It excels at processing long contexts.
qwen/qwen3.6-flash
Qwen3.5-Flash
Qwen3.5 Flash models are efficient vision-language models with a hybrid architecture, excelling at fast inference for multimodal tasks.
qwen/qwen3.5-flash
Qwen3.6 Plus
Qwen 3.6 Plus is a powerful, scalable LLM with a hybrid architecture, excelling at high-performance inference.
qwen/qwen3.6-plus
Qwen3 Max
Qwen3-Max is an advanced LLM with enhanced reasoning, instruction following, and multilingual capabilities. It excels at understanding and generating text across diverse topics and languages.
qwen/qwen3-max