DarkRouter DarkRouter
Catalogue · 52 models

Models and providers

Every model has several unofficial providers, and each has its own price for input, output, and cache. DarkRouter unifies them behind one API and always picks the cheapest working route. Expand any model to see all providers and their prices.

A

Anthropic

10
A

Claude Fable 5

context 1 M

Claude Fable 5 is a Mythos-class model for autonomous knowledge work and coding. It handles text, image, and file inputs to produce text outputs with reasoning support.

INPUT / 1M
$2.43 $10
OUTPUT / 1M
$12.17 $50
model identifier
anthropic/claude-fable-5
Wcnbai · ClaudeCode −76%
input$2.43
output$12.17
cache$0.243
Wcnbai · ClaudeCode-Max −74%
input$2.6
output$13
cache$0.26
Wcnbai · ClaudeCode-Max(Pro) −74%
input$2.6
output$13
cache$0.26
Apixly · claude code −65%
input$3.5
output$17.5
cache$0.35
Apixly · claude −60%
input$4
output$20
cache$0.4
Apixly · claude hybrid channel −51%
input$4.9
output$24.5
cache$0.49
Apixly · claude vertex channel −45%
input$5.5
output$27.5
cache$0.55
Wcnbai · Offical-Claude −15%
input$8.53
output$42.65
cache$0.853
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Opus 5

context 1 M

Claude Opus 5 is a flagship model excelling at demanding reasoning, coding, and long-horizon agentic tasks, including software development and visual analysis.

INPUT / 1M
$0.6 $5
OUTPUT / 1M
$0.6 $25
model identifier
anthropic/claude-opus-5
Wellflow · direct −96%
input$0.6
output$0.6
cache$0.6
Apixly · [discount] claude −80%
input$1
output$5
cache$0.1
Wcnbai · ClaudeCode-Max(Pro) −74%
input$1.3
output$6.5
cache$0.13
Wcnbai · ClaudeCode-Max −74%
input$1.3
output$6.5
cache$0.13
Apixly · claude code −65%
input$1.75
output$8.75
cache$0.175
Apixly · claude −60%
input$2
output$10
cache$0.2
Apixly · claude hybrid channel −51%
input$2.45
output$12.25
cache$0.245
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Sonnet 5

context 1 M

Sonnet 5 is Anthropic's most capable Sonnet model, excelling in coding, agents, and professional tasks with selectable reasoning effort levels.

INPUT / 1M
$0.4 $2
OUTPUT / 1M
$0.4 $10
model identifier
anthropic/claude-sonnet-5
Wellflow · direct −93%
input$0.4
output$0.4
cache$0.4
Apixly · [discount] claude −80%
input$0.4
output$2
cache$0.04
Wcnbai · ClaudeCode −76%
input$0.487
output$2.43
cache$0.049
Wcnbai · ClaudeCode-Max(Pro) −74%
input$0.52
output$2.6
cache$0.052
Apixly · claude code −65%
input$0.7
output$3.5
cache$0.07
Apixly · claude −60%
input$0.8
output$4
cache$0.08
Apixly · claude hybrid channel −51%
input$0.98
output$4.9
cache$0.098
Apixly · claude vertex channel −45%
input$1.1
output$5.5
cache$0.11
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Haiku 4.5

context 200 K

Fastest, most affordable Claude for everyday tasks.

INPUT / 1M
$0.2 $1
OUTPUT / 1M
$1 $5
model identifier
anthropic/claude-haiku-4.5
Apixly · [discount] claude −80%
input$0.2
output$1
cache$0.02
Wcnbai · ClaudeCode −76%
input$0.244
output$1.22
cache$0.024
Wcnbai · ClaudeCode-Max(Pro) −74%
input$0.26
output$1.3
cache$0.026
Wcnbai · ClaudeCode-Max −74%
input$0.26
output$1.3
cache$0.026
Apixly · claude code −65%
input$0.35
output$1.75
cache$0.035
Apixly · claude −60%
input$0.4
output$2
cache$0.04
Apixly · claude hybrid channel −51%
input$0.49
output$2.45
cache$0.049
Apixly · claude vertex channel −45%
input$0.55
output$2.75
cache$0.055
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Sonnet 4.6

context 200 K

Balanced Claude with strong quality at a moderate price.

INPUT / 1M
$0.4 $3
OUTPUT / 1M
$0.4 $15
model identifier
anthropic/claude-sonnet-4.6
Wellflow · direct −96%
input$0.4
output$0.4
cache$0.4
Apixly · [discount] claude −80%
input$0.6
output$3
cache$0.06
Wcnbai · ClaudeCode −76%
input$0.731
output$3.66
cache$0.073
Wcnbai · ClaudeCode-Max(Pro) −74%
input$0.78
output$3.9
cache$0.078
Wcnbai · ClaudeCode-Max −74%
input$0.78
output$3.9
cache$0.078
Apixly · claude code −65%
input$1.05
output$5.25
cache$0.105
Apixly · claude −60%
input$1.2
output$6
cache$0.12
Apixly · claude hybrid channel −51%
input$1.47
output$7.35
cache$0.147
Apixly · claude vertex channel −45%
input$1.65
output$8.25
cache$0.165
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Opus 4.8

context 200 K

Newest Opus — top-tier coding, analysis, and long-horizon tasks.

INPUT / 1M
$0.6 $5
OUTPUT / 1M
$0.6 $25
model identifier
anthropic/claude-opus-4.8
Wellflow · direct −96%
input$0.6
output$0.6
cache$0.6
Apixly · [discount] claude −80%
input$1
output$5
cache$0.1
Wcnbai · ClaudeCode −76%
input$1.22
output$6.09
cache$0.122
Apixly · claude −60%
input$2
output$10
cache$0.2
Apixly · claude hybrid channel −51%
input$2.45
output$12.25
cache$0.245
Apixly · claude vertex channel −45%
input$2.75
output$13.75
cache$0.275
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Opus 4.7

context 200 K

Most capable Claude for deep reasoning and agentic work.

INPUT / 1M
$0.6 $5
OUTPUT / 1M
$0.6 $25
model identifier
anthropic/claude-opus-4.7
Wellflow · direct −96%
input$0.6
output$0.6
cache$0.6
Apixly · [discount] claude −80%
input$1
output$5
cache$0.1
Wcnbai · ClaudeCode −76%
input$1.22
output$6.09
cache$0.122
Wcnbai · ClaudeCode-Max −74%
input$1.3
output$6.5
cache$0.13
Apixly · claude code −65%
input$1.75
output$8.75
cache$0.175
Apixly · claude −60%
input$2
output$10
cache$0.2
Apixly · claude hybrid channel −51%
input$2.45
output$12.25
cache$0.245
Apixly · claude vertex channel −45%
input$2.75
output$13.75
cache$0.275
Wcnbai · Offical-Claude −15%
input$4.26
output$21.32
cache$0.426
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Opus 4.6

INPUT / 1M
$0.6 $5
OUTPUT / 1M
$0.6 $25
model identifier
anthropic/claude-opus-4.6
Wellflow · direct −96%
input$0.6
output$0.6
cache$0.6
Apixly · [discount] claude −80%
input$1
output$5
cache$0.1
Wcnbai · ClaudeCode −76%
input$1.22
output$6.09
cache$0.122
Wcnbai · ClaudeCode-Max(Pro) −74%
input$1.29
output$6.47
cache$0.129
Wcnbai · ClaudeCode-Max −74%
input$1.3
output$6.5
cache$0.13
Apixly · claude code −65%
input$1.75
output$8.75
cache$0.175
Apixly · claude −60%
input$2
output$10
cache$0.2
Apixly · claude hybrid channel −51%
input$2.45
output$12.25
cache$0.245
Apixly · claude vertex channel −45%
input$2.75
output$13.75
cache$0.275
Wcnbai · Offical-Claude −15%
input$4.26
output$21.32
cache$0.426
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Opus 4.5

INPUT / 1M
$1.22 $5
OUTPUT / 1M
$6.09 $25
model identifier
anthropic/claude-opus-4.5
Wcnbai · ClaudeCode −76%
input$1.22
output$6.09
cache$0.122
Wcnbai · ClaudeCode-Max −74%
input$1.3
output$6.5
cache$0.13
Wcnbai · ClaudeCode-Max(Pro) −74%
input$1.3
output$6.5
cache$0.13
Apixly · claude code −65%
input$1.75
output$8.75
cache$0.175
Apixly · claude −60%
input$2
output$10
cache$0.2
Apixly · claude hybrid channel −51%
input$2.45
output$12.25
cache$0.245
Apixly · claude vertex channel −45%
input$2.75
output$13.75
cache$0.275
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
A

Claude Sonnet 4.5

context 1 M

Claude Sonnet 4.5 is an advanced model optimized for agentic workflows and coding. It excels at complex coding tasks and benchmarks.

INPUT / 1M
$0.6 $3
OUTPUT / 1M
$3 $15
model identifier
anthropic/claude-sonnet-4.5
Apixly · [discount] claude −80%
input$0.6
output$3
cache$0.06
Wcnbai · ClaudeCode −76%
input$0.731
output$3.66
cache$0.073
Wcnbai · ClaudeCode-Max −74%
input$0.78
output$3.9
cache$0.078
Wcnbai · ClaudeCode-Max(Pro) −74%
input$0.78
output$3.9
cache$0.078
Apixly · claude −60%
input$1.2
output$6
cache$0.12
Apixly · claude hybrid channel −51%
input$1.47
output$7.35
cache$0.147
Apixly · claude vertex channel −45%
input$1.65
output$8.25
cache$0.165
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

OpenAI

15
O

GPT-5.6 Sol

context 1.1 M

This model excels at complex reasoning, coding, and agentic workflows, particularly with command-line and multi-step coding tasks.

INPUT / 1M
$1 $5
OUTPUT / 1M
$1 $30
model identifier
openai/gpt-5.6-sol
Subscriptions · direct −94%
input$1
output$1
cache
Apixly · gpt −90%
input$0.5
output$3
cache$0.05
Wcnbai · Codex −49%
input$3.24
output$14.6
cache$0.324
Wcnbai · AZ +7%
input$6.81
output$30.64
cache$0.681
Wcnbai · az定制 +53%
input$9.73
output$43.77
cache$0.973
Wcnbai · oai +150%
input$15.89
output$71.52
cache$1.59
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.6 Terra

context 1.1 M

A balanced model for everyday coding, reasoning, and agentic tasks.

INPUT / 1M
$0.4 $2.5
OUTPUT / 1M
$0.4 $15
model identifier
openai/gpt-5.6-terra
Subscriptions · direct −95%
input$0.4
output$0.4
cache
Apixly · gpt −90%
input$0.25
output$1.5
cache$0.025
Wcnbai · Codex −49%
input$1.62
output$7.3
cache$0.162
Wcnbai · AZ +7%
input$3.4
output$15.32
cache$0.34
Wcnbai · az定制 +53%
input$4.86
output$21.89
cache$0.486
Wcnbai · oai +150%
input$7.95
output$35.76
cache$0.795
Wcnbai · AZ蒸馏 +155%
input$8.12
output$36.55
cache$0.812
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.6 Luna

context 1.1 M

A fast, cost-efficient model for high-volume, latency-sensitive tasks like chat and classification. It offers capable reasoning for lightweight agentic workflows.

INPUT / 1M
$0.1 $1
OUTPUT / 1M
$0.1 $6
model identifier
openai/gpt-5.6-luna
Subscriptions · direct −97%
input$0.1
output$0.1
cache
Apixly · gpt −90%
input$0.1
output$0.6
cache$0.01
Wcnbai · Codex −49%
input$0.649
output$2.92
cache$0.065
Wcnbai · AZ +7%
input$1.36
output$6.13
cache$0.136
Wcnbai · az定制 +53%
input$1.95
output$8.75
cache$0.195
Wcnbai · oai +150%
input$3.18
output$14.3
cache$0.318
Wcnbai · AZ蒸馏 +155%
input$3.25
output$14.62
cache$0.325
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.5

context 1 M

Latest frontier model with the strongest general capabilities.

INPUT / 1M
$0.3 $5
OUTPUT / 1M
$0.3 $30
model identifier
openai/gpt-5.5
Wellflow · direct −98%
input$0.3
output$0.3
cache$0.3
Subscriptions · direct −98%
input$0.39
output$0.39
cache$0.39
Wcnbai · Codex −49%
input$3.25
output$14.62
cache$0.325
Wcnbai · AZ +7%
input$6.82
output$30.71
cache$0.682
Wcnbai · oai +150%
input$15.92
output$71.65
cache$1.59
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.4 Pro

context 1.1 M

GPT-5.4 Pro is OpenAI's most advanced model, excelling at complex, high-stakes reasoning tasks with its enhanced capabilities.

INPUT / 1M
$24.37 $30
OUTPUT / 1M
$146.22 $180
model identifier
openai/gpt-5.4-pro
Wcnbai · AZ蒸馏 −19%
input$24.37
output$146.22
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.4

context 1 M

Higher-quality reasoning for complex, multi-step tasks.

INPUT / 1M
$0.19 $2.5
OUTPUT / 1M
$0.19 $15
model identifier
openai/gpt-5.4
Subscriptions · direct −98%
input$0.19
output$0.19
cache$0.19
Wcnbai · Codex −49%
input$1.62
output$7.31
cache$0.162
Wcnbai · oai +150%
input$7.96
output$35.82
cache$0.796
Wcnbai · AZ蒸馏 +155%
input$8.12
output$36.55
cache$0.812
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.4 Mini

context 400 K

Fast and cheap for high-volume, latency-sensitive workloads.

INPUT / 1M
$0.2 $0.75
OUTPUT / 1M
$0.2 $4.5
model identifier
openai/gpt-5.4-mini
Subscriptions · direct −92%
input$0.2
output$0.2
cache$0.2
Apixly · gpt −90%
input$0.075
output$0.45
cache$0.008
Wcnbai · Codex −84%
input$0.122
output$0.731
cache$0.012
Wcnbai · AZ −66%
input$0.256
output$1.54
cache$0.026
Wcnbai · az定制 −51%
input$0.366
output$2.19
cache$0.037
Wcnbai · oai −20%
input$0.597
output$3.58
cache$0.06
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.4 Nano

context 400 K

A fast, cost-efficient model for high-volume tasks, supporting text and image inputs with low latency.

INPUT / 1M
$0.068 $0.2
OUTPUT / 1M
$0.426 $1.25
model identifier
openai/gpt-5.4-nano
Wcnbai · AZ −66%
input$0.068
output$0.426
cache$0.007
Wcnbai · az定制 −51%
input$0.097
output$0.609
cache$0.01
Wcnbai · oai −20%
input$0.159
output$0.995
cache$0.016
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.4 Image 2

context 270 K

This model combines advanced reasoning with image generation for multimodal workflows. It excels at tasks involving text, code, and visual content.

INPUT / 1M
$0.609 $8
OUTPUT / 1M
$3.66 $15
model identifier
openai/gpt-5.4-image-2
Wcnbai · anti −81%
input$0.609
output$3.66
cache
Wcnbai · anti per request
per request$0.015 /req
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5.2

context 400 K

Balanced flagship for everyday reasoning, chat, and tool use.

INPUT / 1M
$1.39 $1.75
OUTPUT / 1M
$11.14 $14
model identifier
openai/gpt-5.2
Wcnbai · oai −20%
input$1.39
output$11.14
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-5 Mini

context 400 K

GPT-5 Mini is a smaller, faster, and cheaper version of GPT-5, ideal for lighter reasoning tasks while retaining instruction-following and safety features.

INPUT / 1M
$0.085 $0.25
OUTPUT / 1M
$0.682 $2
model identifier
openai/gpt-5-mini
Wcnbai · AZ −66%
input$0.085
output$0.682
cache$0.009
Wcnbai · az定制 −51%
input$0.122
output$0.975
cache$0.012
Wcnbai · oai −21%
input$0.198
output$1.59
cache$0.02
Wcnbai · AZ蒸馏 −19%
input$0.203
output$1.62
cache$0.02
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-4.1

context 1 M

This advanced LLM excels at complex instruction following, software engineering tasks, and reasoning over very long contexts.

INPUT / 1M
$1.59 $2
OUTPUT / 1M
$6.37 $8
model identifier
openai/gpt-4.1
Wcnbai · oai −20%
input$1.59
output$6.37
cache$0.398
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-4.1 Mini

context 1 M

A mid-sized model offering GPT-4o-level performance with lower latency and cost. It excels with a large context window and strong performance on challenging tasks.

INPUT / 1M
$0.136 $0.4
OUTPUT / 1M
$0.546 $1.6
model identifier
openai/gpt-4.1-mini
Wcnbai · AZ −66%
input$0.136
output$0.546
cache$0.034
Wcnbai · az定制 −51%
input$0.195
output$0.78
cache$0.049
Wcnbai · oai −20%
input$0.318
output$1.27
cache$0.08
Wcnbai · AZ蒸馏 −19%
input$0.325
output$1.3
cache$0.081
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-4o

context 130 K

This is OpenAI's latest multimodal model, accepting text and image inputs to generate text outputs. It offers GPT-4 Turbo intelligence with enhanced speed.

INPUT / 1M
$0.853 $2.5
OUTPUT / 1M
$3.41 $10
model identifier
openai/gpt-4o
Wcnbai · AZ −66%
input$0.853
output$3.41
cache$0.426
Wcnbai · az定制 −51%
input$1.22
output$4.87
cache$0.609
Wcnbai · oai −20%
input$1.99
output$7.96
cache$0.995
Wcnbai · AZ蒸馏 −19%
input$2.03
output$8.12
cache$1.02
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
O

GPT-4o-mini

context 130 K

This advanced small model handles text and image inputs, generating text outputs. It offers powerful capabilities at a significantly reduced cost.

INPUT / 1M
$0.051 $0.15
OUTPUT / 1M
$0.205 $0.6
model identifier
openai/gpt-4o-mini
Wcnbai · AZ −66%
input$0.051
output$0.205
cache$0.026
Wcnbai · az定制 −51%
input$0.073
output$0.292
cache$0.037
Wcnbai · oai −20%
input$0.119
output$0.478
cache$0.06
Wcnbai · AZ蒸馏 −19%
input$0.122
output$0.487
cache$0.061
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Google Gemini

9
G

Gemini 3.6 Flash

context 1 M

Gemini 3.6 Flash is a fast, efficient model for coding and agentic workflows, producing polished outputs with minimal edits.

INPUT / 1M
$0.3 $1.5
OUTPUT / 1M
$0.3 $7.5
model identifier
google/gemini-3.6-flash
Subscriptions · direct −93%
input$0.3
output$0.3
cache
Wcnbai · Vertex −71%
input$0.439
output$2.19
cache$0.044
Wcnbai · Gemini优质文本 −51%
input$0.731
output$3.66
cache$0.073
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 3.5 Flash

context 1 M

Latest Flash generation, optimised for speed and scale.

INPUT / 1M
$0.244 $1.5
OUTPUT / 1M
$1.46 $9
model identifier
google/gemini-3.5-flash
Wcnbai · Gemini CLI −84%
input$0.244
output$1.46
cache$0.024
Wcnbai · 临时Banana画图 −73%
input$0.409
output$2.46
cache$0.041
Wcnbai · Vertex −71%
input$0.437
output$2.62
cache$0.044
Wcnbai · Banana优质画图 −59%
input$0.609
output$3.66
cache$0.061
Wcnbai · 企业Banana画图 −51%
input$0.731
output$4.39
cache$0.073
Wcnbai · Gemini优质文本 −51%
input$0.731
output$4.39
cache$0.073
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 3.1 Pro

context 2 M

Refined Gemini 3 Pro with an even larger 2M-token context.

INPUT / 1M
$0.5 $2
OUTPUT / 1M
$4.5 $12
model identifier
google/gemini-3.1-pro-preview
Subscriptions · direct −64%
input$0.5
output$4.5
cache
Apixly · gemini cli −60%
input$0.8
output$4.8
cache$0.08
Apixly · gemini ai studio −50%
input$1
output$6
cache$0.1
Wcnbai · Gemini CLI −49%
input$1.3
output$5.85
cache$0.13
Wcnbai · Gemini优质文本 +53%
input$3.9
output$17.55
cache$0.39
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 3.1 Flash Lite

context 1 M

Gemini 3.1 Flash Lite is a high-efficiency multimodal model for low-latency, high-volume tasks. It handles text, image, video, audio, and PDF inputs.

INPUT / 1M
$0.04 $0.25
OUTPUT / 1M
$0.4 $1.5
model identifier
google/gemini-3.1-flash-lite
Subscriptions · direct −75%
input$0.04
output$0.4
cache
Wcnbai · Gemini优质文本 −51%
input$0.122
output$0.731
cache$0.012
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 3.1 Flash Lite Preview

context 1 M

A high-efficiency model for high-volume use cases, offering improved quality and performance.

INPUT / 1M
$0.04 $0.25
OUTPUT / 1M
$0.243 $1.5
model identifier
google/gemini-3.1-flash-lite-preview
Wcnbai · Gemini CLI −84%
input$0.04
output$0.243
cache$0.004
Wcnbai · 临时Banana画图 −73%
input$0.068
output$0.409
cache$0.007
Wcnbai · Banana优质画图 −59%
input$0.102
output$0.609
cache$0.01
Wcnbai · 企业Banana画图 −51%
input$0.121
output$0.728
cache$0.012
Wcnbai · Gemini优质文本 −51%
input$0.122
output$0.731
cache$0.012
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 3 Flash Preview

context 1 M

A fast, cost-effective model for agentic workflows, chat, and coding, offering near Pro-level reasoning.

INPUT / 1M
$0.081 $0.5
OUTPUT / 1M
$0.487 $3
model identifier
google/gemini-3-flash-preview
Wcnbai · Gemini CLI −84%
input$0.081
output$0.487
cache$0.008
Apixly · gemini cli −60%
input$0.2
output$1.2
cache$0.02
Wcnbai · Gemini优质文本 −51%
input$0.244
output$1.46
cache$0.024
Apixly · gemini ai studio −50%
input$0.25
output$1.5
cache$0.025
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 2.5 Pro

context 1 M

High-quality reasoning across text, images, and code.

INPUT / 1M
$0.5 $1.25
OUTPUT / 1M
$4 $10
model identifier
google/gemini-2.5-pro
Apixly · gemini cli −60%
input$0.5
output$4
cache$0.05
Apixly · gemini ai studio −50%
input$0.625
output$5
cache$0.062
Wcnbai · Gemini CLI −49%
input$0.812
output$4.87
cache$0.081
Wcnbai · Gemini优质文本 +52%
input$2.44
output$14.62
cache$0.244
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 2.5 Flash

context 1 M

Fast, low-cost multimodal model with a huge context window.

INPUT / 1M
$0.049 $0.3
OUTPUT / 1M
$0.406 $2.5
model identifier
google/gemini-2.5-flash
Wcnbai · Gemini CLI −84%
input$0.049
output$0.406
cache$0.005
Apixly · gemini cli −60%
input$0.12
output$1
cache$0.012
Wcnbai · Gemini优质文本 −51%
input$0.146
output$1.22
cache$0.015
Apixly · gemini ai studio −50%
input$0.15
output$1.25
cache$0.015
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
G

Gemini 2.5 Flash Lite

context 1 M

A fast, cost-efficient reasoning model for low-latency applications. It excels at quick text generation and high throughput.

INPUT / 1M
$0.016 $0.1
OUTPUT / 1M
$0.065 $0.4
model identifier
google/gemini-2.5-flash-lite
Wcnbai · Gemini CLI −84%
input$0.016
output$0.065
cache$0.002
Wcnbai · Gemini优质文本 −51%
input$0.049
output$0.195
cache$0.005
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
D

DeepSeek

3
D

DeepSeek V4 Pro

context 1 M

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model excelling at advanced reasoning and coding tasks.

INPUT / 1M
$0.392 $0.435
OUTPUT / 1M
$0.783 $0.87
model identifier
deepseek/deepseek-v4-pro
Apixly · ChineseLLM −10%
input$0.392
output$0.783
cache$0.003
Wcnbai · 国产 +169%
input$1.17
output$2.34
cache$0.097
Wcnbai · 国产模型 +191%
input$1.27
output$2.53
cache$0.106
Wcnbai · KIMI +214%
input$1.36
output$2.73
cache$0.114
Wcnbai · DeepSeek +348%
input$1.95
output$3.9
cache$0.162
Wcnbai · 阿里百炼 +348%
input$1.95
output$3.9
cache$0.162
Wcnbai · GLM +348%
input$1.95
output$3.9
cache$0.162
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
D

DeepSeek V4 Flash

context 1 M

An efficiency-optimized Mixture-of-Experts model designed for fast inference. It excels at handling long contexts and complex reasoning tasks.

INPUT / 1M
$0.088 $0.098
OUTPUT / 1M
$0.176 $0.196
model identifier
deepseek/deepseek-v4-flash
Apixly · ChineseLLM −10%
input$0.088
output$0.176
cache$0.018
Wcnbai · 国产 −1%
input$0.097
output$0.194
cache$0.002
Wcnbai · 国产模型 +7%
input$0.105
output$0.21
cache$0.002
Wcnbai · KIMI +16%
input$0.114
output$0.227
cache$0.002
Wcnbai · DeepSeek +65%
input$0.162
output$0.324
cache$0.003
Wcnbai · 阿里百炼 +65%
input$0.162
output$0.324
cache$0.003
Wcnbai · GLM +66%
input$0.162
output$0.325
cache$0.003
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
D

DeepSeek V3.2

context 130 K

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance.

INPUT / 1M
$0.194 $0.25
OUTPUT / 1M
$0.291 $0.35
model identifier
deepseek/deepseek-v3.2
Wcnbai · 国产 −19%
input$0.194
output$0.291
cache
Wcnbai · 国产模型 −12%
input$0.21
output$0.316
cache
Wcnbai · DeepSeek +35%
input$0.325
output$0.487
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
K

Kimi

4
K

Kimi K3

context 1 M

Kimi K3 is an open-weight multimodal reasoning model adept at complex coding, knowledge work, and long-horizon agentic workflows.

INPUT / 1M
$2.27 $3
OUTPUT / 1M
$11.37 $15
model identifier
moonshotai/kimi-k3
Wcnbai · KIMI −24%
input$2.27
output$11.37
cache$0.227
Apixly · ChineseLLM −10%
input$2.7
output$13.5
cache$0.27
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
K

Kimi K2.7 Code

context 260 K

Kimi K2.7 Code is a coding-focused model designed for reliable, end-to-end programming tasks over long contexts. It utilizes a multimodal mixture-of-experts architecture.

INPUT / 1M
$0.634 $0.74
OUTPUT / 1M
$2.63 $3.5
model identifier
moonshotai/kimi-k2.7-code
Wcnbai · 国产 −23%
input$0.634
output$2.63
cache$0.127
Wcnbai · 国产模型 −17%
input$0.686
output$2.85
cache$0.137
Wcnbai · KIMI −10%
input$0.739
output$3.07
cache$0.148
Apixly · ChineseLLM −10%
input$0.666
output$3.15
cache$0.135
Wcnbai · 阿里百炼 +28%
input$1.05
output$4.38
cache$0.211
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
K

Kimi K2.6

context 260 K

Kimi K2.6 is a multimodal model excelling at long-horizon coding, UI/UX generation, and multi-agent orchestration across Python, Rust, and Go.

INPUT / 1M
$0.632 $0.67
OUTPUT / 1M
$2.63 $3.41
model identifier
moonshotai/kimi-k2.6
Wcnbai · 国产 −20%
input$0.632
output$2.63
cache$0.107
Wcnbai · 国产模型 −13%
input$0.685
output$2.85
cache$0.116
Wcnbai · KIMI −7%
input$0.739
output$3.07
cache$0.125
Wcnbai · GLM +33%
input$1.05
output$4.38
cache$0.178
Wcnbai · DeepSeek +33%
input$1.05
output$4.38
cache$0.178
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
K

Kimi K2.5

context 260 K

Kimi K2.5 is a multimodal model excelling at visual coding and agent swarm coordination. It leverages extensive pretraining for advanced capabilities.

INPUT / 1M
$0.649 $0.4
OUTPUT / 1M
$3.41 $1.9
model identifier
moonshotai/kimi-k2.5
Wcnbai · 阿里百炼 +76%
input$0.649
output$3.41
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
GL

GLM

2
GL

GLM 5.2

context 1 M

GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering, and complex multi-step automation.

INPUT / 1M
$0.19 $1.4
OUTPUT / 1M
$0.19 $4.4
model identifier
z-ai/glm-5.2
China Subscriptions · direct −93%
input$0.19
output$0.19
cache
Wcnbai · 国产 −39%
input$0.78
output$2.73
cache$0.195
Wcnbai · 国产模型 −34%
input$0.845
output$2.96
cache$0.211
Wcnbai · KIMI −29%
input$0.91
output$3.18
cache$0.227
Apixly · ChineseLLM −10%
input$1.26
output$3.96
cache
Wcnbai · 阿里百炼 +1%
input$1.3
output$4.54
cache$0.324
Wcnbai · GLM +1%
input$1.3
output$4.55
cache$0.325
Wcnbai · DeepSeek +1%
input$1.3
output$4.55
cache$0.325
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
GL

GLM 5.1

context 200 K

GLM-5.1 excels at long-horizon coding tasks, offering independent and continuous operation for complex development.

INPUT / 1M
$0.777 $0.98
OUTPUT / 1M
$2.72 $3.08
model identifier
z-ai/glm-5.1
Wcnbai · 国产 −14%
input$0.777
output$2.72
cache
Apixly · ChineseLLM −10%
input$0.882
output$2.77
cache$0.164
Wcnbai · 国产模型 −7%
input$0.841
output$2.94
cache
Wcnbai · GLM +43%
input$1.29
output$4.53
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
M

MiniMax

2
M

MiniMax M3

context 1 M

This multimodal foundation model accepts text, image, and video inputs to generate text outputs, excelling at long-horizon agentic tasks and coding.

INPUT / 1M
$0.3 $0.3
OUTPUT / 1M
$0.3 $1.2
model identifier
minimax/minimax-m3
China Subscriptions · direct −60%
input$0.3
output$0.3
cache
Apixly · ChineseLLM −10%
input$0.27
output$1.08
cache$0.054
Apixly · ChineseLLM −10%
input$0.27
output$1.08
cache$0.054
Wcnbai · M3 +262%
input$1.09
output$4.35
cache$0.217
Wcnbai · M3-cc接口 +355%
input$1.36
output$5.46
cache$0.273
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
M

MiniMax M2.7

context 200 K

MiniMax-M2.7 is a large language model designed for autonomous productivity and continuous improvement. It excels at active participation in its own evolution through advanced agentic capabilities.

INPUT / 1M
$0.221 $0.25
OUTPUT / 1M
$0.883 $1
model identifier
minimax/minimax-m2.7
Wcnbai · 国产模型 −12%
input$0.221
output$0.883
cache$0.044
Apixly · ChineseLLM −10%
input$0.225
output$0.9
cache$0.045
Wcnbai · minimax +9%
input$0.273
output$1.09
cache$0.055
Wcnbai · minimax-cc接口 +9%
input$0.273
output$1.09
cache$0.055
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
xA

xAI

1
xA

Grok 4.5

context 500 K

Grok 4.5 is SpaceXAI's most advanced model, excelling in coding, knowledge work, and STEM tasks.

INPUT / 1M
$0.2 $2
OUTPUT / 1M
$0.6 $6
model identifier
x-ai/grok-4.5
Apixly · gpt −90%
input$0.2
output$0.6
cache$0.05
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
Q

Qwen

6
Q

Qwen3.7 Max

context 1 M

Qwen3.7-Max is a flagship model for agent-centric workloads, excelling at coding and office productivity tasks.

INPUT / 1M
$0.2 $1.25
OUTPUT / 1M
$0.2 $3.75
model identifier
qwen/qwen3.7-max
Claude Hub · direct −92%
input$0.2
output$0.2
cache$0.2
Wellflow · direct −90%
input$0.25
output$0.25
cache$0.25
Apixly · ChineseLLM −10%
input$1.12
output$3.38
cache$0.225
Wcnbai · 国产模型 +1%
input$1.26
output$3.79
cache$0.252
Wcnbai · 阿里百炼 +56%
input$1.95
output$5.85
cache$0.39
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
Q

Qwen3.7 Plus

context 1 M

A multimodal model that excels at text generation and understanding, with enhanced capabilities for image processing.

INPUT / 1M
$0.15 $0.32
OUTPUT / 1M
$0.15 $1.28
model identifier
qwen/qwen3.7-plus
Wellflow · direct −81%
input$0.15
output$0.15
cache$0.15
Apixly · ChineseLLM −10%
input$0.288
output$1.15
cache$0.058
Wcnbai · 国产模型 +296%
input$1.27
output$5.07
cache$0.084
Wcnbai · 阿里百炼 +509%
input$1.95
output$7.8
cache$0.13
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
Q

Qwen3.6 Flash

context 1 M

Qwen3.6 Flash is a fast, efficient multimodal model supporting text, image, and video input. It excels at processing long contexts.

INPUT / 1M
$0.195 $0.188
OUTPUT / 1M
$1.17 $1.12
model identifier
qwen/qwen3.6-flash
Wcnbai · 阿里百炼 +4%
input$0.195
output$1.17
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
Q

Qwen3.5-Flash

context 1 M

Qwen3.5 Flash models are efficient vision-language models with a hybrid architecture, excelling at fast inference for multimodal tasks.

INPUT / 1M
$0.1 $0.065
OUTPUT / 1M
$0.1 $0.26
model identifier
qwen/qwen3.5-flash
Wellflow · direct −38%
input$0.1
output$0.1
cache$0.1
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
Q

Qwen3.6 Plus

context 1 M

Qwen 3.6 Plus is a powerful, scalable LLM with a hybrid architecture, excelling at high-performance inference.

INPUT / 1M
$0.21 $0.325
OUTPUT / 1M
$1.26 $1.95
model identifier
qwen/qwen3.6-plus
Wcnbai · 国产模型 −35%
input$0.21
output$1.26
cache$0.021
Wcnbai · 阿里百炼
input$0.325
output$1.95
cache$0.032
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure
Q

Qwen3 Max

context 260 K

Qwen3-Max is an advanced LLM with enhanced reasoning, instruction following, and multilingual capabilities. It excels at understanding and generating text across diverse topics and languages.

INPUT / 1M
$0.631 $0.78
OUTPUT / 1M
$2.52 $3.9
model identifier
qwen/qwen3-max
Wcnbai · 国产模型 −33%
input$0.631
output$2.52
cache
Wcnbai · 阿里百炼 +4%
input$0.975
output$3.9
cache
Discount shown against the model's official price · prices per 1M tokens · DarkRouter takes the cheapest working route and switches over on failure