Large Model API Pricing

Explore pricing for our model APIs. Find the right plan for your needs with transparent rates and flexible options.

Anthropic logo

Anthropic

Anthropic's Claude model offers advanced AI safety capabilities, focusing on useful, harmless, and honest AI assistants with powerful reasoning and conversational abilities.

Model NameInput Token RangeContextInput(/Mt)Cache Write(/Mt)Cache Read(/Mt)Output(/Mt)Actions
claude-fable-5-1,000,000$10$12.5(5m)·$20(1h)$1$50Try Now
claude-opus-4-7-1,000,000
$4.75 $5
$5.9375 (5 min) · $9.50 (1 hr) $6.25 (5 min) × $10 (1 hr)
$0.475 $0.5
$23.75 $25
Try Now
claude-sonnet-5-1,000,000$2$2.5(5m)·$4(1h)$0.2$10Try Now
claude-opus-4-8-1,000,000
$4.75 $5
$5.9375 (5 min) · $9.50 (1 hr) $6.25 (5 min) × $10 (1 hr)
$0.475 $0.5
$23.75 $25
Try Now
claude-opus-4-8-r-1,000,000
$1$5
$1.25(5m)·$2(1h)$6.25(5m)·$10(1h)
$0.1$0.5
$5$25
Try Now
claude-opus-4-7-r-1,000,000
$1$5
$1.25(5m)·$2(1h)$6.25(5m)·$10(1h)
$0.1$0.5
$5$25
Try Now
claude-opus-4-6-dd-1,000,000
$2.75$5
$3.4375(5m)·$5.5(1h)$6.25(5m)·$10(1h)
$0.275$0.5
$13.75$25
Try Now
Claude, Op. 4, No. 61–200,0001,000,000$5$6.25 (5 min) · $10 (1 hr)$0.5$25Try Now
200,000–1,000,0001,000,000$5$6.25 (5 min) · $10 (1 hr)$0.5$25Try Now
claude-opus-4-6-r-1,000,000
$1$5
$1.25(5m)·$2(1h)$6.25(5m)·$10(1h)
$0.1$0.5
$5$25
Try Now
claude-sonnet-4-61–200,0001,000,000$3$3.75 (5 min) · $6 (1 hr)$0.3$15Try Now
200,000–1,000,0001,000,000$3$3.75 (5 min) · $6 (1 hr)$0.3$15Try Now
claude-sonnet-4-6-dd-1,000,000
$1.65$3
$2.0625(5m)·$3.3(1h)$3.75(5m)·$6(1h)
$0.165$0.3
$8.25$15
Try Now
claude-sonnet-4-6-r-1,000,000
$0.6$3
$0.75(5m)·$1.2(1h)$3.75(5m)·$6(1h)
$0.06$0.3
$3$15
Try Now
claude-opus-4-5-20251101-200,000
$4.75 $5
$5.9375 (5 min) · $9.50 (1 hr) $6.25 (5 min) × $10 (1 hr)
$0.475 $0.5
$23.75 $25
Try Now
claude-opus-4-5-20251101-dd-200,000
$2.75$5
$3.4375(5m)$6.25(5m)
$0.275$0.5
$13.75$25
Try Now
claude-sonnet-4-5-202509291–200,000200,000$3$3.75 (5 min) · $6 (1 hr)$0.3$15Try Now
200,000–1,000,000200,000$6$7.5 (5 min) × $12 (1 hr)$0.6$22.5Try Now
claude-sonnet-4-5-20250929-dd-200,000
$1.65$3
$2.0625(5m)$3.75(5m)
$0.165$0.3
$8.25$15
Try Now
claude-haiku-4-5-20251001-20,000$1$1.25 (5 min) · $2 (1 hr)$0.1$5Try Now
claude-haiku-4-5-20251001-dd-200,000
$0.55$1
$0.6875(5m)·$1.1(1h)$1.25(5m)·$2(1h)
$0.055$0.1
$2.75$5
Try Now
claude-haiku-4-5-20251001-r-200,000
$0.2$1
$0.25(5m)·$0.4(1h)$1.25(5m)·$2(1h)
$0.02$0.1
$1$5
Try Now
claude-opus-4-8-cc-1,000,000
$1.85$5
$2.3125(5m)·$3.7(1h)$6.25(5m)·$10(1h)
$0.185$0.5
$9.25$25
Try Now
claude-opus-4-6-cc-1,000,000
$1.85$5
$2.3125(5m)·$3.7(1h)$6.25(5m)·$10(1h)
$0.185$0.5
$9.25$25
Try Now
claude-opus-4-7-cc-1,000,000
$1.85$5
$2.3125(5m)·$3.7(1h)$6.25(5m)·$10(1h)
$0.185$0.5
$9.25$25
Try Now
claude-sonnet-4-6-cc-1,000,000
$1.11$3
$1.3875(5m)·$2.22(1h)$3.75(5m)·$6(1h)
$0.111$0.3
$5.55$15
Try Now
claude-haiku-4-5-20251001-cc-200,000
$0.37$1
$0.4625(5m)·$0.74(1h)$1.25(5m)·$2(1h)
$0.037$0.1
$1.85$5
Try Now
OpenAI

OpenAI

OpenAI's GPT series of models offer state-of-the-art language understanding and generation capabilities, delivering outstanding performance across a wide range of tasks, and are among the industry's leading AI models.

Model NameInput Token RangeContextInput(/Mt)Cache Write(/Mt)Cache Read(/Mt)Output(/Mt)Actions
gpt-5.6-terra-es1–272,0003,720,000
$0.25$2.5
$0.3125(30m)$3.125(30m)
$0.025$0.25
$1.5$15
Try Now
272,000-372,0003,720,000
$0.5$5
$0.625(30m)$6.25(30m)
$0.05$0.5
$2.25$22.5
Try Now
gpt-5.6-luna-es1–272,000372,000
$0.1$1
$0.125(30m)$1.25(30m)
$0.01$0.1
$0.6$6
Try Now
272,000-372,000372,000
$0.2$2
$0.25(30m)$2.5(30m)
$0.02$0.2
$0.9$9
Try Now
gpt-5.6-sol-es1–272,000372,000
$0.5$5
$0.625(30m)$6.25(30m)
$0.05$0.5
$3$30
Try Now
272,000-372,000372,000
$1$10
$1.25(30m)$12.5(30m)
$0.1$1
$4.5$45
Try Now
gpt-5.6-terra1–272,0001,050,000$2.5$3.125(30m)$0.25$15Try Now
272,000–1,050,0001,050,000$5$6.25(30m)$0.5$22.5Try Now
gpt-5.6-luna1–272,0001,050,000$1$1.25(30m)$0.1$6Try Now
272,000–1,050,0001,050,000$2$2.5(30m)$0.2$9Try Now
gpt-5.6-sol1–272,0001,050,000$5$6.25(30m)$0.5$30Try Now
272,000–1,050,0001,050,000$10$12.5(30m)$1$45Try Now
gpt-5.5-r-1,050,000
$0.5$5
-
$0.05$0.5
$3$30
Try Now
gpt-5.51–272,0001,050,000
$4.75 $5
-
$0.475 $0.5
$28.5$30
Try Now
272,000–1,050,0001,050,000
$9.5 $10
-
$0.95$1
$42.75$45
Try Now
gpt-4.1-nano-1,047,576
$0.095 $0.1
-
$0.0237 $0.025
$0.38 $0.4
Try Now
gpt-5.5-pro1–272,0001,050,000$30--$180Try Now
272,000–1,050,0001,050,000$60--$270Try Now
gpt-5.4-pro1–272,0001,050,000$30--$180Try Now
272,000–1,050,0001,050,000$60--$270Try Now
GPT-5.41–272,0001,050,000$2.5-$0.25$15Try Now
272,000–1,050,0001,050,000$5-$0.5$22.5Try Now
gpt-5.4-mini-400,000
$0.7125 $0.75
-
$0.0712 $0.075
$4.275 $4.5
Try Now
gpt-5.4-nano-400,000
$0.19 $0.2
-
$0.019 $0.02
$1.1875 $1.25
Try Now
gpt-5.3-codex-400,000
$1.6625 $1.75
-
$0.1662 $0.175
$13.3 $14
Try Now
gpt-5.2-pro-400,000
$19.95 $21
--
$159.6 $168
Try Now
gpt-5.2-codex-400,000$1.75-$0.175$14Try Now
GPT-5.2-400,000
$1.6625 $1.75
-
$0.1662 $0.175
$13.3 $14
Try Now
gpt-5.1-codex-max-400,000
$1.1875 $1.25
-
$0.1187 $0.125
$9.5 $10
Try Now
gpt-5.1-codex-400,000
$1.1875 $1.25
-
$0.1187 $0.125
$9.5 $10
Try Now
gpt-5.1-codex-mini-400,000
$0.2375 $0.25
-
$0.0237 $0.025
$1.9 $2
Try Now
gpt-5.1-400,000
$1.1875 $1.25
-
$0.1187 $0.125
$9.5 $10
Try Now
gpt-5-pro-400,000
$14.25 $15
--
$114 $120
Try Now
gpt-5-codex-400,000
$1.1875 $1.25
-
$0.1187 $0.125
$9.5 $10
Try Now
GPT-5-400,000
$1.1875 $1.25
-
$0.1187 $0.125
$9.5 $10
Try Now
gpt-5-mini-400,000
$0.2375 $0.25
-
$0.0237 $0.025
$1.9 $2
Try Now
gpt-5-nano-400,000
$0.0475 $0.05
-
$0.0047 $0.005
$0.38 $0.4
Try Now
gpt-4.1-1,047,576$2-$0.5$8Try Now
gpt-4.1-mini-1,047,576$0.4-$0.1$1.6Try Now
GPT-4O-131,072
$2.375 $2.5
-
$1.1875 $1.25
$9.5 $10
Try Now
GPT-4o-mini-128,000
$0.1425 $0.15
-
$0.0712 $0.075
$0.57 $0.6
Try Now
OpenAI: GPT OSS 20B-131,072$0.05--$0.2Try Now
OpenAI GPT OSS 120B-131,072$0.1--$0.5Try Now
Gemini logo

Gemini

Google's Gemini model offers high-quality natural language processing capabilities, performs exceptionally well across a wide range of NLP tasks, and boasts powerful multimodal capabilities.

Model NameInput Token RangeContextInput(/Mt)Cache Write(/Mt)Cache Read(/Mt)Output(/Mt)Actions
gemini-3.1-flash-lite-1,048,576$0.25$0.083(5m)·$1(1h)$0.025$1.5Try Now
gemini-3.5-flash-1,048,576
$1.425 $1.5
$0.0788 (5 min) × $0.95 (1 hr) $0.083 (5 min) × $1 (1 hr)
$0.1425 $0.15
$8.55$9
Try Now
gemini-3.1-pro-preview1–204,8001,048,576$2$0.375 (5 min) × $4.5 (1 hr)$0.2$12Try Now
204,800–1,048,5761,048,576$4$0.375 (5 min) × $4.5 (1 hr)$0.4$18Try Now
gemini-3-flash-preview-1,048,576
$0.475 $0.5
$0.0788 (5 min) × $0.95 (1 hr) $0.083 (5 min) × $1 (1 hr)
$0.0475 $0.05
$2.85 $3
Try Now
gemini-2.5-pro-1,048,576
$1.1875 $1.25
$0.3562 (5-month) × $4.275 (1-hour) $0.375 (5 min) × $4.5 (1 hr)
$0.1187 $0.125
$9.5 $10
Try Now
gemini-2.5-pro-preview-06-05-1,048,576
$1.1875 $1.25
$0.3562 (5-month) × $4.275 (1-hour) $0.375 (5 min) × $4.5 (1 hr)
$0.1187 $0.125
$9.5 $10
Try Now
gemini-2.5-flash-preview-05-20-1,048,576
$0.1425 $0.15
$0.0788 (5 min) × $0.95 (1 hr) $0.083 (5 min) × $1 (1 hr)
$0.0285 $0.03
$3.325 $3.5
Try Now
gemini-2.5-flash-1,048,576
$0.285 $0.3
$0.0788 (5 min) × $0.95 (1 hr) $0.083 (5 min) × $1 (1 hr)
$0.0285 $0.03
$2.375 $2.5
Try Now
gemini-2.5-flash-lite-preview-09-2025-1,048,576
$0.095 $0.1
$0.0788 (5 min) × $0.95 (1 hr) $0.083 (5 min) × $1 (1 hr)
$0.0095 $0.01
$0.38 $0.4
Try Now
gemini-2.5-flash-lite-1,048,576
$0.095 $0.1
$0.0788 (5 min) × $0.95 (1 hr) $0.083 (5 min) × $1 (1 hr)
$0.0095 $0.01
$0.38 $0.4
Try Now
Gemma 3 27B-32,768$0.119--$0.2Try Now
Gemma3 12B-131,072$0.05--$0.1Try Now
gemini-3.6-flash-1,048,576$1.5$0.0833(5m)·$1(1h)$0.15$7.5Try Now
gemini-3.5-flash-lite-1,048,576$0.3$0.0833(5m)·$1(1h)$0.03$2.5Try Now
Llama logo

Llama

Meta's Llama model offers state-of-the-art language understanding capabilities and features an open architecture, making it suitable for a wide range of applications.

Model NameContextInput(/Mt)Output(/Mt)Operation
Llama 3.1 8B Instruct16,384$0.02$0.05Try Now
Llama 3.2 3B Instruct32,768$0.03$0.05Try Now
Llama 3.3 70B Instruct131,072$0.13$0.39Try Now
Llama 4 Maverick Instructions1,048,576$0.17$0.85Try Now
Llama 4 Scout Instructor131,072$0.1$0.5Try Now
Qwen logo

Qwen

The Qwen series of models offers powerful natural language processing capabilities and is available in a range of parameter sizes, from lightweight to enterprise-grade solutions.

Model NameInput Token RangeContextInput(/Mt)Output(/Mt)Actions
Qwen3.5-Plus1-256,0001,000,000$0.4$2.4Try Now
256,000-1,000,0001,000,000$1.2$7.2Try Now
Qwen3 235B A22B Instruct 2507-131,072$0.15$0.8Try Now
Qwen 2.5 72B Instruct-32,000$0.38$0.4Try Now
Qwen MT Plus-4,096$0.25$0.75Try Now
Qwen 2.5 7B Instruct-32,000$0.07$0.07Try Now
Qwen 2.5 VL 72B Instruction Manual-32,768$0.8$0.8Try Now
Qwen3 30B A3B-40,960$0.09$0.45Try Now
Qwen3 32B-40,960$0.1$0.45Try Now
Qwen3 235B A22B-40,960$0.2$0.8Try Now
Qwen3 235B A22b Thinking 2507-131,072$0.3$3Try Now
Qwen3 Coder 480B A35B Instructions-262,144$0.29$1.2Try Now
Qwen3 Coder Next FP8-262,144$0.2$1.5Try Now
Qwen3 Next 80B A3B Instruct-65,536$0.15$1.5Try Now
Qwen3 Next 80B A3B Thinking-65,536$0.15$1.5Try Now
Qwen3.5-27B-262,144$0.3$2.4Try Now
Qwen3.5-122B-A10B-262,144$0.4$3.2Try Now
Qwen3.5-35B-A3B-262,144$0.25$2Try Now
Qwen3.5-397B-A17B-262,144$0.6$3.6Try Now
Wenxin

Baidu

Baidu's ERNIE model offers advanced Chinese language understanding and multimodal capabilities, is optimized for Chinese applications, and is competitively priced.

Model NameContextInput(/Mt)Output(/Mt)Operation
ERNIE 4.5 VL 424B A47B123,000$0.42$1.25Try Now
ERNIE 4.5 300B A47B123,000$0.28$1.1Try Now
ChatGLM

THUDM

The GLM series of models from Tsinghua University feature advanced Chinese language understanding and generation capabilities.

Model NameContextInput(/Mt)Cache Read(/Mt)Output(/Mt)Operation
GLM 5.21,048,576$1.4$0.26$4.4Try Now
GLM-5.1204,800$1.38$0.26$4.4Try Now
GLM-5V-Turbo204,800$1.2$0.24$4Try Now
GLM 4.5V65,536$0.6-$1.8Try Now
GLM-4.5131,072$0.6-$2.2Try Now
GLM-4.7204,800$0.6-$2.2Try Now
GLM-4.7-Flash200,000$0.07$0.01$0.4Try Now
GLM-5204,800$1$0.2$3.2Try Now
GLM-5-Turbo202,800$1.2$0.24$4Try Now
Sao10K logo

Sao10K

A fine-tuned model specifically optimized for creative and role-playing applications, featuring enhanced storytelling capabilities.

Model NameContextInput(/Mt)Output(/Mt)Operation
L3 8B Stheno V3.28,192$0.05$0.05Try Now
L3 70B Euryale V2.1 8,192$1.48$1.48Try Now
L31 70B Euryale V2.28,192$1.48$1.48Try Now
Sao10k L3 8B Lunaris 8,192$0.05$0.05Try Now
Mistralai logo

Mistralai

A powerful and efficient language model from Mistral AI, designed for both commercial and open-source applications.

Model NameContextInput(/Mt)Output(/Mt)Operation
Mistral Nemo60,288$0.04$0.17Try Now
Mistral 7B Instruct32,768$0.029$0.059Try Now
Deepseek logo

Deepseek

Advanced AI models from DeepSeek, offering cutting-edge inference capabilities and competitive pricing for enterprise and research applications.

Model NameContextInput(/Mt)Cache Write(/Mt)Cache Read(/Mt)Output(/Mt)Operation
Deepseek V4 Flash1,048,576$0.14-$0.028$0.28Try Now
Deepseek V4 Pro1,048,576$0.435-$0.145$0.87Try Now
DeepSeek R1 0528163,840$0.7-$0.35$2.5Try Now
DeepSeek V3 0324163,840$0.28$0.14 (5m)$0.14$1.14Try Now
DeepSeek V3.1163,840$0.27--$1Try Now
DeepSeek-OCR 28,192$0.03--$0.03Try Now
MiniMax logo

MiniMax

MiniMax AI's advanced language model delivers powerful conversational AI capabilities, excelling in customer service, content generation, and creative applications, with robust multilingual support and enterprise-grade scalability.

Model NameContextInput(/Mt)Output(/Mt)Operation
MiniMax M11,000,000$0.55$2.2Try Now
Gryphe logo

Gryphe

An innovative AI model from Gryphe that offers professional-grade language understanding capabilities, with a focus on efficiency and adaptability, making it ideal for niche applications.

Model NameContextInput(/Mt)Output(/Mt)Operation
Mythomax L2 13B4,096$0.09$0.09Try Now

Mixture of Experts

A sophisticated collection of state-of-the-art AI models, featuring advanced reasoning and mathematical proof capabilities, as well as cutting-edge language understanding across multiple domains.

Model NameInput Token RangeContextInput(/Mt)Cache Write(/Mt)Cache Read(/Mt)Output(/Mt)Actions
Qwen3.5-Plus1-256,0001,000,000$0.4--$2.4Try Now
256,000-1,000,0001,000,000$1.2--$7.2Try Now
GLM 5.2-1,048,576$1.4-$0.26$4.4Try Now
Deepseek V4 Flash-1,048,576$0.14-$0.028$0.28Try Now
Deepseek V4 Pro-1,048,576$0.435-$0.145$0.87Try Now
MiniMax M2.7-204,800$0.3-$0.03$1.2Try Now
Kimi K2.5-262,144$0.6-$0.1$3Try Now
Kimi K2 Instruct-131,072$0.57--$2.3Try Now
GLM-5.1-204,800$1.38-$0.26$4.4Try Now
GLM-5V-Turbo-204,800$1.2-$0.24$4Try Now
ERNIE 4.5 VL 424B A47B-123,000$0.42--$1.25Try Now
OpenAI: GPT OSS 20B-131,072$0.05--$0.2Try Now
OpenAI GPT OSS 120B-131,072$0.1--$0.5Try Now
DeepSeek R1 0528-163,840$0.7-$0.35$2.5Try Now
DeepSeek V3 0324-163,840$0.28$0.14 (5m)$0.14$1.14Try Now
DeepSeek V3.1-163,840$0.27--$1Try Now
ERNIE 4.5 300B A47B-123,000$0.28--$1.1Try Now
GLM 4.5V-65,536$0.6--$1.8Try Now
GLM-4.5-131,072$0.6--$2.2Try Now
GLM-4.7-204,800$0.6--$2.2Try Now
GLM-4.7-Flash-200,000$0.07-$0.01$0.4Try Now
GLM-5-204,800$1-$0.2$3.2Try Now
GLM-5-Turbo-202,800$1.2-$0.24$4Try Now
Llama 4 Maverick Instructions-1,048,576$0.17--$0.85Try Now
Llama 4 Scout Instructor-131,072$0.1--$0.5Try Now
MiniMax M1-1,000,000$0.55--$2.2Try Now
Minimax M2.1-204,800$0.3$0.375 (5m)$0.03$1.2Try Now
MiniMax M2.5-204,800$0.3-$0.03$1.2Try Now
MiniMax M2.7-highspeed-204,800$0.6-$0.06$2.4Try Now
MiniMax M2.5-highspeed-204,800$0.6-$0.03$2.4Try Now
Qwen3 30B A3B-40,960$0.09--$0.45Try Now
Qwen3 32B-40,960$0.1--$0.45Try Now
Qwen3 235B A22B-40,960$0.2--$0.8Try Now
Qwen3 235B A22b Thinking 2507-131,072$0.3--$3Try Now
XiaomiMiMo/MiMo-V2.5-Pro1-262,1441,048,576$1-$0.2$3Try Now
262,144-1,048,5761,048,576$2-$0.4$6Try Now
Qwen3.5-122B-A10B-262,144$0.4--$3.2Try Now
Qwen3.5-35B-A3B-262,144$0.25--$2Try Now
Qwen3.5-397B-A17B-262,144$0.6--$3.6Try Now
Contact Us