Skip to main content
Octavus
Back to pricing

Model pricing

When using Octavus-managed API keys, the provider cost for each request is passed through at the rates below. These prices are automatically synced from provider APIs. With BYOK (Bring Your Own Keys), provider costs are not charged.

OpenAI(70 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
computer-use-preview
$3.00$12.00---
OpenAI: GPT-3.5 Turbo
$0.50$1.50--16K
OpenAI: GPT-3.5 Turbo (older v0613)
$1.00$2.00--4K
OpenAI: GPT-3.5 Turbo 16k
$3.00$4.00--16K
OpenAI: GPT-3.5 Turbo Instruct
$1.50$2.00--4K
OpenAI: GPT-4
$30.00$60.00--8K
OpenAI: GPT-4 Turbo
$10.00$30.00--128K
OpenAI: GPT-4 Turbo Preview
$10.00$30.00--128K
OpenAI: GPT-4.1
$2.00$8.00$0.50-1M
OpenAI: GPT-4.1 Mini
$0.40$1.60$0.10-1M
OpenAI: GPT-4.1 Nano
$0.10$0.40$0.025-1M
OpenAI: GPT-4o
$2.50$10.00$1.25-128K
OpenAI: GPT-4o (2024-05-13)
$5.00$15.00--128K
OpenAI: GPT-4o (2024-08-06)
$2.50$10.00$1.25-128K
OpenAI: GPT-4o (2024-11-20)
$2.50$10.00$1.25-128K
OpenAI: GPT-4o-mini
$0.15$0.60$0.075-128K
OpenAI: GPT-4o-mini (2024-07-18)
$0.15$0.60$0.075-128K
gpt-4o-mini-transcribe
$3.00$5.00---
gpt-4o-transcribe
$6.00$10.00---
OpenAI: GPT-5
$1.25$10.00$0.125-400K
OpenAI: GPT-5 Image
$10.00$10.00$1.25-400K
OpenAI: GPT-5 Image Mini
$2.50$2.00$0.25-400K
OpenAI: GPT-5 Mini
$0.25$2.00$0.025-400K
OpenAI: GPT-5 Nano
$0.05$0.40$0.005-400K
OpenAI: GPT-5 Pro
$15.00$120.00--400K
OpenAI: GPT-5.1
$1.25$10.00$0.125-400K
OpenAI: GPT-5.1-Codex
$1.25$10.00$0.13-400K
OpenAI: GPT-5.1-Codex-Max
$1.25$10.00$0.125-400K
OpenAI: GPT-5.1-Codex-Mini
$0.25$2.00$0.03-400K
OpenAI: GPT-5.2
$1.75$14.00$0.175-400K
OpenAI: GPT-5.2 Chat
$1.75$14.00$0.175-128K
OpenAI: GPT-5.2-Codex
$1.75$14.00$0.175-400K
OpenAI: GPT-5.2 Pro
$21.00$168.00--400K
OpenAI: GPT-5.3-Codex
$1.75$14.00$0.175-400K
OpenAI: GPT-5.4
$2.50$5.00 >272K
$15.00$22.50 >272K
$0.25$0.50 >272K
-1M
OpenAI: GPT-5.4 Image 2
$8.00$15.00$2.00-272K
OpenAI: GPT-5.4 Mini
$0.75$4.50$0.075-400K
OpenAI: GPT-5.4 Nano
$0.20$1.25$0.02-400K
OpenAI: GPT-5.4 Pro
$30.00$60.00 >272K
$180.00$270.00 >272K
--1M
OpenAI: GPT-5.5
$5.00$10.00 >272K
$30.00$45.00 >272K
$0.50$1.00 >272K
-1M
OpenAI: GPT-5.5 Pro
$30.00$60.00 >272K
$180.00$270.00 >272K
--1M
gpt-5.6
$5.00$10.00 >272K
$30.00$45.00 >272K
$0.50$1.00 >272K
--
OpenAI: GPT-5.6 Luna
$0.20$0.40 >272K
$1.20$1.80 >272K
$0.02$0.04 >272K
-1M
OpenAI: GPT-5.6 Luna Pro
$0.20$1.20$0.02-1M
OpenAI: GPT-5.6 Sol
$5.00$10.00 >272K
$30.00$45.00 >272K
$0.50$1.00 >272K
-1M
OpenAI: GPT-5.6 Sol Pro
$2.00$10.00$0.20-1M
OpenAI: GPT-5.6 Terra
$2.00$4.00 >272K
$12.00$18.00 >272K
$0.20$0.40 >272K
-1M
OpenAI: GPT-5.6 Terra Pro
$2.00$12.00$0.20-1M
OpenAI: GPT Audio
$2.50$10.00--128K
OpenAI: GPT Audio Mini
$0.60$2.40--128K
OpenAI: GPT Chat Latest
$5.00$30.00$0.50-400K
gpt-image-1
$5.00$40.00$1.25--
gpt-image-1-mini
$2.50$8.00$0.25--
gpt-image-1.5
$5.00$10.00$1.25--
OpenAI: gpt-oss-120b
$0.037$0.17--131K
OpenAI: gpt-oss-120b (batch)
$0.15$0.60--131K
OpenAI: gpt-oss-20b
$0.03$0.13$0.03-131K
OpenAI: gpt-oss-20b (batch)
$0.05$0.20--131K
OpenAI: gpt-oss-safeguard-20b
$0.075$0.30$0.0375-131K
OpenAI: o1
$15.00$60.00$7.50-200K
o1-mini
$1.10$4.40$0.55--
OpenAI: o1-pro
$150.00$600.00--200K
OpenAI: o3
$2.00$8.00$0.50-200K
o3-deep-research
$10.00$40.00$2.50--
OpenAI: o3 Mini
$1.10$4.40$0.55-200K
OpenAI: o3 Mini High
$1.10$4.40$0.55-200K
OpenAI: o3 Pro
$20.00$80.00--200K
OpenAI: o4 Mini
$1.10$4.40$0.275-200K
o4-mini-deep-research
$2.00$8.00$0.50--
OpenAI: o4 Mini High
$1.10$4.40$0.275-200K

Anthropic(28 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Anthropic: Claude 3 Haiku
$0.25$1.25$0.03-200K
Anthropic: Claude Fable 5
$10.00$50.00$1.00-1M
Anthropic: Claude Fable 5 (batch)
$5.00$25.00$0.50-1M
Anthropic: Claude Haiku 4.5
$1.00$5.00$0.10-200K
Anthropic: Claude Haiku 4.5 (batch)
$0.50$2.50$0.05-200K
Anthropic: Claude Opus 4
$15.00$75.00$1.50-200K
Anthropic: Claude Opus 4.1
$15.00$75.00$1.50-200K
Anthropic: Claude Opus 4.5
$5.00$25.00$0.50-200K
Anthropic: Claude Opus 4.6
$5.00$25.00$0.50-1M
Anthropic: Claude Opus 4.7
$5.00$25.00$0.50-1M
Anthropic: Claude Opus 4.8
$5.00$25.00$0.50-1M
Anthropic: Claude Opus 4.1 (batch)
$7.50$37.50$0.75-200K
Anthropic: Claude Opus 4.5 (batch)
$2.50$12.50$0.25-200K
Anthropic: Claude Opus 4.6 (batch)
$2.50$12.50$0.25-1M
Anthropic: Claude Opus 4.7 (Fast)
$30.00$150.00$3.00-1M
Anthropic: Claude Opus 4.7 (batch)
$2.50$12.50$0.25-1M
Anthropic: Claude Opus 4.8 (Fast)
$10.00$50.00$1.00-1M
Anthropic: Claude Opus 4.8 (batch)
$2.50$12.50$0.25-1M
Claude Opus 5
$5.00$25.00$0.50-1M
Claude Opus 5 (Fast)
$10.00$50.00$1.00-1M
Claude Opus 5 (batch)
$2.50$12.50$0.25-1M
Anthropic: Claude Sonnet 4
$3.00$15.00$0.30-1M
Anthropic: Claude Sonnet 4.5
$3.00$15.00$0.30-1M
Anthropic: Claude Sonnet 4.6
$3.00$15.00$0.30-1M
Anthropic: Claude Sonnet 4.5 (batch)
$1.50$7.50$0.15-1M
Anthropic: Claude Sonnet 4.6 (batch)
$1.50$7.50$0.15-1M
Anthropic: Claude Sonnet 5
$2.00$10.00$0.20-1M
Anthropic: Claude Sonnet 5 (batch)
$1.00$5.00$0.10-1M

Google(40 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
gemini-2.0-flash
$0.10$0.40$0.025--
gemini-2.0-flash-lite
$0.075$0.30---
Google: Gemini 2.5 Flash
$0.30$2.50$0.03$2.501M
Google: Nano Banana (Gemini 2.5 Flash Image)
$0.30$30.00$0.03$2.5033K
Google: Gemini 2.5 Flash Lite
$0.10$0.40$0.01$0.401M
Google: Gemini 2.5 Flash Lite (batch)
$0.05$0.20$0.01$0.201M
Google: Gemini 2.5 Flash (batch)
$0.15$1.25$0.03$1.251M
Google: Gemini 2.5 Pro
$1.25$2.50 >200K
$10.00$15.00 >200K
$0.125$0.25 >200K
$10.001M
Google: Gemini 2.5 Pro Preview 06-05
$1.25$10.00$0.125$10.001M
Google: Gemini 2.5 Pro Preview 05-06
$1.25$10.00$0.125$10.001M
Google: Gemini 2.5 Pro (batch)
$0.625$5.00$0.125$5.001M
Google: Gemini 3 Flash Preview
$0.50$3.00$0.05$3.001M
Google: Gemini 3 Flash Preview (batch)
$0.25$1.50-$1.501M
Google: Nano Banana Pro (Gemini 3 Pro Image)
$2.00$12.00$0.20$12.00131K
Google: Nano Banana Pro (Gemini 3 Pro Image Preview)
$2.00$12.00$0.20$12.0066K
Google: Nano Banana 2 (Gemini 3.1 Flash Image)
$0.50$60.00--131K
Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)
$0.50$3.00--66K
Google: Gemini 3.1 Flash Lite
$0.25$1.50$0.025$1.501M
Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)
$0.25$1.50--66K
Google: Gemini 3.1 Flash Lite Preview
$0.25$1.50$0.025$1.501M
Google: Gemini 3.1 Flash Lite (batch)
$0.125$0.75$0.0125$0.751M
Google: Gemini 3.1 Pro Preview
$2.00$4.00 >200K
$12.00$18.00 >200K
$0.20$0.40 >200K
$12.001M
Google: Gemini 3.1 Pro Preview Custom Tools
$2.00$12.00$0.20$12.001M
Google: Gemini 3.1 Pro Preview (batch)
$1.00$6.00-$6.001M
Google: Gemini 3.5 Flash
$1.50$9.00$0.15$9.001M
Google: Gemini 3.5 Flash Lite
$0.30$2.50$0.03$2.501M
Google: Gemini 3.5 Flash Lite (batch)
$0.15$1.25$0.015$1.251M
Google: Gemini 3.5 Flash (batch)
$0.75$4.50$0.075$4.501M
Google: Gemini 3.6 Flash
$0.75$3.75$0.075$3.751M
Google: Gemini 3.6 Flash (batch)
$0.375$1.88$0.0375$1.881M
Google: Gemini 3.7 Flash
$0.75$3.75$0.075$3.751M
Google: Gemini 3.7 Flash (batch)
$0.1875$0.9375$0.0187$0.93751M
Google: Gemma 2 27B
$0.65$0.65--8K
Google: Gemma 3 12B
$0.05$0.15--131K
Google: Gemma 3 27B
$0.08$0.45$0.04-131K
Google: Gemma 3 4B
$0.05$0.10--131K
Google: Gemma 4 26B A4B
$0.07$0.34--262K
Google: Gemma 4 31B
$0.09$0.34$0.05-262K
Google: Gemma 4 31B (batch)
$0.39$0.97--262K
imagen-4.0-generate-001
$0.04$0.04---

~anthropic(4 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Anthropic: Claude Fable Latest
$10.00$50.00$1.00-1M
Anthropic Claude Haiku Latest
$1.00$5.00$0.10-200K
Anthropic: Claude Opus Latest
$5.00$25.00$0.50-1M
Anthropic Claude Sonnet Latest
$2.00$10.00$0.20-1M

~deepseek(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
DeepSeek V4 Flash Latest
$0.03$0.16$0.01-1M

~google(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Google Gemini Flash Latest
$0.75$3.75$0.075$3.751M
Google Gemini Pro Latest
$2.00$12.00$0.20$12.001M

~moonshotai(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
MoonshotAI Kimi Latest
$2.55$12.75$0.256-1M

~openai(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
OpenAI GPT Latest
$2.00$10.00$0.20-1M
OpenAI GPT Mini Latest
$0.75$4.50$0.075-400K

~x-ai(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
xAI: Grok Latest
$2.00$6.00$0.50-500K

~z-ai(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Z.ai: GLM Latest
$1.19$4.18$0.247-1M

aion-labs(4 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
AionLabs: Aion-2.0
$0.80$1.60$0.20-131K
AionLabs: Aion-3.0
$3.00$6.00$0.75-131K
AionLabs: Aion-3.0-Mini
$0.70$1.40$0.18-131K
AionLabs: Aion-RP 1.0 (8B)
$0.80$1.60--33K

Amazon(5 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Amazon: Nova 2 Lite
$0.30$2.50--1M
Amazon: Nova Lite 1.0
$0.06$0.24--300K
Amazon: Nova Micro 1.0
$0.035$0.14--128K
Amazon: Nova Premier 1.0
$2.50$12.50$0.625-1M
Amazon: Nova Pro 1.0
$0.80$3.20--300K

anthracite-org(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Magnum v4 72B
$3.00$5.00--33K

arcee-ai(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Arcee AI: Trinity Large Thinking
$0.25$0.80$0.06-262K

baidu(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Baidu: ERNIE 4.5 VL 424B A47B
$0.42$1.25--123K

bytedance(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
ByteDance: UI-TARS 7B
$0.10$0.20$0.10-128K

bytedance-seed(6 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
ByteDance Seed: Seed 1.6
$0.25$2.00--262K
ByteDance Seed: Seed 1.6 Flash
$0.075$0.30--262K
ByteDance Seed: Seed 2.1 Turbo
$0.50$2.50--262K
ByteDance Seed: Seed-2.0-Code
$0.50$3.00--262K
ByteDance Seed: Seed-2.0-Lite
$0.25$2.00--262K
ByteDance Seed: Seed-2.0-Mini
$0.10$0.40--262K

cognitivecomputations(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Venice: Uncensored
$0.20$0.90--128K

Cohere(4 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Cohere: Command A
$2.50$10.00--256K
Cohere: Command R (08-2024)
$0.15$0.60--128K
Cohere: Command R+ (08-2024)
$2.50$10.00--128K
Cohere: Command R7B (12-2024)
$0.0375$0.15--128K

DeepSeek(16 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
DeepSeek: DeepSeek V3
$0.2574$1.03--164K
DeepSeek: DeepSeek V3 0324
$0.25$1.00--164K
DeepSeek: DeepSeek V3.1
$0.55$1.65$0.55-164K
DeepSeek: R1
$0.70$2.50--64K
DeepSeek: R1 0528
$0.50$2.15$0.35-164K
DeepSeek: R1 Distill Llama 70B
$0.80$0.80--8K
DeepSeek: DeepSeek V3.1 Terminus
$0.27$1.00$0.135-164K
DeepSeek: DeepSeek V3.2
$0.269$0.40$0.1345-164K
DeepSeek: DeepSeek V3.2 Exp
$0.27$0.41--164K
DeepSeek: DeepSeek V4 Flash 0423
$0.079$0.1579$0.0158-1M
DeepSeek: DeepSeek V4 Flash 0731
$0.065$0.18$0.016-1M
DeepSeek: DeepSeek V4 Flash 0731 (batch)
$0.14$0.28$0.03-1M
DeepSeek: DeepSeek V4 Flash Vision Exp
$0.22$0.66$0.007-1M
DeepSeek: DeepSeek V4 Pro 0423
$0.4173$0.8345$0.0348-1M
DeepSeek: DeepSeek V4 Pro 0813
$0.66$1.98$0.022-1M
DeepSeek: DeepSeek V4 Pro 0813 (batch)
$1.32$3.96$0.13-1M

gryphe(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
MythoMax 13B
$0.06$0.06--8K

ibm-granite(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
IBM: Granite 4.0 Micro
$0.017$0.112--131K
IBM: Granite 4.1 8B
$0.05$0.10$0.05-131K

inception(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Inception: Mercury 2
$0.25$0.75$0.025-128K

inclusionai(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Ling-3.0-flash
$0.021$0.063$0.0042-262K

kwaipilot(3 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Kwaipilot: KAT-Coder-Air V2.5
$0.15$0.60$0.03-256K
Kwaipilot: KAT-Coder-Pro V2
$0.30$1.20$0.06-262K
Kwaipilot: KAT-Coder-Pro V2.5
$0.74$2.96$0.15-262K

mancer(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Mancer: Weaver (alpha)
$0.50$0.75--8K

meituan(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Meituan: LongCat 2.0
$0.30$1.20$0.006-1M

Meta(5 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Meta: Muse Glimmer 30B
$0.30$1.20$0.04-131K
Meta: Muse Glimmer 30B (batch)
$0.35$1.50$0.04-131K
Meta: Muse Spark 1.1
$1.25$4.25$0.15-1M
Meta: Muse Spark 1.2
$1.25$4.25$0.15-1M
Meta: Muse Spark 1.2 Contributor
$0.10$0.20$0.002-1M

Meta(8 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Meta: Llama 3.1 70B Instruct
$0.40$0.40--131K
Meta: Llama 3.1 8B Instruct
$0.05$0.08$0.025-131K
Meta: Llama 3.2 1B Instruct
$0.027$0.201--60K
Meta: Llama 3.2 3B Instruct
$0.05$0.33--131K
Meta: Llama 3.3 70B Instruct
$0.71$0.71$0.71-131K
Meta: Llama 4 Maverick
$0.20$0.80--1M
Meta: Llama 4 Scout
$0.11$0.34$0.055-1M
Meta: Llama Guard 4 12B
$0.18$0.18--164K

microsoft(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Microsoft: Phi 4
$0.07$0.14--16K
WizardLM-2 8x22B
$0.62$0.62--66K

MiniMax(9 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
MiniMax: MiniMax-01
$0.20$1.10--1M
MiniMax: MiniMax M1
$0.55$2.20--1M
MiniMax: MiniMax M2
$0.255$1.02--205K
MiniMax: MiniMax M2-her
$0.30$1.20$0.03-66K
MiniMax: MiniMax M2.1
$0.30$1.20$0.03-205K
MiniMax: MiniMax M2.5
$0.27$1.08$0.027-205K
MiniMax: MiniMax M2.7
$0.30$1.20$0.06-205K
MiniMax: MiniMax M3
$0.30$1.20$0.06-1M
MiniMax: MiniMax M3 (batch)
$0.30$1.20$0.06-524K

Mistral(25 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Mistral: Codestral 2508
$0.30$0.90$0.03-256K
Mistral: Codestral 2508 (batch)
$0.30$0.90$0.03-256K
Mistral: Devstral 2 2512
$0.40$2.00$0.04-262K
Mistral: Ministral 3 14B 2512
$0.20$0.20$0.02-262K
Mistral: Ministral 3 3B 2512
$0.10$0.10$0.01-131K
Mistral: Ministral 3 8B 2512
$0.15$0.15$0.015-262K
Mistral: Ministral 3 8B 2512 (batch)
$0.15$0.15$0.015-262K
Mistral Large
$2.00$6.00$0.20-128K
Mistral Large 2407
$2.00$6.00$0.20-131K
Mistral: Mistral Large 3 2512
$0.50$1.50$0.05-262K
Mistral: Mistral Large 3 2512 (batch)
$0.50$1.50$0.05-262K
Mistral: Mistral Medium 3
$0.40$2.00$0.04-131K
Mistral: Mistral Medium 3.5
$1.50$7.50--262K
Mistral: Mistral Medium 3.5 (batch)
$0.75$3.75--262K
Mistral: Mistral Medium 3.1
$0.40$2.00$0.04-131K
Mistral: Mistral Medium 3.1 (batch)
$0.40$2.00$0.04-131K
Mistral: Mistral Nemo
$0.019$0.03--131K
Mistral: Saba
$0.20$0.60$0.02-33K
Mistral: Mistral Small 3
$0.05$0.08--33K
Mistral: Mistral Small 4
$0.15$0.60$0.015-262K
Mistral: Mistral Small 4 (batch)
$0.15$0.60$0.015-262K
Mistral: Mistral Small 3.1 24B
$0.351$0.555--128K
Mistral: Mistral Small 3.2 24B
$0.075$0.20--131K
Mistral: Mixtral 8x22B Instruct
$2.00$6.00$0.20-66K
Mistral: Voxtral Small 24B 2507
$0.10$0.30$0.01-33K

Moonshot(8 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
MoonshotAI: Kimi K2 0711
$0.57$2.30--131K
MoonshotAI: Kimi K2 0905
$0.60$2.50--262K
MoonshotAI: Kimi K2 Thinking
$0.60$2.50$0.15-262K
MoonshotAI: Kimi K2.5
$0.60$3.00$0.10-262K
MoonshotAI: Kimi K2.6
$0.95$4.00$0.16-262K
MoonshotAI: Kimi K2.7 Code
$0.66$3.40$0.18-262K
MoonshotAI: Kimi K3
$3.00$15.00$0.30-1M
MoonshotAI: Kimi K3 (batch)
$3.00$15.00$0.30-1M

morph(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Morph: Morph V3 Fast
$0.80$1.20--82K
Morph: Morph V3 Large
$0.90$1.90--262K

nex-agi(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Nex AGI: Nex-N2-Mini
$0.025$0.10$0.0025-262K
Nex AGI: Nex-N2-Pro
$0.25$1.00$0.025-262K

nousresearch(4 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Nous: Hermes 3 405B Instruct
$1.00$1.00--131K
Nous: Hermes 3 70B Instruct
$0.70$0.70--131K
Nous: Hermes 4 405B
$1.00$3.00--131K
Nous: Hermes 4 70B
$0.13$0.40--131K

NVIDIA(5 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
NVIDIA: Nemotron 3 Nano 30B A3B
$0.05$0.20$0.025-262K
NVIDIA: Nemotron 3 Super
$0.085$0.40--1M
NVIDIA: Nemotron 3 Ultra
$0.50$2.20$0.10-262K
NVIDIA: Nemotron 3 Ultra (batch)
$0.60$3.60$0.20-512K
NVIDIA: Nemotron 3.5 Lightning
$0.08$0.20$0.04-262K

openrouter(5 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Auto Router
$-1000000.00$-1000000.00--2M
Auto Router (Beta)
$-1000000.00$-1000000.00--2M
Body Builder (beta)
$-1000000.00$-1000000.00--128K
OpenRouter: Fusion
$-1000000.00$-1000000.00--1M
Pareto Code Router
$-1000000.00$-1000000.00--2M

perceptron(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Perceptron: Perceptron Mk1
$0.15$1.50--33K

perplexity(5 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Perplexity: Sonar
$1.00$1.00--127K
Perplexity: Sonar Deep Research
$2.00$8.00-$3.00128K
Perplexity: Sonar Pro
$3.00$15.00--200K
Perplexity: Sonar Pro Search
$3.00$15.00--200K
Perplexity: Sonar Reasoning Pro
$2.00$8.00--128K

poolside(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Poolside: Laguna S 2.1
$0.09$0.18$0.009-1M
Poolside: Laguna XS 2.1
$0.06$0.12$0.03-262K

Qwen(53 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Qwen2.5 72B Instruct
$0.36$0.40--33K
Qwen: Qwen2.5 7B Instruct
$0.10$0.20--33K
Qwen2.5 Coder 32B Instruct
$0.66$1.00--33K
Qwen: Qwen-Plus
$0.26$0.78$0.052-1M
Qwen: Qwen Plus 0728
$0.26$0.78--1M
Qwen: Qwen2.5 VL 72B Instruct
$0.25$0.75--128K
Qwen: Qwen3 14B
$0.12$0.24--131K
Qwen: Qwen3 235B A22B
$0.455$1.82--131K
Qwen: Qwen3 235B A22B Instruct 2507
$0.0875$0.35$0.0175-262K
Qwen: Qwen3 235B A22B Thinking 2507
$0.23$2.30--131K
Qwen: Qwen3 30B A3B
$0.12$0.50--131K
Qwen: Qwen3 30B A3B Instruct 2507
$0.0481$0.193--262K
Qwen: Qwen3 30B A3B Thinking 2507
$0.20$2.40--82K
Qwen: Qwen3 32B
$0.08$0.28--131K
Qwen: Qwen3 8B
$0.117$0.455--131K
Qwen: Qwen3 Coder 480B A35B
$0.30$1.00$0.10-262K
Qwen: Qwen3 Coder 30B A3B Instruct
$0.07$0.28--262K
Qwen: Qwen3 Coder Flash
$0.195$0.975$0.039-1M
Qwen: Qwen3 Coder Next
$0.12$0.80$0.07-262K
Qwen: Qwen3 Coder Plus
$0.65$3.25$0.13-1M
Qwen: Qwen3 Max
$0.78$3.90$0.156-262K
Qwen: Qwen3 Max Thinking
$0.78$3.90--262K
Qwen: Qwen3 Next 80B A3B Instruct
$0.09$1.10--262K
Qwen: Qwen3 Next 80B A3B Thinking
$0.15$1.20--262K
Qwen: Qwen3 VL 235B A22B Instruct
$0.21$1.90$0.10-262K
Qwen: Qwen3 VL 235B A22B Thinking
$0.40$4.00--131K
Qwen: Qwen3 VL 30B A3B Instruct
$0.15$0.60--262K
Qwen: Qwen3 VL 30B A3B Thinking
$0.20$2.40--262K
Qwen: Qwen3 VL 32B Instruct
$0.104$0.416--131K
Qwen: Qwen3 VL 8B Instruct
$0.117$0.455--262K
Qwen: Qwen3 VL 8B Thinking
$0.18$2.10--131K
Qwen: Qwen3.5-122B-A10B
$0.29$2.40--262K
Qwen: Qwen3.5-27B
$0.195$1.56--262K
Qwen: Qwen3.5-35B-A3B
$0.25$1.25$0.25-262K
Qwen: Qwen3.5 397B A17B
$0.39$2.34--262K
Qwen: Qwen3.5-9B
$0.10$0.15--262K
Qwen: Qwen3.5-9B (batch)
$0.17$0.25--262K
Qwen: Qwen3.5-Flash
$0.065$0.26--1M
Qwen: Qwen3.5 Plus 2026-02-15
$0.26$1.56--1M
Qwen: Qwen3.5 Plus 2026-04-20
$0.30$1.80--1M
Qwen: Qwen3.6 27B
$0.60$3.60$0.12-262K
Qwen: Qwen3.6 35B A3B
$0.10$0.90$0.05-262K
Qwen: Qwen3.6 Flash
$0.1875$1.13--1M
Qwen: Qwen3.6 Max Preview
$1.03$6.16--262K
Qwen: Qwen3.6 Plus
$0.325$1.95--1M
Qwen: Qwen3.7 Flash
$0.03$0.13$0.006-1M
Qwen: Qwen3.7 Max
$1.48$4.42$0.295-1M
Qwen: Qwen3.7 Plus
$0.32$1.28$0.064-1M
Qwen: Qwen3.8 2.4T A95B
$2.00$6.00$0.25-1M
Qwen: Qwen3.8 2.4T A95B (batch)
$2.50$6.25$0.50-1M
Qwen: Qwen3.8 27B
$0.425$2.55$0.085-1M
Qwen: Qwen3.8 Flash
$0.15$0.47$0.016-1M
Qwen: Qwen3.8 Max
$2.00$6.00$0.25-1M

rekaai(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Reka Edge
$0.10$0.10--16K
Reka Flash 3
$0.10$0.20--66K

relace(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Relace: Relace Apply 3
$0.85$1.25--256K
Relace: Relace Search
$1.00$3.00--256K

sakana(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Sakana: Fugu Ultra
$5.00$30.00$0.50-1M
Sakana: Sakana Namazu
$0.95$4.00$0.15-262K

sao10k(3 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Sao10K: Llama 3 8B Lunaris
$0.04$0.05--8K
Sao10K: Llama 3.1 Euryale 70B v2.2
$0.85$0.85--131K
Sao10K: Llama 3.3 Euryale 70B
$0.65$0.75--131K

stepfun(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
StepFun: Step 3.5 Flash
$0.10$0.30--262K
StepFun: Step 3.7 Flash
$0.20$1.15$0.04-262K

tencent(7 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Tencent: Hunyuan A13B Instruct
$0.14$0.57--131K
Tencent: Hy-MT2-1.8B
$0.044$0.177--8K
Tencent: Hy-MT2-30B-A3B
$0.074$0.295--8K
Tencent: Hy-MT2-7B
$0.074$0.295--8K
Tencent: Hy3
$0.0825$0.33$0.0206-262K
Tencent: Hy3 preview
$0.18$0.60$0.06-262K
Tencent: Hy4 preview
$0.834$2.50$0.042-1M

thedrummer(3 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
TheDrummer: Cydonia 24B V4.1
$0.30$0.50$0.15-131K
TheDrummer: Skyfall 36B V2
$0.55$0.80$0.25-33K
TheDrummer: UnslopNemo 12B
$0.40$0.40--1M

thinkingmachines(4 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Thinking Machines: Inkling
$1.00$4.05$0.17-1M
Thinking Machines: Inkling Small
$0.45$1.20$0.10-1M
Thinking Machines: Inkling Small (batch)
$0.50$1.20$0.10-524K
Thinking Machines: Inkling (batch)
$1.00$4.05$0.17-524K

undi95(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
ReMM SLERP 13B
$0.45$0.65--6K

upstage(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Upstage: Solar Pro 3
$0.15$0.60$0.015-131K
Upstage: Solar Pro 4
$0.03$0.12$0.006-524K

writer(1 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Writer: Palmyra X5
$0.60$6.00--1M

xAI(6 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
SpaceXAI: Grok 4.20
$1.25$2.50$0.20-2M
SpaceXAI: Grok 4.20 Multi-Agent
$1.25$2.50$0.20-2M
SpaceXAI: Grok 4.3
$1.25$2.50$0.20-1M
SpaceXAI: Grok 4.5
$2.00$4.00 >200K
$6.00$12.00 >200K
$0.30$0.60 >200K
-500K
SpaceXAI: Grok 4.6
$2.00$4.00 >200K
$6.00$12.00 >200K
$0.50$1.00 >200K
-500K
SpaceXAI: Grok Build 0.1
$1.00$2.00$0.20-256K

xiaomi(2 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Xiaomi: MiMo-V2.5
$0.14$0.28$0.0028-1M
Xiaomi: MiMo-V2.5-Pro
$0.435$0.87$0.0036-1M

Z.ai(15 models)

ModelInput / 1M tokensOutput / 1M tokensCache Read / 1MReasoning / 1MContext
Z.ai: GLM 4.5
$0.60$2.20$0.11-131K
Z.ai: GLM 4.5 Air
$0.13$0.85$0.025-131K
Z.ai: GLM 4.5V
$0.60$1.80$0.11-66K
Z.ai: GLM 4.6
$0.43$1.75$0.08-205K
Z.ai: GLM 4.6V
$0.30$0.90$0.055-131K
Z.ai: GLM 4.7
$0.40$1.75$0.08-205K
Z.ai: GLM 4.7 Flash
$0.06$0.40$0.01-203K
Z.ai: GLM 5
$0.60$1.92$0.12-205K
Z.ai: GLM 5 Turbo
$1.20$4.00$0.24-203K
Z.ai: GLM 5.1
$0.966$3.04$0.1794-205K
Z.ai: GLM 5.2
$1.19$3.74$0.221-1M
Z.ai: GLM 5.3
$1.40$4.40$0.26-1M
Z.ai: GLM 5.3 Flash
$0.075$0.25$0.015-1M
Z.ai: GLM 5.3 Flash (batch)
$0.15$0.50$0.03-1M
Z.ai: GLM 5V Turbo
$1.20$4.00$0.24-203K