Just shipped

Kimi K3 to Claude Opus 5,up to 20% below list.

Open weights and closed frontier in one OpenAI-compatible API — Kimi K3, DeepSeek V4 Pro, Claude Opus 5, GPT-5.6, Grok 4.5, Gemini 3.5. One key, one bill, no subscription, and up to 20% below the standard per-token rate. Change two lines of code.

Try it in your browserfree trial balance, no card

  • No subscription
  • OpenAI compatible
  • Pay as you go
  • Text + images
Drop-in with
  • OpenAI SDK
  • Aider
  • Cursor
  • Cline
  • Continue.dev
  • LangChain
  • Vercel AI SDK
python
1# Two lines. That is the whole migration.
2from openai import OpenAI
3 
4client = OpenAI(
5 base_url="https://api.quicksilverpro.io/v1",
6 api_key="your-api-key",
7)
Pricing

Frontier and open models. One API. Priced below list.

Everything about price, in one place — no scrolling required.

Model
Context
Input
Output
vs. list
deepseek-v4-pro
premium reasoning
1M
$0.435
$0.87
kimi-k3
Multimodal reasoning, agentic coding
1M
$2.40$3.00
$12.00$15.00
−20%
deepseek-v4-flash
fast chat & coding, thinking on by default
1M
$0.112$0.14
$0.224$0.28
−20%
glm-5.3
complex software engineering, long-horizon agents
1M
$1.12$1.40
$3.52$4.40
−20%
glm-5.2
long-horizon agents, project-level coding
1M
$1.12$1.40
$3.52$4.40
−20%
nemotron-3.5-lightning
fast, tool-heavy agents and high-volume automation
262K
$0.08
$0.20
minimax-m3
long-horizon agentic coding, tool use
1M
$0.24
$0.96
qwen3.8-max
Qwen 3.8 flagship, autonomous coding
1M
$2.00
$6.00
qwen3.7-max
Qwen 3.7 flagship, agent / coding
1M
$1.25$1.475
$3.75$4.425
−15%
qwen3.7-flash
fast multimodal agents, visual coding, search
1M
$0.024$0.03
$0.104$0.13
−20%
kimi-k2.6
Opus-class agentic / planning
256K
$0.5472
$2.728
Claude Opus 5New
claude-opus-5
demanding reasoning, end-to-end coding, visual analysis, long-horizon agents
1M
$4.00$5.00
$20.00$25.00
−20%
GPT-5.6 SolNew
gpt-5.6-sol
complex reasoning, agentic coding, long-horizon tasks
1M
$4.00$5.00
$24.00$30.00
−20%
Hy3New
hy3
general-purpose coding, agentic workflows
262K
$0.1056$0.132
$0.4224$0.528
−20%
mimo-v2.5
cost-efficient everyday coding, agentic workflows
1M
$0.112
$0.224
qwen3.6-plus
thinks-by-default flagship
1M
$0.26$0.325
$1.56$1.95
−20%
qwen3.7-plus
Qwen 3.7 agent flagship, long-horizon coding
1M
$0.256$0.32
$1.024$1.28
−20%
qwen3.6-35b
long-context RAG, 35B MoE
262K
$0.112$0.14
$0.80$1.00
−20%
kimi-k2.7-code
Long-horizon agentic coding
256K
$0.584$0.73
$2.80$3.50
−20%
GPT-5.6 TerraNew
gpt-5.6-terra
everyday coding, reasoning, balanced agentic
1M
$0.80$1.00
$4.80$6.00
−20%
GPT-5.6 LunaNew
gpt-5.6-luna
high-volume chat, classification, lightweight agentic
1M
$0.08$0.10
$0.48$0.60
−20%
Grok 4.5New
grok-4.5
coding, knowledge work, STEM
500K
$1.60$2.00
$4.80$6.00
−20%
grok-4.6
frontier coding, knowledge work, STEM, visual analysis
500K
$2.00
$6.00
Claude Fable 5New
claude-fable-5
most capable reasoning, long-horizon agentic
1M
$8.00$10.00
$40.00$50.00
−20%
Claude Opus 4.8New
claude-opus-4-8
top-tier reasoning, coding, agentic
1M
$4.00$5.00
$20.00$25.00
−20%
Claude Opus 4.6New
claude-opus-4-6
deep reasoning and coding
1M
$4.00$5.00
$20.00$25.00
−20%
Claude Sonnet 4.6New
claude-sonnet-4-6
balanced mid-tier, fast & capable
1M
$2.40$3.00
$12.00$15.00
−20%
Claude Sonnet 5New
claude-sonnet-5
newest Sonnet, stronger reasoning & coding, below list price
1M
$2.00$3.00
$10.00$15.00
−33%
Claude Haiku 4.5New
claude-haiku-4-5
fast, low-cost, high-volume tasks
200K
$0.80$1.00
$4.00$5.00
−20%
gemini-3.6-flash
current general-purpose Flash GA
1M
$1.275$1.50
$6.375$7.50
−15%
gemini-3.5-flash-lite
current low-cost, high-volume workloads
1M
$0.255$0.30
$2.125$2.50
−15%
gemini-3.5-flash
next-gen Flash GA
1M
$1.275$1.50
$7.65$9.00
−15%
gemini-3.1-pro-preview
flagship reasoning
1M
$1.70$2.00
$10.20$12.00
−15%
flux.2-pro
flagship image generation
$0.027/img$0.031/img
Gemini 3 Pro ImageNew
gemini-3-pro-image
GA pro-grade image generation
1M
$1.70$2.00
$10.20$12.00
per image $0.114/img$0.134/img
−15%

Prices are exact per-token rates, not rounded — some carry more decimals than others. Each struck-through figure is that row's reference list price and links to the page that publishes it, so every line can be checked. "At list" means we sell at the reference price; "mixed" means we are below it on one side and above on the other. Every row checkable, every model verifiable →

Building an agent? Fund a key programmatically with USDC — no account, no card. x402 docs →

FAQ

Common questions

QuickSilver Pro is an OpenAI-compatible inference API. The current catalog is DeepSeek V4 Flash, DeepSeek V4 Pro, Qwen3.8 Max, Qwen3.7 Max, Qwen3.7 Plus, Qwen3.7 Flash, Qwen3.6 Plus, Qwen3.6-35B-A3B, Kimi K2.6, Kimi K2.7 Code, Kimi K3, Muse Spark 1.2, Muse Glimmer 30B, GLM 5.3, GLM 5.2, Nemotron 3.5 Lightning, GPT-5.6 Luna, GPT-5.6 Terra, GPT-5.6 Sol, Grok 4.5, Grok 4.6, MiniMax M3, MiMo-V2.5, Hy3, Claude Opus 5, Claude Fable 5, Claude Opus 4.8, Claude Opus 4.6, Claude Sonnet 4.6, Claude Sonnet 5, Claude Haiku 4.5, Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash, Gemini 3.1 Pro Preview, Gemini 3 Pro Image, Gemini 3 Flash Preview, Gemini 3.1 Flash Lite and FLUX.2 Pro, served through one endpoint and one API key.

V4 Flash revision 0731 is the current agent build. It has 1M context and supports Chat Completions and Responses.

Up to 20% below the standard published per-token list rate, on most of the catalog. DeepSeek V4 Flash: $0.112 / $0.224. DeepSeek V4 Pro: $0.435 / $0.87. Kimi K3: $2.40 / $12.00. GLM 5.2: $1.12 / $3.52. Claude Opus 5: $4.00 / $20.00. GPT-5.6 Sol: $4.00 / $24.00. Grok 4.5: $1.60 / $4.80. Closed frontier models run through the same endpoint and the same key as the open-weight ones — Claude, GPT-5.6, Gemini and Grok included.

Yes. Change base_url to https://api.quicksilverpro.io/v1 in the official openai Python / Node / Swift SDKs. Streaming, tool calling, json_schema strict mode, and usage.cost accounting all work out of the box.

Yes, two ways. Every account gets a free trial balance to use in the browser chat at /chat — no card, no API key. And on your first credit purchase we match 100% of it, up to $50 in bonus credits: pay $5 and get $10, pay $50 and get $100. A larger first payment still receives the $50 maximum. One-time, applied automatically; standard pay-as-you-go after that.

Get your API key

Create an account, get your API key in 30 seconds.

Get API Key

Model launches & updates. A few emails a month.