Frontier and open models. One API. Priced below list.
Everything about price, in one place — no scrolling required.
Pay with
- Card
- Apple Pay
- Google Pay
- Alipay
- WeChat Pay
- Crypto (USDC)
Options at checkout depend on your country. Minimum top-up $5.
71 models. Sorted by Popular
gpt-6.1-solclaude-opus-5-5claude-fable-5-1gpt-6-astraclaude-sonnet-5-5claude-haiku-5-5gemini-3.8-flashgrok-4.7deepseek-v4.1-flashgpt-6-solgpt-6-lunamuse-spark-1.3qwen3.8-max-primeglm-5.3-primemimo-v2.6-proglm-5.3-flashdeepseek-v4-flashgpt-oss-120bqwen3.8-flash-nextqwen3.8-omni-flashnemotron-3-ultranemotron-3.5-lightningmuse-spark-1.2muse-glimmer-30bclaude-opus-5gpt-5.6-solmistral-large-4mimo-v2.6-flashgpt-5.6-terragpt-5.6-lunagrok-4.5grok-4.6claude-opus-4-8claude-sonnet-5gemini-3.5-flash-litejev-1.13gemini-3-pro-imagegemini-3-flash-previewgemini-3.1-flash-litePrices are exact per-token rates, not rounded — some carry more decimals than others. Each struck-through figure is that row's reference list price and links to the page that publishes it, so every line can be checked. "At list" means we sell at the reference price; "mixed" means we are below it on one side and above on the other. Every row checkable, every model verifiable →
Building an agent? Fund a key programmatically with USDC — no account, no card. x402 docs →
Plug in your monthly usage to see what it costs here, next to published list rates.
Monthly estimate for DeepSeek V4.1 Flash
Estimates use the selected model’s base input and output rates and your monthly token totals; adjust the sliders to your workload.
Side-by-side pricing vs every competitor
qspBuilt for terminals and AI agents. --json output with stable exit codes — Claude Code, Cursor, Aider can call it without parsing HTML.
brew install machinefi/qspro/qspro
qsp init
qsp chat "Write a haiku about debugging." -m deepseek-v4.1-flash1# Two lines. That is the whole migration.2from openai import OpenAI34client = OpenAI(5 base_url="https://api.quicksilverpro.io/v1",6 api_key="your-api-key",7)
Common questions
QuickSilver Pro is an OpenAI-compatible inference API with 71 models, one endpoint and one API key. Browse models →
Yes. To turn reasoning off for direct chat, send reasoning.enabled=false in the request body; in Python use extra_body.
Up to 20% below the standard published per-token list rate on most of the catalog, and 50–67% below on Claude. DeepSeek V4.1 Flash: $0.125 / $0.55. DeepSeek V4 Pro: $0.70 / $2.10. Kimi K3: $2.55 / $12.75. GLM 5.3: $1.12 / $3.52. Claude Opus 5.5: $1.60 / $8.00. GPT-6.1 Sol: $2.00 / $10.00. Grok 4.7: $2.00 / $6.00. Closed frontier models run through the same endpoint and the same key as the open-weight ones — Claude, GPT-6.1, Gemini and Grok included.
Yes. Change base_url to https://api.quicksilverpro.io/v1 in the official openai Python / Node / Swift SDKs. Streaming, tool calling, and usage.cost accounting all work out of the box. json_schema strict mode is model-dependent: the Claude models do not support it, so a schema there is advisory and the enforced path is a tool with strict: true.
New accounts get $0.05 free credit with no card, usable in browser chat and through the API with your account key on 5 starter models (GPT-6 Luna, Claude Haiku 5.5, DeepSeek V4.1 Flash, Qwen3.8 Flash Next and MiMo-V2.6-Flash); top up from $5 to unlock the rest. Browser chat also has a separate allowance of 5 free messages. And your first credit purchase is matched 100%, up to $50 in bonus credits. The bonus is added to your balance the next time you top up (any amount from $5): pay $20, then top up $5 later, and $20 of bonus arrives with it. A larger first payment still receives the $50 maximum. One per customer and payment card; standard pay-as-you-go after that.