Frequently asked questions
Everything you might want to know about QuickSilver Pro — the OpenAI-compatible inference API for Kimi K3, DeepSeek V4 Flash & Pro, GLM 5.2, Claude Opus 5, GPT-5.6, Grok 4.5 and Gemini 3.5.
Frequently asked questions
QuickSilver Pro is an OpenAI-compatible inference API. The current catalog is DeepSeek V4 Flash, DeepSeek V4 Pro, Qwen3.8 Max, Qwen3.7 Max, Qwen3.7 Plus, Qwen3.7 Flash, Qwen3.6 Plus, Qwen3.6-35B-A3B, Kimi K2.6, Kimi K2.7 Code, Kimi K3, Muse Spark 1.2, Muse Glimmer 30B, GLM 5.3, GLM 5.2, Nemotron 3.5 Lightning, GPT-5.6 Luna, GPT-5.6 Terra, GPT-5.6 Sol, Grok 4.5, Grok 4.6, MiniMax M3, MiMo-V2.5, Hy3, Claude Opus 5, Claude Fable 5, Claude Opus 4.8, Claude Opus 4.6, Claude Sonnet 4.6, Claude Sonnet 5, Claude Haiku 4.5, Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash, Gemini 3.1 Pro Preview, Gemini 3 Pro Image, Gemini 3 Flash Preview, Gemini 3.1 Flash Lite and FLUX.2 Pro, served through one endpoint and one API key.
The live catalog is DeepSeek V4 Flash, DeepSeek V4 Pro, Qwen3.8 Max, Qwen3.7 Max, Qwen3.7 Plus, Qwen3.7 Flash, Qwen3.6 Plus, Qwen3.6-35B-A3B, Kimi K2.6, Kimi K2.7 Code, Kimi K3, Muse Spark 1.2, Muse Glimmer 30B, GLM 5.3, GLM 5.2, Nemotron 3.5 Lightning, GPT-5.6 Luna, GPT-5.6 Terra, GPT-5.6 Sol, Grok 4.5, Grok 4.6, MiniMax M3, MiMo-V2.5, Hy3, Claude Opus 5, Claude Fable 5, Claude Opus 4.8, Claude Opus 4.6, Claude Sonnet 4.6, Claude Sonnet 5, Claude Haiku 4.5, Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, Gemini 3.5 Flash, Gemini 3.1 Pro Preview, Gemini 3 Pro Image, Gemini 3 Flash Preview, Gemini 3.1 Flash Lite and FLUX.2 Pro. Model facts and availability are updated from the same catalog used by the API.
V4 Flash revision 0731 is the current agent build. It has 1M context and supports Chat Completions and Responses.
Up to 20% below the standard published per-token list rate, on most of the catalog. DeepSeek V4 Flash: $0.112 / $0.224. DeepSeek V4 Pro: $0.435 / $0.87. Kimi K3: $2.40 / $12.00. GLM 5.2: $1.12 / $3.52. Claude Opus 5: $4.00 / $20.00. GPT-5.6 Sol: $4.00 / $24.00. Grok 4.5: $1.60 / $4.80. Closed frontier models run through the same endpoint and the same key as the open-weight ones — Claude, GPT-5.6, Gemini and Grok included.
Yes. Change base_url to https://api.quicksilverpro.io/v1 in the official openai Python / Node / Swift SDKs. Streaming, tool calling, json_schema strict mode, and usage.cost accounting all work out of the box.
Yes. Any tool that accepts an OpenAI base_url and API key can use QuickSilver Pro. Start with deepseek-v4-flash, deepseek-v4-pro, qwen3.8-max, qwen3.7-max, qwen3.7-plus and qwen3.7-flash.
Change base_url to api.quicksilverpro.io/v1, swap the API key, and apply these model-ID mappings: deepseek/deepseek-v4-flash-0731 -> deepseek-v4-flash, deepseek/deepseek-v4-pro -> deepseek-v4-pro, qwen/qwen3.7-max -> qwen3.7-max, qwen/qwen3.7-plus -> qwen3.7-plus, qwen/qwen3.7-flash -> qwen3.7-flash, qwen/qwen3.6-plus -> qwen3.6-plus, qwen/qwen3.6-35b-a3b -> qwen3.6-35b, moonshotai/kimi-k2.6 -> kimi-k2.6, moonshotai/kimi-k2.7-code -> kimi-k2.7-code, moonshotai/kimi-k3 -> kimi-k3, z-ai/glm-5.3 -> glm-5.3 and z-ai/glm-5.2 -> glm-5.2.
Yes, two ways. Every account gets a free trial balance to use in the browser chat at /chat — no card, no API key. And on your first credit purchase we match 100% of it, up to $50 in bonus credits: pay $5 and get $10, pay $50 and get $100. A larger first payment still receives the $50 maximum. One-time, applied automatically; standard pay-as-you-go after that.
Yes. Create an account and you get a free trial balance to chat with several models in your browser at /chat — no card and no API key needed. Every reply shows what it cost, at the same per-token rate the API charges.
MachineFi Inc., a Delaware corporation based in Menlo Park, CA. Payments are processed by Stripe and appear on your statement as MACHINEFI INC.
Still have questions?
hello@quicksilverpro.ioBuilt by MachineFi Labs.