FAQ

Frequently asked questions

Everything you might want to know about QuickSilver Pro — the OpenAI-compatible inference API for GPT-6.1 Sol, Claude Opus 5.5, Sonnet 5.5 and Haiku 5.5, Gemini 3.8 Flash, Grok 4.7, Qwen3.8 Max, GLM-5.3, DeepSeek V4.1, Mistral Large 4 and Kimi K3.

Frequently asked questions

QuickSilver Pro is an OpenAI-compatible inference API with 71 models, one endpoint and one API key. Browse models →

The live catalog contains 71 models. Browse their capabilities, context windows and current prices. Browse models →

Yes. To turn reasoning off for direct chat, send reasoning.enabled=false in the request body; in Python use extra_body.

Up to 20% below the standard published per-token list rate on most of the catalog, and 50–67% below on Claude. DeepSeek V4.1 Flash: $0.125 / $0.55. DeepSeek V4 Pro: $0.70 / $2.10. Kimi K3: $2.55 / $12.75. GLM 5.3: $1.12 / $3.52. Claude Opus 5.5: $1.60 / $8.00. GPT-6.1 Sol: $2.00 / $10.00. Grok 4.7: $2.00 / $6.00. Closed frontier models run through the same endpoint and the same key as the open-weight ones — Claude, GPT-6.1, Gemini and Grok included.

Yes. Change base_url to https://api.quicksilverpro.io/v1 in the official openai Python / Node / Swift SDKs. Streaming, tool calling, and usage.cost accounting all work out of the box. json_schema strict mode is model-dependent: the Claude models do not support it, so a schema there is advisory and the enforced path is a tool with strict: true.

Yes. Any tool that accepts an OpenAI base_url and API key can use QuickSilver Pro. Start with gpt-6.1-sol, gpt-6-luna, claude-opus-5-5, claude-sonnet-5-5, claude-haiku-5-5, gemini-3.8-flash, mistral-large-4, deepseek-v4.1-flash, qwen3.8-max, kimi-k3, glm-5.3 and grok-4.7.

Change base_url to api.quicksilverpro.io/v1, swap the API key, and apply these model-ID mappings: deepseek/deepseek-v4-flash-0731 -> deepseek-v4-flash, deepseek/deepseek-v4-pro -> deepseek-v4-pro, qwen/qwen3.7-max -> qwen3.7-max, qwen/qwen3.7-plus -> qwen3.7-plus, qwen/qwen3.7-flash -> qwen3.7-flash, qwen/qwen3.6-plus -> qwen3.6-plus, qwen/qwen3.6-35b-a3b -> qwen3.6-35b, moonshotai/kimi-k2.6 -> kimi-k2.6, moonshotai/kimi-k2.7-code -> kimi-k2.7-code, moonshotai/kimi-k3 -> kimi-k3, z-ai/glm-5.3 -> glm-5.3, z-ai/glm-5.3-flash -> glm-5.3-flash, openai/gpt-oss-120b -> gpt-oss-120b, qwen/qwen3.8-27b -> qwen3.8-27b, qwen/qwen3.8-flash -> qwen3.8-flash-next, qwen/qwen3.8-omni-flash -> qwen3.8-omni-flash and z-ai/glm-5.2 -> glm-5.2.

New accounts get $0.05 free credit with no card, usable in browser chat and through the API with your account key on 5 starter models (GPT-6 Luna, Claude Haiku 5.5, DeepSeek V4.1 Flash, Qwen3.8 Flash Next and MiMo-V2.6-Flash); top up from $5 to unlock the rest. Browser chat also has a separate allowance of 5 free messages. And your first credit purchase is matched 100%, up to $50 in bonus credits. The bonus is added to your balance the next time you top up (any amount from $5): pay $20, then top up $5 later, and $20 of bonus arrives with it. A larger first payment still receives the $50 maximum. One per customer and payment card; standard pay-as-you-go after that.

New accounts get $0.05 free credit with no card, usable in browser chat and through the API with your account key on 5 starter models (GPT-6 Luna, Claude Haiku 5.5, DeepSeek V4.1 Flash, Qwen3.8 Flash Next and MiMo-V2.6-Flash); top up from $5 to unlock the rest. Browser chat also has a separate allowance of 5 free messages.

MachineFi Inc., a Delaware corporation based in Menlo Park, CA. Payments are processed by Stripe and appear on your statement as MACHINEFI INC.

Still have questions?

hello@quicksilverpro.io

Built by MachineFi Labs.