QuickSilver Pro system status

Status unknown
Checking...

Services

90-day history · client-measured
Inference API
api.quicksilverpro.io
Checking...
·
Dashboard backend
pay.quicksilverpro.io
Checking...
·
Website
quicksilverpro.io
Checking...
·

Model availability

Synthetic probe · updates every 3 min
DeepSeek V4 Flash
deepseek-v4-flash
official 0731 agent model, 1M context, Responses API
Checking...
·
DeepSeek V4 Pro
deepseek-v4-pro
premium reasoning, 1M context
Checking...
·
Qwen3.8 Max
qwen3.8-max
Qwen 3.8 flagship, 2.4T MoE, 1M context, autonomous coding
Checking...
·
Qwen3.7 Max
qwen3.7-max
Qwen 3.7 flagship, 1M context, thinks by default
Checking...
·
Qwen3.7 Plus
qwen3.7-plus
Qwen 3.7 agent flagship, 1M context, long-horizon coding
Checking...
·
Qwen3.7 Flash
qwen3.7-flash
fast multimodal agents, vision and tool use, 1M context
Checking...
·
Qwen3.6 Plus
qwen3.6-plus
1T-MoE flagship, 1M context, thinks by default
Checking...
·
Qwen3.6-35B-A3B
qwen3.6-35b
262K long-context, MoE, drop-in 3.5 upgrade
Checking...
·
Kimi K2.6
kimi-k2.6
Opus-class reasoning, 256K context
Checking...
·
Kimi K2.7 Code
kimi-k2.7-code
Agentic-coding K2, 256K context, coding-tuned K2.6 successor
Checking...
·
Kimi K3
kimi-k3
2.8T multimodal reasoning flagship, 1M context
Checking...
·
Muse Spark 1.2
muse-spark-1.2
Coding-focused reasoning model, 1M context
Checking...
·
Muse Glimmer 30B
muse-glimmer-30b
Compact agentic multimodal model, distilled from Muse Spark
Checking...
·
GLM 5.3
glm-5.3
Z.ai latest reasoning flagship, 1M context, complex software engineering & long-horizon agents
Checking...
·
GLM 5.2
glm-5.2
Z.ai reasoning flagship, 1M context, long-horizon agents & coding
Checking...
·
Nemotron 3.5 Lightning
nemotron-3.5-lightning
30B-A3B open MoE for fast, tool-heavy agents
Checking...
·
GPT-5.6 Luna
gpt-5.6-luna
OpenAI's fast, cost-efficient GPT-5.6 tier, 1M context
Checking...
·
GPT-5.6 Terra
gpt-5.6-terra
OpenAI's balanced GPT-5.6 mid-tier, 1M context
Checking...
·
GPT-5.6 Sol
gpt-5.6-sol
OpenAI's flagship GPT-5.6, complex reasoning & agentic coding, 1M context
Checking...
·
Grok 4.5
grok-4.5
xAI's smartest model, frontier coding, knowledge work & STEM, 500K context
Checking...
·
Grok 4.6
grok-4.6
frontier coding, knowledge work and STEM, vision, 500K context
Checking...
·
MiniMax M3
minimax-m3
MiniMax open-weight frontier model, 1M context, agentic coding
Checking...
·
MiMo-V2.5
mimo-v2.5
Xiaomi open-weight MiMo-V2.5, cost-efficient coding & agentic, 1M context
Checking...
·
Hy3
hy3
Tencent Hy3, selectable reasoning & agentic coding, 262K context
Checking...
·
Claude Opus 5
claude-opus-5
Anthropic flagship for demanding reasoning, coding & long-horizon agentic work, 1M context
Checking...
·
Claude Fable 5
claude-fable-5
Anthropic Mythos-class flagship, most capable reasoning
Checking...
·
Claude Opus 4.8
claude-opus-4-8
Anthropic flagship, top-tier reasoning & agentic
Checking...
·
Claude Opus 4.6
claude-opus-4-6
Anthropic flagship, deep reasoning & coding
Checking...
·
Claude Sonnet 4.6
claude-sonnet-4-6
balanced Anthropic mid-tier, fast & capable
Checking...
·
Claude Sonnet 5
claude-sonnet-5
Anthropic's newest Sonnet tier, 1M context
Checking...
·
Claude Haiku 4.5
claude-haiku-4-5
fast, low-cost Anthropic, high-volume
Checking...
·
Gemini 3.7 Flash
gemini-3.7-flash
most capable Flash for agents; promotional pricing through 2026-12-31
Checking...
·
Gemini 3.6 Flash
gemini-3.6-flash
current general-purpose Flash GA, 1M context
Checking...
·
Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
current low-cost Flash-Lite GA, 1M context
Checking...
·
Gemini 3.5 Flash
gemini-3.5-flash
Google's next-gen Flash GA, 1M context, thinking
Checking...
·
Gemini 3.1 Pro Preview
gemini-3.1-pro-preview
Google's flagship reasoning, 1M context, thinking
Checking...
·
Gemini 3 Pro Image
gemini-3-pro-image
GA pro-grade image generation
Checking...
·
Gemini 3 Flash Preview
gemini-3-flash-preview
legacy compatibility; migrate to Gemini 3.6 Flash
Checking...
·
Gemini 3.1 Flash Lite
gemini-3.1-flash-lite
legacy compatibility; migrate to Gemini 3.5 Flash-Lite
Checking...
·
FLUX.2 Pro
flux.2-pro
flagship image generation, billed per image
Checking...
·

Roadmap - how we become a real inference company

1

Now - launched on a curated catalog

Live

Customers save 20% today on a curated catalog at low list prices. The narrow operational surface is what keeps the gap honest, and the same surface scales straight into Phase 2.

2

Q2 2026 - our own inference stack on H100/H200

Planned

Self-hosted serving on dedicated GPUs using SGLang + continuous batching, EAGLE-3 speculative decoding, FP8 quantization via DeepGEMM, and SageAttention / ThunderMLA custom kernels. At that point system_fingerprint becomes stable (it changes only when we rev the stack), and repeatable-seed workflows start working properly. Target: 30-50% below current prices on the DeepSeek V4 wave.

3

H2 2026 - colocated data center + AIDC partnerships

Future

Move from rented (Vast.ai) to self-owned or colocated racks. Partner with AI-datacenter operators where that makes sense. The goal is the lowest-cost reliable inference for open-source models on the planet - full stack, our engineering.

About this page

Service rows run client-side probes from your browser. Model rows reflect a real 1-token probe sent server-side every 3 minutes from our backend. Historical bars show the results of recent probes stored in this browser's localStorage; cleared if you switch devices.

Public uptime tracking began 2026-04-16. For a contractual SLA and third-party-monitored history, contact us.