QuickSilver Pro system status

Status unknown
Checking...

Services

Browser service checks · reported API model checks
Listed API model checks
pay.quicksilverpro.io/v1/status
Checking...
— · —
Dashboard backend
pay.quicksilverpro.io
Checking...
— · —
Website
quicksilverpro.io
Checking...
— · —

Model availability

Synthetic probe · updates every 3 min
DeepSeek V4 Flash
deepseek-v4-flash
official 0731 agent model, 1M context, Responses API
Checking...
— · —
DeepSeek V4.1 Flash
deepseek-v4.1-flash
official DeepSeek V4.1 Flash — new architecture, faster, 1M context, Responses API
Checking...
— · —
DeepSeek V4 Pro
deepseek-v4-pro
premium reasoning, 1M context
Checking...
— · —
Qwen3.8 Max
qwen3.8-max
Qwen 3.8 flagship, 2.4T MoE, 1M context, autonomous coding
Checking...
— · —
Qwen3.8 Max Prime
qwen3.8-max-prime
Qwen 3.8 top-tier flagship, deepest reasoning & autonomous coding, 1M context
Checking...
— · —
Qwen3.8 Omni Flash
qwen3.8-omni-flash
Qwen's first agentic omni model: native image, audio and video understanding with tool use and reasoning, 1M context, at Alibaba's first-party price
Checking...
— · —
Qwen3.7 Max
qwen3.7-max
Qwen 3.7 flagship, 1M context, thinks by default
Checking...
— · —
Qwen3.7 Plus
qwen3.7-plus
Qwen 3.7 agent flagship, 1M context, long-horizon coding
Checking...
— · —
Qwen3.7 Flash
qwen3.7-flash
fast multimodal agents, vision and tool use, 1M context
Checking...
— · —
Qwen3.6 Plus
qwen3.6-plus
1T-MoE flagship, 1M context, thinks by default
Checking...
— · —
Qwen3.6-35B-A3B
qwen3.6-35b
262K long-context, MoE, drop-in 3.5 upgrade
Checking...
— · —
Qwen3.8 27B
qwen3.8-27b
Qwen's fast open 27B reasoning model with a 1M-token context window
Checking...
— · —
Qwen3.8 Flash Next
qwen3.8-flash-next
Qwen's open Qwen4-architecture preview: 6B-active MoE, fast and cheap, with a 1M-token context window
Checking...
— · —
Kimi K2.6
kimi-k2.6
Opus-class reasoning, 256K context
Checking...
— · —
Kimi K2.7 Code
kimi-k2.7-code
Agentic-coding K2, 256K context, coding-tuned K2.6 successor
Checking...
— · —
Kimi K3
kimi-k3
2.8T multimodal reasoning flagship, 1M context
Checking...
— · —
Muse Spark 1.3
muse-spark-1.3
Coding & agentic reasoning model, 1M context
Checking...
— · —
Muse Spark 1.2
muse-spark-1.2
Coding-focused reasoning model, 1M context
Checking...
— · —
Muse Glimmer 30B
muse-glimmer-30b
Compact agentic multimodal model, distilled from Muse Spark
Checking...
— · —
GLM 5.3
glm-5.3
Z.ai latest reasoning flagship, 1M context, complex software engineering & long-horizon agents
Checking...
— · —
GLM 5.3 Prime
glm-5.3-prime
Z.ai GLM 5.3 Prime, top reasoning flagship, 1M context, complex software engineering
Checking...
— · —
GLM 5.3 Flash
glm-5.3-flash
Z.ai fast open-weights reasoning model, 1M context at workhorse pricing
Checking...
— · —
GLM 5.2
glm-5.2
Z.ai reasoning flagship, 1M context, long-horizon agents & coding
Checking...
— · —
Nemotron 3 Ultra
nemotron-3-ultra
550B-A55B open MoE for frontier reasoning and agent orchestration
Checking...
— · —
Nemotron 3.5 Lightning
nemotron-3.5-lightning
30B-A3B open MoE for fast, tool-heavy agents
Checking...
— · —
GPT-OSS 120B
gpt-oss-120b
OpenAI's open-weight 117B MoE, high-volume agents & reasoning, 131K context
Checking...
— · —
GPT-6 Astra
gpt-6-astra
OpenAI's GPT-6 flagship, long-horizon agentic coding & deep research, 1M context
Checking...
— · —
GPT-6 Luna
gpt-6-luna
OpenAI's cost-efficient GPT-6, high-volume chat & lightweight agentic work, 1M context
Checking...
— · —
GPT-6.1 Sol
gpt-6.1-sol
OpenAI's GPT-6.1 Sol, near-Astra agentic coding at Sol pricing, 1M context
Checking...
— · —
GPT-6 Sol
gpt-6-sol
OpenAI's GPT-6 mid-flagship, strong reasoning & agentic coding, 1M context
Checking...
— · —
GPT-5.6 Luna
gpt-5.6-luna
OpenAI's fast, cost-efficient GPT-5.6 tier, 1M context
Checking...
— · —
GPT-5.6 Terra
gpt-5.6-terra
OpenAI's balanced GPT-5.6 mid-tier, 1M context
Checking...
— · —
GPT-5.6 Sol
gpt-5.6-sol
OpenAI's flagship GPT-5.6, complex reasoning & agentic coding, 1M context
Checking...
— · —
Grok 4.5
grok-4.5
xAI's smartest model, frontier coding, knowledge work & STEM, 500K context
Checking...
— · —
Grok 4.6
grok-4.6
frontier coding, knowledge work and STEM, vision, 500K context
Checking...
— · —
Grok 4.7
grok-4.7
frontier coding, knowledge work and STEM, vision, 500K context
Checking...
— · —
MiniMax M3
minimax-m3
MiniMax open-weight frontier model, 1M context, agentic coding
Checking...
— · —
MiMo-V2.6-Pro
mimo-v2.6-pro
Xiaomi flagship open-weight MiMo-V2.6-Pro, top open model on AA Intelligence Index, coding & agentic, vision, 1M context
Checking...
— · —
MiMo-V2.6-Flash
mimo-v2.6-flash
Xiaomi open-weight MiMo-V2.6-Flash, fast low-cost coding & agentic, vision, 1M context
Checking...
— · —
MiMo-V2.5
mimo-v2.5
Xiaomi open-weight MiMo-V2.5, cost-efficient coding & agentic, 1M context
Checking...
— · —
Hy4 Preview
hy4-preview
Tencent Hy4 (preview), reasoning & agentic coding, 1M context
Checking...
— · —
Hy3
hy3
Tencent Hy3, selectable reasoning & agentic coding, 262K context
Checking...
— · —
Mistral Large 4
mistral-large-4
Mistral's frontier multimodal model for coding and agents, 524K context
Checking...
— · —
Jev 1.13
jev-1.13
Typed decisions with calibrated probabilities from a single forward pass: routing, classification and scoring without text generation
Checking...
— · —
Claude Opus 5.5
claude-opus-5-5
Anthropic's most capable Claude: agentic coding, knowledge work & computer use, 1M context
Checking...
— · —
Claude Opus 5
claude-opus-5
Anthropic Opus 5 for demanding reasoning, coding & long-horizon agentic work, 1M context
Checking...
— · —
Claude Fable 5.1
claude-fable-5-1
Anthropic Mythos-class model for deep reasoning and long-horizon agentic work (5.1)
Checking...
— · —
Claude Fable 5
claude-fable-5
Anthropic Mythos-class flagship, most capable reasoning
Checking...
— · —
Claude Opus 4.8
claude-opus-4-8
Anthropic flagship, top-tier reasoning & agentic
Checking...
— · —
Claude Sonnet 5.5
claude-sonnet-5-5
Anthropic's newest Sonnet, faster and fewer tokens per task, 1M context
Checking...
— · —
Claude Sonnet 5
claude-sonnet-5
Anthropic's newest Sonnet tier, 1M context
Checking...
— · —
Claude Haiku 5.5
claude-haiku-5-5
Anthropic's newest small, fast model for subagents and high-volume work, 1M context
Checking...
— · —
Claude Haiku 4.5
claude-haiku-4-5
fast, low-cost Anthropic, high-volume
Checking...
— · —
Gemini 3.7 Flash
gemini-3.7-flash
fast multimodal Flash for agents; promotional pricing through 2026-12-31
Checking...
— · —
Gemini 3.8 Flash
gemini-3.8-flash
most capable Flash for agents; promotional pricing through 2026-12-31
Checking...
— · —
Gemini 3.6 Flash
gemini-3.6-flash
current general-purpose Flash GA, 1M context
Checking...
— · —
Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
current low-cost Flash-Lite GA, 1M context
Checking...
— · —
Gemini 3.5 Flash
gemini-3.5-flash
Google's next-gen Flash GA, 1M context, thinking
Checking...
— · —
Gemini 3.1 Pro Preview
gemini-3.1-pro-preview
Google's flagship reasoning, 1M context, thinking
Checking...
— · —
Gemini 3 Pro Image
gemini-3-pro-image
GA pro-grade image generation
Checking...
— · —
Gemini 3 Flash Preview
gemini-3-flash-preview
legacy compatibility; migrate to Gemini 3.6 Flash
Checking...
— · —
Gemini 3.1 Flash Lite
gemini-3.1-flash-lite
legacy compatibility; migrate to Gemini 3.5 Flash-Lite
Checking...
— · —
FLUX.2 Pro
flux.2-pro
flagship image generation, billed per image
Checking...
— · —
FLUX.1 Schnell
flux.1-schnell
ultra-fast open image generation, billed per image
Checking...
— · —
SDXL Turbo
sdxl-turbo
fast open image generation (SDXL Turbo), billed per image
Checking...
— · —
FLUX.2 Klein
flux.2-klein
open FLUX.2 image generation, billed per image
Checking...
— · —
Qwen-Image Max
qwen-image-max
Alibaba's Qwen-Image flagship: high-fidelity generation with strong typography, billed per image
Checking...
— · —
Seedream 5.0 Pro
seedream-5.0-pro
ByteDance's Seedream 5.0 Pro: flagship photorealistic image generation, billed per image
Checking...
— · —
Seedream 4
seedream-4
ByteDance Seedream 4: fast, high-quality image generation, billed per image
Checking...
— · —
Bria FIBO 1.5
bria-fibo-1.5
Bria FIBO 1.5: image generation trained on fully licensed data (commercially safe), billed per image
Checking...
— · —
GPT Image 2
gpt-image-2
OpenAI's GPT Image 2 at a fixed ~1.5MP standard output, billed per image
Checking...
— · —

Roadmap - how we become a real inference company

1

Now - launched on a curated catalog

Live

Customers save 20% today on a curated catalog at low list prices. The narrow operational surface is what keeps the gap honest, and the same surface scales straight into Phase 2.

2

Q2 2026 - our own inference stack on H100/H200

Planned

Self-hosted serving on dedicated GPUs using SGLang + continuous batching, EAGLE-3 speculative decoding, FP8 quantization via DeepGEMM, and SageAttention / ThunderMLA custom kernels. At that point system_fingerprint becomes stable (it changes only when we rev the stack), and repeatable-seed workflows start working properly. Target: 30-50% below current prices on the DeepSeek V4 wave.

3

H2 2026 - colocated data center + AIDC partnerships

Future

Move from rented (Vast.ai) to self-owned or colocated racks. Partner with AI-datacenter operators where that makes sense. The goal is the lowest-cost reliable inference for open-source models on the planet - full stack, our engineering.

About this page

Backend and website rows run reachability checks from your browser. The API row summarizes the listed model checks reported by our backend. Model rows reflect a real 1-token probe sent server-side every 3 minutes from our backend. Historical bars show the results of recent probes stored in this browser's localStorage; cleared if you switch devices.

Public uptime tracking began 2026-04-16. For a contractual SLA and third-party-monitored history, contact us.