QuickSilver Pro 系统状态

状态未知
检查中...

服务

浏览器服务检测 · API 模型检测报告
所列 API 模型检测
pay.quicksilverpro.io/v1/status
检查中...
— · —
仪表盘后端
pay.quicksilverpro.io
检查中...
— · —
官网
quicksilverpro.io
检查中...
— · —

模型可用性

合成探测 · 每 3 分钟更新
DeepSeek V4 Flash
deepseek-v4-flash
official 0731 agent model, 1M context, Responses API
检查中...
— · —
DeepSeek V4.1 Flash
deepseek-v4.1-flash
official DeepSeek V4.1 Flash — new architecture, faster, 1M context, Responses API
检查中...
— · —
DeepSeek V4 Pro
deepseek-v4-pro
premium reasoning, 1M context
检查中...
— · —
Qwen3.8 Max
qwen3.8-max
Qwen 3.8 flagship, 2.4T MoE, 1M context, autonomous coding
检查中...
— · —
Qwen3.8 Max Prime
qwen3.8-max-prime
Qwen 3.8 top-tier flagship, deepest reasoning & autonomous coding, 1M context
检查中...
— · —
Qwen3.8 Omni Flash
qwen3.8-omni-flash
Qwen's first agentic omni model: native image, audio and video understanding with tool use and reasoning, 1M context, at Alibaba's first-party price
检查中...
— · —
Qwen3.7 Max
qwen3.7-max
Qwen 3.7 flagship, 1M context, thinks by default
检查中...
— · —
Qwen3.7 Plus
qwen3.7-plus
Qwen 3.7 agent flagship, 1M context, long-horizon coding
检查中...
— · —
Qwen3.7 Flash
qwen3.7-flash
fast multimodal agents, vision and tool use, 1M context
检查中...
— · —
Qwen3.6 Plus
qwen3.6-plus
1T-MoE flagship, 1M context, thinks by default
检查中...
— · —
Qwen3.6-35B-A3B
qwen3.6-35b
262K long-context, MoE, drop-in 3.5 upgrade
检查中...
— · —
Qwen3.8 27B
qwen3.8-27b
Qwen's fast open 27B reasoning model with a 1M-token context window
检查中...
— · —
Qwen3.8 Flash Next
qwen3.8-flash-next
Qwen's open Qwen4-architecture preview: 6B-active MoE, fast and cheap, with a 1M-token context window
检查中...
— · —
Kimi K2.6
kimi-k2.6
Opus-class reasoning, 256K context
检查中...
— · —
Kimi K2.7 Code
kimi-k2.7-code
Agentic-coding K2, 256K context, coding-tuned K2.6 successor
检查中...
— · —
Kimi K3
kimi-k3
2.8T multimodal reasoning flagship, 1M context
检查中...
— · —
Muse Spark 1.3
muse-spark-1.3
Coding & agentic reasoning model, 1M context
检查中...
— · —
Muse Spark 1.2
muse-spark-1.2
Coding-focused reasoning model, 1M context
检查中...
— · —
Muse Glimmer 30B
muse-glimmer-30b
Compact agentic multimodal model, distilled from Muse Spark
检查中...
— · —
GLM 5.3
glm-5.3
Z.ai latest reasoning flagship, 1M context, complex software engineering & long-horizon agents
检查中...
— · —
GLM 5.3 Prime
glm-5.3-prime
Z.ai GLM 5.3 Prime, top reasoning flagship, 1M context, complex software engineering
检查中...
— · —
GLM 5.3 Flash
glm-5.3-flash
Z.ai fast open-weights reasoning model, 1M context at workhorse pricing
检查中...
— · —
GLM 5.2
glm-5.2
Z.ai reasoning flagship, 1M context, long-horizon agents & coding
检查中...
— · —
Nemotron 3 Ultra
nemotron-3-ultra
550B-A55B open MoE for frontier reasoning and agent orchestration
检查中...
— · —
Nemotron 3.5 Lightning
nemotron-3.5-lightning
30B-A3B open MoE for fast, tool-heavy agents
检查中...
— · —
GPT-OSS 120B
gpt-oss-120b
OpenAI's open-weight 117B MoE, high-volume agents & reasoning, 131K context
检查中...
— · —
GPT-6 Astra
gpt-6-astra
OpenAI's GPT-6 flagship, long-horizon agentic coding & deep research, 1M context
检查中...
— · —
GPT-6 Luna
gpt-6-luna
OpenAI's cost-efficient GPT-6, high-volume chat & lightweight agentic work, 1M context
检查中...
— · —
GPT-6.1 Sol
gpt-6.1-sol
OpenAI's GPT-6.1 Sol, near-Astra agentic coding at Sol pricing, 1M context
检查中...
— · —
GPT-6 Sol
gpt-6-sol
OpenAI's GPT-6 mid-flagship, strong reasoning & agentic coding, 1M context
检查中...
— · —
GPT-5.6 Luna
gpt-5.6-luna
OpenAI's fast, cost-efficient GPT-5.6 tier, 1M context
检查中...
— · —
GPT-5.6 Terra
gpt-5.6-terra
OpenAI's balanced GPT-5.6 mid-tier, 1M context
检查中...
— · —
GPT-5.6 Sol
gpt-5.6-sol
OpenAI's flagship GPT-5.6, complex reasoning & agentic coding, 1M context
检查中...
— · —
Grok 4.5
grok-4.5
xAI's smartest model, frontier coding, knowledge work & STEM, 500K context
检查中...
— · —
Grok 4.6
grok-4.6
frontier coding, knowledge work and STEM, vision, 500K context
检查中...
— · —
Grok 4.7
grok-4.7
frontier coding, knowledge work and STEM, vision, 500K context
检查中...
— · —
MiniMax M3
minimax-m3
MiniMax open-weight frontier model, 1M context, agentic coding
检查中...
— · —
MiMo-V2.6-Pro
mimo-v2.6-pro
Xiaomi flagship open-weight MiMo-V2.6-Pro, top open model on AA Intelligence Index, coding & agentic, vision, 1M context
检查中...
— · —
MiMo-V2.6-Flash
mimo-v2.6-flash
Xiaomi open-weight MiMo-V2.6-Flash, fast low-cost coding & agentic, vision, 1M context
检查中...
— · —
MiMo-V2.5
mimo-v2.5
Xiaomi open-weight MiMo-V2.5, cost-efficient coding & agentic, 1M context
检查中...
— · —
Hy4 Preview
hy4-preview
Tencent Hy4 (preview), reasoning & agentic coding, 1M context
检查中...
— · —
Hy3
hy3
Tencent Hy3, selectable reasoning & agentic coding, 262K context
检查中...
— · —
Mistral Large 4
mistral-large-4
Mistral's frontier multimodal model for coding and agents, 524K context
检查中...
— · —
Jev 1.13
jev-1.13
Typed decisions with calibrated probabilities from a single forward pass: routing, classification and scoring without text generation
检查中...
— · —
Claude Opus 5.5
claude-opus-5-5
Anthropic's most capable Claude: agentic coding, knowledge work & computer use, 1M context
检查中...
— · —
Claude Opus 5
claude-opus-5
Anthropic Opus 5 for demanding reasoning, coding & long-horizon agentic work, 1M context
检查中...
— · —
Claude Fable 5.1
claude-fable-5-1
Anthropic Mythos-class model for deep reasoning and long-horizon agentic work (5.1)
检查中...
— · —
Claude Fable 5
claude-fable-5
Anthropic Mythos-class flagship, most capable reasoning
检查中...
— · —
Claude Opus 4.8
claude-opus-4-8
Anthropic flagship, top-tier reasoning & agentic
检查中...
— · —
Claude Sonnet 5.5
claude-sonnet-5-5
Anthropic's newest Sonnet, faster and fewer tokens per task, 1M context
检查中...
— · —
Claude Sonnet 5
claude-sonnet-5
Anthropic's newest Sonnet tier, 1M context
检查中...
— · —
Claude Haiku 5.5
claude-haiku-5-5
Anthropic's newest small, fast model for subagents and high-volume work, 1M context
检查中...
— · —
Claude Haiku 4.5
claude-haiku-4-5
fast, low-cost Anthropic, high-volume
检查中...
— · —
Gemini 3.7 Flash
gemini-3.7-flash
fast multimodal Flash for agents; promotional pricing through 2026-12-31
检查中...
— · —
Gemini 3.8 Flash
gemini-3.8-flash
most capable Flash for agents; promotional pricing through 2026-12-31
检查中...
— · —
Gemini 3.6 Flash
gemini-3.6-flash
current general-purpose Flash GA, 1M context
检查中...
— · —
Gemini 3.5 Flash-Lite
gemini-3.5-flash-lite
current low-cost Flash-Lite GA, 1M context
检查中...
— · —
Gemini 3.5 Flash
gemini-3.5-flash
Google's next-gen Flash GA, 1M context, thinking
检查中...
— · —
Gemini 3.1 Pro Preview
gemini-3.1-pro-preview
Google's flagship reasoning, 1M context, thinking
检查中...
— · —
Gemini 3 Pro Image
gemini-3-pro-image
GA pro-grade image generation
检查中...
— · —
Gemini 3 Flash Preview
gemini-3-flash-preview
legacy compatibility; migrate to Gemini 3.6 Flash
检查中...
— · —
Gemini 3.1 Flash Lite
gemini-3.1-flash-lite
legacy compatibility; migrate to Gemini 3.5 Flash-Lite
检查中...
— · —
FLUX.2 Pro
flux.2-pro
flagship image generation, billed per image
检查中...
— · —
FLUX.1 Schnell
flux.1-schnell
ultra-fast open image generation, billed per image
检查中...
— · —
SDXL Turbo
sdxl-turbo
fast open image generation (SDXL Turbo), billed per image
检查中...
— · —
FLUX.2 Klein
flux.2-klein
open FLUX.2 image generation, billed per image
检查中...
— · —
Qwen-Image Max
qwen-image-max
Alibaba's Qwen-Image flagship: high-fidelity generation with strong typography, billed per image
检查中...
— · —
Seedream 5.0 Pro
seedream-5.0-pro
ByteDance's Seedream 5.0 Pro: flagship photorealistic image generation, billed per image
检查中...
— · —
Seedream 4
seedream-4
ByteDance Seedream 4: fast, high-quality image generation, billed per image
检查中...
— · —
Bria FIBO 1.5
bria-fibo-1.5
Bria FIBO 1.5: image generation trained on fully licensed data (commercially safe), billed per image
检查中...
— · —
GPT Image 2
gpt-image-2
OpenAI's GPT Image 2 at a fixed ~1.5MP standard output, billed per image
检查中...
— · —

路线图:我们如何成为真正的推理基础设施公司

1

当前:精选模型目录上线

已上线

客户今天就能省下 20%。精选模型目录让运营面收得很窄,差距来自工程精度,同一套接口直接延续到 Phase 2。

2

2026 年 Q2:在 H100/H200 上运行自有推理栈

计划中

使用专用 GPU 自托管推理服务,结合 SGLang + continuous batching、EAGLE-3 speculative decoding、通过 DeepGEMM 做 FP8 quantization,以及 SageAttention / ThunderMLA 自定义内核。到那时 system_fingerprint 会稳定下来(只有我们升级推理栈时才变化),可重复 seed 工作流也会真正可用。目标:DeepSeek V4 系列的价格比当前再低 30-50%。

3

2026 年 H2:机柜托管数据中心 + AIDC 合作

未来

从租用(Vast.ai)转向自有或托管机柜,并在合适的地方与 AI 数据中心运营商合作。目标是做出地球上最便宜且可靠的开源模型推理服务:从全栈到底层工程都由我们自己掌控。

关于此页面

后端和官网行显示从你的浏览器发出的连通性检测。API 行汇总后端报告的所列模型检测结果。 模型行反映的是我们后端每 3 分钟发起一次、真实 1-token 的服务端探测。 历史条展示的是近期探测结果,数据存储在当前浏览器的 localStorage;如果你切换设备,这些历史会被清空。

公开可用性追踪始于 2026-04-16. 如果你需要带合同约束的 SLA 和第三方监控历史,请 联系我们。