71 models in the catalog

Explore AI models

Compare prices, capabilities and context windows across the QuickSilver Pro catalog.

All models71 models found. Sorted by Featured

Model / author
Official
Our price
Discount
Context
Latency
gpt-image-2OpenAI
New
—
per image
Output$0.01
—
—
—
claude-haiku-5-5Anthropic
New
per 1M tokens
Input$0.05
Output$0.25
Cache read$0.005
Long-context rates apply
−50%
1M
—
mistral-large-4Mistral AI
New
at list
per 1M tokens
Input$0.68
Output$2.09
Cache read$0.07
at list
524K
—
gpt-6.1-solOpenAI
New
at list
per 1M tokens
Input$2.00
Output$10.00
Cache read$0.10
Long-context rates apply
at list
1M
—
claude-sonnet-5-5Anthropic
New
per 1M tokens
Input$1.00
Output$5.00
Cache read$0.10
−50%
1M
—
qwen3.8-max-primeAlibaba / Qwen
New
at list
per 1M tokens
Input$4.00
Output$12.00
Cache read$0.50
at list
1M
—
per 1M tokens
Input$2.24
Output$7.04
Cache read$0.448
−20%
1M
—
gpt-6-lunaOpenAI
New
at list
per 1M tokens
Input$0.10
Output$0.50
Cache read$0.01
Long-context rates apply
at list
1M
—
gpt-6-solOpenAI
New
at list
per 1M tokens
Input$2.00
Output$10.00
Cache read$0.20
Long-context rates apply
at list
1M
—
claude-opus-5-5Anthropic
New
per 1M tokens
Input$1.60
Output$8.00
Cache read$0.08
−60%
1M
—
grok-4.7xAI
New
—
per 1M tokens
Input$2.00
Output$6.00
Cache read$0.50
Long-context rates apply
—
500K
—
mimo-v2.6-proXiaomi
New
per 1M tokens
Input$0.348
Output$0.696
Cache read$0.00288
−20%
1M
—
per 1M tokens
Input$0.112
Output$0.224
Cache read$0.00224
−20%
1M
—
qwen3.8-omni-flashAlibaba / Qwen
New
at list
per 1M tokens
Input$0.15
Output$0.47
Cache read$0.016
at list
1M
—
qwen-image-maxAlibaba / Qwen
New
—
per image
Output$0.10
—
—
—
seedream-5.0-proByteDance
New
—
per image
Output$0.07
—
—
—
seedream-4ByteDance
New
—
per image
Output$0.055
—
—
—
—
per image
Output$0.055
—
—
—
jev-1.13TypeSafe
New
at list
per 1M tokens
Input$0.042
OutputFree
Cache read—
at list
32K
—
per 1M tokens
Input$0.125
Output$0.55
Cache read$0.0125
−54–58%
1M
—
at list
per 1M tokens
Input$10.00
Output$50.00
Cache read$1.00
Long-context rates apply
at list
1M
—
per 1M tokens
Input$1.00
Output$3.40
Cache read$0.12
−20%
1M
—
per 1M tokens
Input$0.6375
Output$3.1875
Cache read$0.06375
−15%
1M
—
per 1M tokens
Input$4.00
Output$20.00
Cache read$0.10
−60%
1M
—
per 1M tokens
Input$0.6672
Output$2.0008
Cache read$0.0336
−20%
1M
—
qwen3.8-flash-nextAlibaba / Qwen
per 1M tokens
Input$0.12
Output$0.376
Cache read$0.0128
−20%
1M
—
per 1M tokens
Input$0.06
Output$0.20
Cache read$0.012
−20%
1M
—
per 1M tokens
Input$1.12
Output$3.52
Cache read$0.208
−20%
1M
—
qwen3.8-27bAlibaba / Qwen
per 1M tokens
Input$0.34
Output$2.04
Cache read$0.068
−20%
1M
—
per 1M tokens
Input$0.6375
Output$3.1875
Cache read$0.06375
−15%
1M
—
—
per 1M tokens
Input$2.00
Output$6.00
Cache read$0.50
Long-context rates apply
—
500K
—
—
per 1M tokens
Input$0.066
Output$0.176
Cache read$0.033
—
262K
—
per 1M tokens
Input$0.28
Output$1.20
Cache read$0.032
−20%
131K
—
per 1M tokens
Input$1.00
Output$5.00
Cache read—
−67%
1M
—
per 1M tokens
Input$1.00
Output$3.40
Cache read$0.12
−20%
1M
—
qwen3.8-maxAlibaba / Qwen
at list
per 1M tokens
Input$2.00
Output$6.00
Cache read$0.25
at list
1M
—
—
per 1M tokens
Input$0.086
Output$0.173
Cache read$0.0086
—
1M
—
at list
per 1M tokens
Input$0.50
Output$2.20
Cache read$0.10
at list
262K
—
per 1M tokens
Input$0.40
Output$2.00
Cache read—
−60%
200K
—
per 1M tokens
Input$0.041
Output$0.187
Cache read$0.041
−69–73%
131K
—
per 1M tokens
Input$0.70
Output$2.10
Cache read$0.10
−47%
1M
—
qwen3.7-maxAlibaba / Qwen
per 1M tokens
Input$1.25
Output$3.75
Cache read$0.25
−15%
1M
—
qwen3.7-plusAlibaba / Qwen
per 1M tokens
Input$0.256
Output$1.024
Cache read$0.0512
−20%
1M
—
qwen3.7-flashAlibaba / Qwen
per 1M tokens
Input$0.024
Output$0.104
Cache read$0.0048
−20%
1M
—
qwen3.6-plusAlibaba / Qwen
per 1M tokens
Input$0.26
Output$1.56
Cache read—
−20%
1M
—
qwen3.6-35bAlibaba / Qwen
per 1M tokens
Input$0.112
Output$0.80
Cache read$0.04
−20%
262K
—
kimi-k2.6Moonshot AI
—
per 1M tokens
Input$0.5472
Output$2.728
Cache read$0.292
mixed
256K
—
kimi-k2.7-codeMoonshot AI
per 1M tokens
Input$0.584
Output$2.80
Cache read$0.1278
−20%
256K
—
kimi-k3Moonshot AI
per 1M tokens
Input$2.55
Output$12.75
Cache read$0.255
−15%
1M
—
per 1M tokens
Input$1.12
Output$3.52
Cache read$0.208
−20%
1M
—
per 1M tokens
Input$0.16
Output$0.96
Cache read$0.016
−20%
1M
—
per 1M tokens
Input$1.60
Output$9.60
Cache read$0.16
−20%
1M
—
per 1M tokens
Input$1.60
Output$8.00
Cache read$0.16
−20%
1M
—
per 1M tokens
Input$1.60
Output$4.80
Cache read$0.40
−20%
500K
—
minimax-m3MiniMax
per 1M tokens
Input$0.24
Output$0.96
Cache read$0.048
−20%
1M
—
mimo-v2.5Xiaomi
at list
per 1M tokens
Input$0.112
Output$0.224
Cache read$0.00224
at list
1M
—
hy3Tencent
—
per 1M tokens
Input$0.091
Output$0.363
Cache read$0.0227
—
262K
—
claude-opus-5Anthropic
per 1M tokens
Input$2.00
Output$10.00
Cache read—
−60%
1M
—
per 1M tokens
Input$4.00
Output$20.00
Cache read—
−60%
1M
—
per 1M tokens
Input$2.00
Output$10.00
Cache read—
−60%
1M
—
per 1M tokens
Input$0.6375
Output$3.1875
Cache read—
−15%
1M
—
per 1M tokens
Input$0.255
Output$2.125
Cache read—
−15%
1M
—
per 1M tokens
Input$1.275
Output$7.65
Cache read—
−15%
1M
—
per 1M tokens
Input$1.70
Output$10.20
Cache read—
−15%
1M
—
per 1M tokens
Input$1.70
Output$10.20
Cache read—
Image · per image$0.11424
−15%
1M
—
—
per 1M tokens
Input$0.425
Output$2.55
Cache read—
—
1M
—
—
per 1M tokens
Input$0.2125
Output$1.275
Cache read—
—
1M
—
flux.2-proBlack Forest Labs
per image
Output$0.027
−13%
—
—
flux.1-schnellBlack Forest Labs
at market
per image
Output$0.003
at market
—
—
sdxl-turboStability AI
at market
per image
Output$0.003
at market
—
—
flux.2-kleinBlack Forest Labs
at market
per image
Output$0.02
at market
—
—

USD per 1M tokens, except image models billed per image. Official prices appear only where a published reference exists. Models marked for long-context rates charge different rates above their prompt-token threshold.

Latency comes from standardized status probes with a 16-token output limit.