What shipped, and when
Models added, prices changed, surfaces shipped. Only things you can see from outside — a model that appeared in the catalog, a rate that moved, a page that started existing. Current prices are always on the pricing table, and machine-readable at pricing.json.
- Models
Nemotron 3.5 Lightning
NVIDIA's open 30B-A3B efficiency model joins the catalog for fast, tool-heavy agents at $0.08 input / $0.20 output per 1M tokens, with cached input at $0.04 and a stable 262K context window.
- Product
Browser chat, no API key required
Every account now gets a free trial balance to spend at /chat — no card and no key. Four models are available on it, and each reply shows the exact cost of that turn at the same per-token rate the API charges.
- Models
Qwen3.8 Max
Alibaba's Qwen3.8 Max joins the catalog at $2.00 input / $6.00 output per 1M tokens, with implicit caching at $0.25, passed through at Alibaba's own published price.
- Pricing
Qwen3.7 Max repriced to Alibaba's own list
Now $1.25 input / $3.75 output per 1M tokens, down from $1.475 / $4.425. Alibaba is the only provider of this model, so we sell it at their published price rather than marking it up.
- Models
DeepSeek V4 Flash 0731, refreshed Gemini, audited open catalog
V4 Flash 0731 brings a 1M context window and native Responses API support. The Gemini line was refreshed and the open-weight catalog re-audited against each publisher's current list; retired models were removed rather than left to rot.
- Models
Claude Opus 5, and Google sign-in
Claude Opus 5 is available through the same endpoint and the same key as the open-weight catalog. Signing in with Google now works alongside email and password.
- Models
Kimi K3
Moonshot's K3 lands on release day at $2.40 input / $12.00 output per 1M tokens, 20% below Moonshot's own published rate.
- Models
MiniMax M3, MiMo-V2.5 and Hy3
Three additions to the open-weight side of the catalog, each with its reference list price published alongside ours.