Coming to CheapestInference — models under review
New models don’t land in a pool by surprise. Every candidate goes through the same review — Reviewing → Confirmed → Capacity secured → Live — and this page is its live status board. Check back or watch the changelog for the moment a model goes live.
Current pipeline
Section titled “Current pipeline”| Model | State | Pool candidate | Notes | Updated |
|---|---|---|---|---|
| GLM-5.3-Flash | Reviewing | TBD | Z.ai’s 320B/18B-active MoE — 57 on the AA index at $0.10/M blended, MIT weights, native multimodal, 1M context. Quality and fit under evaluation. Analysis | 2026-08-31 |
Recently shipped
Section titled “Recently shipped”Models that completed this pipeline and went live:
- 2026-09-10 — DeepSeek V4.1 Flash — upgraded the Core Pool’s DeepSeek slot in place, the day DeepSeek published the open weights
- 2026-08-30 — GLM 5.3 — upgraded the Frontier Pool’s GLM slot in place, two days after Z.ai published the open weights
- 2026-08-14 — Qwen3.8 Max — joined the Flagship Pool next to Kimi K3, the day after its open-weight variant shipped and lifted the licensing gate
- 2026-07-31 — DeepSeek V4 Flash (0731 build) — upgraded in the Core Pool
- 2026-07-28 — Kimi K3 — launched the Flagship Pool
- Full history in the changelog
Want a model reviewed?
Section titled “Want a model reviewed?”Tell us what you’d pay a flat monthly fee for: support@cheapestinference.com. Real demand moves models up this list.