Skip to content

Coming to CheapestInference — models under review

New models don’t land in a pool by surprise. Every candidate goes through the same review — Reviewing → Confirmed → Capacity secured → Live — and this page is its live status board. Check back or watch the changelog for the moment a model goes live.

ModelStatePool candidateNotesUpdated
GLM-5.3-FlashReviewingTBDZ.ai’s 320B/18B-active MoE — 57 on the AA index at $0.10/M blended, MIT weights, native multimodal, 1M context. Quality and fit under evaluation. Analysis2026-08-31

Models that completed this pipeline and went live:

  • 2026-09-10DeepSeek V4.1 Flash — upgraded the Core Pool’s DeepSeek slot in place, the day DeepSeek published the open weights
  • 2026-08-30GLM 5.3 — upgraded the Frontier Pool’s GLM slot in place, two days after Z.ai published the open weights
  • 2026-08-14Qwen3.8 Max — joined the Flagship Pool next to Kimi K3, the day after its open-weight variant shipped and lifted the licensing gate
  • 2026-07-31DeepSeek V4 Flash (0731 build) — upgraded in the Core Pool
  • 2026-07-28Kimi K3 — launched the Flagship Pool
  • Full history in the changelog

Tell us what you’d pay a flat monthly fee for: support@cheapestinference.com. Real demand moves models up this list.