Ask
Funding

Cerebras completes IPO, priced above range and 20x oversubscribed

Cerebras pricing

Cerebras Systems (wafer-scale AI inference cloud, Nvidia challenger) completed its IPO in the week of 2026-08-03, pricing above its initial range and reportedly 20x oversubscribed on debut, before shares fell over the following days. Confirmed via Stocktwits, The Motley Fool, TradingKey, and renaissancecapital.com. Cerebras's usage-based per-token API pricing (Free Trial, $10 Developer, Enterprise) plus the fixed-price Cerebras Code subscriptions were unchanged at capture time.

Cerebras Systems, the wafer-scale AI inference cloud and hardware company positioned as an Nvidia challenger, completed its long-anticipated IPO in the week of 2026-08-03. Coverage (Stocktwits, The Motley Fool, TradingKey, renaissancecapital.com) reports the offering priced above its initial range and was roughly 20x oversubscribed on its first trading day, though shares subsequently pulled back from the IPO-day pop (The Motley Fool notes an over-58% decline from the first-day high within the same week).

No pricing changes accompanied the listing — Cerebras’s per-token inference API (GPT-OSS-120B, Gemma, etc.), $5 free-trial credits, $10 self-serve Developer tier, and the fixed-price Cerebras Code plans ($50/$200) remain as last captured. This is purely a capital-markets/liquidity event.

Detection note: caught via Layer-0 Google-News corporate-event scanning — an IPO does not touch a company’s live pricing page, so Layer-1 diffing was structurally blind to it.

From Cerebras's pricing timeline
ZAI-GLM-4.7 deprecated and removed from the public rate card

Consistent with the deprecation date footnoted on the pricing page a month earlier, ZAI-GLM-4.7 ($2.25 input / $2.75 output per million tokens) was removed from Cerebras's public "Developer Tier Pricing" rate card and the docs Model Catalog. Confirmed via a fresh capture on 2026-08-26, which shows the rate card and Model Catalog both down to two models — GPT-OSS-120B (production) and Google Deepmind Gemma 4 31B (Preview). Z.AI's GLM 4.X and GLM 5.X model families remain reachable only through Dedicated Endpoints on custom reserved-capacity pricing; no per-token price moved on the two remaining public-card models.

About Cerebras
cerebras.ai ↗

Cerebras operates a per-token inference API (Cerebras Inference) powered by its proprietary Wafer Scale Engine (WSE) chips — the only major LLM inference platform that runs entirely on non-Nvidia hardware at inference scale.

Free tier
No
Commits
Available
Transparency
public

Cerebras pricing history

  1. Aug 2026
    Cerebras announces CS-4, a 4th-generation wafer-scale system
  2. Aug 2026
    ZAI-GLM-4.7 deprecated and removed from the public rate card
  3. Jul 2026
    Free tier becomes a $5 credit trial; Gemma 4 31B joins the rate card
  4. May 2026
    Public rate card narrows; Cerebras Code subscriptions launch
  5. Jul 2025
    Qwen-3-32B and ZAI-GLM-4.x Models Added
Full Cerebras timeline

More Cerebras activity

All pricing activity