Ask
Packaging

Groq removes its public pricing page — rates survive only in developer docs

Groq pricing

groq.com/pricing/ now 308-redirects to Groq's homepage, dropping the dedicated marketing pricing page entirely; per-token, per-hour, and per-character rates are unchanged but now live only inside the GroqCloud developer docs model catalog.

Before

Dedicated marketing page at groq.com/pricing/ published the full rate card (per-token LLM pricing, Whisper transcription, Orpheus TTS, built-in tool pricing) alongside 'linear and predictable' pricing messaging.

After

groq.com/pricing/ 308-redirects to the Groq homepage, which carries no pricing content and no 'Pricing' nav link. Per-token/hour/character rates for LLMs, Whisper, and Orpheus TTS are unchanged and still published in the GroqCloud docs model catalog (console.groq.com/docs/models), but built-in agentic tool pricing (web search, website visits, code execution) has no live public source anywhere, since the docs for those tools defer to the now-dead pricing page instead of printing a rate.

Groq’s homepage now leads with an “ANNOUNCING OUR $650 MILLION FUNDRAISE” banner and repositions the company as the “Premier NeoCloud for fast inference,” alongside a new LPX platform that pairs Groq’s LPU with NVIDIA’s next-generation GPUs. No published per-token, per-hour, or per-character rate moved — every figure on the surviving developer-docs rate card (Llama 3.1 8B $0.05/$0.08, Llama 3.3 70B $0.59/$0.79, GPT OSS 120B $0.15/$0.60, GPT OSS 20B $0.075/$0.30, Whisper $0.111/$0.04 per hour, Orpheus $22/$40 per 1M characters) is identical to the last capture on 2026-07-21. The practical impact is a transparency gap: three built-in agentic tools (web search, website visits, code execution) previously had a public per-use rate on the marketing pricing page, and the docs pages for those tools explicitly point back to that page rather than stating a number — so as of this capture their pricing cannot be verified from any live Groq source.

From Groq's pricing timeline
Public Pricing Page Removed — Rates Survive Only in Developer Docs

groq.com/pricing/ now 308-redirects to Groq's homepage, which carries no pricing content, no Pricing nav link, and instead promotes a $650M raise and a "Premier NeoCloud for fast inference" repositioning. Per-token, per-hour, and per-character rates did not move — the entire published rate card (Llama 3.1 8B $0.05/$0.08, Llama 3.3 70B $0.59/$0.79, GPT OSS 120B $0.15/$0.60, GPT OSS 20B/Safeguard $0.075/$0.30, Qwen 3.6 27B $0.60/$3.00, Whisper $0.111/$0.04 per hour, Orpheus $22/$40 per 1M characters) is identical to 2026-07-21 — but it now lives exclusively in the GroqCloud developer-docs model catalog (console.groq.com/docs/models), not on any public marketing page.

About Groq
groq.com ↗

Groq runs a pure-usage per-token serverless inference API on its proprietary LPU silicon. As of 2026-08-11 the dedicated marketing pricing page (groq.com/pricing/) has been removed and redirects to the homepage; rates survive only in the GroqCloud developer docs. As of 2026-08-26, Llama 3.1 8B Instant and Llama 3.3 70B Versatile — previously $0.05/$0.08 and $0.59/$0.79 per 1M tokens — moved to Enterprise-only "Contact Sales" pricing and dropped off both the Free and Developer rate-limit tables. The self-serve catalog is now GPT OSS 20B and Safety GPT OSS 20B (formerly "GPT OSS Safeguard 20B") at $0.075/$0.30 (1,000 T/SEC), GPT OSS 120B at $0.15/$0.60 (500 T/SEC), Qwen 3.6 27B at $0.60/$3.00, and two Preview moderation models — Llama Prompt Guard 2 22M and Prompt Guard 2 86M — at $0.03/$0.03 and $0.04/$0.04 per 1M tokens. As of 2026-08-27, Qwen 3.8-27B joined the Preview catalog with a published price of $0.80/$4.00 per 1M tokens (450 T/SEC) — the only change versus the prior day's capture.

Free tier
Yes
Commits
Available
Transparency
public

Groq pricing history

  1. Aug 2026
    Qwen 3.8-27B Gains a Published Preview-Tier Price
  2. Aug 2026
    Llama 3.1 8B Instant and Llama 3.3 70B Versatile Moved to Enterprise-Only Pricing
  3. Aug 2026
    Public Pricing Page Removed — Rates Survive Only in Developer Docs
  4. Jul 2026
    Catalog Contraction: Two Models and the Browser Automation Tool Withdrawn
  5. Jul 2026
    Text-to-Speech SKU + Browser Automation Tool + New Models
Full Groq timeline

More Groq activity

All pricing activity