Funding

Fireworks AI announces a Series D and $1B ARR

Fireworks AI pricing

Fireworks AI's site banner now announces a Series D round alongside a $1B annual recurring revenue milestone, four years after the company was founded in 2022.

Before

Series C reported in 2025 at a $4B+ valuation; no public ARR figure.

After

Series D announced with a stated $1B ARR milestone.

Fireworks AI is running a site-wide banner announcing its Series D round together with a $1B ARR milestone. The round amount and valuation are not stated on the pricing surface itself.

The ARR figure is notable for a pure-usage inference platform: Fireworks charges no subscription at all, so the entire run-rate is metered consumption — per-token serverless inference, per-GPU-second on-demand deployments, per-1M-token fine-tuning and per-1M-token embeddings — drawn against prepaid credits.

From Fireworks AI's pricing timeline
Three serving paths + Fire Pass; Series D and $1B ARR

Serverless inference is now documented as three named serving paths — Standard (default), Priority (higher reliability under peak traffic, set via service_tier) and Fast (100+ tokens/sec, selected by model ID) — replacing the earlier "Turbo + Priority" framing, with per-model input / cached-input / output rates published for each. Fireworks also shipped Fire Pass, an experimental promo-code pass that removes per-token charges on included open-weight models for personal agentic coding, and the site banner now announces a Series D and $1B ARR. Headline rate card (H100/H200 $7.00/hr, B200 $10.00/hr, B300 $12.00/hr, fine-tuning from $0.50 per 1M training tokens, batch at 50%) is unchanged.

About Fireworks AI
fireworks.ai ↗

Fireworks AI runs a pure-usage inference platform: serverless per-token endpoints for popular open-weight models (Llama, DeepSeek, Mixtral, Qwen) plus on-demand dedicated GPU deployments at $7/hr H100/H200, $10/hr B200, $12/hr B300.

Free tier
Yes
Commits
Available
Transparency
public

Fireworks AI pricing history

  1. Jul 2026
    Three serving paths + Fire Pass; Series D and $1B ARR
  2. Feb 2026
    Embeddings Pricing by Parameter Size
  3. Aug 2025
    B200 and B300 GPU Availability — Frontier Pricing
  4. Mar 2025
    Batch API at 50% Discount Across All Models
  5. Nov 2024
    Turbo + Priority Tiers + Cached Input Discount
Full Fireworks AI timeline

More Fireworks AI activity

All pricing activity