Ask
Packaging

Novita AI drops RTX 6000 Ada 48GB from self-serve GPUs, adds a base RTX 5090 tier

Novita AI pricing

Novita AI removed the RTX 6000 Ada 48GB instance from its self-serve GPU lineup and added a new base RTX 5090 32GB on-demand tier at $0.73/hr, while its model catalog grew from 172 to 177 listings.

Before

Self-serve `/en/gpus` lineup: RTX 4090 24GB, RTX 4090 24GB (HF), RTX 5090 32GB (HF) at 0.72/hr on-demand (0.36/hr spot), and RTX 6000 Ada 48GB at 0.77/hr on-demand (0.39/hr spot); 172 models listed on the catalog.

After

Self-serve `/en/gpus` lineup: RTX 4090 24GB, RTX 4090 24GB (HF), RTX 5090 32GB (HF) at 0.72/hr, plus a new base RTX 5090 32GB tier at 0.73/hr on-demand (0.37/hr spot); RTX 6000 Ada 48GB no longer listed; 177 models on the catalog.

Novita AI’s self-serve GPU-instance page (/en/gpus) swapped one SKU for another between the 2026-07-29 and 2026-08-04 captures. RTX 6000 Ada 48GB — the highest-VRAM card in the self-serve lineup at $0.77/hr on-demand — was removed entirely, continuing a pattern of narrowing the self-serve page to consumer-class RTX 4090/5090 silicon (L40S and H100 SXM were pulled from the same page back in June). In its place, a second RTX 5090 32GB listing appeared: a base on-demand tier at $0.73/hr ($0.37/hr spot), sitting alongside the existing “High frequency” RTX 5090 variant at $0.72/hr ($0.36/hr spot) — so the base SKU is now priced a cent above the high-frequency one, an inversion of the RTX 4090 pair, where the high-frequency variant costs roughly double the base rate.

Dedicated endpoints, bare-metal nodes, and the Agent Sandbox rate card were all unchanged in this capture. The model catalog grew modestly from 172 to 177 listings, adding SKUs including Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL.

From Novita AI's pricing timeline
RTX 6000 Ada dropped from self-serve GPUs; base RTX 5090 tier added at $0.73/hr

The /en/gpus self-serve lineup lost RTX 6000 Ada 48GB (was $0.77/hr on-demand, $0.39 spot) and gained a new base RTX 5090 32GB on-demand tier at $0.73/hr ($0.37 spot) — priced a cent above the existing 'High frequency' RTX 5090 variant at $0.72/hr, an inversion of the RTX 4090 pair where the HF variant costs roughly double the base rate. Dedicated endpoints, bare-metal, and Agent Sandbox rate cards held steady; the model catalog grew from 172 to 177 with additions including Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL.

About Novita AI
novita.ai ↗

Novita AI is a pay-as-you-go AI cloud offering inference across 174 listed models as of the 2026-08-14 capture (down from 177 at 2026-08-04/2026-08-11, after that 2026-08-04 refresh added Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL, following a 2026-07-29 delisting of legacy video/image SKUs including Hunyuan Video Fast, PixVerse V4.5, and the Kling-o1 lineup), on-demand and bare-metal GPUs, and secure per-second agent sandboxes under a single API; a 2026-08-25 capture cut the Image catalog from 14 to 5 SKUs and pruned most legacy Video lines while the site's own model-catalog counter read 144, a divergence from the 174 figure this page has tracked that remains unreconciled.

Free tier
Yes
Commits
None
Transparency
public

Novita AI pricing history

  1. Aug 2026
    H100 SXM and L40S return to self-serve GPU instances; RTX 6000 Ada and RTX 5090 HF variant removed
  2. Aug 2026
    Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard
  3. Aug 2026
    Ling 3.0 Tiny delisted after 3 days, emptying the free-LLM shelf
  4. Aug 2026
    Three free LLMs graduate to paid pricing; Ling 3.0 Tiny takes the $0 slot
  5. Aug 2026
    RTX 6000 Ada dropped from self-serve GPUs; base RTX 5090 tier added at $0.73/hr
Full Novita AI timeline

More Novita AI activity

All pricing activity