Ask
Price change

Novita AI cuts self-serve GPU prices ~48% and trims its instance lineup to RTX-class

Novita AI pricing

Novita on-demand RTX 4090 GPU instances drop to $0.35/hr ($0.18 spot) from $0.67/$0.34; L40S and H100 instances removed, H100 moves to dedicated/bare-metal only.

Before

RTX 4090 24GB on-demand $0.67/hr ($0.34 spot); instance lineup included L40S 48GB ($0.55/hr) and H100 SXM 80GB ($2.59/hr).

After

RTX 4090 24GB on-demand $0.35/hr ($0.18 spot); lineup narrowed to RTX-class (RTX 4090, RTX 4090 HF $0.69, RTX 5090 HF $0.72, new RTX 6000 Ada 48GB $0.77); L40S and H100 instances removed.

On 2026-06-24 Novita AI repriced its self-serve GPU-instance page (/en/gpus). The on-demand RTX 4090 24GB rate fell roughly 48% to $0.35/hr (spot $0.18, down from $0.34), and the instance catalog was narrowed to RTX-class GPUs only - RTX 4090, RTX 4090 (High frequency) at $0.69, RTX 5090 (High frequency) at $0.72, and a newly added RTX 6000 Ada 48GB at $0.77. The previously listed L40S 48GB ($0.55/hr) and H100 SXM 80GB ($2.59/hr) instances are gone; H100 capacity is now sold only via dedicated endpoints ($1.99/GPU-hr) and 8-GPU bare-metal nodes ($1.70/GPU/hr).

Two adjacent moves landed in the same capture: dedicated endpoints added an RTX-5090 SKU at $0.73/GPU-hr, and the Agent Sandbox began publishing its underlying per-unit rate card ($0.0000098/vCPU-second, $0.0000032/GiB-second, $0.00009/GB-hour with the first 60 GB included) alongside a new $100 / 90-day free-credit offer.

From Novita AI's pricing timeline
GPU-instance repricing + published sandbox rate card

Self-serve GPU instances (/en/gpus) cut and re-lined-up to RTX-class only: RTX 4090 24GB drops to $0.35/hr on-demand ($0.18 spot) from $0.67/$0.34, with RTX 4090 HF $0.69, RTX 5090 HF $0.72, and a new RTX 6000 Ada 48GB at $0.77 — L40S and H100 SXM instances removed (H100 now sold via dedicated endpoints $1.99/GPU-hr and bare-metal $1.70/GPU/hr). Dedicated endpoints add RTX-5090 at $0.73/GPU-hr. Agent Sandbox now publishes per-unit rates ($0.0000098/vCPU-second, $0.0000032/GiB-second, $0.00009/GB-hour with first 60 GB free) and a $100 / 90-day free-credit offer.

About Novita AI
novita.ai ↗

Novita AI is a pay-as-you-go AI cloud offering inference across 174 listed models as of the 2026-08-14 capture (down from 177 at 2026-08-04/2026-08-11, after that 2026-08-04 refresh added Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL, following a 2026-07-29 delisting of legacy video/image SKUs including Hunyuan Video Fast, PixVerse V4.5, and the Kling-o1 lineup), on-demand and bare-metal GPUs, and secure per-second agent sandboxes under a single API; a 2026-08-25 capture cut the Image catalog from 14 to 5 SKUs and pruned most legacy Video lines while the site's own model-catalog counter read 144, a divergence from the 174 figure this page has tracked that remains unreconciled.

Free tier
Yes
Commits
None
Transparency
public

Novita AI pricing history

  1. Aug 2026
    H100 SXM and L40S return to self-serve GPU instances; RTX 6000 Ada and RTX 5090 HF variant removed
  2. Aug 2026
    Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard
  3. Aug 2026
    Ling 3.0 Tiny delisted after 3 days, emptying the free-LLM shelf
  4. Aug 2026
    Three free LLMs graduate to paid pricing; Ling 3.0 Tiny takes the $0 slot
  5. Aug 2026
    RTX 6000 Ada dropped from self-serve GPUs; base RTX 5090 tier added at $0.73/hr
Full Novita AI timeline

More Novita AI activity

All pricing activity