Ask
Price change

Novita AI cuts self-serve GPU prices ~48% and trims its instance lineup to RTX-class

Novita AI pricing

Novita on-demand RTX 4090 GPU instances drop to $0.35/hr ($0.18 spot) from $0.67/$0.34; L40S and H100 instances removed, H100 moves to dedicated/bare-metal only.

Before

RTX 4090 24GB on-demand $0.67/hr ($0.34 spot); instance lineup included L40S 48GB ($0.55/hr) and H100 SXM 80GB ($2.59/hr).

After

RTX 4090 24GB on-demand $0.35/hr ($0.18 spot); lineup narrowed to RTX-class (RTX 4090, RTX 4090 HF $0.69, RTX 5090 HF $0.72, new RTX 6000 Ada 48GB $0.77); L40S and H100 instances removed.

Proof of change

Novita AI's pricing pages, as we captured them on two dates.

Capture only
Captured Jun 2, 2026 Captured Jun 24, 2026 · 22 days apart
models
$1.56 $6 $0.51 $0.74 $0.95 $0.19 $4 $1.15 $1.13 $0.52
Captured Jun 2, 2026
Captured Jun 24, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

Show 4 other pages we compared
main-serverless-endpoints
$1.56 $0.74 $1.13 $0.95 $4 $1.15 $0.52 $1.04
Captured Jun 2, 2026
Captured Jun 24, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

agent-sandbox
$100 $0
Captured Jun 2, 2026
Captured Jun 24, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

main-dedicated-endpoints
$0.73
Captured Jun 2, 2026
Captured Jun 24, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

gpu-instances
Captured Jun 2, 2026
Captured Jun 24, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

What these images do and don't show
  • These prices come from our capture alone — they were not confirmed against an independent second source.

Both images are our own captures, taken on the dates shown. The pixels are unmodified — nothing is retouched, and where a region is highlighted the marker is drawn over the image, not into it. Open either capture to see it at full resolution.

On 2026-06-24 Novita AI repriced its self-serve GPU-instance page (/en/gpus). The on-demand RTX 4090 24GB rate fell roughly 48% to $0.35/hr (spot $0.18, down from $0.34), and the instance catalog was narrowed to RTX-class GPUs only - RTX 4090, RTX 4090 (High frequency) at $0.69, RTX 5090 (High frequency) at $0.72, and a newly added RTX 6000 Ada 48GB at $0.77. The previously listed L40S 48GB ($0.55/hr) and H100 SXM 80GB ($2.59/hr) instances are gone; H100 capacity is now sold only via dedicated endpoints ($1.99/GPU-hr) and 8-GPU bare-metal nodes ($1.70/GPU/hr).

Two adjacent moves landed in the same capture: dedicated endpoints added an RTX-5090 SKU at $0.73/GPU-hr, and the Agent Sandbox began publishing its underlying per-unit rate card ($0.0000098/vCPU-second, $0.0000032/GiB-second, $0.00009/GB-hour with the first 60 GB included) alongside a new $100 / 90-day free-credit offer.

From Novita AI's pricing timeline
GPU-instance repricing + published sandbox rate card

Self-serve GPU instances (/en/gpus) cut and re-lined-up to RTX-class only: RTX 4090 24GB drops to $0.35/hr on-demand ($0.18 spot) from $0.67/$0.34, with RTX 4090 HF $0.69, RTX 5090 HF $0.72, and a new RTX 6000 Ada 48GB at $0.77 — L40S and H100 SXM instances removed (H100 now sold via dedicated endpoints $1.99/GPU-hr and bare-metal $1.70/GPU/hr). Dedicated endpoints add RTX-5090 at $0.73/GPU-hr. Agent Sandbox now publishes per-unit rates ($0.0000098/vCPU-second, $0.0000032/GiB-second, $0.00009/GB-hour with first 60 GB free) and a $100 / 90-day free-credit offer.

About Novita AI
novita.ai ↗

Novita AI is a pay-as-you-go AI cloud offering inference across 174 listed models as of the 2026-08-14 capture (down from 177 at 2026-08-04/2026-08-11, after that 2026-08-04 refresh added Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL, following a 2026-07-29 delisting of legacy video/image SKUs including Hunyuan Video Fast, PixVerse V4.5, and the Kling-o1 lineup), on-demand and bare-metal GPUs, and secure per-second agent sandboxes under a single API; a 2026-08-25 capture cut the Image catalog from 14 to 5 SKUs and pruned most legacy Video lines while the site's own model-catalog counter read 144, a divergence from the 174 figure this page has tracked that remains unreconciled.

Free tier
Yes
Commits
None
Transparency
public

Novita AI pricing history

  1. Sep 2026
    RTX 5090 cut 23%; H100 pulled from self-serve again; Image Endpoints and sandbox credit removed
  2. Sep 2026
    Macaron V1 Venti and Tall cut 45% across input, output, and cache-read
  3. Aug 2026
    H100 SXM and L40S return to self-serve GPU instances; RTX 6000 Ada and RTX 5090 HF variant removed
  4. Aug 2026
    Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard
  5. Aug 2026
    Ling 3.0 Tiny delisted after 3 days, emptying the free-LLM shelf
Full Novita AI timeline

More Novita AI activity

All pricing activity