Ask
Price change

Novita AI hikes Deepseek V4 Flash 0731 pricing and prunes its media catalog

Novita AI pricing

Novita AI raised Deepseek V4 Flash 0731 token rates over 3x, added a new DeepSeek V4 Pro 0813 SKU, swapped an RTX GPU tier for the returning RTX 6000 Ada 48GB, and delisted most of its Image and Video model catalog.

Before

Deepseek V4 Flash 0731 at $0.14/M input and $0.28/M output; self-serve GPU lineup included RTX 4090 24GB 'High frequency' at $0.69/hr; Image catalog carried 14 SKUs (Z Image Turbo, Seedream 4.0, FLUX 2 family, Image Eraser/Remove Background/Upscaler, Flux.1 Kontext variants, Qwen-Image); Video catalog included Veo 3.1, Vidu Q2/Q3, older Kling versions, Wan 2.1/2.2, and Seedance 1.5 Pro.

After

Deepseek V4 Flash 0731 at $0.44/M input and $1.32/M output (cache read unchanged at $0.028/M); DeepSeek V4 Pro 0813 added at $1.32/M input and $3.96/M output; RTX 6000 Ada 48GB back on self-serve GPUs at $0.77/hr in place of the delisted RTX 4090 HF tier; Image catalog down to 5 SKUs; Video catalog now centered on Kling v3.0 and Wan 2.5–2.7 only.

Proof of change

Novita AI's pricing pages, as we captured them on two dates.

Second source confirmed
Captured Aug 14, 2026 Captured Aug 25, 2026 · 11 days apart
main-serverless-endpoints
$0.54 $0.46 $0.92 $0.70 $1.40 $0.30 $1.32 $3.96 $0.44
Captured Aug 14, 2026
Captured Aug 25, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

Show 2 other pages we compared
models
$0.46 $0.51 $1.32 $3.96 $0.44 $0.12 $0.5
Captured Aug 14, 2026
Captured Aug 25, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

gpu-instances
Captured Aug 14, 2026
Captured Aug 25, 2026

Showing the whole page as captured — scroll either panel, or open it at full size.

Both images are our own captures, taken on the dates shown. The pixels are unmodified — nothing is retouched, and where a region is highlighted the marker is drawn over the image, not into it. Open either capture to see it at full resolution.

Novita AI’s serverless rate card saw its sharpest single-SKU repricing since the July 2026 Flux.1 Kontext Pro 10x jump. Deepseek V4 Flash 0731 — a dated variant sitting alongside the unchanged base “Deepseek V4 Flash” model — rose from $0.14/M input and $0.28/M output to $0.44/M input and $1.32/M output (input +214%, output +371%), while its cache-read rate held at $0.028/M. A new flagship dated SKU, DeepSeek V4 Pro 0813, was added at $1.32/M input ($0.132/M cache read) and $3.96/M output.

On the GPU side, the self-serve /en/gpus lineup swapped RTX 4090 24GB “High frequency” ($0.69/hr on-demand, $0.35 spot) for the return of RTX 6000 Ada 48GB ($0.77/hr on-demand, $0.39 spot) — the same GPU that was pulled from this exact page on 2026-08-04, back three weeks later at an unchanged rate. Dedicated-endpoint, bare-metal, and Agent Sandbox pricing were confirmed unchanged via a re-capture of the Dedicated Endpoints tab.

The model catalog also contracted sharply on the media side: the Image category fell from 14 to 5 listed SKUs as Z Image Turbo, Seedream 4.0, Image Eraser, Image Remove Background, Image Upscaler, and the FLUX 2 Dev/Flex/Pro family were all delisted, leaving only the Flux.1 Kontext and Qwen-Image lines. The Video category dropped Veo 3.1, the entire Vidu Q2/Q3 lineup, older Kling versions (V1.6, V2.5 Turbo, V2.6 Pro), Wan 2.1/2.2, and Seedance 1.5 Pro, consolidating around the newer Kling v3.0 and Wan 2.5–2.7 families. Separately, the /en/models catalog counter (All Models tab) now reads 144, versus the 174 figure this page logged for the 2026-08-14 capture — a discrepancy this capture could not reconcile and flags for the next research pass.

From Novita AI's pricing timeline
Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard

Deepseek V4 Flash 0731 jumped from $0.14/$0.28 to $0.44/$1.32 per M tokens (input +214%, output +371%) as a new DeepSeek V4 Pro 0813 SKU ($1.32/$3.96) landed alongside it. On the self-serve GPU page, RTX 6000 Ada 48GB returned at $0.77/hr in place of the delisted RTX 4090 'High frequency' tier ($0.69/hr). The Image catalog shrank from 14 to 5 SKUs (Z Image Turbo, Seedream 4.0, Image Eraser/Remove Background/Upscaler, and the FLUX 2 family all delisted) and the Video catalog dropped Veo 3.1, the Vidu Q2/Q3 lineup, older Kling versions (V1.6/V2.5 Turbo/V2.6 Pro), Wan 2.1/2.2, and Seedance 1.5 Pro in favor of the newer Kling v3.0 and Wan 2.5–2.7 lines. The `/en/models` catalog counter reads 144, versus 174 logged on 2026-08-14 (methodology not reconciled by this capture).

About Novita AI
novita.ai ↗

Novita AI is a pay-as-you-go AI cloud offering inference across 174 listed models as of the 2026-08-14 capture (down from 177 at 2026-08-04/2026-08-11, after that 2026-08-04 refresh added Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL, following a 2026-07-29 delisting of legacy video/image SKUs including Hunyuan Video Fast, PixVerse V4.5, and the Kling-o1 lineup), on-demand and bare-metal GPUs, and secure per-second agent sandboxes under a single API; a 2026-08-25 capture cut the Image catalog from 14 to 5 SKUs and pruned most legacy Video lines while the site's own model-catalog counter read 144, a divergence from the 174 figure this page has tracked that remains unreconciled.

Free tier
Yes
Commits
None
Transparency
public

Novita AI pricing history

  1. Sep 2026
    RTX 5090 cut 23%; H100 pulled from self-serve again; Image Endpoints and sandbox credit removed
  2. Sep 2026
    Macaron V1 Venti and Tall cut 45% across input, output, and cache-read
  3. Aug 2026
    H100 SXM and L40S return to self-serve GPU instances; RTX 6000 Ada and RTX 5090 HF variant removed
  4. Aug 2026
    Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard
  5. Aug 2026
    Ling 3.0 Tiny delisted after 3 days, emptying the free-LLM shelf
Full Novita AI timeline

More Novita AI activity

All pricing activity