Ask
Price change

Novita AI hikes Deepseek V4 Flash 0731 pricing and prunes its media catalog

Novita AI pricing

Novita AI raised Deepseek V4 Flash 0731 token rates over 3x, added a new DeepSeek V4 Pro 0813 SKU, swapped an RTX GPU tier for the returning RTX 6000 Ada 48GB, and delisted most of its Image and Video model catalog.

Before

Deepseek V4 Flash 0731 at $0.14/M input and $0.28/M output; self-serve GPU lineup included RTX 4090 24GB 'High frequency' at $0.69/hr; Image catalog carried 14 SKUs (Z Image Turbo, Seedream 4.0, FLUX 2 family, Image Eraser/Remove Background/Upscaler, Flux.1 Kontext variants, Qwen-Image); Video catalog included Veo 3.1, Vidu Q2/Q3, older Kling versions, Wan 2.1/2.2, and Seedance 1.5 Pro.

After

Deepseek V4 Flash 0731 at $0.44/M input and $1.32/M output (cache read unchanged at $0.028/M); DeepSeek V4 Pro 0813 added at $1.32/M input and $3.96/M output; RTX 6000 Ada 48GB back on self-serve GPUs at $0.77/hr in place of the delisted RTX 4090 HF tier; Image catalog down to 5 SKUs; Video catalog now centered on Kling v3.0 and Wan 2.5–2.7 only.

Novita AI’s serverless rate card saw its sharpest single-SKU repricing since the July 2026 Flux.1 Kontext Pro 10x jump. Deepseek V4 Flash 0731 — a dated variant sitting alongside the unchanged base “Deepseek V4 Flash” model — rose from $0.14/M input and $0.28/M output to $0.44/M input and $1.32/M output (input +214%, output +371%), while its cache-read rate held at $0.028/M. A new flagship dated SKU, DeepSeek V4 Pro 0813, was added at $1.32/M input ($0.132/M cache read) and $3.96/M output.

On the GPU side, the self-serve /en/gpus lineup swapped RTX 4090 24GB “High frequency” ($0.69/hr on-demand, $0.35 spot) for the return of RTX 6000 Ada 48GB ($0.77/hr on-demand, $0.39 spot) — the same GPU that was pulled from this exact page on 2026-08-04, back three weeks later at an unchanged rate. Dedicated-endpoint, bare-metal, and Agent Sandbox pricing were confirmed unchanged via a re-capture of the Dedicated Endpoints tab.

The model catalog also contracted sharply on the media side: the Image category fell from 14 to 5 listed SKUs as Z Image Turbo, Seedream 4.0, Image Eraser, Image Remove Background, Image Upscaler, and the FLUX 2 Dev/Flex/Pro family were all delisted, leaving only the Flux.1 Kontext and Qwen-Image lines. The Video category dropped Veo 3.1, the entire Vidu Q2/Q3 lineup, older Kling versions (V1.6, V2.5 Turbo, V2.6 Pro), Wan 2.1/2.2, and Seedance 1.5 Pro, consolidating around the newer Kling v3.0 and Wan 2.5–2.7 families. Separately, the /en/models catalog counter (All Models tab) now reads 144, versus the 174 figure this page logged for the 2026-08-14 capture — a discrepancy this capture could not reconcile and flags for the next research pass.

From Novita AI's pricing timeline
Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard

Deepseek V4 Flash 0731 jumped from $0.14/$0.28 to $0.44/$1.32 per M tokens (input +214%, output +371%) as a new DeepSeek V4 Pro 0813 SKU ($1.32/$3.96) landed alongside it. On the self-serve GPU page, RTX 6000 Ada 48GB returned at $0.77/hr in place of the delisted RTX 4090 'High frequency' tier ($0.69/hr). The Image catalog shrank from 14 to 5 SKUs (Z Image Turbo, Seedream 4.0, Image Eraser/Remove Background/Upscaler, and the FLUX 2 family all delisted) and the Video catalog dropped Veo 3.1, the Vidu Q2/Q3 lineup, older Kling versions (V1.6/V2.5 Turbo/V2.6 Pro), Wan 2.1/2.2, and Seedance 1.5 Pro in favor of the newer Kling v3.0 and Wan 2.5–2.7 lines. The `/en/models` catalog counter reads 144, versus 174 logged on 2026-08-14 (methodology not reconciled by this capture).

About Novita AI
novita.ai ↗

Novita AI is a pay-as-you-go AI cloud offering inference across 174 listed models as of the 2026-08-14 capture (down from 177 at 2026-08-04/2026-08-11, after that 2026-08-04 refresh added Deepseek V3.2, Deepseek V4 Flash, DeepSeek-OCR 2, and PaddleOCR-VL, following a 2026-07-29 delisting of legacy video/image SKUs including Hunyuan Video Fast, PixVerse V4.5, and the Kling-o1 lineup), on-demand and bare-metal GPUs, and secure per-second agent sandboxes under a single API; a 2026-08-25 capture cut the Image catalog from 14 to 5 SKUs and pruned most legacy Video lines while the site's own model-catalog counter read 144, a divergence from the 174 figure this page has tracked that remains unreconciled.

Free tier
Yes
Commits
None
Transparency
public

Novita AI pricing history

  1. Aug 2026
    H100 SXM and L40S return to self-serve GPU instances; RTX 6000 Ada and RTX 5090 HF variant removed
  2. Aug 2026
    Deepseek V4 Flash 0731 repriced sharply; RTX 6000 Ada returns; Image/Video catalog cut hard
  3. Aug 2026
    Ling 3.0 Tiny delisted after 3 days, emptying the free-LLM shelf
  4. Aug 2026
    Three free LLMs graduate to paid pricing; Ling 3.0 Tiny takes the $0 slot
  5. Aug 2026
    RTX 6000 Ada dropped from self-serve GPUs; base RTX 5090 tier added at $0.73/hr
Full Novita AI timeline

More Novita AI activity

All pricing activity