Novita cuts Macaron V1 Venti and Tall token rates 45%
Mind Lab's Macaron V1 Venti and Macaron V1 Tall both dropped 45% on Novita's serverless rate card, less than a month after converting from a free launch tag to paid per-token pricing.
Macaron V1 Venti $1.5/M in ($0.3/M cache read) · $4.5/M out; Macaron V1 Tall $0.45/M in ($0.08/M cache read) · $2.6/M out (rates set when both models converted off their 'Time Limited Free' tag on 2026-08-11, still in effect as of the 2026-08-26 capture).
Macaron V1 Venti $0.825/M in ($0.165/M cache read) · $2.475/M out; Macaron V1 Tall $0.2475/M in ($0.044/M cache read) · $1.43/M out — every field on both models scaled to exactly 55% of its prior value.
Both Macaron models moved by the identical 55% multiplier across input, output, and cache-read rates alike — input, output, and cache-read all landing at exactly 55% of their 2026-08-11 conversion price — which points to a deliberate blanket repricing of the Mind Lab family rather than a per-field adjustment. It is the fastest reprice-after-launch cycle tracked for these two SKUs: converted from free to paid on 2026-08-11, then cut 45% again less than four weeks later, on 2026-09-07.
Just 12 days after H100 returned to self-serve, the /en/gpus page removed it again — H100 SXM 80GB stays available only via the $1.99/GPU-hr dedicated endpoint and $1.70/GPU/hr bare-metal routes. The self-serve RTX 5090 32GB base tier was cut ~23% to $0.56/hr (no spot rate shown), while a separate 'High frequency' RTX 5090 tier was reinstated at $0.72/hr on-demand / $0.36/hr spot. The same pass dropped Dedicated Endpoints' Image Endpoint subscriptions ($559/mo Standard, $1,199/mo Pro) and Agent Sandbox's $100/90-day signup-credit offer.