verifier.org

Novita AI

near-cheapest prices with the full frontier catalogue behind them, and a batch discount the cheap rivals don't offer.

$0.135 / 1m in#most-teams-it-s-within-a#it-actually-carries-what

the best combination on this list: deepinfra's pricing without deepinfra's catalogue gaps, plus an explicit half-price batch tier.

on the anchor model novita is $0.135 per million in — 35% above the cheapest price in the category and a seventh of the most expensive — and it undercuts the reference price on deepseek v4 pro at $1.60/$3.20 where three larger vendors all charge $1.74/$3.48. on glm-5.2 it prints the same $1.40/$4.40 as everyone else.

the catalogue is the reason it ranks first rather than second. multiple deepseek generations, multiple glm generations including 5.2, the qwen family and llama — so the model you standardise on today and the one you migrate to in october are both here. deepinfra is cheaper and carries no glm at all, which is a worse trade than $0.035 per million.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
$0.135 / 1m in
website
novita.ai
our verdict

the best combination on this list: deepinfra's pricing without deepinfra's catalogue gaps, plus an explicit half-price batch tier.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more llm inference providers

Anyscale

ray-based platform for serving and scaling your own models

#ray#self-serve

Hyperbolic

open-model inference and rented gpus on one account

#open-models#gpu

Parasail

serverless and dedicated endpoints for open-weight models

#serverless#dedicated

DeepInfra

we checked this$0.10 / 1m in

the cheapest tokens in the category by a distance — with no glm models at all.

#high-volume-workloads-on#deepseek-where-the-model#the-bill-is-the-problem