verifier.org

inference.net

the cheapest llama 4 scout price we found, wrapped in a plan structure that meters something other than tokens.

$0.08 / 1m in#cost-driven-workloads-on

the headline prices are excellent and the catalogue we could verify is three models deep — a good deal you can't fully evaluate.

llama 4 scout at $0.08/$0.15 and maverick at $0.35/$0.40 are the cheapest rates we verified for those models anywhere in this category, and deepseek v3.2 at $0.14/$0.28 is likewise strong.

the plan layer is unusual and needs reading twice. on top of per-token usage sit tiers metered by 'gateway requests': pay-as-you-go includes a million a month at 30 requests per minute, growth costs $250/month for fifty million at 250 per minute. that request ceiling, not the token price, is what will actually constrain a busy application.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
$0.08 / 1m in
our verdict

the headline prices are excellent and the catalogue we could verify is three models deep — a good deal you can't fully evaluate.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more llm inference providers

Anyscale

ray-based platform for serving and scaling your own models

#ray#self-serve

Hyperbolic

open-model inference and rented gpus on one account

#open-models#gpu

Parasail

serverless and dedicated endpoints for open-weight models

#serverless#dedicated

Novita AI

we checked this$0.135 / 1m in

near-cheapest prices with the full frontier catalogue behind them, and a batch discount the cheap rivals don't offer.

#most-teams-it-s-within-a#it-actually-carries-what