verifier.org

SambaNova

custom silicon built for speed, from a vendor that publishes no speed figures and stopped at deepseek v3.2.

$0.60 / 1m in#teams-already-committed-

fair llama pricing on interesting hardware, undone by a catalogue that has fallen a generation behind and documentation with holes in it.

the llama price is competitive — $0.60/$1.20 on the anchor model, roughly matching groq — and the rdu architecture is a genuine third approach alongside gpus and groq's lpu. openai-sdk compatibility is confirmed in its docs.

the catalogue is where it falls behind. six models, no glm, no qwen, and deepseek stops at v3.1 and v3.2 while most of this list has moved to v4 — both v3 variants priced at $3.00/$4.50, which is more than fireworks charges for v4 pro. a cached-input rate on minimax at $0.06 against $0.60 standard is the one striking discount.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
$0.60 / 1m in
our verdict

fair llama pricing on interesting hardware, undone by a catalogue that has fallen a generation behind and documentation with holes in it.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more llm inference providers

Anyscale

ray-based platform for serving and scaling your own models

#ray#self-serve

Hyperbolic

open-model inference and rented gpus on one account

#open-models#gpu

Parasail

serverless and dedicated endpoints for open-weight models

#serverless#dedicated

Novita AI

we checked this$0.135 / 1m in

near-cheapest prices with the full frontier catalogue behind them, and a batch discount the cheap rivals don't offer.

#most-teams-it-s-within-a#it-actually-carries-what