verifier.org

Baseten

the compliance-first option — soc 2 type ii and hipaa, hybrid deployment, and no public price for half its catalogue.

$0.60 / 1m in#regulated-teams-who-need#the-option-to-run-the-sa

the right answer when compliance drives the decision — priced at the market reference on what it does publish, and silent on the rest.

baseten is the only entry here leading with soc 2 type ii and hipaa, and one of the few offering self-hosted and hybrid deployment alongside its cloud. for teams whose blocker is a security review rather than a price, that combination decides it.

published pricing sits exactly on the market reference — glm-5.2 at $1.40/$4.40, deepseek v4 at $1.74/$3.48 — with glm 4.7 at $0.60/$2.20 as a cheaper step down and a 'fast' glm-5.2 variant at $2.10/$6.60 for latency-sensitive work. billing is compute-time only, with no idle charge.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
$0.60 / 1m in
website
baseten.co
our verdict

the right answer when compliance drives the decision — priced at the market reference on what it does publish, and silent on the rest.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more llm inference providers

Anyscale

ray-based platform for serving and scaling your own models

#ray#self-serve

Hyperbolic

open-model inference and rented gpus on one account

#open-models#gpu

Parasail

serverless and dedicated endpoints for open-weight models

#serverless#dedicated

Novita AI

we checked this$0.135 / 1m in

near-cheapest prices with the full frontier catalogue behind them, and a batch discount the cheap rivals don't offer.

#most-teams-it-s-within-a#it-actually-carries-what