verifier.org

Fireworks AI

the most legible pricing in the category, and the cheapest route to a frontier deepseek by a wide margin.

$0.14 / 1m in#teams-who-want-the-front#caching-spelled-out-befo

publishes more of its own pricing structure than anyone here, and deepseek v4 flash at $0.14/$0.28 is the standout value in the whole category.

the pricing docs are unusually honest about complexity instead of hiding it: every model has an explicit standard and priority rate, cached input is priced separately and steeply discounted — deepseek v4 pro's cached input is $0.145 against $1.74 standard — and models without a named row fall into published parameter-count tiers rather than a quote form.

deepseek v4 flash is the find. at $0.14 in and $0.28 out it costs a twelfth of v4 pro, and for the large share of workloads that don't need the pro tier that is the cheapest frontier-adjacent capability on this page.

the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.

pricing
$0.14 / 1m in
our verdict

publishes more of its own pricing structure than anyone here, and deepseek v4 flash at $0.14/$0.28 is the standout value in the whole category.

we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.

more llm inference providers

Anyscale

ray-based platform for serving and scaling your own models

#ray#self-serve

Hyperbolic

open-model inference and rented gpus on one account

#open-models#gpu

Parasail

serverless and dedicated endpoints for open-weight models

#serverless#dedicated

Novita AI

we checked this$0.135 / 1m in

near-cheapest prices with the full frontier catalogue behind them, and a batch discount the cheap rivals don't offer.

#most-teams-it-s-within-a#it-actually-carries-what