inference.net
the cheapest llama 4 scout price we found, wrapped in a plan structure that meters something other than tokens.
the headline prices are excellent and the catalogue we could verify is three models deep — a good deal you can't fully evaluate.
llama 4 scout at $0.08/$0.15 and maverick at $0.35/$0.40 are the cheapest rates we verified for those models anywhere in this category, and deepseek v3.2 at $0.14/$0.28 is likewise strong.
the plan layer is unusual and needs reading twice. on top of per-token usage sit tiers metered by 'gateway requests': pay-as-you-go includes a million a month at 30 requests per minute, growth costs $250/month for fifty million at 250 per minute. that request ceiling, not the token price, is what will actually constrain a busy application.
the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.
- category
- llm inference providers
- pricing
- $0.08 / 1m in
- website
- inference.net
the headline prices are excellent and the catalogue we could verify is three models deep — a good deal you can't fully evaluate.
we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.