Requesty
5% on model spend, stated plainly and applied consistently
the most transparent pricing in the category and the only one that charges for inference, which makes it costlier than the pass-through gateways on identical traffic.
requesty charges 5% on base model cost and puts a worked example on its own pricing page: 'a model costing $10 per 1m tokens from openai costs $10.50 through requesty'. so claude sonnet 4.5 lands at $3.15 per million input tokens against $3.00 direct. no subscription, no seat fees, no minimum spend. in a category where everyone else's real cost is buried in platform fees, log counts or credit-purchase charges, that is genuinely the easiest number to reason about.
it is also, on identical traffic, more expensive than openrouter with byok, cloudflare or vercel. the trade is that routing, caching and eu data residency are included on every plan, and fallbacks, spend limits and advanced observability arrive at the first paid tier rather than the third.
the description above is ours, condensed from the ranking. pricing moves — check it on the vendor's own page before you rely on it.
- category
- llm gateways
- pricing
- $0 free tier — pay-as-you-go is 5% on model cost
- website
- requesty.ai
the most transparent pricing in the category and the only one that charges for inference, which makes it costlier than the pass-through gateways on identical traffic.
we researched this category against vendors' own pricing pages and licence files. that is where this line comes from — not from the vendor, and not from anything they paid for.