Inference platform (open models) · LLM APIs
Nearly identical hosted rates to BaseTen on the same models; the real differentiator between the two is tooling and latency, not price.
Only one tier verified so far — more being added over time.
Want to calculate exact usage instead of just the representative rate? Open the full calculator.