Inference platform (open models) · LLM APIs
Not its own model — you're paying for very fast inference on open models like Llama. Smaller models on the same platform run far cheaper than this flagship rate.
Only one tier verified so far — more being added over time.
Want to calculate exact usage instead of just the representative rate? Open the full calculator.