Inference platform (open models) · LLM APIs
Same idea as Groq — wafer-scale speed on open models. Actual cost depends entirely on which model you pick, not a single flat rate.
Only one tier verified so far — more being added over time.
Want to calculate exact usage instead of just the representative rate? Open the full calculator.