Inference platform (open models) · LLM APIs
Custom silicon (not GPUs) is the pitch — genuinely fast inference on open models, priced competitively with Groq and Cerebras.
Only one tier verified so far — more being added over time.
Want to calculate exact usage instead of just the representative rate? Open the full calculator.