Groq
Shared by midudev↗The fastest inference on the market. No credit card.
1,000 requests per day131K context
OpenAI
OpenAI’s open-weights release. Runs on your own hardware, or free on Groq. Cerebras still serves it on a $5 signup credit.
The fastest inference on the market. No credit card.
1,000 requests per day131K context
One key, hundreds of models, $5 of credits every month. No markup on tokens.
$5 of credits per month1M context
Wafer-scale inference. Thousands of tokens per second.
$5 one-time credit65K context
Models at the edge, one fetch away from your Worker.
10,000 neurons per day262K context
No sign-up: grab the base URL and start calling the model.
10 requests/minute (40 with a token)5 turbo models
Llama and DeepSeek at high speed on custom hardware.
20 requests per day1M context
The same models and commands as local Ollama, but running on someone else’s GPU.
Monthly starter credits1M context
European hosting. Seven models at €0 on the catalog; the rest you can still call without an account at 2 requests a minute.
7 models at €0262K context