rate-limits
Everything on Ground Truth tagged “rate-limits” — 1 item.
Cerebras is serving an open 27B model at 1,500 tokens a second, and the free tier caps it exactly News
Cerebras now serves Qwen 3.8 27B at roughly 1,500 output tokens per second, but its own rate-limit page caps free-tier users at 90,000 tokens per minute -- almost precisely the model's raw output rate -- so the headline speed only becomes usable on the paid tier.