Ground Truth.
AI, checked against the source.

← All topics

cerebras

Everything on Ground Truth tagged “cerebras” — 2 items.

Cerebras is serving an open 27B model at 1,500 tokens a second, and the free tier caps it exactly News

Cerebras now serves Qwen 3.8 27B at roughly 1,500 output tokens per second, but its own rate-limit page caps free-tier users at 90,000 tokens per minute -- almost precisely the model's raw output rate -- so the headline speed only becomes usable on the paid tier.

OpenAI put its most intelligent model on Cerebras chips at 750 tokens a second News

OpenAI is previewing Ultrafast, a service tier that runs GPT-5.6 Sol on Cerebras hardware at up to 14 times the speed of standard processing and up to 750 output tokens per second.