Qwen 3.8 27B at 1500 tok/s on Cerebras vs my garage rack. Ran the same Go refactor on both.
Cerebras put Qwen 3.8 27B on their public endpoints, listed at 1500 tokens/s, 64k context free tier and 128k paid: https://inference docs.cerebras.ai/models/overview The catalog page also says they d…
