Qwen 3.8 27B open-weight model runs on Cerebras at 1500 tokens/sec
Qwen 3.8 27B, a new open-weight model, is now available on Cerebras hardware, achieving a throughput of 1500 tokens per second. The deployment leverages Cerebras's specialized architecture to deliver high-speed inference for the model.
Coverage timeline
Hacker Newsaltertable