Cerebras adds Qwen 3.8 27B model to public API at ~1500 tokens/sec
Cerebras has added the 27-billion-parameter Qwen 3.8 model to its public API endpoints, offering inference speeds of roughly 1500 tokens per second. The model supports 64k context on free tier and 128k on paid tier, and is available under Cerebras's free trial and pay-as-you-go pricing, subject to rate limits.