Pinned
The world's fastest AI inference and training.
Try the latest open models at: inference.cerebras.ai
- "You could imagine yourself actually being 10x more productive." @DeyNolan from the Cerebras Core ML team speaks with @alyciazcary on why speed contributes to intelligence. In agentic workloads like Claude Code, queries can run for hours. Faster inference means more test time
- “We think not only in terms of training workloads, but also inference workloads.” @jeffreygwang of @OpenAI explains why reinforcement learning has become an inference workload, and why training and inference must be co-designed.
- “Getting the maximum amount that you can out of a given chip is a super high priority for us.” @jeffreygwang of @OpenAI explains why training performance depends on moving parameters efficiently and reducing communication overhead, not just arithmetic.

