
How a Trillion-Parameter Coding Model Just Got 23x Faster Than the Competition
Cerebras launches Kimi K2.6, a trillion-parameter open-weight model hitting 981 tokens per second. It’s 6.7x faster than GPU cloud services and 23x faster than the median provider.