I remember when people were saying "It's useless to open-source big models because nobody will be ab...
Cerebras is now running the trillion-parameter Kimi K2.6 model in enterprise trials at ~1,000 tokens/s, shattering the old belief that open-source large models are impractical due to hardware limitations.
入选理由:Cerebras 在企业测试中以约1000 tokens/s的速度运行Kimi K2.6(千亿参数模型),创当前最快推理记录。



![[AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable)](https://substackcdn.com/image/fetch/$s_!1XAa!,w_1456,c_limit,f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F34020777-3ff4-437c-af38-c915ca21b7fd_2464x1352.png)






