Kimi / Cerebras Cerebras发布Kimi-Linear剪枝版本 将Kimi-Linear从48B剪枝到35B,性能反而提升 在LiveCodeBench、AIME25、HumanEval等基准测试中表现更好 输出速度可达2000 token/s,适合本地部署