Grok just broke the trend
Grok 4.5在编码基准测试中以83.3分超越Opus 4.8,Cursor公司正基于600亿美元收购数据训练下一代模型。
入选理由:Grok 4.5在编码基准测试中得分83.3,领先Opus 4.8五分
模型
别名:Composer 2.5、Composer 3
Cursor公司开发的编码模型系列
已跟踪 6 条高相关材料
最近变化
2026-07-09 · Grok 4.5在编码基准测试中得分83.3,领先Opus 4.8五分
为什么值得关注
Composer 被反复提及时,通常意味着它正在影响产品路线、开发者工作流或 AI 产业判断。这个页面把分散材料合并成一个可持续更新的观察入口。
Grok just broke the trend
Matthew Berman · 8.5 分
Grok 4.5在编码基准测试中以83.3分超越Opus 4.8,Cursor公司正基于600亿美元收购数据训练下一代模型。
https://t.co/WEy47rIccs
Yangyi(@Yangyixxxx) · 8.5 分
Cursor 在3年内实现200亿美元估值,ARR从0到20亿,几乎零营销投入,其增长核心在于将产品设计与用户心智迁移深度结合:通过fork VS Code降低迁移成本、用Tab和Composer消除编码摩擦、借Karpathy等KOL自然传播,形成‘试用一周回不去’的强留存闭...
Cursor | The Hidden Bug in Every Large-Scale RL Run
Sequoia Capital · 7.5 分
在大规模强化学习训练中,由于模型版本不一致和数值计算差异,导致推理阶段的对数概率值出现不匹配,进而引发训练偏差。该问题被称为‘数值不匹配’,是当前大模型训练中的隐性缺陷。
已收录 6 条与 Composer 相关的内容,按评分排序。
Grok 4.5在编码基准测试中以83.3分超越Opus 4.8,Cursor公司正基于600亿美元收购数据训练下一代模型。
入选理由:Grok 4.5在编码基准测试中得分83.3,领先Opus 4.8五分
Cursor achieved a $20 billion valuation in 3 years, scaling ARR from 0 to $2 billion with almost no marketing spend. Its growth engine combines product design and user habit migration: forking VS Code to reduce friction, using Tab and Composer to eliminate coding friction, and leveraging KOLs like Karpathy to spread 'vibe coding' virally, creating a 'can't go back after one week' retention loop.
入选理由:Cursor 通过 fork VS Code 实现近乎零迁移成本,借势1亿开发者心智,获客效率极高。
In large-scale RL training, numerical mismatches arise due to model version drift and floating-point precision differences, causing inconsistent log probabilities during inference and introducing training bias.
入选理由:在异步训练中,需重运行前向传播以生成对数概率,但相同模型版本下结果可能不同。
Eric Zakariasson 请求对 Composer 2.5 的行为、速度和质量提供反馈,以改进下一个模型。
入选理由:Composer 2.5 需要在行为、速度和质量方面进行改进。
Cursor's official Twitter account posted a brief statement announcing Composer 2.5 is built on the same open-source base as Composer 2, referencing Moonshot's Kimi K2.5, but provided no technical details or architectural explanations.
入选理由:Composer 2.5 与 Composer 2 共享同一开源代码基础
This tweet is merely a brief recommendation to try Grok Composer 2.5, lacking technical details, benchmark data, or architectural analysis; its information density is too low for engineering reference.
入选理由:原文仅含“试试 Grok 的 Composer 2.5”一句推荐语,无任何功能说明或技术指标。