Scott Wu(@ScottWu46)

Benchmark scores are exciting but more importantly we are seeing incredible results so far using thi...

6.6内容质量
Benchmark scores are exciting but more importantly we are seeing incredible results so far using thi...

TL;DR · AI 摘要

Scott Wu on X: "Benchmark scores are exciting but more importantly we are seeing incredible results so far using this mo...

核心要点

  • 主题聚焦:Benchmark scores are exciting but more important
  • 来源:Scott Wu(@ScottWu46),建议结合原文判断细节。
  • AI 分析暂不可用,本条为保底评分与摘要。

Scott Wu 在 X 上的发言:"基准测试成绩令人兴奋,但更重要的是,我们在 Devin 中使用该模型已经取得了惊人的成果!试试看:"

Scott Wu

@ScottWu46

基准测试成绩令人兴奋,但更重要的是,我们在 Devin 中使用该模型已经取得了惊人的成果!试试看:

Cognition

@cognition

7月8日

我们推出 SWE-1.7,这是迄今为止训练出的最强大的模型。它在成本仅为最前沿模型几分之一的情况下,得分差距仅在几个点之内,现在处理速度已达到 1000 tok/s。强化学习尚未触及极限:在优化我们的方案后,随着规模扩大,我们仍在持续看到性能提升。

2026年7月8日 下午4:14

25.9K

次浏览

2

0

20

9

8

280

阅读20条回复