T
traeai
Sign in

模型

Claude-4-Sonnet

别名:Claude 4 Sonnet

Anthropic公司发布的大型语言模型,常作为基准比较对象。

已跟踪 1 条高相关材料

TraeAI 观察

最近变化

2026-05-31 · ToolCUA在OSWorld-MCP上达46.85%准确率,超越Claude-4-Sonnet,接近Claude-4.5-Sonnet。

为什么值得关注

Claude-4-Sonnet 被反复提及时,通常意味着它正在影响产品路线、开发者工作流或 AI 产业判断。这个页面把分散材料合并成一个可持续更新的观察入口。

AgentCUAOpen SourceReinforcement LearningTool Selection

相关材料

已收录 1 条与 Claude-4-Sonnet 相关的内容,按评分排序。

别光给Agent加Tool了,它根本选不明白!复旦×通义提出全新CUA训练范式

Fudan and Tongyi introduce ToolCUA, solving Agent’s inability to select between GUI and Tool actions; achieves 46.85% accuracy on OSWorld-MCP, surpassing Claude-4-Sonnet, via synthetic trajectory generation and trajectory-level reward design.

入选理由:ToolCUA在OSWorld-MCP上达46.85%准确率,超越Claude-4-Sonnet,接近Claude-4.5-Sonnet。

FeaturedArticle#Agent#CUA#Tool Selection#Reinforcement Learning#Open Source中文

跨材料问答 · Claude-4-Sonnet

回答基于:Claude-4-Sonnet 相关 1 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.