Opus 5 (Fully Tested): A MID-MODEL for a BIG PRICE that still UNDERPERFORMS K3?!
AICodeKing2862 字 (约 12 分钟)
85
Anthropic的Claude Opus 5在基准测试中超越Fable 5,但实际测试显示其表现存在争议。
入选理由:Opus 5在Frontier Bench得分43.3%,是Fable 5的两倍
FeaturedVideo#Anthropic#Claude#模型评测#AI性能对比中英混合
