T
traeai
Sign in

公司

什么是 Arena

也叫:ml_angelopoulos

AI评估平台,实现1亿美元ARR

为什么现在值得关注?

最近变化

2026-06-30 · Meta Brain2Qwerty v2实现61%词级准确率,接近侵入式BCI效果

Arena 被反复提及时,通常意味着它正在影响产品路线、开发者工作流或 AI 产业判断。这个页面把分散材料合并成一个可持续更新的观察入口。

📰 Arena 最新动态

已收录 6 篇与「Arena」相关的 AI 资讯和分析。

Latent Space 图标

[AINews] not much happened today

Latent Space1934 字 (约 8 分钟)
85

本文汇总了2026年6月AI领域的重要进展,涵盖脑机接口、产品发布和商业化动态。

入选理由:Meta Brain2Qwerty v2实现61%词级准确率,接近侵入式BCI效果

FeaturedArticle#AI#脑机接口#产品发布#商业化中英混合
GLM-5.2 is the step change for open agents

GLM-5.2 is the step change for open agents

Interconnects AI1879 字 (约 8 分钟)
85

GLM-5.2 在开放模型中表现突出,超越了多个主流模型,成为开放代理的重要进展。

入选理由:GLM-5.2 在 Arena 的代理排行榜中表现优于 OpenAI 和 Anthropic 的最新模型。

FeaturedArticle#GLM-5.2#AI 模型#开放代理#Z.ai英文
Millions of votes a week. One tagging system.

Arena researchers Guanglei Song and I-Hung Hsu walk t...

Millions of votes a week. One tagging system.

lmarena.ai(@lmarena_ai)176 字 (约 1 分钟)
85

Arena.ai uses a unified tagging system to process millions of votes per week, with a data pipeline built on Databricks and Spark.

入选理由:Arena.ai 每周处理数百万次用户投票,依赖统一标签系统进行分类。

FeaturedTweet#Arena#LLM#Data Pipeline英文
Qwen最新3.7 Max预览版空降!两代超大杯并行迭代,林俊旸走了但还在加速

Qwen 3.7 Max Preview ranks 13th in Arena's text domain and 16th in vision domain, both topping Chinese models. Alibaba's LLM iteration pace has significantly accelerated since 2025, with release cycles shortened from 4-6 months to 2-3 months, and nearly monthly updates in 2026, demonstrating sustained acceleration.

入选理由:Qwen 3.7-Max-Preview在Arena文本榜排名第13,是全球前十五唯一中国模型

FeaturedArticle#Qwen#Large Language Model#Alibaba#Arena Leaderboard#Model Iteration中文
Qwen3.7预览版登陆竞技场,阿里视觉排名升至第五

Qwen3.7 Preview Lands on Arena, Alibaba Vision Ranks Fifth

AI HOT 精选111 字 (约 1 分钟)
60

Qwen3.7 preview version is now on Arena, Alibaba's vision ranking rises to fifth, and the model series will be released soon.

入选理由:Qwen3.7-Plus-Preview在Arena视觉竞技场排名第五,整体排名第十六

FeaturedArticle#Qwen#Vision Model#Alibaba Cloud中文
Gemma 4 shifts Pareto Frontier on Code @arena.🔥

Among open models, Gemma-4-31b ranks #13 and Gemma...

Gemma 4 Shifts Pareto Frontier in Code Arena

Philipp Schmid(@_philschmid)193 字 (约 1 分钟)
55

The Gemma-4 series of open models ranks #13 (31b) and #17 (26b-a4b) in code performance, marking strong efficiency and local deployability on devices like MBP.

入选理由:Gemma-4-31b在开源代码模型中排名全球第13。

FeaturedTweet#Gemma#Code Model#Open Source AI中英混合

与「Arena」经常一起出现的 AI 术语。

💡 想追踪「Arena」的长期趋势?去 实体雷达 · Arena 查看详细分析和跨材料问答。

AI may generate inaccurate information. Please verify important content.