Qwen(@Alibaba_Qwen)
VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimo...
5.5内容质量

TL;DR · AI 摘要
Qwen3.6-35B-A3B 是原生多模态模型,仅激活约30亿参数,在多个视觉语言基准上媲美甚至超越 Claude Sonnet 4.5。
核心要点
- Qwen3.6-35B-A3B 为原生多模态架构,非后期对齐
- 激活参数仅约30亿,但性能接近Claude Sonnet 4.5
- 在RefCOCO(92.0)和ODInW13(50.8)等空间理解任务表现突出
#Qwen#多模态#大模型#视觉语言模型#阿里巴巴
打开原文Qwen on X: "VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exceed what its size would suggest, with only around 3 billion activated parameters. Across most vision-language benchmarks, its performance https://t.co/nOVBNlVfzW" / X
Don’t miss what’s happening

VLM Performance:Qwen3.6 is natively multimodal, and Qwen3.6-35B-A3B showcases perception and multimodal reasoning capabilities that far exceed what its size would suggest, with only around 3 billion activated parameters. Across most vision-language benchmarks, its performance matches Claude Sonnet 4.5, and even surpasses it on several tasks. Its strengths are particularly evident in spatial intelligence, where it achieves 92.0 on RefCOCO and 50.8 on ODInW13.
·
6
24
393
50
Read 6 replies