T
traeai
Sign in

产品

Unsloth

提供量化模型的AI工具

已跟踪 3 条高相关材料

TraeAI 观察

相关材料

已收录 3 条与 Unsloth 相关的内容,按评分排序。

Gemma 4 12B: The Developer Guide

Gemma 4 12B: The Developer Guide

Google Developers Blog1171 字 (约 5 分钟)
92

Gemma 4 12B features an encoder-free multimodal architecture that runs locally on 16GB VRAM devices with native audio support. By eliminating separate vision and audio encoders, it reduces latency and pairs with a dedicated MTP model for faster inference, marking the first mid-sized multimodal model with a macOS desktop app for fully offline interaction.

入选理由:Gemma 4 12B移除独立编码器,视觉仅用35M参数嵌入层,音频直接线性投影至LLM输入空间

FeaturedArticle#Gemma 4#Multimodal LLM#Encoder-Free Architecture#Local AI#Google英文
Qwen3.8-Flash-Next

Qwen3.8-Flash-Next

Simon Willison's Weblog297 字 (约 2 分钟)
85

Qwen3.8-Flash-Next是通义千问推出的多模态MoE模型,具有125B参数和显著性能提升。

入选理由:Qwen3.8-Flash-Next采用多模态MoE架构,是Qwen4的早期预览版本

FeaturedArticle#AI#LLM#模型架构#量化技术中英混合
Google DeepMind Blog 图标

DiffusionGemma: 4x faster text generation

Google DeepMind Blog1006 字 (约 5 分钟)
85

DiffusionGemma 模型通过并行生成文本块,实现高达 4 倍的文本生成速度,适用于需要高速处理的本地交互场景。

入选理由:DiffusionGemma 在 NVIDIA H100 上每秒生成 1000+ tokens,速度比传统模型快 4 倍。

FeaturedArticle#DiffusionGemma#文本生成#AI模型#Google DeepMind英文

跨材料问答 · Unsloth

回答基于:Unsloth 相关 3 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.