We evaluated Nemotron-3-Embed from @NVIDIAAI for our @mem0ai memory retrieval pipeline Tested on lo...
mem0团队采用Nemotron-3-Embed模型后,长文本检索准确率提升1.67个百分点,其开放权重和NVFP4加速特性成为关键选择因素。
入选理由:Nemotron-3-Embed在longmemeval测试中实现80.38的检索准确率,超越qwen-3-600m模型
概念
别名:NVFP
NVIDIA第四代浮点运算加速技术
已跟踪 6 条高相关材料
最近变化
2026-07-16 · Nemotron-3-Embed在longmemeval测试中实现80.38的检索准确率,超越qwen-3-600m模型
为什么值得关注
NVFP4 被反复提及时,通常意味着它正在影响产品路线、开发者工作流或 AI 产业判断。这个页面把分散材料合并成一个可持续更新的观察入口。
We evaluated Nemotron-3-Embed from @NVIDIAAI for our @mem0ai memory retrieval pipeline Tested on lo...
mem0(@mem0ai) · 8.5 分
mem0团队采用Nemotron-3-Embed模型后,长文本检索准确率提升1.67个百分点,其开放权重和NVFP4加速特性成为关键选择因素。
See the benchmarks, full recipe breakdown, and MaxText example 👇 https://t.co/ClmeZCb2SF
NVIDIA AI(@NVIDIAAI) · 8.5 分
NVIDIA 使用 JAX 和 MaxText 在 Blackwell 上训练模型,显著提升训练速度。
Holo3.1: Fast & Local Computer Use Agents
Hugging Face Blog · 8.5 分
Holo3.1 是 Hugging Face 推出的全新计算机使用代理模型,支持跨桌面、移动端与多框架部署,并首次提供 FP8/Q4 GGUF/NVFP4 量化权重以实现本地高效推理。
已收录 6 条与 NVFP4 相关的内容,按评分排序。
mem0团队采用Nemotron-3-Embed模型后,长文本检索准确率提升1.67个百分点,其开放权重和NVFP4加速特性成为关键选择因素。
入选理由:Nemotron-3-Embed在longmemeval测试中实现80.38的检索准确率,超越qwen-3-600m模型
NVIDIA 使用 JAX 和 MaxText 在 Blackwell 上训练模型,显著提升训练速度。
入选理由:使用 JAX 和 MaxText 可以在 NVIDIA Blackwell 上显著提升模型训练速度。
Holo3.1 is Hugging Face's new computer-use agent model supporting cross-platform, multi-framework deployment and first releasing quantized weights (FP8/Q4 GGUF/NVFP4) for local inference.
入选理由:Holo3.1 在 AndroidWorld 上 35B-A3B 模型准确率从 67% 提升至 79.3%
NVIDIA Nemotron 3 Ultra is now available on Amazon SageMaker JumpStart with one-click deployment. This 550B-parameter MoE model is designed for long-running agents, delivering 5x faster inference, 30% lower cost, and 1M token context support.
入选理由:Nemotron 3 Ultra采用混合Transformer-Mamba MoE架构,550B总参仅激活55B,显著降低Agent任务计算开销。
NVIDIA Research releases LongLive-2.0 system that adopts end-to-end NVFP4 training and inference architecture to solve long video generation problems, eliminating model deployment gaps through unified training-inference precision alignment while improving speed and memory efficiency.
入选理由:LongLive-2.0采用NVFP4低精度训练推理架构
Nvidia releases LongLive-2.0, an NVFP4 parallel infrastructure for long video generation, but the tweet only announces the product name without disclosing any technical implementation details.
入选理由:Nvidia发布LongLive-2.0长视频生成基础设施