T
traeai
Sign in

公司

NVIDIA

别名:英伟达

图形处理器制造商,持续支持方

已跟踪 30 条高相关材料

TraeAI 观察

相关材料

已收录 30 条与 NVIDIA 相关的内容,按评分排序。

#567. 黄仁勋:Agent 时代普通人和企业的新生产力,AI 基础设施竞赛下的计算革命

Jensen Huang announced at GTC Taipei 2026 that the Agentic AI era has arrived, shifting AI from content generation to autonomous task execution. NVIDIA launched infrastructure products like Vera Rubin and Vera CPU, driving a computing paradigm shift where AI becomes a direct generator of profit and GDP.

入选理由:NVIDIA发布Vera Rubin超级计算系统,专为Agent设计,支持解耦、异构和分布式AI工作负载。

FeaturedPodcast#AI Agent#NVIDIA#Vera Rubin#Agentic AI#AI Infrastructure中文
Welcome NVIDIA Cosmos 3: The First Open Omni-model for Physical AI Reasoning and Action

NVIDIA Cosmos 3 is the first open-source omni-model for physical AI, integrating world generation, physical reasoning, and action generation into one unified system. Built on MoT architecture, it supports robotics, autonomous driving, and synthetic data pipelines via Hugging Face and Diffusers.

入选理由:Cosmos 3 是首个统一物理AI能力的开源模型,融合世界生成、物理推理与动作生成于单模型。

FeaturedArticle#NVIDIA#Physical AI#Omni-model#Hugging Face#MoT Architecture英文
Meet Cosmos 3: Our Latest Frontier Model for Physical AI

Meet Cosmos 3: Our Latest Frontier Model for Physical AI

NVIDIA Developer482 字 (约 2 分钟)
92

NVIDIA releases Cosmos 3, the first Omni model integrating vision, language, sound, and action, built on Mixture-of-Transformer architecture, achieving top scores across multiple physical AI benchmarks with open weights for customization and edge deployment.

入选理由:Cosmos 3 是首个融合语言/视频/声音/动作的Omni模型,基于Mixture-of-Transformer架构。

FeaturedVideo#NVIDIA#Physical AI#Omni Model#Mixture-of-Transformer#Open Model英文
Introducing NVIDIA Cosmos 3

Introducing NVIDIA Cosmos 3: Unified Multimodal Model for Physical AI

NVIDIA Developer543 字 (约 3 分钟)
92

NVIDIA launches Cosmos 3, the first unified multimodal model integrating language, video, sound, and action inputs/outputs, built on Mixture of Transformer architecture, open-sourced with weights available on Hugging Face, achieving top scores across physical AI benchmarks including Robo Lab, PiBench, and Vintage.

入选理由:Cosmos 3 是首个整合语言/视频/声音/动作输入输出的 omni 模型,基于 Mixture of Transformer 架构。

FeaturedVideo#NVIDIA#Physical AI#Multimodal Model#Mixture of Transformers#Open Source英文
英伟达清华团队提出Gamma-World:世界模型从「一个人玩」到「多人共处」

Gamma-World systematically solves architectural gaps in multi-agent world modeling via simplex agent encoding and sparse hub attention, achieving >40% average FVD reduction, zero-shot generalization from 2 to 4 agents, and 24 FPS real-time rollout.

入选理由:采用正单纯形(regular simplex)编码玩家身份,实现任意玩家间几何等距,支持零样本扩展至更多玩家且无需重训

FeaturedArticle#World Model#Multi-Agent#Transformer#NVIDIA#Tsinghua中文
英伟达清华团队提出Gamma-World:世界模型从「一个人玩」到「多人共处」

Gamma-World systematically solves multi-agent world modeling via simplex agent encoding and sparse hub attention, enabling zero-shot generalization from 2-player training to 4-player inference and 24 FPS real-time rollout, with average FVD reduction >40%.

入选理由:采用正单纯形(regular simplex)编码实现玩家身份等距、无参数、可扩展,支持训练时2人→推理时4人零样本泛化

FeaturedArticle#World Model#Multi-Agent#Transformer#NVIDIA#Tsinghua中文
Towards Speed-of-Light Text Generation with Nemotron-Labs Diffusion Language Models

NVIDIA introduces Nemotron-Labs Diffusion, a diffusion-based language model family supporting AR, diffusion, and self-speculation modes, enabling parallel multi-token generation and iterative refinement for higher throughput and flexibility.

入选理由:Nemotron-Labs Diffusion 提供 3B/8B/14B 三档模型,支持商业友好许可证(NVIDIA Nemotron Open Model License)。

FeaturedArticle#Diffusion LM#NVIDIA#Nemotron#LLM Inference#Text Generation英文
#546. 电力、晶圆与 AI 基础设施的未来

#546. Power, Wafers, and the Future of AI Infrastructure

跨国串门儿计划3114 字 (约 13 分钟)
92

AI infrastructure is undergoing an unprecedented systemic重构 in capitalist history, with power and wafers as the core bottlenecks; Anthropic's $11B monthly ARR surge reveals explosive demand, while TSMC, NVIDIA, and SpaceX are reshaping the global geopolitics of compute.

入选理由:Anthropic单月ARR增长110亿美元,远超市场预期,证明AI基础设施需求远超资本定价能力。

FeaturedPodcast#AI Infrastructure#Semiconductor#TSMC#NVIDIA#Compute Bottleneck中文
在AWS上进行基础模型训练与推理的核心构建模块

Building Blocks for Foundation Model Training and Inference on AWS

AI HOT 精选4633 字 (约 19 分钟)
92

AWS provides a comprehensive technical stack for large-scale foundation model training and inference, integrating high-performance compute, networking, storage, and open-source software to support NVIDIA's 'three scaling laws'.

入选理由:NVIDIA 的三大缩放定律包括预训练、后训练(如 SFT 和 RL)和推理时计算,需统一基础设施支持。

FeaturedArticle#AWS#Foundation Models#Distributed Training#PyTorch#Scalability英文
Unlocking large scale AI training networks with MRC (Multipath Reliable Connection)

OpenAI, in collaboration with AMD, NVIDIA, and others, introduces MRC—a new networking protocol that enhances performance and reliability for large-scale AI training, now open-sourced via OCP to advance industry standards.

入选理由:MRC通过多路径传输和静态源路由设计,有效规避网络拥塞与单点故障,保障AI训练稳定性。

FeaturedArticle#MRC#AI training network#OpenAI#supercomputer architecture#OCP英文
英伟达重新思考AI TCO:为何每Token成本才是唯一重要的指标

NVIDIA advocates for cost per token as the core economic metric for AI infrastructure, replacing traditional measures like compute cost or FLOPS per dollar, emphasizing full-stack optimization to reduce inference costs and enhance business value.

入选理由:每Token成本是衡量AI基础设施经济效益的核心指标,直接反映实际产出效率。

FeaturedArticle#NVIDIA#AI TCO#Inference Optimization#Cost Per Token中文
OpenAI 把训练 ChatGPT 用的网络协议开源了。https://t.co/s2euzfedsb

这套协议叫 MRC(Multipath Reliable Connection,多路径可靠连...

OpenAI Open-Sources the Networking Protocol Used to Train ChatGPT

宝玉(@dotey)666 字 (约 3 分钟)
92

OpenAI, together with AMD, Intel, NVIDIA and others, has open-sourced MRC, a new networking protocol that improves reliability and efficiency in large-scale AI training clusters.

入选理由:MRC 实现微秒级故障绕过,避免传统网络中断导致训练重启。

FeaturedTweet#MRC#OpenAI#Large Model Training#Networking Protocol#SRv6中文
Cosmos 3 is here.

Cosmos 3 is here

NVIDIA Developer268 字 (约 2 分钟)
90

NVIDIA launches Cosmos 3, an open omni-model for physical AI based on a novel mixture-of-transformers architecture, capable of generating physics-accurate synthetic video, serving as a world model and simulator, and enabling training for robotic and mobile intelligent systems.

入选理由:Cosmos 3 使用新型混合 Transformer 架构,结合自回归和扩散 Transformer 实现感知、推理与生成。

FeaturedVideo#NVIDIA#AI#Physical AI#Transformer#World Model英文
The Counterintuitive Networking Decisions Behind OpenAI’s 131,000-GPU Training Fabric

The Counterintuitive Networking Decisions Behind OpenAI’s 131,000-GPU Training Fabric

Towards Data Science4587 字 (约 19 分钟)
90

OpenAI built a 131,000 GPU training network using the MRC protocol with counterintuitive design choices that significantly improve training efficiency.

入选理由:MRC协议消除传统三层控制平面,提升网络可靠性

FeaturedArticle#Network Architecture#AI Training#OpenAI中文
AI That Designs Its Own Chips: Ricursive's Anna Goldie and Azalia Mirhoseini

AI That Designs Its Own Chips: Ricursive's Anna Goldie and Azalia Mirhoseini

Sequoia Capital2294 字 (约 10 分钟)
89

Anna Goldie and Azalia Mirhoseini founded Recursive Intelligence, which uses AI to automate chip design, applied in multiple generations of Google TPUs, and plans to democratize chip design through three phases.

入选理由:Recursive Intelligence 的 AlphaChip 已经在 Google TPU 上应用了四代,显著提高了芯片设计效率。

FeaturedVideo#AI#chip design#automation英文
Introducing NVIDIA Nemotron 3 Ultra: An Open 550B Model for Long-Running Agents

NVIDIA today launches Nemotron 3 Ultra, a 550B-parameter open model built on the same architecture as Nemotron 3 Super, optimized for long-running AI agents. It employs LatentMoE to quadruple the number of experts at the same inference cost, introduces multi-token prediction to boost single-user inference speed, and is released under the Linux Foundation’s Open MDW license to enable enterprise deployment.

入选理由:Nemotron 3 Ultra 为 550B 参数模型,基于与 Nemotron 3 Super 相同架构,面向长时运行的智能代理场景。

FeaturedVideo#NVIDIA#Nemotron#AI Agent#LatentMoE#OpenMDW英文
Nemotron 3 Ultra NVIDIA's 550B Open Model

Nemotron 3 Ultra: NVIDIA's 550B Open Agent Model

Sam Witteveen3906 字 (约 16 分钟)
87

NVIDIA introduces the 550B-parameter Neotron 3 Ultra, a mixture-of-experts agent model trained for task orchestration, outperforming many trillion-parameter open agents on benchmarks, with full data and recipe transparency to enable enterprise on-prem deployment and fine-tuning.

入选理由:Neotron 3 Ultra 为 550B 参数混合专家模型,活跃参数约 55B,专为代理任务训练。

FeaturedVideo#Nemotron3Ultra#550B#Mixture-of-Experts#Agent Benchmarks#Open Models英文
#559. All-in:SpaceX、AI 递归自我进化、Nvidia 巨额利润、美国为何开始害怕 AI?

Recursive self-improvement is driving AI models into a 'new Moore’s Law' era; SpaceX is building a trillion-dollar 'Elon Web Services' ecosystem via Starlink and Colossus compute; American fear of AI stems from job displacement anxiety, CEO communication failures, and regulatory misalignment—not the technology itself.

入选理由:Anthropic已实现LLM ARR盈利,递归式自我改进(如Claude优化自身)可能使AI迭代速度超越人类工程师

FeaturedPodcast#AI#SpaceX#Nvidia#Recursive Self-Improvement#Tech Ethics中文
https://t.co/HjRCjERRGY

Launching Recursive: Building Recursive Self-Improving Superintelligence

Richard Socher(@RichardSocher)1158 字 (约 5 分钟)
87

Richard Socher announces Recursive, a new company focused on recursive self-improving superintelligence to automate scientific discovery and accelerate progress toward superintelligence.

入选理由:Recursive由Richard Socher与Tim Rocktaeschel等七位联合创始人共同创立,团队包含Peter Norvig等顶级AI专家。

FeaturedTweet#Recursive#Superintelligence#Self-Improving AI#Automated Science#Richard Socher英文
https://t.co/bey4uOXXHk

Analysis of the 100 Most Popular Hardware Setups on Hugging Face

clem 🤗(@ClementDelangue)811 字 (约 4 分钟)
87

Based on 297k Hugging Face users, this analysis reveals top AI hardware configurations, showing NVIDIA, Apple, and Intel dominance in their respective categories and the importance of VRAM for AI workloads.

入选理由:AI开发者更看重显存容量而非算力,RTX 3060 12GB成为最流行独立GPU。

FeaturedTweet#Hugging Face#AI Hardware#NVIDIA#Apple Silicon#Local AI英文
Stratechery 图标

An Interview with OpenAI President Greg Brockman About Astra and Alignment

Stratechery12976 字 (约 52 分钟)
85

OpenAI推出新模型Astra并强调对齐重要性,讨论其在AI价值链中的定位及与微软、Nvidia的竞争关系。

入选理由:Astra是OpenAI最新模型,专注于对齐与安全性。

FeaturedArticle#OpenAI#Astra#AI对齐#AI价值链#微软#Nvidia英文
NVIDIA AI(@NVIDIAAI) 图标

Read the technical blog for the five guidelines: https://t.co/HaXwdiTYYB

NVIDIA AI(@NVIDIAAI)54 字 (约 1 分钟)
85

NVIDIA 技术博客提出通过 Speculative Decoding 优化大模型推理的五项指南,可使推理速度提升 2-5 倍。

入选理由:Speculative Decoding 技术可使 LLM 推理速度提升 2-5 倍

FeaturedTweet#LLM#Speculative Decoding#NVIDIA#AI模型优化英文
AI News: OpenAI Made a Massive Move Against NVIDIA

AI News: OpenAI Made a Massive Move Against NVIDIA

Matt Wolfe7568 字 (约 31 分钟)
85

OpenAI发布自研Jalapeno芯片性能超NVIDIA 104倍,同时NVIDIA被曝收购Hugging Face,AI芯片与开源生态竞争加剧。

入选理由:OpenAI Jalapeno芯片在推理任务上实现NVIDIA 104倍性能提升

FeaturedVideo#AI芯片#OpenAI#NVIDIA#开源模型中英混合
FeaturedTweet#虚拟化#GPU管理#Helix#QEMU#NVIDIA英文
Martin Fowler 图标

Fragments: September 1

Martin Fowler1123 字 (约 5 分钟)
85

人类识别AI生成文本能力有限,NVIDIA的AVO工具在长期任务中表现良好,AI正在重塑CI流程假设。

入选理由:2025年研究显示人类识别AI文本准确率仅57%-64%

FeaturedArticle#AI检测#NVIDIA架构#持续集成#技术趋势英文
How to extract meaning from charts and tables in PDFs

How to extract meaning from charts and tables in PDFs

Weaviate Blog2871 字 (约 12 分钟)
85

Weaviate提出Late Interaction RAG方法,通过多向量模型直接解析PDF图表,解决传统RAG无法提取图表信息的缺陷。

入选理由:传统RAG处理PDF图表时,30%的查询会因图表信息缺失导致错误

FeaturedArticle#RAG#PDF处理#Weaviate#OCR#多向量模型英文

跨材料问答 · NVIDIA

回答基于:NVIDIA 相关 30 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.