T
traeai
Sign in

公司

DeepSeek

别名:deepseek_ai

提供新型自扩展AI系统的科技公司

已跟踪 30 条高相关材料

TraeAI 观察

相关材料

已收录 30 条与 DeepSeek 相关的内容,按评分排序。

科技爱好者周刊(第 399 期):中国 AI 大厂访问记

Weekly for Tech Enthusiasts (Issue 399): Visiting China's AI Giants

阮一峰的网络日志4694 字 (约 19 分钟)
92

US analysts' visit reveals China's AI compute is 1/8th of the US, but a 4-7x efficiency gain bridges the hardware gap.

入选理由:2025年底美国AI算力约为中国8倍,中国当前总算力仅相当于美国2023年水平。

FeaturedArticle#AI Infrastructure#Compute Efficiency#LLM Open Source#US-China AI中文
https://t.co/nw0GoHamCI

DeepSeek's $10 Trillion Grand Strategy [Translation]

宝玉(@dotey)5655 字 (约 23 分钟)
92

DeepSeek builds a low-cost, high-efficiency model system through multiple foundational innovations to drive China's $10 trillion AI hardware ecosystem and achieve its own $1 trillion valuation.

入选理由:DeepSeek V4 Pro在100万上下文中仅需5.48GB HBM显存,远低于竞品的60-89GB。

FeaturedTweet#DeepSeek#AI Model#MoE#KV Cache Optimization#Hardware Ecosystem中文
DeepSeek 的 10 万亿美元大战略

DeepSeek's 10 Trillion USD Grand Strategy

宝玉的分享5756 字 (约 24 分钟)
92

DeepSeek reduces KV cache requirements through innovations, driving China's AI hardware ecosystem toward a $10 trillion industry.

入选理由:DeepSeek V4 Pro仅需5.48GB HBM,相比GLM5的60GB和Qwen3-235B-A22B的89GB显著节省显存

FeaturedArticle#AI Model#Hardware Ecosystem#KV Cache#DeepSeek#China AI中文
#546. 电力、晶圆与 AI 基础设施的未来

#546. Power, Wafers, and the Future of AI Infrastructure

跨国串门儿计划3114 字 (约 13 分钟)
92

AI infrastructure is undergoing an unprecedented systemic重构 in capitalist history, with power and wafers as the core bottlenecks; Anthropic's $11B monthly ARR surge reveals explosive demand, while TSMC, NVIDIA, and SpaceX are reshaping the global geopolitics of compute.

入选理由:Anthropic单月ARR增长110亿美元,远超市场预期,证明AI基础设施需求远超资本定价能力。

FeaturedPodcast#AI Infrastructure#Semiconductor#TSMC#NVIDIA#Compute Bottleneck中文
Latent Space 图标

Scaling Past Informal AI - Carina Hong, Axiom Math

Latent Space1535 字 (约 7 分钟)
87

Axiom Math advances Verified AI to scale brilliance and compound it through formal proofs with Lean, achieving 12/12 on Putnam and 99% (187/189) on Verina Codegen, far exceeding OpenAI o3’s 4.9%, providing critical capability verification and knowledge propagation for AGI.

入选理由:Axiom在Putnam考试中取得12/12,优于顶尖本科生与当时最接近的AI系统DeepSeek(103/120)。

FeaturedArticle#Verified AI#Formal Verification#Lean#AGI#Putnam Exam英文
#558.AI时代的个人革命:Garry Tan 谈开源 AI、创业信仰、创伤动力

Garry Tan argues that AI is triggering the next personal computing revolution, where open-source Agents and personal AI will empower ordinary people with unprecedented creative capacity; YC’s core tenet is “make something people want”; entrepreneurs must convert trauma into creativity through authentic perception and strong agency.

入选理由:Garry Tan 提出‘个人AI必须由自己拥有和控制’,并正在开发 G Brain——整合邮件、日历、联系人与笔记的个人知识记忆系统。

FeaturedPodcast#AI#Open Source#Startup#YC#Personal Computing中文
DeepSeek’s New AI Is A Game Changer

DeepSeek’s New AI Is A Game Changer

Two Minute Papers1580 字 (约 7 分钟)
87

DeepSeek’s visual pointing lets open-source VLMs slash visual tokens by 90 % while matching or beating GPT-4V on seven public benchmarks and delivering traceable reasoning paths.

入选理由:视觉指针机制将视觉 token 用量压缩 90%,仍保持 SOTA 精度

FeaturedVideo#DeepSeek#Vision-Language Models#Visual Pointing#Token Efficiency#Open Research英文
DeepSeek’s New AI System Shouldn’t Be Possible

DeepSeek’s New AI System Shouldn’t Be Possible

Two Minute Papers937 字 (约 4 分钟)
85

DeepSeek推出可自我扩展的AI系统,通过动态重写程序实现功能定制,但缺乏详细技术原理说明。

入选理由:自扩展AI系统可通过用户指令动态生成功能模块

FeaturedVideo#AI#DeepSeek#自扩展系统#开源英文
DeepSeek is back... and Silicon Valley is terrified

DeepSeek is back... and Silicon Valley is terrified

Fireship1559 字 (约 7 分钟)
85

DeepSeek发布新模型DeepSeek Coder引发行业关注,OpenAI因安全风险暂停训练,Claude代码泄露事件被深度分析。

入选理由:DeepSeek Coder成为GitHub最快获得星标的历史性项目

FeaturedVideo#AI模型#DeepSeek#OpenAI#代码安全#GitHub中英混合
DeepSeek(@deepseek_ai) 图标

DeepSeek推出支持多模态的API,允许通过图像和文本混合输入进行交互,适用于需要视觉和文本处理的应用场景。

入选理由:DeepSeek的'deepseek-v4-flash-vision-exp'模型支持图像和文本混合输入,每个图像最多384个token计费。

FeaturedTweet#API#多模态#DeepSeek#图像处理中英混合
KDnuggets 图标

Speed Up LLM Inference with DSpark Speculative Decoding

KDnuggets1718 字 (约 7 分钟)
85

DSpark推测解码通过结合并行和顺序组件,可提升LLM生成速度60-85%。

入选理由:DSpark结合并行与顺序组件,提升生成速度达60-85%

FeaturedArticle#LLM#DSpark#推测解码#CUDA#llama.cpp英文
AI HOT 精选 图标

DeepSeek 开源多模态模型 V4-Flash-Vision-Exp,支持图片输入且能力接近 Opus-4.8,采用 MIT 许可证。

入选理由:DeepSeek-V4-Flash-Vision-Exp 支持 JPEG/PNG/GIF/WebP 四种图片格式输入

FeaturedArticle#多模态模型#开源#DeepSeek#Agent#MIT License中文
Another DeepSeek Moment Has Arrived

Another DeepSeek Moment Has Arrived

Two Minute Papers863 字 (约 4 分钟)
85

DeepSeek通过后训练技术使模型性能提升7倍,超越更大规模模型,展示了AI训练范式的突破性进展。

入选理由:后训练步骤使DeepSeek Flash模型性能提升7倍,超越5倍参数量的Pro版本

FeaturedVideo#DeepSeek#AI训练#后训练#模型优化#技术突破中英混合
Pi Agent + GPT-5.6 Luna, V4 Flash & Every Model: THIS IS THE BEST!

Pi Agent + GPT-5.6 Luna, V4 Flash & Every Model: THIS IS THE BEST!

AICodeKing2365 字 (约 10 分钟)
85

Pi Agent凭借简洁系统提示和实际测试数据,成为使用DeepSeek等开源模型的最佳工具,且Anthropic已验证其有效性。

入选理由:Pi系统提示仅数百token,成本降低80%且性能无损

FeaturedVideo#AI模型#系统提示#基准测试#开源模型英文
宝玉的分享 图标

关于 Agent 的几个判断

宝玉的分享1237 字 (约 5 分钟)
85

通用AI Agent将主导市场,垂直Agent和封闭系统面临挑战,模型能力是关键竞争因素。

入选理由:通用Agent将取代垂直Agent,Codex和Claude等通用型产品占据优势

FeaturedArticle#AI Agent#技术趋势#模型能力#插件生态中文
How Frontier Labs Are Building Subtle Developer Lock-In

How Frontier Labs Are Building Subtle Developer Lock-In

Mozilla AI Blog862 字 (约 4 分钟)
85

前沿AI实验室通过状态持久性和专有执行环境构建开发者锁定,影响开放系统开发。

入选理由:OpenAI GPT-5.6 Sol通过状态持久性使ARC-AGI-3得分提升3倍

FeaturedArticle#AI#开发者锁定#OpenAI#专有执行环境英文
Latent Space 图标

[AINews] not much happened today

Latent Space1843 字 (约 8 分钟)
85

DeepSeek V4-Flash 0731通过微调实现性能跃升,成本降低60%,但未改变架构。

入选理由:DeepSeek V4-Flash 0731在Terminal-Bench测试中提升25.8个百分点至82.7

FeaturedArticle#DeepSeek#AI模型#API#技术更新中英混合
China’s Open AI Models Are Challenging Silicon Valley’s Playbook

China’s Open AI Models Are Challenging Silicon Valley’s Playbook

Wired AI1447 字 (约 6 分钟)
85

中国开源AI模型性能接近西方顶尖模型,挑战硅谷封闭策略。

入选理由:中国开源模型Kimi K3性能接近Anthropic的Fable,且开放权重。

FeaturedArticle#AI#开源#中美竞争#技术政策英文
Moonshot is Chinese But Its AI Models Are From Another Planet

Moonshot is Chinese But Its AI Models Are From Another Planet

The Algorithmic Bridge3034 字 (约 13 分钟)
85

Moonshot的Kimi K3模型达到美国前沿AI模型水平,可能改变全球AI格局。

入选理由:Kimi K3性能与Anthropic的Mythos/Fable及OpenAI的GPT-5.6相当。

FeaturedArticle#AI模型#开源#地缘政治#Moonshot英文
Interconnects AI 图标

6 months to live for open models

Interconnects AI1968 字 (约 8 分钟)
85

开源AI模型可能在6个月内面临政策限制,中美竞争与监管行动将重塑技术格局。

入选理由:美国可能通过行政命令限制超过GPT-5.5能力的开源模型

FeaturedArticle#AI政策#开源模型#监管科技#中美竞争英文
𝗔𝗱𝗱𝗶𝗻𝗴 𝗮𝗻 𝗔𝗜 𝗮𝗴𝗲𝗻𝘁 𝘁𝗼 𝗱𝗮𝘁𝗮𝗯𝗮𝘀𝗲 𝘁𝗼𝗼𝗹𝗶𝗻𝗴 𝗿𝗮𝗶𝘀𝗲𝘀 𝗮𝗻 ...

Milvus在Attu 3.0 beta中引入自带LLM方法,用户可自主控制LLM提供商和数据路径,无需依赖托管服务。

入选理由:Attu 3.0 beta支持OpenAI/Anthropic等5大LLM提供商的自定义接入

FeaturedTweet#Milvus#数据库工具#LLM集成#Attu中英混合
第三方服务商也因此能把成本做到 DeepSeek 官方 API 的五分之一。

SGLang + Prefill-Decode 解耦 + 专家并行 + AMD MI300——这就是整套技术栈。
htt...

通过SGLang与AMD MI300等技术组合,第三方服务商将推理成本降至DeepSeek官方API的五分之一。

入选理由:使用AMD MI300显卡可降低50%以上推理成本

FeaturedTweet#AI推理优化#成本控制#AMD MI300#模型架构中英混合
SuperTechFans 图标

2026 06 28 HackerNews

SuperTechFans10427 字 (约 42 分钟)
85

美国政府监管GPT-5.6引发争议,DSpark推理加速技术提升60-85%效率,OpenAI安全防护升级成焦点。

入选理由:DSpark通过半自回归架构使大模型生成速度提升60%-85%

FeaturedArticle#AI监管#大模型#推理加速#OpenAI#DSpark中文
DeepSeeK 突然发布 DSpark,让 AI 的回答不再「挤牙膏」

DeepSeeK 突然发布 DSpark,让 AI 的回答不再「挤牙膏」

爱范儿3218 字 (约 13 分钟)
85

DSpark通过半自回归架构和置信度调度验证,使大模型生成速度提升60%-85%。该框架解决了推测解码中的后缀衰减问题,已在DeepSeek-V4系列模型中落地。

入选理由:DSpark采用半自回归架构,结合并行生成与顺序校验模块,减少后缀衰减问题

FeaturedArticle#大模型推理加速#推测解码#DeepSeek#半自回归架构中文

跨材料问答 · DeepSeek

回答基于:DeepSeek 相关 30 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.