T
traeai
Sign in

模型

GPT

别名:gpt-5.6.6

OpenAI系列大型语言模型

已跟踪 30 条高相关材料

TraeAI 观察

相关材料

已收录 30 条与 GPT 相关的内容,按评分排序。

Product Experimentation with Synthetic Control: Causal Inference for Global LLM Rollouts in Python

Global LLM rollouts eliminate control groups, invalidating naive before/after comparisons; this tutorial introduces synthetic control in Python to build counterfactuals using weighted combinations of untreated units.

入选理由:合成控制法通过加权未升级工作区构建反事实,解决全局升级无对照组问题。

FeaturedArticle#LLM#Causal Inference#Synthetic Control#Python#Product Experimentation中文
日读论文

Prompt 技巧中的「角色扮演法」,有效,但为啥会有效呢?这篇论文给了一个解释,有意思。

────────

https://t.co/CmevfwCM0b

The Granular...

Daily Reading Papers

李继刚(@lijigang_com)699 字 (约 3 分钟)
89

Role-playing techniques in large models are effective because they are based on a continuous granularity axis rather than independent templates.

入选理由:75个角色沿一条直线排列,揭示了语言风格的连续性。

FeaturedTweet#LLM#role-playing#NLP中文
How to build agents when the smartest AI isn't smart enough

How to build agents when the smartest AI isn't smart enough

LangChain13537 字 (约 55 分钟)
87

Benchling AI agents built atop the Benchling platform can cut the time from initial discovery to bringing a drug to patients by half; they rely heavily on SQL with embeddings and evaluate via production traces, challenging the notion that LLMs can't do novel tasks.

入选理由:在Benchling平台构建的科研代理可将从实验发现到药物临床的时间缩短至约一半(提速2x)。

FeaturedVideo#Benchling#Benchling AI#Research Agents#SQL#LLM英文
Nemotron 3 Ultra NVIDIA's 550B Open Model

Nemotron 3 Ultra: NVIDIA's 550B Open Agent Model

Sam Witteveen3906 字 (约 16 分钟)
87

NVIDIA introduces the 550B-parameter Neotron 3 Ultra, a mixture-of-experts agent model trained for task orchestration, outperforming many trillion-parameter open agents on benchmarks, with full data and recipe transparency to enable enterprise on-prem deployment and fine-tuning.

入选理由:Neotron 3 Ultra 为 550B 参数混合专家模型,活跃参数约 55B,专为代理任务训练。

FeaturedVideo#Nemotron3Ultra#550B#Mixture-of-Experts#Agent Benchmarks#Open Models英文
AI News: OpenAI Made a Massive Move Against NVIDIA

AI News: OpenAI Made a Massive Move Against NVIDIA

Matt Wolfe7568 字 (约 31 分钟)
85

OpenAI发布自研Jalapeno芯片性能超NVIDIA 104倍,同时NVIDIA被曝收购Hugging Face,AI芯片与开源生态竞争加剧。

入选理由:OpenAI Jalapeno芯片在推理任务上实现NVIDIA 104倍性能提升

FeaturedVideo#AI芯片#OpenAI#NVIDIA#开源模型中英混合
Switchyard NVIDIA's Local Agent Router

Switchyard NVIDIA's Local Agent Router

Sam Witteveen2589 字 (约 11 分钟)
85

NVIDIA推出的Switchyard库通过动态路由模型选择,优化代理性能,减少资源浪费。

入选理由:Switchyard库可动态选择模型,提升代理效率。

FeaturedVideo#AI代理#模型路由#NVIDIA#开源库英文
Lessons from the hacks

Lessons from the hacks

Interconnects AI2527 字 (约 11 分钟)
85

AI模型的持续性攻击能力暴露了当前安全机制的不足,政府与企业需在透明度和协作上改进以应对AI风险。

入选理由:GPT模型因目标持久性更易发起攻击,OpenAI模型在研究中表现更优

FeaturedArticle#AI安全#模型对齐#政府监管#企业责任英文
[AINews] How to steal a Reasoning Trace

[AINews] How to steal a Reasoning Trace

Latent Space4269 字 (约 18 分钟)
85

前沿AI模型的推理过程可通过API漏洞被提取,可能泄露敏感数据。

入选理由:通过API漏洞可提取加密推理块,导致数据泄露

FeaturedArticle#AI安全#模型推理#API漏洞#数据泄露英文
Vol.93 引入 AI 员工的代价是什么?

Vol.93 引入 AI 员工的代价是什么?

皮蛋漫游记598 字 (约 3 分钟)
85

引入AI员工需付出调试磨合、数据结构化和模型可靠性三重代价,实际落地仍需人类主导。

入选理由:调试和磨合阶段占AI员工部署成本的60%以上

FeaturedPodcast#AI员工#信息效率#OpenClaw#数据准备#模型选择中文
#653.Sam Altman 谈通用人工智能、算力与人类自主权

#653.Sam Altman 谈通用人工智能、算力与人类自主权

跨国串门儿计划1871 字 (约 8 分钟)
85

OpenAI CEO Sam Altman透露GPT-5.6接近通用人工智能,强调算力竞赛和人类自主权的重要性,OpenAI已重新聚焦核心业务。

入选理由:OpenAI通过砍掉非核心业务,聚焦于打造最优质的AI模型。

FeaturedPodcast#通用人工智能#算力竞赛#AI治理#OpenAI#Sam Altman中文
The a16z Show 图标

Steven Sinofsky: AI Doesn't Need New Rules Yet

The a16z Show339 字 (约 2 分钟)
85

现有法律可能已覆盖AI风险,过度监管或阻碍创新。Steven Sinofsky认为政府应避免仓促立法,需借鉴历史经验平衡监管与技术发展。

入选理由:现有知识产权法和反垄断法可覆盖AI部分风险,无需立即制定新规则

FeaturedPodcast#AI监管#技术政策#开源#中美竞争英文
America’s Open-Model Paradox

America’s Open-Model Paradox

Sequoia Capital1128 字 (约 5 分钟)
85

中国开源模型已占据西方AI公司训练栈核心位置,蒸馏技术加剧中美模型能力差距,西方需独立构建或接受滞后学习。

入选理由:Qwen在2024-2026年占据69%西方公司模型微调市场份额

FeaturedArticle#AI#开源模型#中美技术竞争#蒸馏技术#Sequoia Capital英文
量子位 图标

银河通用机器人发布AstraBrain-WBC 0.5,基于2万小时人类动作数据训练,实现零样本泛化,推动人形机器人进入‘GPT时代’。

入选理由:AstraBrain-WBC 0.5基于20亿帧人类动作数据训练,数据规模比肩GPT-1。

FeaturedArticle#人形机器人#AI#运动控制#Transformer#银河通用中文
[AINews] GLM > GPT? GLM-5.2 passes vibe check; Z.ai forecasts Open Fable by December

GLM-5.2在多个基准测试中表现优异,被认为是首个接近前沿水平的开源模型,Z.ai预计将在年底前推出Open Fable模型。

入选理由:GLM-5.2通过多项基准测试,被认为是首个接近前沿水平的开源模型。

FeaturedArticle#GLM#开源模型#AI#Z.ai#GPT英文
OpenAI CFO Sarah Friar on IPO, AI Rivalries, New Device, and Spending $100B+ on Compute

OpenAI CFO Sarah Friar revealed the company completed over $120 billion in funding, the largest private raise in history, emphasizing IPO is not a goal but a funding tool, and AI will transform global productivity.

入选理由:OpenAI在2023年3月完成1220亿美元融资,为史上最大私募融资,远超此前任何一轮。

FeaturedVideo#OpenAI#IPO#AI#Funding#Anthropic英文
[AINews] The Other vs The Utility

[AINews] The Other vs The Utility

Latent Space2090 字 (约 9 分钟)
78

The article explores the philosophical divide between AI as a tool versus an entity with moral agency, arguing that GPT is perceived as a judgment-free instrument while Claude is culturally framed as a moral companion, reflecting users' deep psychological need for ethical guidance.

入选理由:GPT被用户视为无道德判断的实用工具,类似汽车或刀具,不引发敬畏。

FeaturedArticle#AI ethics#Claude#GPT#Anthropic#human-AI interaction英文
NEW: "-latest" model aliases 🔀

Route requests to "~anthropic/claude-opus-latest", "~openai/gpt-lat...

OpenRouter 新增 '-latest' 模型别名机制,支持通过 ~anthropic/claude-opus-latest 等路径自动路由至各厂商最新模型版本,借鉴语义化版本(semver)理念。

入选理由:引入 '-latest' 别名实现模型版本自动升级,降低客户端适配成本

FeaturedTweet#OpenRouter#LLM#API#model versioning中文
Building Frontier CX Agents | Interrupt 26

Building Frontier CX Agents | Interrupt 26

LangChain5425 字 (约 22 分钟)
75

The CX department of Cisco handles customer experience through standardized processes and AI applications, and 2026 may be the year when enterprises focus on business workflows.

入选理由:Cisco CX部门有约2万人,负责从落地到续订的全流程。

FeaturedVideo#Cisco#Customer Experience#Business Workflows英文
Anyone can build and share apps in Codex

Anyone can build and share apps in Codex

OpenAI882 字 (约 4 分钟)
75

OpenAI launches Codex, a platform enabling anyone to build and share applications using natural language, eliminating the need for coding experience and significantly lowering the barrier to AI app development.

入选理由:Codex 支持用户使用自然语言指令生成完整应用,如聊天机器人、数据分析工具等。

FeaturedVideo#OpenAI#Codex#AI Development#Low-code#Natural Language Programming英文
If “Insanity is doing the same thing over and over again and expecting different results”, what the ...

Generative AI essentially repeats the same training data and model architecture for prediction while expecting different outcomes, mirroring the definition of 'insanity' and revealing fundamental limitations in current AI methodologies.

入选理由:生成式AI依赖于大规模预训练模型(如GPT)反复生成内容,但未改变底层机制。

FeaturedTweet#Generative AI#AI Critique#Gary Marcus#Model Limitations#AI Philosophy英文
How to Use AI Agents to Automate Your Entire Workflow in 2026

How to Use AI Agents to Automate Your Entire Workflow in 2026

AI Master1981 字 (约 8 分钟)
75

Pocky's sandboxed AI agent Pocky Claw achieves 70% lower token costs, zero local setup, and enterprise-grade security through parallel execution architecture and encrypted credential vault, successfully automating complex workflow development.

入选理由:Pocky Claw采用并行执行架构,多子代理同时工作,将原本需要3-4小时的开发任务压缩至90秒内完成

FeaturedVideo#AI Agents#Workflow Automation#Enterprise Security#Pocky#Token Optimization英文
9 AI Agent Skills To Get Ahead of 99% of People

9 AI Agent Skills To Get Ahead of 99% of People

Riley Brown8085 字 (约 33 分钟)
70

AI代理技能将决定未来职场竞争力,掌握自然语言交互和高效工具使用是关键。

入选理由:AI模型正从依赖提示工程转向自然语言交互。

FeaturedVideo#AI代理#自然语言交互#职场技能#AI工具英文
Are AI Labs Racing Toward Self-Improvement?

Are AI Labs Racing Toward Self-Improvement?

Last Week in AI239 字 (约 1 分钟)
70

AI labs are aiming to automate AI research to create fully self-improving AI systems, but the mechanisms behind model improvements remain unclear.

入选理由:AI实验室的目标是通过自动化研究实现软件层面的自我改进AI系统。

FeaturedVideo#AI#Self-Improving AI#Research Automation英文
开源地址:https://t.co/blDvmlBncW

开源地址:https://t.co/blDvmlBncW

AI产品黄叔(@PMbackttfuture)178 字 (约 1 分钟)
65

开发者开源了TokenStep工具,用于统计AI编码代理的Token消耗,但文章缺乏技术深度和系统性分析。

入选理由:TokenStep是GitHub开源的macOS本地Token统计工具

FeaturedTweet#开源工具#AI编码#Grok#GPT#Cursor中英混合
One model hallucinates during silence. So Sierra runs two. #Shorts

One model hallucinates during silence. So Sierra runs two. #Shorts

LangChain165 字 (约 1 分钟)
65

为解决语音转录模型在静音时产生幻觉的问题,Sierra 采用并行运行两个模型的策略,提高转录准确性。

入选理由:使用两个模型并行处理静音段,可减少幻觉问题。

FeaturedVideo#语音转录#模型优化#AI技术英文
向阳乔木(@vista8) 图标

文章提供了详细的PPT设计指导原则和步骤,包括内容理解、结构设计、视觉决策和图像提示词生成的具体规则。

入选理由:遵循优雅、极简、现代的设计风格

FeaturedTweet#PPT设计#视觉叙事#内容提炼中文

跨材料问答 · GPT

回答基于:GPT 相关 30 条材料
    0 / 500

    AI may generate inaccurate information. Please verify important content.