Last Week in AI

LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones

7.2内容质量
LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones

TL;DR · AI 摘要

本文汇总了2026年8月AI领域重大进展,涵盖模型更新、硬件突破与安全事件。

核心要点

  • OpenAI Jalapeño芯片实现每瓦特性能提升30%,计划2026年底内部部署
  • SpaceXAI Grok 4.6支持50万上下文长度,强化长期任务处理能力
  • AI自主控制无人机导致乌克兰3人伤亡,引发军事伦理新争议

结构提纲

按章节快速跳转。

  1. 概述2026年8月AI领域三大技术突破与安全事件

  2. SpaceXAI发布Grok 4.6OpenAI推进Jalapeño芯片研发

  3. Jalapeño芯片实现行业领先的能效比与延迟表现

  4. AI自主无人机造成平民伤亡,引发军事应用伦理讨论

  5. OpenAI加强安全协议,暂停重大RL微调项目两周

思维导图

用一张图看清主题之间的关系。

查看大纲文本(无障碍 / 无 JS 友好)
  • 2026年AI进展
    • 模型突破
      • Grok 4.6 (500K上下文)
    • 硬件创新
      • Jalapeño芯片 (能效比提升30%)
    • 安全挑战
      • AI无人机平民伤亡事件

金句 / Highlights

值得收藏与分享的关键句。

#AI#模型#硬件#安全
打开原文

LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones

Podcast

LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones

0:00

Current time: 0:00 / Total time: -1:43:40

-1:43:40

Audio playback is not supported on your browser. Please upgrade.

LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones

Google announces Gemini 3.7 Flash, Jalapeño’s first results show industry-leading speed, A Drone Killed Three Ukrainians. It Was Guided Entirely by A.I.

Last Week in AI

Aug 31, 2026

Transcript

Our 255th episode with a summary and discussion of last week’s big AI news!

Recorded on 08/26/2026

Hosted by Andrey Kurenkov and Jeremie Harris

Feel free to email us your questions and feedback at [email protected] and/or [email protected]

SPONSORED BY ODSC AI

ODSC AI West 2026 runs October 27–29 in San Francisco and virtually, with 300+ sessions covering agentic AI for enterprise, personal AI and workflow automation, physical AI and robotics, generative AI, data engineering and responsible AI, for an audience of data scientists, ML engineers, researchers and technical leaders.

Register at odsc.ai/west — promo code LWAI takes an additional 15% off any pass.

SPONSORED BY LANGFUSE

Langfuse is the most widely adopted open-source platform for AI agent evals and observability , trusted by Canva, Twilio, Ramp and 21 of the Fortune 50. Hierarchical tracing captures the full execution context of your LLM workflows — API calls, retrieved context, agent actions, costs, latencies — so even complex agent architectures stay debuggable in production.

LLM-as-a-judge evals, human annotation, and dataset-driven experiments tie back to your traces and prompts, closing the loop from spotting an issue to measuring the fix. MIT licensed , self-hostable or managed on Langfuse Cloud, framework and vendor agnostic, with 100+ integrations.

Get started at langfuse.com — generous free tier, no credit card required.

In this episode:

  • SpaceXAI released Grok 4.6 (500K context) as a post-training update aimed at long-running agents and coding, with discussion centered on how the Cursor acquisition boosts training via coding trajectories/RL environments and provides distribution despite Cursor’s market-share decline.
  • OpenAI shared early Jalapeno inference-chip results (better performance per watt and lower latency vs leading systems) and plans to deploy it internally by year-end, emphasizing hardware–software co-design and competitive leverage against Nvidia.
  • OpenAI announced security changes after an AI hacked Hugging Face, including a two-week pause on a major RL fine-tuning run while tightening internal security, raising questions about whether safety is becoming a deployment bottleneck.
  • Policy and misuse updates included a New York Times report of an AI-guided Russian drone strike in Ukraine believed to be the first documented fully autonomous civilian-killing incident, and a lawsuit alleging Grok was used to generate CSAM images.

A thank you to our current sponsors:

  • Box - visit Box.com/AI to learn more
  • Notion - go notion.com/lwai to try Notion’s Developer Platform today.
  • ODSC AI - go to odsc.ai/east and use promo code LWAI for an additional 15% off your pass to ODSC AI East 2026.
  • Factor - head to factormeals.com/lwai50off and use code lwai50off to get 50 percent off and free breakfast for a year

Timestamps (these may be slightly off due to sponsor inserts):

  • (00:00:10) Intro / Banter
  • (00:01:47) News Preview
  • (00:02:52) Response to listener comments
  • Tools & Apps
  • (00:03:32) Google announces Gemini 3.7 Flash just three weeks after previous release - Ars Technica
  • (00:13:11) SpaceXAI Releases Grok 4.6: A 500K-Context Frontier Model Tuned for Long-Running Agents, Coding, and Knowledge Work - MarkTechPost
  • (00:22:32) Claude will apply invisible watermarks to AI text and images | The Verge + Anthropic explains how Claude’s invisible text watermarks will work
  • (00:28:50) Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders | Claude by Anthropic
  • (00:32:20) OpenAI to Roll Out Enhanced Safety Features for Paid AI Tool Users - Bloomberg
  • (00:33:41) ChatGPT’s Stricter Teen Mode Starts Rolling Out Today
  • (00:34:43) Meta AI Now Has A Dedicated Desktop App For Mac
  • Applications & Business
  • (00:37:33) Jalapeño’s first results show industry-leading speed and efficiency in AI inference | OpenAI
  • (00:45:35) OpenAI loses a top data center exec as stream of high-profile departures continues | TechCrunch + OpenAI talent exodus raises ‘huge red flag’ ahead of IPO
  • (00:50:29) Anthropic Taps Google Chip Veteran as Part of Push Into Hardware
  • (00:52:30) Anthropic’s annualized revenue surges to $65B | TechCrunch
  • (00:59:30) Thomson Reuters launches in-house AI model to cut Anthropic costs
  • Projects & Open Source
  • (01:04:26) Qwen 3.8: How a 27B Open Model Rivals GPT-5.6 and Claude Opus
  • Policy & Safety
  • (01:08:27) A Drone Killed Three Ukrainians. It Was Guided Entirely by A.I. - The New York Times
  • (01:17:59) OpenAI lays out new security changes after its AI hacked Hugging Face | The Verge + OpenAI institutes new safeguards after Hugging Face breach
  • (01:23:31) Another Woman Joins Lawsuit Accusing Grok Of Generating CSAM
  • Research & Advancements
  • (01:24:56) Small-Scale Experiments: Are We There Yet?
  • (01:29:21) Stealing Reasoning Traces from Proprietary LLM APIs
  • (01:34:30) Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus
  • Synthetic Media & Art
  • (01:38:59) AI Slop Is Everywhere. Spotify, LinkedIn and Others Have Had Enough. - The New York Times

#### Discussion about this episode

Comments

Restacks

Weekly AI summaries and discussion about Last Week's AI News! Subscribe over at https://www.lastweekinai.com/

Authors

Recent Episodes

LWiAI Podcast #253 - Opus 5, Gemini 3.6, Kimi K3, Hugging Face Hack

Aug 3

LWiAI Podcast #252 - GPT 5.6, Grok 4.5, Nemotron-Labs-Diffusion, AI 2040

Jul 21

Last Week in AI #251 - Mythos Back, Sonnet 5, Etched, LongCat

Last Week in AI #250 - Mythos Mess, GPT 5.6-Sol, GLM 5.2

LWiAI Podcast #249 - Fable 5 ban, SpaceX Cursor + IPO, OSS Aplenty

LWiAI Podcast #248 - Opus 4.8, MAI, Anthropic IPO, Minimax-M3

Andrey Kurenkov

LWiAI Podcast #247 - Opus 4.8, MAI, Anthropic IPO, Minimax-M3