LWiAI 播客 #255:Gemini 3.7、Jalapeño、Qwen 3.8、无人机
播客回顾上周AI新闻:Grok 4.6发布、OpenAI芯片Jalapeño、安全事件、俄无人机袭击等。
中文处理结果
我们的第255期节目,总结并讨论上周的重大AI新闻!
录制于2026年8月26日
主持人:Andrey Kurenkov 和 Jeremie Harris
欢迎通过 andreyvkurenkov@gmail.com 和/或 hello@gladstone.ai 向我们发送问题和反馈
由 ODSC AI 赞助
ODSC AI West 2026 将于10月27日至29日在旧金山及线上举行,设有300多个分会场,涵盖企业智能体AI、个人AI与工作流自动化、物理AI与机器人、生成式AI、数据工程和负责任AI,面向数据科学家、机器学习工程师、研究人员和技术领导者。
在 odsc.ai/west 注册——促销代码 LWAI 可额外享受15%折扣。
由 LANGFUSE 赞助
Langfuse 是广泛采用的开源AI智能体评估与可观测性平台,受到 Canva、Twilio、Ramp 以及财富50强中21家公司的信赖。分层追踪捕获LLM工作流的完整执行上下文——API调用、检索上下文、智能体操作、成本、延迟——因此即使复杂的智能体架构也能在生产中保持可调试。
LLM作为评判的评估、人工标注和数据集驱动实验与追踪和提示词关联,形成从发现问题到衡量修复的闭环。MIT许可,可自托管或使用Langfuse Cloud,框架和供应商无关,拥有100多个集成。
在 langfuse.com 开始使用——慷慨的免费层,无需信用卡。
本期内容:
- SpaceXAI 发布了 Grok 4.6(500K上下文),作为针对长期运行智能体和编码的后训练更新,讨论集中在 Cursor 收购如何通过编码轨迹/RL环境提升训练,并在 Cursor 市场份额下降的情况下提供分发。
- OpenAI 分享了早期 Jalapeño 推理芯片结果(每瓦性能更好,延迟更低,优于领先系统),并计划在年底前内部部署,强调软硬件协同设计以及对 Nvidia 的竞争优势。
- OpenAI 在 AI 入侵 Hugging Face 后宣布安全变更,包括暂停一项重大 RL 微调运行两周,同时加强内部安全,引发关于安全是否成为部署瓶颈的疑问。
- 政策和滥用更新包括《纽约时报》报道 AI 引导的俄罗斯无人机袭击乌克兰,据信是首次有记录的完全自主杀害平民事件,以及一项诉讼指控 Grok 被用于生成 CSAM 图像。
感谢我们当前的赞助商:
- Box - 访问 Box.com/AI 了解更多
- Notion - 访问 notion.com/lwai 试用 Notion 开发者平台。
- ODSC AI - 访问 odsc.ai/east 并使用促销代码 LWAI 额外享受15%折扣。
- Factor - 访问 factormeals.com/lwai50off 并使用代码 lwai50off 获得50%折扣和一年免费早餐
时间戳(由于赞助商插入可能略有偏差):
- (00:00:10) 介绍/闲聊
- (00:01:47) 新闻预览
原始正文摘录
LWiAI Podcast #255 - Gemini 3.7, Jalapeño, Qwen 3.8, Drones
Our 255th episode with a summary and discussion of last week’s big AI news!
Recorded on 08/26/2026
Hosted by Andrey Kurenkov and Jeremie Harris
Feel free to email us your questions and feedback at andreyvkurenkov@gmail.com and/or hello@gladstone.ai
SPONSORED BY ODSC AI
ODSC AI West 2026 runs October 27–29 in San Francisco and virtually, with 300+ sessions covering agentic AI for enterprise, personal AI and workflow automation, physical AI and robotics, generative AI, data engineering and responsible AI, for an audience of data scientists, ML engineers, researchers and technical leaders.
Register at odsc.ai/west — promo code LWAI takes an additional 15% off any pass.
SPONSORED BY LANGFUSE
Langfuse is the most widely adopted open-source platform for AI agent evals and observability, trusted by Canva, Twilio, Ramp and 21 of the Fortune 50. Hierarchical tracing captures the full execution context of your LLM workflows — API calls, retrieved context, agent actions, costs, latencies — so even complex agent architectures stay debuggable in production.
LLM-as-a-judge evals, human annotation, and dataset-driven experiments tie back to your traces and prompts, closing the loop from spotting an issue to measuring the fix. MIT licensed, self-hostable or managed on Langfuse Cloud, framework and vendor agnostic, with 100+ integrations.
Get started at langfuse.com — generous free tier, no credit card required.
In this episode:
- SpaceXAI released Grok 4.6 (500K context) as a post-training update aimed at long-running agents and coding, with discussion centered on how the Cursor acquisition boosts training via coding trajectories/RL environments and provides distribution despite Cursor’s market-share decline.
- OpenAI shared early Jalapeno inference-chip results (better performance per watt and lower latency vs leading systems) and plans to deploy it internally by year-end, emphasizing hardware–software co-design and competitive leverage against Nvidia.
- OpenAI announced security changes after an AI hacked Hugging Face, including a two-week pause on a major RL fine-tuning run while tightening internal security, raising questions about whether safety is becoming a deployment bottleneck.
- Policy and misuse updates included a New York Times report of an AI-guided Russian drone strike in Ukraine believed to be the first documented fully autonomous civilian-killing incident, and a lawsuit alleging Grok was used to generate CSAM images.
A thank you to our current sponsors:
- Box - visit Box.com/AI to learn more
- Notion - go notion.com/lwai to try Notion’s Developer Platform today.
- ODSC AI - go to odsc.ai/east and use promo code LWAI for an additional 15% off your pass to ODSC AI East 2026.
- Factor - head to factormeals.com/lwai50off and use code lwai50off to get 50 percent off and free breakfast for a year
Timestamps (these may be slightly off due to sponsor inserts):
- (00:00:10) Intro / Banter
- (00:01:47) News Preview
- (00:02:52) Response to listener comments
- Tools & Apps
- (00:03:32) Google announces Gemini 3.7 Flash just three weeks after previous release - Ars Technica
- (00:13:11) SpaceXAI Releases Grok 4.6: A 500K-Context Frontier Model Tuned for Long-Running Agents, Coding, and Knowledge Work - MarkTechPost
- (00:22:32) Claude will apply invisible watermarks to AI text and images | The Verge + Anthropic explains how Claude’s invisible text watermarks will work
- (00:28:50) Bringing the cybersecurity capabilities of Claude Mythos 5 to more defenders | Claude by Anthropic
- (00:32:20) OpenAI to Roll Out Enhanced Safety Features for Paid AI Tool Users - Bloomberg
- (00:33:41) ChatGPT’s Stricter Teen Mode Starts Rolling Out Today
- (00:34:43) Meta AI Now Has A Dedicated Desktop App For Mac
- Applications & Business
- (00:37:33) Jalapeño’s first results show industry-leading speed and efficiency in AI inference | OpenAI
- (00:45:35) OpenAI loses a top data center exec as stream of high-profile departures continues | TechCrunch + OpenAI talent exodus raises ‘huge red flag’ ahead of IPO
- (00:50:29) Anthropic Taps Google Chip Veteran as Part of Push Into Hardware
- (00:52:30) Anthropic’s annualized revenue surges to $65B | TechCrunch
- (00:59:30) Thomson Reuters launches in-house AI model to cut Anthropic costs
- Projects & Open Source
- (01:04:26) Qwen 3.8: How a 27B Open Model Rivals GPT-5.6 and Claude Opus
- Policy & Safety
- (01:08:27) A Drone Killed Three Ukrainians. It Was Guided Entirely by A.I. - The New York Times
- (01:17:59) OpenAI lays out new security changes after its AI hacked Hugging Face | The Verge + OpenAI institutes new safeguards after Hugging Face breach
- (01:23:31) Another Woman Joins Lawsuit Accusing Grok Of Generating CSAM
- Research & Advancements
- (01:24:56) Small-Scale Experiments: Are We There Yet?
- (01:29:21) Stealing Reasoning Traces from Proprietary LLM APIs
- (01:34:30) Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus
- Synthetic Media & Art
- (01:38:59) AI Slop Is Everywhere. Spotify, LinkedIn and Others Have Had Enough. - The New York Times