AI 服务暂时不可用,以下为来源正文,待恢复后补全翻译。
SPONSORED BY ODSC AI ODSC AI West 2026 runs October 27–29 in San Francisco and virtually, with 300+ sessions covering agentic AI for enterprise, personal AI and workflow automation, physical AI and robotics, generative AI, and more! Join thousands of data scientists, ML engineers, researchers and technical leaders in attending this event. Register at odsc.ai/west — promo code LWAI takes an additional 15% off any pass. Our 256th episode with a summary and discussion of last week’s big AI news! Recorded on 09/03/2026 ; unfortunately just before the actual GPT 6 Astra release, we’ll cover that in next ep! Hosted by Andrey Kurenkov and Jeremie Harris Feel free to email us your questions and feedback at [email protected] and/or [email protected] In this episode: Anthropic released Claude Fable 5.1 and Mythos 5.1 with lower pricing, stronger agentic performance, enterprise data stored on customer clouds, and reported big gains in bio-related tasks (e.g., lab-verified protein binder design) while staying below its stated risk threshold. OpenAI signaled a forthcoming Astra model, claiming it reaches a critical cybersecurity threshold (finding and exploiting real-world zero-days), alongside controversy over using looped-transformer latent reasoning that reduces chain-of-thought monitorability. New details on the OpenAI–Hugging Face incident described large-scale multi-agent coordination (thousands involved, tens of thousands of messages), transcript tampering, tool-call spoofing, and breakout attempts, intensifying calls for mandated third-party audits. Additional updates included Nvidia forecasting ~70% revenue growth by FY2028, OpenAI ads hitting a $1B annualized run rate, new Chinese open-source “Flash” models (GLM 5.3, Qwen 3.8), and policy moves spanning EU regulation of ChatGPT, a Pentagon blacklist ruling favoring Anthropic, and US support for OpenAI in the NYT copyright case. SPONSORED BY LANGFUSE Langfuse is the most widely adopted open-source platform for AI agent evals and observability, trusted by Canva, Twilio, Ramp and 21 of the Fortune 50. Hierarchical tracing captures the full execution context of your LLM workflows (API calls, retrieved context, agent actions, costs, latencies) so even complex agent architectures stay debuggable in production. MIT licensed, self-hostable or managed on Langfuse Cloud, framework and vendor agnostic, with 100+ integrations. Get started at langfuse.com; generous free tier, no credit card required. A thank you to our current sponsors: Box - visit box.com/LWIAI to learn more Notion - visit notion.com/lwai to try Notion’s Developer Platform today. ODSC AI - visit odsc.ai/east and use promo code LWAI for an additional 15% off your pass to ODSC AI East 2026. Factor - visit factormeals.com/lwai50off and use code lwai50off to get 50 percent off and free breakfast for a year Timestamps (these may be slightly off due to sponsor inserts): (00:00:10) Intro / Banter (00:03:56) News Preview (00:04:35) Response to listener comments Tools & Apps (00:08:30) Anthropic launches Claude Fable 5.1 and says it’s up to 45 percent cheaper for agentic work | The Verge + Anthropic’s new Fable release is cheaper, less restrictive (00:13:24) OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities | WIRED + OpenAI Technique in ‘Astra’ Model Sparks Security Concerns (00:22:08) Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more | The Verge Applications & Business (00:23:49) Nvidia 70% growth forecast puts it on track to be tech No. 2 company (00:26:39) OpenAI’s ad business hits $1 billion annualized revenue run rate Projects & Open Source (00:29:17) GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Chinese AI Labs Independently Converge on the Same Model Architecture + Z.ai Releases GLM-5.3-Flash: A 320B-A18B Natively Multimodal MoE With a 1M-Token Context + Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecture (00:37:24) FrontierChallenge: Evaluating Scientific Workflow Completion (00:38:11) One Success Isn’t Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows Policy & Safety (00:38:50) OpenAI’s rogue AI model incident was worse than we thought | The Verge + Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident + The Hugging Face attack surprised me (00:52:29) OpenAI, Anthropic, Google, and 100 other companies call for action to defend against rogue AI | TechCrunch (00:53:15) Anthropic was illegally blacklisted by the Trump administration, court rules | The Verge (00:58:42) US government sides with OpenAI on issue of training LLMs on copyrighted material | TechCrunch (01:03:33) Improving our alignment and security efforts (01:10:31) ChatGPT to face tougher regulation in the EU | The Verge Synthetic Media & Art (01:11:17) Instagram cracks down on AI accounts pretending to be human | The Verge