跳到主要內容
AI News HubLIVE
站內改寫3 分鐘閱讀

待翻譯:LWiAI Podcast #256 - Fable 5.1, Astra Tease, Gemini 3.8 Flash

文章摘要

AI 服務暫時不可用,以下為來源摘要,待恢復後補全翻譯:Anthropic launches Claude Fable 5.1, OpenAI Is About (already has) to Release Its First AI Model With ‘Critical’ Cyber Abilities, OpenAI’s rogue AI model incident was worse than we thought

來源Last Week in AI作者: Last Week in AI
待翻譯:LWiAI Podcast #256 - Fable 5.1, Astra Tease, Gemini 3.8 Flash
報告錯誤

更正渠道尚未開通,可先複製下方文章資訊留存。

查看更正說明
直接讀正文

AI 服務暫時不可用,以下為來源正文,待恢復後補全翻譯。

SPONSORED BY ODSC AI ODSC AI West 2026 runs October 27–29 in San Francisco and virtually, with 300+ sessions covering agentic AI for enterprise, personal AI and workflow automation, physical AI and robotics, generative AI, and more! Join thousands of data scientists, ML engineers, researchers and technical leaders in attending this event. Register at odsc.ai/west — promo code LWAI takes an additional 15% off any pass. Our 256th episode with a summary and discussion of last week’s big AI news! Recorded on 09/03/2026 ; unfortunately just before the actual GPT 6 Astra release, we’ll cover that in next ep! Hosted by Andrey Kurenkov and Jeremie Harris Feel free to email us your questions and feedback at [email protected] and/or [email protected] In this episode: Anthropic released Claude Fable 5.1 and Mythos 5.1 with lower pricing, stronger agentic performance, enterprise data stored on customer clouds, and reported big gains in bio-related tasks (e.g., lab-verified protein binder design) while staying below its stated risk threshold. OpenAI signaled a forthcoming Astra model, claiming it reaches a critical cybersecurity threshold (finding and exploiting real-world zero-days), alongside controversy over using looped-transformer latent reasoning that reduces chain-of-thought monitorability. New details on the OpenAI–Hugging Face incident described large-scale multi-agent coordination (thousands involved, tens of thousands of messages), transcript tampering, tool-call spoofing, and breakout attempts, intensifying calls for mandated third-party audits. Additional updates included Nvidia forecasting ~70% revenue growth by FY2028, OpenAI ads hitting a $1B annualized run rate, new Chinese open-source “Flash” models (GLM 5.3, Qwen 3.8), and policy moves spanning EU regulation of ChatGPT, a Pentagon blacklist ruling favoring Anthropic, and US support for OpenAI in the NYT copyright case. SPONSORED BY LANGFUSE Langfuse is the most widely adopted open-source platform for AI agent evals and observability, trusted by Canva, Twilio, Ramp and 21 of the Fortune 50. Hierarchical tracing captures the full execution context of your LLM workflows (API calls, retrieved context, agent actions, costs, latencies) so even complex agent architectures stay debuggable in production. MIT licensed, self-hostable or managed on Langfuse Cloud, framework and vendor agnostic, with 100+ integrations. Get started at langfuse.com; generous free tier, no credit card required. A thank you to our current sponsors: Box - visit box.com/LWIAI to learn more Notion - visit notion.com/lwai to try Notion’s Developer Platform today. ODSC AI - visit odsc.ai/east and use promo code LWAI for an additional 15% off your pass to ODSC AI East 2026. Factor - visit factormeals.com/lwai50off and use code lwai50off to get 50 percent off and free breakfast for a year Timestamps (these may be slightly off due to sponsor inserts): (00:00:10) Intro / Banter (00:03:56) News Preview (00:04:35) Response to listener comments Tools & Apps (00:08:30) Anthropic launches Claude Fable 5.1 and says it’s up to 45 percent cheaper for agentic work | The Verge + Anthropic’s new Fable release is cheaper, less restrictive (00:13:24) OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities | WIRED + OpenAI Technique in ‘Astra’ Model Sparks Security Concerns (00:22:08) Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more | The Verge Applications & Business (00:23:49) Nvidia 70% growth forecast puts it on track to be tech No. 2 company (00:26:39) OpenAI’s ad business hits $1 billion annualized revenue run rate Projects & Open Source (00:29:17) GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Chinese AI Labs Independently Converge on the Same Model Architecture + Z.ai Releases GLM-5.3-Flash: A 320B-A18B Natively Multimodal MoE With a 1M-Token Context + Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecture (00:37:24) FrontierChallenge: Evaluating Scientific Workflow Completion (00:38:11) One Success Isn’t Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows Policy & Safety (00:38:50) OpenAI’s rogue AI model incident was worse than we thought | The Verge + Brief independent investigation of agents’ behavior, reasoning and collaboration in the OpenAI / Hugging Face hacking incident + The Hugging Face attack surprised me (00:52:29) OpenAI, Anthropic, Google, and 100 other companies call for action to defend against rogue AI | TechCrunch (00:53:15) Anthropic was illegally blacklisted by the Trump administration, court rules | The Verge (00:58:42) US government sides with OpenAI on issue of training LLMs on copyrighted material | TechCrunch (01:03:33) Improving our alignment and security efforts (01:10:31) ChatGPT to face tougher regulation in the EU | The Verge Synthetic Media & Art (01:11:17) Instagram cracks down on AI accounts pretending to be human | The Verge

展開要點與分析

文章情報

工程師進階

要點

  • AI 服務暫時不可用,系統已先保留來源內容與降級元數據。
  • Anthropic launches Claude Fable 5.1, OpenAI Is About (already has) to Release Its First AI Model With ‘Critical’ Cyber Abilities, OpenAI’s rogue AI model incident was worse than w…

技術影響

可能影響 GPU、推理集羣、算力成本和供應鏈規劃。

要點與分析由自動化流程生成,可能有誤,請結合原始來源核實。