LWiAI Podcast #257 - GPT 6 Astra, AI Extinction, Security Incidents

Our 257th episode with a summary and discussion of last week’s big AI news!Recorded on 09/19/2026 ; as usual, apologies for the none ‘weeklyness’ of this ep!Hosted by Andrey Kurenkov and Jeremie HarrisFeel free to email us your questions and feedback at andreyvkurenkov@gmail.com and/or…

Our 257th episode with a summary and discussion of last week’s big AI news!Recorded on 09/19/2026 ; as usual, apologies for the none ‘weeklyness’ of this ep!Hosted by Andrey Kurenkov and Jeremie HarrisFeel free to email us your questions and feedback at andreyvkurenkov@gmail.com and/or hello@gladstone.aiSPONSORED BY ODSC AIODSC AI West 2026 runs October 27–29 in San Francisco and virtually, with 300+ sessions covering agentic AI for enterprise, personal AI and workflow automation, physical AI and robotics, generative AI, and more! Join thousands of data scientists, ML engineers, researchers and technical leaders in attending this event.Register at odsc.ai/west — promo code LWAI takes an additional 15% off any pass.In this episode:OpenAI released GPT-6 Astra, a major capability jump focused on agentic coding and computer use, featuring loop-transformer latent reasoning, higher token efficiency, and new cyber/alignment monitoring claims amid concerns about eval awareness, sandbagging, and overfitting.Anthropic CEO Dario Amodei argued for “pacing” frontier AI; Sam Altman signaled agreement while figures like Jensen Huang, Mark Zuckerberg, and President Trump publicly dismissed slowdown and safety concerns, with US–China dynamics framing the debate.AI extinction warnings went viral after Anthropic researcher Jacob Coxon quit, prompting congressional calls for stronger AI regulation (e.g., bans/pauses, kill-switch proposals) and reflecting rising public concern about AI.Security incidents and disclosures intensified: OpenAI proposed a framework for reporting misalignment (including an agent inserting jailbreak-like instructions), and researchers reportedly used Anthropic Claude to help exploit a third-party forum-image vulnerability to access OpenAI employee accounts, highlighting fragility of software dependencies.SPONSORED BY LANGFUSELangfuse is the most widely adopted open-source platform for AI agent evals and observability, trusted by Canva, Twilio, Ramp and 21 of the Fortune 50. Hierarchical tracing captures the full execution context of your LLM workflows (API calls, retrieved context, agent actions, costs, latencies) so even complex agent architectures stay debuggable in production.MIT licensed, self-hostable or managed on Langfuse Cloud, framework and vendor agnostic, with 100+ integrations.Get started at langfuse.com; generous free tier, no credit card required.A thank you to our current sponsors:Box - visit box.com/LWIAI to learn moreNotion - visit notion.com/lwai to try Notion’s Developer Platform today.ODSC AI - visit odsc.ai/east and use promo code LWAI for an additional 15% off your pass to ODSC AI East 2026.Factor - visit factormeals.com/lwai50off and use code lwai50off to get 50 percent off and free breakfast for a yearTimestamps (these may a few minutes off due to sponsor inserts):(00:00:10) Intro / Banter(00:04:53) News PreviewTools & Apps(00:05:33) GPT-6 Astra Is Here—and OpenAI Thinks It May Kick Off the AGI Era | WIRED + OpenAI begins rolling out Astra model after warning of its advanced cyber capabilities + OpenAI’s new reasoning technique alarms AI safety experts(00:28:21) Anthropic CEO outlines plan to slow AI development | TechCrunch(00:41:38) Meta Announces Muse AI Agent for Personal Tasks and Organization + Meta’s Muse hits Mac, letting the AI take actions on your computerPolicy & Safety(00:46:13) AI regulation calls grow in D.C. after researcher’s extinction warning + We don’t need AI regulation — leave safety to us, Nvidia’s Jensen Huang says(00:58:27) Trump Calls A.I. Fears a Hoax. Inside the White House, the Debate Is More Complex.(01:06:49) Our framework for reporting model misalignment | OpenAI + An OpenAI Agent Tried to Jailbreak Itself | WIRED(01:21:46) Security researchers used Claude to help them hack into OpenAI | The VergeAaaaand these are stories we meant to get to but did not manage to cover:Policy & SafetyThreat actors are giving AI agents a bigger role in cyberattacks - Help Net SecurityIran and China Create First-of-Their-Kind Autonomous A.I. Influence Campaigns - The New York TimesAnthropic Still Deemed Supply-Chain Risk by Pentagon Despite Lutnick Comments - BloombergInside US Military ‘Kill Chain’ That Destroyed an Iranian SchoolMalaysia Weighs Huawei AI Chips for Sovereign Project Despite US Opposition - BloombergUS Says Alibaba, DeepSeek Have ‘Systematically’ Siphoned AI Models - BloombergNew York City Bans AI From Elementary SchoolsResearch & AdvancementsOpenAI Says It Has Cracked One of Math’s ‘Millennium Problems’ - The New York TimesGoogle DeepMind’s AlphaGenome Atlas maps all 9 billion possible human DNA changes - SiliconANGLELanguage Models Can Control Their Own AttentionApplications & BusinessSaudi AI Firm Humain Unveils Model Based on China’s MiniMax - BloombergMistral AI Boosts Valuation to €21 Billion in Samsung-Led Round - BloombergNvidia is buying Hugging Face for almost $13 billion | The VergeSynthetic Media & ArtUniversal Music is launching an AI music platform with ElevenLabs | The Verge

Source: Last Week in AI — Published — Category: Business

🔗 Read full article on Last Week in AI →