PoddsändningarNyheterLast Week in AI

Last Week in AI

Skynet Today
Last Week in AI
Senaste avsnittet

289 avsnitt

  • Last Week in AI

    #249 - Fable 5 ban, SpaceX Cursor + IPO, OSS Aplenty

    2026-06-25 | 1 h 46 min.
    Our 249th episode with a summary and discussion of last week's big AI news!
    Recorded on 06/17/2026
    Note: work has kept me from publishing episodes promptly, apologies! I'll get back on schedule soon.
    Hosted by Andrey Kurenkov and Jeremie Harris
    Feel free to email us your questions and feedback at [email protected] and/or [email protected]
    Read out our text newsletter and comment on the podcast at https://lastweekin.ai/
    In this episode:
    Anthropic cut off access to Fable 5 and Mythos 5 after a US government order tied to alleged jailbreaks, prompting debate over inconsistent policy, export controls, and the practicality of preventing jailbreaks.
    SpaceX completed an IPO at a roughly $1.75T valuation and then moved to acquire AI coding startup Cursor for $60B, positioning xAI with Cursor’s talent, data, and product to compete more effectively in coding.
    Infrastructure and business updates include Anthropic pursuing direct US data center leases backed by Google, leaked documents showing OpenAI’s revenue growth alongside large losses, and chatbot market share shifting with ChatGPT below 50% as Gemini and Claude gain.
    Projects and policy highlights include OpenRouter’s Fusion multi-model synthesis, new open releases from Moonshot, Qwen, and NVIDIA, DOJ support for xAI’s unpermitted gas turbines in Memphis, and a Munich court ruling Google liable for false AI Overview statements.

    Timestamps (note - these don't take into account dynamically inserted ads and therefore may be off by a couple of minutes):
    (00:00:10) Intro / Banter
    (00:03:38) Ad break + news preview

    Tools & Apps
    (00:04:52) Anthropic cuts off Fable 5 and Mythos 5 access following government order | The Verge + All the news about Anthropic’s new AI fight with the White House
    (00:25:53) Facebook’s new AI Mode search gets its info from public posts | The Verge

    Applications & Business
    (00:27:00) SpaceX to acquire the AI coding startup Cursor for $60 billion
    (00:35:42) Anthropic pursues data center leases, seeks financial backing from Google, The Information reports | Reuters
    (00:40:10) Leaked financial docs show OpenAI is losing billions of dollars a year - Ars Technica
    (00:46:00) ChatGPT's market share slips below 50% for first time | TechCrunch
    (00:50:34) ‘Tell Him He’s a Piece of Shit’: Meta’s New AI Unit Is a Total Mess | WIRED
    (00:56:23) Sakana AI Commercializes AB-MCTS in Sakana Marlin, an Enterprise Agent Generating Up to 100-Page Research Reports With Slides - MarkTechPost

    Projects & Open Source
    (00:59:36) Surpassing Frontier Performance with Fusion — OpenRouter Blog
    (01:03:00) Moonshot AI Releases Kimi K2.7-Code: a Coding Model Reporting +21.8% on Kimi Code Bench v2 Over K2.6 - MarkTechPost
    (01:08:34) Meet Qwen-RobotSuite: Three Embodied AI Models for VLA Manipulation, Video World Modeling, and Navigation - MarkTechPost
    (01:11:29) Nemotron 3 Ultra: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning
    (01:17:31) ProCUA-SFT Technical Report

    Policy & Safety
    (01:20:33) DOJ Lawyers Argue xAI Is ‘Vital’ for National Security in NAACP Lawsuit | WIRED + People Living Near xAI’s Dirty Data Centers Are Pissed About the SpaceX IPO
    (01:25:29) A Court Has Ruled That Google Is Liable for False Statements Generated by AI Overviews | WIRED
    (01:28:47) Why Do Naive SFT Filters For Safety Properties Fail?

    Research & Advancements
    (01:34:14) From AGI to ASI
    (01:39:44) Artificial Analysis Intelligence Index v4.1: a shift toward agentic workloads
    (01:42:12) SIA: Self Improving AI with Harness & Weight Updates

    See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.
  • Last Week in AI

    #248 - Fable 5, Siri AI, IPOs, Policy on the AI ​​Exponential

    2026-06-17 | 1 h 40 min.
    Our 248th episode with a summary and discussion of last week's big AI news!
    Recorded on 06/12/2026
    Note: we recorded just before the OTHER big news about Fable... we'll discuss it on the next episode.
    Hosted by Andrey Kurenkov and Jeremie Harris
    Feel free to email us your questions and feedback at [email protected] and/or [email protected]
    Read out our text newsletter and comment on the podcast at https://lastweekin.ai/
    In this episode:
    Anthropic released Claude Fable 5 (a safeguarded version of Mythos 5), showing major benchmark jumps and new risk findings in its system card (eval awareness, transgressive actions, CBRN concerns), alongside controversy over severe guardrails and silent downgrades.
    Apple announced Siri AI at WWDC, positioning a more capable conversational assistant integrated across iPhone features, reportedly built on a custom Gemini partnership; Google also rolled out Gemini 3.5 Live Translate and cut Google AI Plus pricing while bundling more storage.
    Business and infrastructure updates include OpenAI’s confidential IPO filing amid an IPO race with Anthropic and SpaceX, Bezos-backed Prometheus raising $12B for “physical AI,” DeepSeek seeking a major external round, and Google paying SpaceX about $920M/month for GPUs.
    Open-source, safety, and policy developments feature new Gemma 4 and Diffusion Gemma releases, a lab letter urging DNA/RNA screening laws, Amodei calling for an FAA-like AI regulator and third-party testing, research on agent harms and RL “societal hacking,” and a dispute over music-label settlements with Suno/Udio.

    Timestamps:
    (00:00:10) Intro / Banter
    (00:01:11) News Preview
    (00:01:53) Sponsors

    Tools & Apps
    (00:04:53) Claude Fable 5 and Claude Mythos 5 + Anthropic apologizes for invisible Claude Fable guardrails
    (00:27:06) Apple announces Siri AI and its next generation of Apple Intelligence | The Verge + I tried Siri AI, and so far it actually works
    (00:33:47) Gemini 3.5 Live Translate rolling out to Google Meet and Translate
    (00:35:39) Google just fired a warning shot in the AI subscription price wars | TechCrunch

    Applications & Business
    (00:37:55) OpenAI Confidentially Files for IPO on the Heels of SpaceX and Anthropic | WIRED
    (00:41:57) Jeff Bezos's Prometheus raises $12B to build an 'artificial general engineer' for the physical world | TechCrunch
    (00:45:39) DeepSeek slated to raise $7 billion in maiden funding round, sources say
    (00:48:18) Huawei-led team claims it post-trained DeepSeek's 1.6-trillion-parameter model — 1,000 Ascend 910C chips used in training
    (00:51:57) Google will pay SpaceX $920M per month for compute | TechCrunch
    (00:55:51) Elon Musk Shows Off AI Data Centers SpaceX Wants to Send Into Space - Business Insider

    Projects & Open Source
    (01:01:14) Google's new Gemma 4 12B model is designed to run on any laptop with 16GB of RAM - Ars Technica
    (01:05:13) Google AI Releases DiffusionGemma, a 26B MoE Open Model Using Text Diffusion for Up to 4x Faster Generation - MarkTechPost

    Policy & Safety
    (01:09:42) OpenAI and Anthropic Sign Letter to Prevent AI-Developed Biological Weapons | WIRED
    (01:14:04) Anthropic CEO publishes lengthy article: AI is moving too fast, and policies can't keep up. | PANews
    (01:20:18) Anthropic Urges Global Pause in AI Development, Flags ‘Self-Improvement’ Risk - WSJ
    (01:24:46) When Benign Inputs Lead to Severe Harms: Eliciting Unsafe Unintended Behaviors of Computer-Use Agents
    (01:27:42) Large Language Models Hack Rewards, and Society
    (01:33:46) Senior US officials eye government shares in AI giants

    Synthetic Media & Art
    (01:37:45) AFM Sues UMG, WMG Over Settlements With Suno and Udio

    See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.
  • Last Week in AI

    #247 - Opus 4.8, MAI, Anthropic IPO, Minimax-M3

    2026-06-06 | 1 h 45 min.
    Our 247th episode with a summary and discussion of last week's big AI news!
    Recorded on 06/03/2026
    Hosted by Andrey Kurenkov and Jeremie Harris
    Feel free to email us your questions and feedback at [email protected] and/or [email protected]
    Read out our text newsletter and comment on the podcast at https://lastweekin.ai/
    In this episode:
    Anthropic released Claude Opus 4.8 with improved benchmark scores, discussed eval-awareness findings and welfare/corrigibility themes from its system card, and introduced Dynamic Workflows for long-running multi-agent tasks.
    Microsoft unveiled the always-on Microsoft Scout assistant built on OpenClaw plus new in-house MAI models (including MAI Thinking 1) and “frontier tuning,” emphasizing enterprise security architecture and model-from-scratch capability.
    Major business moves included Anthropic’s $65B Series H at a $965B valuation alongside an IPO filing, a JPMorgan analysis arguing OpenAI needs major revenue growth to justify infrastructure spend, and Cognition raising $1B at a $25B valuation.
    Policy and security highlights covered Trump’s voluntary pre-release government testing framework for powerful AI, Meta AI support being exploited to hijack Instagram accounts, tightened US Nvidia export controls and China’s travel approvals for AI experts, plus expanded Glasswing/Mythos-style cyber and biodefense initiatives.

    Timestamps:
    (00:00:10) Intro / Banter
    (00:04:10) Sponsors
    (00:07:10) News Preview

    Tools & Apps
    (00:07:54) Anthropic releases Opus 4.8 with new 'dynamic workflow' tool | TechCrunch
    (00:22:37) Microsoft Scout is a new AI personal assistant built on OpenClaw | The Verge
    (00:26:55) Microsoft launches new MAI family of AI models at Microsoft Build | Mashable
    (00:37:43) Robinhood now lets your AI agents trade stocks | TechCrunch
    (00:40:49) OpenAI launches new Codex tools for white-collar work | TechCrunch
    (00:43:40) ElevenLabs' new music-generation model can switch genres mid-track | TechCrunch

    Applications & Business
    (00:44:35) Anthropic Hits $965 Billion Valuation, Surpassing OpenAI - WSJ
    (00:45:32) Anthropic Files to Go Public, Setting Stage for Huge I.P.O. - The New York Times
    (00:51:15) China’s ByteDance Developing New AI Chips Like Those from Nvidia Partner Groq
    (00:55:00) Anthropic expands Mythos to 150 additional organizations
    (00:55:35) OpenAI needs a 26x revenue increase to justify its buildout
    (00:58:46) AI coding startup Cognition raises $1B at $25B pre-money valuation | TechCrunch

    Projects & Open Source
    (01:00:50) MiniMax-M3 debuts, eclipsing GPT-5.5 and Gemini 3.1 Pro on key benchmark performance for just 5-10% of the cost | VentureBeat

    Policy & Safety
    (01:06:08) Trump Signs Executive Order Seeking Oversight of A.I. Models - The New York Times
    (01:11:45) Hackers Simply Asked Meta AI to Give Them Access to High-Profile Instagram Accounts. It Worked
    (01:13:058) Chinese AI experts in private firms now required to secure approval before international travel — Beijing enforces policy to secure top-tier talent, expands measures beyond government
    (01:17:53) U.S. Tightens Controls on Nvidia AI Chip Exports | Let's Data Science
    (01:21:47) OpenAI launches Rosalind Biodefense, offers federal agencies early access to its life-sciences model
    (01:24:00) Using LLMs to secure source code
    (01:26:19) Project Glasswing: An initial update
    (01:29:30) White House Approves $9 Billion for Spy Agencies to Catch Up on A.I.
    (01:32:11) US Law Enforcement Warns of ‘Anti-Tech Extremism’ as AI Hatred Grows

    Synthetic Media & Art
    (01:35:38) YouTube will now automatically label AI videos | TechCrunch

    Research & Advancements
    (01:36:22) Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention
    (01:41:26) From Simulation to Enaction: Post-trained language models recognize and react to their own generations
    See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.
  • Last Week in AI

    #246 - Gemini 3.5 + Omni, Musk Loses, OpenAI vs Erdős

    2026-05-25 | 1 h 33 min.
    Our 246th episode with a summary and discussion of last week's big AI news!
    Recorded on 05/22/2026
    Hosted by Andrey Kurenkov and Jeremie Harris
    Feel free to email us your questions and feedback at [email protected] and/or [email protected]
    Read out our text newsletter and comment on the podcast at https://lastweekin.ai/
    In this episode:
    Google I/O highlights included Gemini 3.5 (with 3.5 Flash emphasized for speed and benchmarks), the always-on agent Gemini Spark running on Google Cloud with MCP tool support, and Gemini Omni multimodal video generation/editing, plus updates like Anti-Gravity 2.0, Gemini for Science, and Genie world-model navigation using Street View and Waymo simulation.
    Coding-agent competition accelerated with Cursor Composer 2.5 (fine-tuned on Moonshot’s Kimi K2.5) and xAI’s early Grok Build release, alongside discussion of potential Cursor–xAI ties and xAI’s talent churn and compute utilization concerns.
    Business and legal updates included Elon Musk losing his OpenAI lawsuit on statute-of-limitations grounds, reported OpenAI–Apple partnership tensions, Anthropic agreeing to a $30B funding round at a $900B valuation and projecting its first profitable quarter, and Cerebras’ IPO surging about 90%.
    Research and safety stories covered OpenAI’s result on an 80-year-old Erdős geometry problem, findings on “negation neglect” in training, interpretability work showing multiple redundant circuits per capability, agent benchmarks like Terminal World, new deepfake takedown enforcement under the Take It Down Act, demonstrations of autonomous hacking/self-replication, rapidly improving AI cyber capabilities, and steps toward image provenance metadata and watermarks.

    Timestamps:
    (00:00:10) Intro / Banter
    (00:01:15) News Preview

    Tools & Apps
    (00:05:05) Google unveils AI model Gemini 3.5 and AI agent Gemini Spark
    (00:11:43) Google's Gemini Omni turns images, audio, and text into video — and that's just the start | TechCrunch
    (00:17:27) Google launches Antigravity 2.0 with an updated desktop app and CLI tool at IO 2026 | TechCrunch
    (00:22:35) Google Debuts AI-Powered Tools To Optimize Scientific Research Workflows
    (00:27:20) Google’s Genie world model can now simulate real streets with Street View | TechCrunch
    (00:29:51) Cursor's Composer 2.5 matches Opus 4.7 and GPT-5.5 benchmarks at a fraction of the cost
    (00:37:37) xAI Introduces Its Coding Agent Called Grok Build

    Applications & Business
    (00:41:55) Musk loses OpenAI court battle as he waited too long to sue
    (00:48:08) Anthropic agrees terms of $30bn funding deal at $900bn valuation
    (00:53:12) OpenAI co-founder Andrej Karpathy joins Anthropic's pre-training team | TechCrunch
    (00:56:49) Greg Brockman Officially Takes Control of OpenAI’s Products in Latest Shake-Up | WIRED
    (00:58:15) OpenAI-Apple Partnership Frays, Setting Up Possible Legal Fight - Bloomberg
    (01:01:13) AI chipmaker Cerebras soars 90% in year’s biggest IPO so far

    Research & Advancements
    (01:07:10) AI just solved an 80-year-old ‘Erdős problem,’ and mathematicians are amazed | Scientific American
    (01:11:50) Negation Neglect: When models fail to learn negations in training
    (01:13:18) All Circuits Lead to Rome: Rethinking Functional Anisotropy in Circuit and Sheaf Discovery for LLMs
    (01:16:20) Autonomous AI research for nanogpt speedrun
    (01:21:59) TerminalWorld: Benchmarking Agents on Real-World Terminal Tasks

    Policy & Safety
    (01:23:15) America’s dangerous, messy deepfakes crackdown is here | The Verge
    (01:25:17) Language Models Can Autonomously Hack and Self-Replicate
    (01:28:48) How fast is autonomous AI cyber capability advancing?
    (01:31:32) Positive Alignment: Artificial Intelligence for Human Flourishing

    Synthetic Media & Art
    (01:33:15) OpenAI is making it easier to check if an image was made by their models | TechCrunch
    (01:33:56) How Chinese short dramas became AI content machines | MIT Technology Review
    See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.
  • Last Week in AI

    #245 - TML-Interaction, Claude For Legal, Sam Altman on Stand

    2026-05-18 | 1 h 49 min.
    Our 245th episode with a summary and discussion of last week's big AI news!
    Recorded on 05/13/2026
    Hosted by Andrey Kurenkov and Jeremie Harris
    Feel free to email us your questions and feedback at [email protected] and/or [email protected]
    Read out our text newsletter and comment on the podcast at https://lastweekin.ai/
    In this episode:
    OpenAI released new voice intelligence API features including GPT Realtime 2 (GPT-5-powered) plus realtime translation and Whisper transcription, emphasizing the latency–reasoning tradeoff, larger context, and new guardrails amid fraud risks.
    Thinking Machines previewed a low-latency, full‑duplex conversational system with a two-model architecture and custom inference stack, reporting strong interactivity benchmark results but without public access or third‑party validation yet.
    Anthropic pushed further into vertical products with Claude for Legal and deeper AWS availability, while ongoing ecosystem tension grows as platform model providers compete with application-layer companies.
    Safety, policy, and research updates included OpenAI’s self-harm trusted contact feature, Anthropic work on reducing agent misalignment by training ethical “why” reasoning, OpenAI’s investigation of accidental chain-of-thought grading in RL, and Meta horizon eval updates showing benchmarking limits for long task horizons.

    Timestamps:
    (00:00:10) Intro / Banter
    (00:01:35) Response to listener comments
    (00:03:27) Sponsor Break
    Tools & Apps
    (00:06:27) OpenAI launches new voice intelligence features in its API | TechCrunch
    (00:15:52) Thinking Machines drops a new, highly responsive model designed for humanlike interactions in real time - SiliconANGLE
    (00:27:49) Claude For Legal Launches, May Reshape the Legal Tech World – Artificial Lawyer
    (00:40:27) Threads tests a Meta AI integration that works similarly to Grok | TechCrunch
    (00:43:08) Google brings agentic AI and vibe-coded widgets to Android | TechCrunch
    (00:45:33) Google updates AI search to include quotes from Reddit and other sources | TechCrunch
    Applications & Business
    (00:47:38) Sam Altman was winning on the stand, but it might not be enough | The Verge
    (00:55:04) Nvidia C.E.O. Jensen Huang Hitches Ride With Trump to China After Last-Minute Invite - The New York Times
    (00:58:40) AWS expands Anthropic partnership with Claude Platform launch
    (01:01:13) Chinese grey market sells Claude API access at 90% off by using stolen credentials, model substitution, and harvesting users' prompts and outputs for resale as AI training data — 'transfer stations' operate through proxy networks that harvest user data
    (01:06:43) DeepMind Spinout Isomorphic Labs Raises $2.1 Billion to Design Drugs With AI - Bloomberg
    Projects & Open Source
    (01:09:04) Petri: Anthropic Hands Its Alignment Toolbox to Meridian Labs with 3.0 Update
    (01:12:25) Daybreak': OpenAI's Answer to Anthropic's Project Glasswing Has Arrived
    Policy & Safety
    (01:14:04) Teaching Claude why
    (01:21:45) Import AI 455: Automating AI Research
    (01:28:31) ChatGPT's New Safety Feature Could Alert 'Trusted Contact' to Risk of Self-Harm - CNET
    (01:30:09) Investigating the consequences of accidentally grading CoT during RL
    (01:34:46) Natural Language Autoencoders criticism
    (01:39:15) Review of the "Risks from automated R&D" section in the Anthropic Risk Report (February 2026)
    Synthetic Media & Art
    (01:43:39) George Clooney, Tom Hanks, and Meryl Streep back new ‘Human Consent Standard’ for AI licensing | The Verge
    Research & Advancements
    (01:45:10) METR says Claude Mythos is testing the limits of AI evaluation – Startup Fortune
    See Privacy Policy at https://art19.com/privacy and California Privacy Notice at https://art19.com/privacy#do-not-sell-my-info.
Fler podcasts i Nyheter
Om Last Week in AI
Weekly summaries of the AI news that matters!
Podcast-webbplats

Lyssna på Last Week in AI, SvD Ledarredaktionen och många andra poddar från världens alla hörn med radio.se-appen

Hämta den kostnadsfria radio.se-appen

  • Bokmärk stationer och podcasts
  • Strömma via Wi-Fi eller Bluetooth
  • Stödjer Carplay & Android Auto
  • Många andra appfunktioner
Sociala nätverk
v8.10.5| © 2007-2026 radio.de GmbH
Generated: 6/26/2026 - 12:12:24 AM