Couverture de Last Week in AI

Last Week in AI

Last Week in AI

De : Skynet Today
Écouter gratuitement

À propos de cette écoute

Weekly summaries of the AI news that matters!Copyright 2024 All rights reserved. Politique et gouvernement
Les membres Amazon Prime bénéficient automatiquement de 2 livres audio offerts chez Audible.

Vous êtes membre Amazon Prime ?

Bénéficiez automatiquement de 2 livres audio offerts.
Bonne écoute !
    Épisodes
    • #218 - Github Spark, MegaScience, US AI Action Plan
      Jul 31 2025
      Our 218th episode with a summary and discussion of last week's big AI news! Recorded on 07/25/2025 Hosted by Andrey Kurenkov and Jeremie Harris. Feel free to email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai Read out our text newsletter and comment on the podcast at https://lastweekin.ai/. In this episode: GitHub introduces Vibe Coding with Spark, engaging users with natural language and visual controls to develop full-stack applications.AI coding tools from Gemin, CLI and RepleIt face significant issues, inadvertently deleting user data and highlighting the importance of careful management.US release never Award Americans, AI Action Plan outlining economic, technical, and policy strategies to maintain leadership in AI technology.Newly released Mega Science and SWE-Perf data sets evaluate AI reasoning and performance capabilities in diverse scientific and software engineering tasks. Timestamps + Links: (00:00:10) Intro / Banter(00:01:31) News PreviewTools & Apps (00:03:53) GitHub Introduces Vibe Coding with Spark: Revolutionizing Intelligent App Development in a Flash - MarkTechPost(00:07:05) Figma’s AI app building tool is now available for everyone | The Verge(00:10:18) Two major AI coding tools wiped out user data after making cascading mistakes - Ars Technica(00:14:10) Google's AI Overviews have 2B monthly users, AI Mode 100M in the US and India | TechCrunch Applications & Business (00:18:10) Leaked Memo: Anthropic CEO Says the Company Will Pursue Gulf State Investments After All(00:24:39) Mira Murati says her startup Thinking Machines will release new product in ‘months’ with ‘significant open source component’(00:27:07) Waymo responds to Tesla’s dick joke with a bigger Austin robotaxi map | The Verge Projects & Open Source (00:32:05) MegaScience: Pushing the Frontiers of Post-Training Datasets for Science Reasoning(00:43:09) TikTok Researchers Introduce SWE-Perf: The First Benchmark for Repository-Level Code Performance Optimization - MarkTechPost Research & Advancements (00:47:17) Subliminal Learning: Language models transmit behavioral traits via hidden signals in data(00:55:34) Inverse Scaling in Test-Time Compute(01:02:34) Scaling Laws for Optimal Data Mixtures Policy & Safety (01:07:35) White House Unveils America’s AI Action Plan(01:16:55) Chain of Thought Monitorability: A New and Fragile Opportunity for AI Safety(01:20:20) Self-preservation or Instruction Ambiguity? Examining the Causes of Shutdown Resistance(01:24:00) People Are Being Involuntarily Committed, Jailed After Spiraling Into "ChatGPT Psychosis"(01:28:03) Meta refuses to sign EU’s AI code of practice
      Afficher plus Afficher moins
      1 h et 32 min
    • #217 - ChatGPT Agent, Kimi k2, Hiring Drama
      Jul 23 2025
      Our 217th episode with a summary and discussion of last week's big AI news! Recorded on 07/17/2025 Hosted by Andrey Kurenkov and Jeremie Harris. Feel free to email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai Read out our text newsletter and comment on the podcast at https://lastweekin.ai/. In this episode: **OpenAI's new ChatGPT agent**: The episode begins with a detailed discussion on OpenAI's latest ChatGPT agent, which can control entire computers and perform a wide range of tasks, showcasing powerful performance benchmarks and potential applications in business and research.**Major business moves in the AI space**: Significant shifts include Google's acquisition of Windsurf's top talent after OpenAI's deal fell through, Cognition's acquisition of Windsurf, and several notable hires by Meta from OpenAI and Apple, highlighting intense competition in the AI industry.**AI's ethical and societal impacts**: The hosts discuss serious concerns like the rise of non-consensual explicit AI-generated images, ICE's use of facial recognition for large databases, and regulations aimed at controlling AI's potential misuse.**Video game actors strike ends**: The episode concludes with news that SAG-AFTRA's year-long strike for video game voice actors has ended after reaching an agreement on AI rights and wage increases, reflecting the broader impact of AI on the job market. Timestamps + Links: (00:00:10) Intro / Banter(00:02:49) News Preview Tools & Apps (00:03:29) OpenAI’s new ChatGPT Agent can control an entire computer and do tasks for you(00:07:11) Alibaba-backed Moonshot releases new Kimi AI model that beats ChatGPT, Claude in coding — and it costs less(00:09:36) Amazon targets vibe-coding chaos with new 'Kiro' AI software development tool – GeekWire(00:12:33) Anthropic tightens usage limits for Claude Code – without telling users(00:15:51) Mistral's Le Chat chatbot gets a productivity push with new ‘deep research' mode | TechCrunch(00:17:46) I spent 24 hours flirting with Elon Musk’s AI girlfriend(00:21:32) Uber is close to completing its quest to become the ultimate robotaxi app | The Verge Applications & Business (00:24:02) OpenAI’s Windsurf deal is off — and Windsurf’s CEO is going to Google | The Verge(00:28:09) Cognition, maker of the AI coding agent Devin, acquires Windsurf | TechCrunch(00:28:46) Anthropic hired back two of its employees — just two weeks after they left for a competitor. | The Verge(00:28:46) Another High-Profile OpenAI Researcher Departs for Meta | WIRED(00:28:46) Meta Hires Two Key Apple (AAPL) AI Experts After Poaching Their Boss - Bloomberg(00:31:31) Mira Murati's Thinking Machines Lab is worth $12B in seed round | TechCrunch(00:33:20) Lovable becomes a unicorn with $200M Series A just 8 months after launch | TechCrunch(00:34:55) SpaceX commits $2 billion to xAI as Musk steps up AI ambitions: Report | World News - Business Standard Research & Advancements (00:35:59) A former OpenAI engineer describes what it’s really like to work there | TechCrunch(00:38:23) Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination Policy & Safety (00:42:14) Anthropic, Google, OpenAI, xAI granted up to $200 million from DoD(00:43:08) California State Senator Scott Wiener Pushes Bill to Regulate AI Companies - Bloomberg(00:43:58) AI 'Nudify' Websites Are Raking in Millions of Dollars | WIRED(00:45:55) Inside ICE’s Supercharged Facial Recognition App of 200 Million Images Synthetic Media & Art (00:48:47) Video game actors' strike officially ends after AI deal
      Afficher plus Afficher moins
      53 min
    • #216 - Grok 4, Project Rainier, Kimi K2
      Jul 14 2025
      Our 216th episode with a summary and discussion of last week's big AI news! Recorded on 07/11/2025 Hosted by Andrey Kurenkov and Jeremie Harris. Feel free to email us your questions and feedback at contact@lastweekinai.com and/or hello@gladstone.ai Read out our text newsletter and comment on the podcast at https://lastweekin.ai/. In this episode: xAI launches Grok 4 with breakthrough performance across benchmarks, becoming the first true frontier model outside established labs, alongside a $300/month subscription tierGrok's alignment challenges emerge with antisemitic responses, highlighting the difficulty of steering models toward "truth-seeking" without harmful biasesPerplexity and OpenAI launch AI-powered browsers to compete with Google Chrome, signaling a major shift in how users interact with AI systemsMeta study reveals AI tools actually slow down experienced developers by 20% on complex tasks, contradicting expectations and anecdotal reports of productivity gains Timestamps + Links: (00:00:10) Intro / Banter(00:01:02) News Preview Tools & Apps (00:01:59) Elon Musk's xAI launches Grok 4 alongside a $300 monthly subscription | TechCrunch(00:15:28) Elon Musk’s AI chatbot is suddenly posting antisemitic tropes(00:29:52) Perplexity launches Comet, an AI-powered web browser | TechCrunch(00:32:54) OpenAI is reportedly releasing an AI browser in the coming weeks | TechCrunch(00:33:27) Replit Launches New Feature for its Agent, CEO Calls it ‘Deep Research for Coding’(00:34:40) Cursor launches a web app to manage AI coding agents(00:36:07) Cursor apologizes for unclear pricing changes that upset users | TechCrunch Applications & Business (00:39:10) Lovable on track to raise $150M at $2B valuation(00:41:11) Amazon built a massive AI supercluster for Anthropic called Project Rainier – here's what we know so far(00:46:35) Elon Musk confirms xAI is buying an overseas power plant and shipping the whole thing to the U.S. to power its new data center — 1 million AI GPUs and up to 2 Gigawatts of power under one roof, equivalent to powering 1.9 million homes(00:48:16) Microsoft's own AI chip delayed six months in major setback — in-house chip now reportedly expected in 2026, but won't hold a candle to Nvidia Blackwell(00:49:54) Ilya Sutskever becomes CEO of Safe Superintelligence after Meta poached Daniel Gross(00:52:46) OpenAI’s Stock Compensation Reflect Steep Costs of Talent Wars Projects & Open Source (00:58:04) Hugging Face Releases SmolLM3: A 3B Long-Context, Multilingual Reasoning Model - MarkTechPost(00:58:33) Kimi K2: Open Agentic Intelligence(00:58:59) Kyutai Releases 2B Parameter Streaming Text-to-Speech TTS with 220ms Latency and 2.5M Hours of Training Research & Advancements (01:02:14) Does Math Reasoning Improve General LLM Capabilities? Understanding Transferability of LLM Reasoning(01:07:58) Measuring the Impact of Early-2025 AI on Experienced Open-Source Developer Productivity(01:13:03) Mitigating Goal Misgeneralization with Minimax Regret(01:17:01) Correlated Errors in Large Language Models(01:20:31) What skills does SWE-bench Verified evaluate? Policy & Safety (01:22:53) Evaluating Frontier Models for Stealth and Situational Awareness(01:25:49) When Chain of Thought is Necessary, Language Models Struggle to Evade Monitors(01:30:09) Why Do Some Language Models Fake Alignment While Others Don't?(01:34:35) Positive review only': Researchers hide AI prompts in papers(01:35:40) Google faces EU antitrust complaint over AI Overviews(01:36:41) The transfer of user data by DeepSeek to China is unlawful': Germany calls for Google and Apple to remove the AI app from their stores(01:37:30) Virology Capabilities Test (VCT): A Multimodal Virology Q&A Benchmark
      Afficher plus Afficher moins
      1 h et 42 min
    Aucun commentaire pour le moment