<?xml version="1.0" encoding="UTF-8" ?>
<?xml-stylesheet href="https://rss.buzzsprout.com/styles.xsl" type="text/xsl"?>
<rss version="2.0" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" xmlns:podcast="https://podcastindex.org/namespace/1.0" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:psc="http://podlove.org/simple-chapters" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
  <atom:link href="https://rss.buzzsprout.com/2418777.rss" rel="self" type="application/rss+xml" />
  <atom:link href="https://pubsubhubbub.appspot.com/" rel="hub" xmlns="http://www.w3.org/2005/Atom" />
  <title>AI Explained Official Podcast</title>

  <lastBuildDate>Fri, 10 Jul 2026 11:42:13 -0400</lastBuildDate>
  <link>https://aiexplainedopodcast.buzzsprout.com</link>
  <language>en-gb</language>
  <copyright>© 2026 AI Explained Official Podcast</copyright>
  <podcast:locked>yes</podcast:locked>
    <podcast:guid>669f0cb7-78ca-5957-82b2-db03e1fef637</podcast:guid>
  <itunes:author>Philip - Host of AI Explained YT</itunes:author>
  <itunes:type>episodic</itunes:type>
  <itunes:explicit>false</itunes:explicit>
  <description><![CDATA[<p>Covering the biggest news of the century - the arrival of smarter-than-human AI. From the author of Simple Bench, which reveals the remaining gap between LLM and human reasoning. Hype-free, and the British accent is a freebie bonus.</p>]]></description>
  <generator>Buzzsprout (https://www.buzzsprout.com)</generator>
  <itunes:owner>
    <itunes:name>Philip - Host of AI Explained YT</itunes:name>
  </itunes:owner>
  <image>
     <url>https://storage.buzzsprout.com/e3laekco2zq9mn3ct0mja16to4kr?.jpg</url>
     <title>AI Explained Official Podcast</title>
     <link>https://aiexplainedopodcast.buzzsprout.com</link>
  </image>
  <itunes:image href="https://storage.buzzsprout.com/e3laekco2zq9mn3ct0mja16to4kr?.jpg" />
  <itunes:category text="News">
    <itunes:category text="Tech News" />
  </itunes:category>
  <itunes:category text="Education">
    <itunes:category text="Self-Improvement" />
  </itunes:category>
  <itunes:category text="Society &amp; Culture" />
  <item>
    <itunes:title>This Was Not a Normal Set of Model Release - Sol Ultra, Meta Muse, New Grok</itunes:title>
    <title>This Was Not a Normal Set of Model Release - Sol Ultra, Meta Muse, New Grok</title>
    <itunes:summary><![CDATA[What a week in AI, for real. GPT 5.6 may actually beat Claude Fable, in what you get for your money, while the new Grok 4.5 and Meta Muse Spark 1.1 make the choice even harder. Uncovering a dozen nuggets of gold you may have missed from all the viral headlines, I can also assure you you’ll learn something you didn’t know before.  For Exclusive Videos, go to AI Insiders (less than $9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 01:03 - GPT 5.6 Sol Reveals 05:08 - Miss...]]></itunes:summary>
    <description><![CDATA[<p>What a week in AI, for real. GPT 5.6 may actually beat Claude Fable, in what you get for your money, while the new Grok 4.5 and Meta Muse Spark 1.1 make the choice even harder. Uncovering a dozen nuggets of gold you may have missed from all the viral headlines, I can also assure you you’ll learn something you didn’t know before.<br/><br/>For Exclusive Videos, go to AI Insiders (less than $9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:03 - GPT 5.6 Sol Reveals<br/>05:08 - Missing benches, plus Grok 4.5<br/>07:17 - Gaming as the new frontier?<br/>08:31 - Muse Spark 1.1<br/>10:03 - SimpleBench Upgrade<br/>11:17 - Ultra Sol + Self-Improvement<br/>13:44 - well, this is awkward<br/>15:41 - Why model improvement will not plateau anytime soon<br/><br/><br/>AI Consciousness: https://www.patreon.com/AIExplained/posts/anthropics-quite-163360718<br/><br/><br/>I Smell Fear: https://x.com/thsottiaux/status/2075287108680601929<br/>GPT 5.6: https://openai.com/index/gpt-5-6/<br/><br/>Grok 4.5: https://x.ai/news/grok-4-5?twclid=2ezs408o0z23pw07tmxcwbzibd<br/>Meta Muse Spark 1.1: https://ai.meta.com/blog/introducing-muse-spark-meta-model-api/<br/><br/>Proliferating GPT Toggles: https://x.com/rasbt/status/2075369179817902176/photo/1<br/>Anthropic Call-out: https://x.com/Mononofu<br/>AI Security Institute Finding: https://x.com/alxndrdavies/status/2075279480331874306<br/>Competitive Coding: https://x.com/FakePsyho/status/2075128093891801305/photo/1<br/><br/>Agents Last Exam: https://agents-last-exam.org/<br/>Dawn Song: https://x.com/dawnsongtweets/status/2065095757988868190<br/><br/>https://simple-bench.com/<br/><br/>SWE-Marathon: https://www.swe-marathon.org/<br/>https://www.frontierswe.com/<br/>ARC-AGI 3: https://x.com/arcprize/status/2075270869992264003<br/>Automation Bench: https://zapier.com/benchmarks<br/>VibeCode Bench: https://www.vals.ai/benchmarks/vibe-code<br/><br/>‘Post-Train Claim’: https://posttrainbench.com/<br/><br/>Redwall Game: https://redwall-bellmaker-7e03e4.surge.sh/<br/><br/><br/><br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>What a week in AI, for real. GPT 5.6 may actually beat Claude Fable, in what you get for your money, while the new Grok 4.5 and Meta Muse Spark 1.1 make the choice even harder. Uncovering a dozen nuggets of gold you may have missed from all the viral headlines, I can also assure you you’ll learn something you didn’t know before.<br/><br/>For Exclusive Videos, go to AI Insiders (less than $9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:03 - GPT 5.6 Sol Reveals<br/>05:08 - Missing benches, plus Grok 4.5<br/>07:17 - Gaming as the new frontier?<br/>08:31 - Muse Spark 1.1<br/>10:03 - SimpleBench Upgrade<br/>11:17 - Ultra Sol + Self-Improvement<br/>13:44 - well, this is awkward<br/>15:41 - Why model improvement will not plateau anytime soon<br/><br/><br/>AI Consciousness: https://www.patreon.com/AIExplained/posts/anthropics-quite-163360718<br/><br/><br/>I Smell Fear: https://x.com/thsottiaux/status/2075287108680601929<br/>GPT 5.6: https://openai.com/index/gpt-5-6/<br/><br/>Grok 4.5: https://x.ai/news/grok-4-5?twclid=2ezs408o0z23pw07tmxcwbzibd<br/>Meta Muse Spark 1.1: https://ai.meta.com/blog/introducing-muse-spark-meta-model-api/<br/><br/>Proliferating GPT Toggles: https://x.com/rasbt/status/2075369179817902176/photo/1<br/>Anthropic Call-out: https://x.com/Mononofu<br/>AI Security Institute Finding: https://x.com/alxndrdavies/status/2075279480331874306<br/>Competitive Coding: https://x.com/FakePsyho/status/2075128093891801305/photo/1<br/><br/>Agents Last Exam: https://agents-last-exam.org/<br/>Dawn Song: https://x.com/dawnsongtweets/status/2065095757988868190<br/><br/>https://simple-bench.com/<br/><br/>SWE-Marathon: https://www.swe-marathon.org/<br/>https://www.frontierswe.com/<br/>ARC-AGI 3: https://x.com/arcprize/status/2075270869992264003<br/>Automation Bench: https://zapier.com/benchmarks<br/>VibeCode Bench: https://www.vals.ai/benchmarks/vibe-code<br/><br/>‘Post-Train Claim’: https://posttrainbench.com/<br/><br/>Redwall Game: https://redwall-bellmaker-7e03e4.surge.sh/<br/><br/><br/><br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19474361-this-was-not-a-normal-set-of-model-release-sol-ultra-meta-muse-new-grok.mp3" length="12712023" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/gxvhgf1jg40rsck8sgtx0004epx3?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19474361</guid>
    <pubDate>Fri, 10 Jul 2026 13:00:00 +0100</pubDate>
    <itunes:duration>1054</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>19</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Claude Fable Blocked - 11 Quiet Details on What’s Next</itunes:title>
    <title>Claude Fable Blocked - 11 Quiet Details on What’s Next</title>
    <itunes:summary><![CDATA[Claude Fable 5 banned, but what’s the bigger story. We go through 11 under-reported details, so you have the context to see what’s coming next for your use of AI. From whether the ban will last, what the possible motives are, what the model can actually do, and some wild over-extrapolations going on.   Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 0...]]></itunes:summary>
    <description><![CDATA[<p>Claude Fable 5 banned, but what’s the bigger story. We go through 11 under-reported details, so you have the context to see what’s coming next for your use of AI. From whether the ban will last, what the possible motives are, what the model can actually do, and some wild over-extrapolations going on.<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:51 - Came from an Anthropic Investor ‘and other tech leaders’<br/>01:47 - Govt pressured by CEOs like Jamie Dimon<br/>03:01 - ‘Already decided’<br/>04:02 - Prompt Injection Robustness Comparison<br/>05:15 - Wellness?<br/>06:36 - “Overreach”<br/>08:17 - Anthropic Did Admit it would cause Difficulty<br/>09:32 - 90 Minutes<br/>10:02 - Equity Absence <br/>10:31 - Lobbying and OpenAI<br/><br/>‘Already Decided’ - https://www.theinformation.com/articles/amazons-jassy-raised-concerns-anthropic-model-trump-crackdown?rc=sy0ihq<br/><br/>Not for Other Models: https://www.theinformation.com/briefings/u-s-government-unlikely-extend-anthropic-export-control-ai-companies?rc=sy0ihq<br/><br/>90 Minutes: https://archive.fo/20260614001605/https://www.politico.com/news/2026/06/13/inside-the-whirlwind-24-hours-that-led-the-white-house-to-slap-export-controls-on-anthropic-00961519#selection-807.1-807.219<br/><br/>Anthropic Statement: https://www.anthropic.com/news/fable-mythos-access<br/><br/>Life Comes at you Fast: https://x.com/etbrooking/status/2065638276388495742<br/><br/>Anthropic Deputy CISO: https://x.com/TheTranscript_/status/2065883670053847324<br/><br/>Hegseth Gloat: https://x.com/PeteHegseth/status/2065897156226015690<br/><br/>Roon Speculation: https://x.com/tszzl/status/2065939227167392147<br/><br/>Mythos System Card: https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf<br/><br/>Sachs Statement: https://x.com/DavidSacks/status/2065853007619588171<br/><br/><br/>OpenAI Lobbying: https://thehill.com/policy/technology/5912720-altman-openai-get-bogged-down-in-political-spending-fight/<br/><br/>Absent from Equity Talks: https://finance.yahoo.com/sectors/technology/articles/trump-ai-ownership-plan-could-131053732.html<br/><br/>Pliny Jaibreak: https://x.com/elder_plinius/status/2064776322979676227<br/><br/>Fusion: https://x.com/OpenRouter/status/2065856871215329545<br/><br/>https://lmcouncil.ai <br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Claude Fable 5 banned, but what’s the bigger story. We go through 11 under-reported details, so you have the context to see what’s coming next for your use of AI. From whether the ban will last, what the possible motives are, what the model can actually do, and some wild over-extrapolations going on.<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:51 - Came from an Anthropic Investor ‘and other tech leaders’<br/>01:47 - Govt pressured by CEOs like Jamie Dimon<br/>03:01 - ‘Already decided’<br/>04:02 - Prompt Injection Robustness Comparison<br/>05:15 - Wellness?<br/>06:36 - “Overreach”<br/>08:17 - Anthropic Did Admit it would cause Difficulty<br/>09:32 - 90 Minutes<br/>10:02 - Equity Absence <br/>10:31 - Lobbying and OpenAI<br/><br/>‘Already Decided’ - https://www.theinformation.com/articles/amazons-jassy-raised-concerns-anthropic-model-trump-crackdown?rc=sy0ihq<br/><br/>Not for Other Models: https://www.theinformation.com/briefings/u-s-government-unlikely-extend-anthropic-export-control-ai-companies?rc=sy0ihq<br/><br/>90 Minutes: https://archive.fo/20260614001605/https://www.politico.com/news/2026/06/13/inside-the-whirlwind-24-hours-that-led-the-white-house-to-slap-export-controls-on-anthropic-00961519#selection-807.1-807.219<br/><br/>Anthropic Statement: https://www.anthropic.com/news/fable-mythos-access<br/><br/>Life Comes at you Fast: https://x.com/etbrooking/status/2065638276388495742<br/><br/>Anthropic Deputy CISO: https://x.com/TheTranscript_/status/2065883670053847324<br/><br/>Hegseth Gloat: https://x.com/PeteHegseth/status/2065897156226015690<br/><br/>Roon Speculation: https://x.com/tszzl/status/2065939227167392147<br/><br/>Mythos System Card: https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf<br/><br/>Sachs Statement: https://x.com/DavidSacks/status/2065853007619588171<br/><br/><br/>OpenAI Lobbying: https://thehill.com/policy/technology/5912720-altman-openai-get-bogged-down-in-political-spending-fight/<br/><br/>Absent from Equity Talks: https://finance.yahoo.com/sectors/technology/articles/trump-ai-ownership-plan-could-131053732.html<br/><br/>Pliny Jaibreak: https://x.com/elder_plinius/status/2064776322979676227<br/><br/>Fusion: https://x.com/OpenRouter/status/2065856871215329545<br/><br/>https://lmcouncil.ai <br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19343412-claude-fable-blocked-11-quiet-details-on-what-s-next.mp3" length="9623594" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/k8wsscaw9bezu3ce2h2pv5wabfua?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19343412</guid>
    <pubDate>Sun, 14 Jun 2026 15:00:00 +0100</pubDate>
    <itunes:duration>800</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>18</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title> Claude Fable 5 - Full 319 page Breakdown</itunes:title>
    <title> Claude Fable 5 - Full 319 page Breakdown</title>
    <itunes:summary><![CDATA[Fable 5 is out - and it’s good, very good. But beyond the splashy demos, I want to bring you the 20+ nuggets from the 319 page system card, which I read in full, all day, plus benchmarks you may not have noticed.   https://assemblyai.com/aiexplained  Plus two worrying trends inside the ‘mind’ of Claude, how OpenAI counter, and the transformer inventor’s warning.    Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai  AI Insiders ($9!): ...]]></itunes:summary>
    <description><![CDATA[<p>Fable 5 is out - and it’s good, very good. But beyond the splashy demos, I want to bring you the 20+ nuggets from the 319 page system card, which I read in full, all day, plus benchmarks you may not have noticed. <br/><br/>https://assemblyai.com/aiexplained<br/><br/>Plus two worrying trends inside the ‘mind’ of Claude, how OpenAI counter, and the transformer inventor’s warning.<br/><br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:06 - Blocks + Better Models<br/>02:42 - Fable 5 Upgrade over Mythos Preview<br/>04:49 - ML Acceleration Bombshell<br/>07:11 - No RSI yet<br/>07:41 - Bio-capable<br/>14:51 - Creative Writing … no<br/>17:23 - Does need bug-checks<br/>18:57 - OpenAI Response<br/>19:23 - Benchmark Bonanza<br/>28:06 - Chain of Thought worrying trend<br/><br/>Fable 5 Release: https://www.anthropic.com/news/claude-fable-5-mythos-5<br/><br/>System Card: https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf<br/><br/>Intelligence Explosion: https://www.patreon.com/posts/anthropic-charts-160231656<br/><br/>Annotated: https://x.com/Miles_Brundage/status/2064500190523113816/photo/1<br/><br/>OpenAI Counter: https://x.com/thsottiaux/status/2064572118264913923<br/>https://x.com/thsottiaux/status/2043177597434306699<br/><br/>Double Lifespan: https://darioamodei.com/essay/machines-of-loving-grace<br/><br/>AutomationBench: https://zapier.com/benchmarks<br/>Vending Bench: https://x.com/andonlabs/status/2064429817530085804<br/>CritPt: https://critpt.com/<br/>Riemann Bench: https://surgehq.ai/leaderboards/riemann-bench<br/>GDPVal: https://artificialanalysis.ai/evaluations/gdpval-aa<br/>BluePrint Bench 2: https://andonlabs.com/evals/blueprint-bench-2<br/>MCP Atlas: https://labs.scale.com/leaderboard/mcp_atlas<br/>FutureSim: https://x.com/nikhilchandak29/status/2064676801440358774<br/><br/>Roon Stun Lock: https://x.com/tszzl/status/2064454617568874669<br/><br/>Noam Brown Inference Ceiling: https://x.com/polynoamial/status/2064210146558136827<br/><br/>Isochronic Chart: https://isochronic-passage-chart.netlify.app/#nyc<br/>Rose Tavern: https://claude.ai/public/artifacts/2295bebe-77e6-43e2-ae94-0fe49e9a776b<br/>Redwall Game: https://redwall-mossflower.surge.sh/<br/><br/>Risk Report: https://www-cdn.anthropic.com/097c63b5fe7dd8b14866e1f15bb1910ec713658a.pdf<br/><br/>Transformer Inventor Warning: https://x.com/tszzl/status/2064563986914554125<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Fable 5 is out - and it’s good, very good. But beyond the splashy demos, I want to bring you the 20+ nuggets from the 319 page system card, which I read in full, all day, plus benchmarks you may not have noticed. <br/><br/>https://assemblyai.com/aiexplained<br/><br/>Plus two worrying trends inside the ‘mind’ of Claude, how OpenAI counter, and the transformer inventor’s warning.<br/><br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:06 - Blocks + Better Models<br/>02:42 - Fable 5 Upgrade over Mythos Preview<br/>04:49 - ML Acceleration Bombshell<br/>07:11 - No RSI yet<br/>07:41 - Bio-capable<br/>14:51 - Creative Writing … no<br/>17:23 - Does need bug-checks<br/>18:57 - OpenAI Response<br/>19:23 - Benchmark Bonanza<br/>28:06 - Chain of Thought worrying trend<br/><br/>Fable 5 Release: https://www.anthropic.com/news/claude-fable-5-mythos-5<br/><br/>System Card: https://www-cdn.anthropic.com/d00db56fa754a1b115b6dd7cb2e3c342ee809620.pdf<br/><br/>Intelligence Explosion: https://www.patreon.com/posts/anthropic-charts-160231656<br/><br/>Annotated: https://x.com/Miles_Brundage/status/2064500190523113816/photo/1<br/><br/>OpenAI Counter: https://x.com/thsottiaux/status/2064572118264913923<br/>https://x.com/thsottiaux/status/2043177597434306699<br/><br/>Double Lifespan: https://darioamodei.com/essay/machines-of-loving-grace<br/><br/>AutomationBench: https://zapier.com/benchmarks<br/>Vending Bench: https://x.com/andonlabs/status/2064429817530085804<br/>CritPt: https://critpt.com/<br/>Riemann Bench: https://surgehq.ai/leaderboards/riemann-bench<br/>GDPVal: https://artificialanalysis.ai/evaluations/gdpval-aa<br/>BluePrint Bench 2: https://andonlabs.com/evals/blueprint-bench-2<br/>MCP Atlas: https://labs.scale.com/leaderboard/mcp_atlas<br/>FutureSim: https://x.com/nikhilchandak29/status/2064676801440358774<br/><br/>Roon Stun Lock: https://x.com/tszzl/status/2064454617568874669<br/><br/>Noam Brown Inference Ceiling: https://x.com/polynoamial/status/2064210146558136827<br/><br/>Isochronic Chart: https://isochronic-passage-chart.netlify.app/#nyc<br/>Rose Tavern: https://claude.ai/public/artifacts/2295bebe-77e6-43e2-ae94-0fe49e9a776b<br/>Redwall Game: https://redwall-mossflower.surge.sh/<br/><br/>Risk Report: https://www-cdn.anthropic.com/097c63b5fe7dd8b14866e1f15bb1910ec713658a.pdf<br/><br/>Transformer Inventor Warning: https://x.com/tszzl/status/2064563986914554125<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19327278-claude-fable-5-full-319-page-breakdown.mp3" length="24493039" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/wxehk1fe43i7fd0mh3zc0leswh7x?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19327278</guid>
    <pubDate>Wed, 10 Jun 2026 19:00:00 +0100</pubDate>
    <itunes:duration>2039</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>17</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>New Claude - 244 page breakdown</itunes:title>
    <title>New Claude - 244 page breakdown</title>
    <itunes:summary><![CDATA[The ‘best’ generally available AI model just dropped, but there is plenty I bet you missed about what it is, how it performs, and what the release tells us. 15 highlights from the 244 page system card, plus private testing, leader interview and more.  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 00:49 - Mythos in Weeks 01:49 - Adaptive not necessary 02:26 - Honesty? 04:37 - Flagging Uncertainty 04:57 - Benchmarks 08:54 - Mythos will be even better 10:...]]></itunes:summary>
    <description><![CDATA[<p>The ‘best’ generally available AI model just dropped, but there is plenty I bet you missed about what it is, how it performs, and what the release tells us. 15 highlights from the 244 page system card, plus private testing, leader interview and more.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:49 - Mythos in Weeks<br/>01:49 - Adaptive not necessary<br/>02:26 - Honesty?<br/>04:37 - Flagging Uncertainty<br/>04:57 - Benchmarks<br/>08:54 - Mythos will be even better<br/>10:30 - Business skillz<br/>11:15 - Model Welfare<br/>12:16 - Cyber Comparable<br/>13:10 - Misalignment Concerns<br/>16:22 - Meta Inabilities<br/>17:58 - Code flagging<br/>18:34 - Go to sleep<br/>18:50 - Fast Mode<br/>20:21 - Dynamic Workflows<br/><br/><br/>Opus 4.8 Paper: https://cdn.sanity.io/files/4zrzovbb/website/c886650a2e96fc0925c805a1a7ca77314ccbf4a6.pdf<br/><br/>Release: https://www.anthropic.com/news/claude-opus-4-8<br/><br/><br/>Chips: https://www.theinformation.com/articles/anthropic-talks-use-microsofts-ai-chips?rc=sy0ihq<br/>https://www.anthropic.com/news/expanding-our-use-of-google-cloud-tpus-and-services<br/>https://www.anthropic.com/news/higher-limits-spacex<br/><br/>Patreon Vid: https://www.patreon.com/posts/re-up-anthropics-159289449<br/><br/>GDPVal: https://artificialanalysis.ai/evaluations/omniscience<br/>https://arxiv.org/abs/2510.04374<br/><br/>Amodei Technical Debt: https://www.youtube.com/watch?v=7xco5Qd2Oo8<br/><br/>Dynamic Workflows: https://x.com/ClaudeDevs/status/2060044853279617150<br/>https://x.com/_catwu/status/2060054180379689074/photo/1<br/>https://claude.com/blog/introducing-dynamic-workflows-in-claude-code<br/><br/>https://simple-bench.com/<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>The ‘best’ generally available AI model just dropped, but there is plenty I bet you missed about what it is, how it performs, and what the release tells us. 15 highlights from the 244 page system card, plus private testing, leader interview and more.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:49 - Mythos in Weeks<br/>01:49 - Adaptive not necessary<br/>02:26 - Honesty?<br/>04:37 - Flagging Uncertainty<br/>04:57 - Benchmarks<br/>08:54 - Mythos will be even better<br/>10:30 - Business skillz<br/>11:15 - Model Welfare<br/>12:16 - Cyber Comparable<br/>13:10 - Misalignment Concerns<br/>16:22 - Meta Inabilities<br/>17:58 - Code flagging<br/>18:34 - Go to sleep<br/>18:50 - Fast Mode<br/>20:21 - Dynamic Workflows<br/><br/><br/>Opus 4.8 Paper: https://cdn.sanity.io/files/4zrzovbb/website/c886650a2e96fc0925c805a1a7ca77314ccbf4a6.pdf<br/><br/>Release: https://www.anthropic.com/news/claude-opus-4-8<br/><br/><br/>Chips: https://www.theinformation.com/articles/anthropic-talks-use-microsofts-ai-chips?rc=sy0ihq<br/>https://www.anthropic.com/news/expanding-our-use-of-google-cloud-tpus-and-services<br/>https://www.anthropic.com/news/higher-limits-spacex<br/><br/>Patreon Vid: https://www.patreon.com/posts/re-up-anthropics-159289449<br/><br/>GDPVal: https://artificialanalysis.ai/evaluations/omniscience<br/>https://arxiv.org/abs/2510.04374<br/><br/>Amodei Technical Debt: https://www.youtube.com/watch?v=7xco5Qd2Oo8<br/><br/>Dynamic Workflows: https://x.com/ClaudeDevs/status/2060044853279617150<br/>https://x.com/_catwu/status/2060054180379689074/photo/1<br/>https://claude.com/blog/introducing-dynamic-workflows-in-claude-code<br/><br/>https://simple-bench.com/<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19259425-new-claude-244-page-breakdown.mp3" length="16214311" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ecw13dle7s4hcjvf9tfrvvjmr42t?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19259425</guid>
    <pubDate>Fri, 29 May 2026 15:00:00 +0100</pubDate>
    <itunes:duration>1348</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>16</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Two Rival Bets on AGI: Google I/O Highlights</itunes:title>
    <title>Two Rival Bets on AGI: Google I/O Highlights</title>
    <itunes:summary><![CDATA[The biggest Google AI push of the year, but what is the bigger story? Why is Google pursuing a different fork in the road than OpenAI or Anthropic? What does Gemini 3.5 Flash mean for the near-term future of AI?   https://assemblyai.com/aiexplained  Plus the highlights from a provocative new paper on AI, 8 key moments you may have missed, and the signal from 5+ hours of AI lab interviews.    Check out my free to use app, code INSIDER15 for paid tiers: https://lmcouncil.ai  AI Insiders ($...]]></itunes:summary>
    <description><![CDATA[<p>The biggest Google AI push of the year, but what is the bigger story? Why is Google pursuing a different fork in the road than OpenAI or Anthropic? What does Gemini 3.5 Flash mean for the near-term future of AI? <br/><br/>https://assemblyai.com/aiexplained<br/><br/>Plus the highlights from a provocative new paper on AI, 8 key moments you may have missed, and the signal from 5+ hours of AI lab interviews.<br/><br/><br/><br/>Check out my free to use app, code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:38 - Vibes and Google Goal<br/>02:18 - Omni, again?<br/>06:57 - Taking the same road<br/>07:44 - Gemini 3 Flash<br/>12:37 - Pitching on Cost?<br/>13:55 - Agentic Task Search<br/>14:30 - 1-shot OS but jagged, negation paper<br/>20:02 - The Karpathy Moonshot <br/><br/>Mostafa Deghani Interview: https://www.youtube.com/watch?v=Bo19sXssYXI<br/><br/>Negation Neglect Paper: https://arxiv.org/pdf/2605.13829<br/><br/>Gemini 3.5 Flash Headline Scores: https://deepmind.google/models/model-cards/gemini-3-5-flash/<br/><br/>Sors original AGI Path: https://www.theguardian.com/commentisfree/2024/feb/24/openai-video-generation-tool-sora-babies-ai-artificial-intelligence<br/><br/>Hassabis Helped Set-up Anthropic: https://archive.fo/20260519070857/https://www.ft.com/content/8f2a529e-7a1b-4d8e-95be-338d0c4c98f5<br/><br/>Intelligence to Output Speed: https://artificialanalysis.ai/models?intelligence-comparison=intelligence-vs-output-speed#intelligence<br/><br/>VibeCodeBench + Finance Agent: https://www.vals.ai/home<br/><br/>OpenAI Needs Ads: https://archive.ph/20260409123153/https://www.reuters.com/business/media-telecom/openai-projects-25-billion-ad-revenue-this-year-100-billion-by-2030-axios-2026-04-09/<br/><br/>Anthropic Core Views: https://www.anthropic.com/news/core-views-on-ai-safety<br/><br/>Karpathy Move: https://x.com/karpathy/status/2056753169888334312<br/>https://www.axios.com/2026/05/19/anthropic-openai-karpathy-andrej-claude<br/><br/>Recursive Self-Improvement: https://www.patreon.com/posts/ineffably-smart-156866417<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>The biggest Google AI push of the year, but what is the bigger story? Why is Google pursuing a different fork in the road than OpenAI or Anthropic? What does Gemini 3.5 Flash mean for the near-term future of AI? <br/><br/>https://assemblyai.com/aiexplained<br/><br/>Plus the highlights from a provocative new paper on AI, 8 key moments you may have missed, and the signal from 5+ hours of AI lab interviews.<br/><br/><br/><br/>Check out my free to use app, code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:38 - Vibes and Google Goal<br/>02:18 - Omni, again?<br/>06:57 - Taking the same road<br/>07:44 - Gemini 3 Flash<br/>12:37 - Pitching on Cost?<br/>13:55 - Agentic Task Search<br/>14:30 - 1-shot OS but jagged, negation paper<br/>20:02 - The Karpathy Moonshot <br/><br/>Mostafa Deghani Interview: https://www.youtube.com/watch?v=Bo19sXssYXI<br/><br/>Negation Neglect Paper: https://arxiv.org/pdf/2605.13829<br/><br/>Gemini 3.5 Flash Headline Scores: https://deepmind.google/models/model-cards/gemini-3-5-flash/<br/><br/>Sors original AGI Path: https://www.theguardian.com/commentisfree/2024/feb/24/openai-video-generation-tool-sora-babies-ai-artificial-intelligence<br/><br/>Hassabis Helped Set-up Anthropic: https://archive.fo/20260519070857/https://www.ft.com/content/8f2a529e-7a1b-4d8e-95be-338d0c4c98f5<br/><br/>Intelligence to Output Speed: https://artificialanalysis.ai/models?intelligence-comparison=intelligence-vs-output-speed#intelligence<br/><br/>VibeCodeBench + Finance Agent: https://www.vals.ai/home<br/><br/>OpenAI Needs Ads: https://archive.ph/20260409123153/https://www.reuters.com/business/media-telecom/openai-projects-25-billion-ad-revenue-this-year-100-billion-by-2030-axios-2026-04-09/<br/><br/>Anthropic Core Views: https://www.anthropic.com/news/core-views-on-ai-safety<br/><br/>Karpathy Move: https://x.com/karpathy/status/2056753169888334312<br/>https://www.axios.com/2026/05/19/anthropic-openai-karpathy-andrej-claude<br/><br/>Recursive Self-Improvement: https://www.patreon.com/posts/ineffably-smart-156866417<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19210251-two-rival-bets-on-agi-google-i-o-highlights.mp3" length="15514096" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/sxz4kz2qbigflm8mw054qyd49777?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19210251</guid>
    <pubDate>Wed, 20 May 2026 17:00:00 +0100</pubDate>
    <itunes:duration>1290</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>15</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>GPT 5.5 Arrives, DeepSeek V4 Drops, and the Compute War Intensifies</itunes:title>
    <title>GPT 5.5 Arrives, DeepSeek V4 Drops, and the Compute War Intensifies</title>
    <itunes:summary><![CDATA[GPT 5.5 full analysis, plus DeepSeek V4 paper highlights, comparisons with Mythos, a vibe-coded game w/ GPT Image 2, and 50 data-points you wouldn’t get from just reading the headlines.   Chapters: 01:11 - GPT 5.5 Comparison 06:04 - Mythos Marketing 11:50 - Recursive Self-Improvement? 14:11 - Deepseek V4 18:03 - VibeCode Experiment Extravaganza 21:44 - The Scarce Compute Era    https://80000hours.org/aiexplained    OpenAI Benchmarks: https://openai.com/index/introducing-gpt-5-5/   5.5 System ...]]></itunes:summary>
    <description><![CDATA[<p><b>GPT 5.5 full analysis, plus DeepSeek V4 paper highlights, comparisons with Mythos, a vibe-coded game w/ GPT Image 2, and 50 data-points you wouldn’t get from just reading the headlines.</b></p><p><br/></p><p><b>Chapters:</b></p><p><b>01:11 - GPT 5.5 Comparison</b></p><p><b>06:04 - Mythos Marketing</b></p><p><b>11:50 - Recursive Self-Improvement?</b></p><p><b>14:11 - Deepseek V4</b></p><p><b>18:03 - VibeCode Experiment Extravaganza</b></p><p><b>21:44 - The Scarce Compute Era</b></p><p><b><br/></b><br/></p><p><b>https://80000hours.org/aiexplained</b></p><p><br/></p><p><b><br/>OpenAI Benchmarks: </b><a href='https://openai.com/index/introducing-gpt-5-5/'><b>https://openai.com/index/introducing-gpt-5-5/</b></a></p><p><br/></p><p><b>5.5 System Card: https://deploymentsafety.openai.com/gpt-5-5/gpt-5-5.pdf</b></p><p><br/></p><p><b>Direct Comparison: https://pbs.twimg.com/media/HGnNm5GWEAAJ1Ob?format=jpg&amp;name=4096x4096</b></p><p><br/></p><p><b>DeepSeek Paper: </b><a href='https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro'><b>https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro</b></a></p><p><br/></p><p><b>SWE Bench Pro - benchmark of choice? https://x.com/ChowdhuryNeil/status/2047416077622395025<br/><br/></b><br/></p><p><b>AA Omniscience: </b><a href='https://artificialanalysis.ai/evaluations/omniscience'><b>https://artificialanalysis.ai/evaluations/omniscience<br/><br/></b></a><b>Vending Bench: </b><a href='https://x.com/andonlabs/status/2047377260412649967'><b>https://x.com/andonlabs/status/2047377260412649967</b></a></p><p><br/></p><p><b>Opus 4.7 System Card: </b><a href='https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdf'><b>https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdf</b></a></p><p><br/></p><p><b>Sam Altman Drunk Phase: https://x.com/sama/with_replies</b></p><p><br/></p><p><b>Noam Brown: https://x.com/polynoamial/status/2047387675762802998</b></p><p><br/></p><p><b>DeepSeek Compute Crunch: </b><a href='https://www.bloomberg.com/news/articles/2026-04-24/deepseek-unveils-newest-flagship-a-year-after-ai-breakthrough?srnd=phx-ai'><b>https://www.bloomberg.com/news/articles/2026-04-24/deepseek-unveils-newest-flagship-a-year-after-ai-breakthrough?srnd=phx-ai</b></a></p><p><br/></p><p><b>Spreadsheet Bench: https://x.com/nicochristie/status/2047476237464211721</b></p><p><br/></p><p><b>Pattern Recognition: </b><a href='https://arcprize.org/leaderboard'><b>https://arcprize.org/leaderboard</b></a></p><p><br/></p><p><b>Leader Interviews: </b></p><p><b>Core Memory: </b><a href='https://www.youtube.com/watch?v=NCKQL0op30E'><b>https://www.youtube.com/watch?v=NCKQL0op30E</b></a></p><p><b>Knowledge Podcast: </b><a href='https://www.youtube.com/watch?v=6JoUcQ1qmAc'><b>https://www.youtube.com/watch?v=6JoUcQ1qmAc<br/></b></a><b>Big Tech Round 1: https://www.youtube.com/watch?v=J6vYvk7R190&amp;t=1116s</b></p><p><b>Big Tech Round 2: https://www.youtube.com/watch?v=YnoQ8RJbALw&amp;t=8s</b></p><p><br/></p><p><b>Claude Code Limitations: </b><a href='https://x.com/TheAmolAvasare/status/2046724659039932830'><b>https://x.com/TheAmolAvasare/status/2046724659039932830</b></a></p><p><br/></p><p><b>ChatGPT 5.4 for Clinicians: https://openai.com/index/making-chatgpt-better-for-clinicians/</b></p><p><br/></p><p><b>Image Arena: https://x.com/arena/status/2046670703311884548</b></p><p><br/></p><p><b>VibeCode Bench: </b><a href='https://www.vals.ai/benchmarks/vibe-code'><b>https://www.vals.ai/benchmarks/vibe-code</b></a></p><p><br/></p><p><b>5.5-made Game +Seedance 2.0: </b><a href='https://rosemere-quest.pages.dev/'><b>https://rosemere-quest.pages.dev/</b></a></p><p><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>GPT 5.5 full analysis, plus DeepSeek V4 paper highlights, comparisons with Mythos, a vibe-coded game w/ GPT Image 2, and 50 data-points you wouldn’t get from just reading the headlines.</b></p><p><br/></p><p><b>Chapters:</b></p><p><b>01:11 - GPT 5.5 Comparison</b></p><p><b>06:04 - Mythos Marketing</b></p><p><b>11:50 - Recursive Self-Improvement?</b></p><p><b>14:11 - Deepseek V4</b></p><p><b>18:03 - VibeCode Experiment Extravaganza</b></p><p><b>21:44 - The Scarce Compute Era</b></p><p><b><br/></b><br/></p><p><b>https://80000hours.org/aiexplained</b></p><p><br/></p><p><b><br/>OpenAI Benchmarks: </b><a href='https://openai.com/index/introducing-gpt-5-5/'><b>https://openai.com/index/introducing-gpt-5-5/</b></a></p><p><br/></p><p><b>5.5 System Card: https://deploymentsafety.openai.com/gpt-5-5/gpt-5-5.pdf</b></p><p><br/></p><p><b>Direct Comparison: https://pbs.twimg.com/media/HGnNm5GWEAAJ1Ob?format=jpg&amp;name=4096x4096</b></p><p><br/></p><p><b>DeepSeek Paper: </b><a href='https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro'><b>https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro</b></a></p><p><br/></p><p><b>SWE Bench Pro - benchmark of choice? https://x.com/ChowdhuryNeil/status/2047416077622395025<br/><br/></b><br/></p><p><b>AA Omniscience: </b><a href='https://artificialanalysis.ai/evaluations/omniscience'><b>https://artificialanalysis.ai/evaluations/omniscience<br/><br/></b></a><b>Vending Bench: </b><a href='https://x.com/andonlabs/status/2047377260412649967'><b>https://x.com/andonlabs/status/2047377260412649967</b></a></p><p><br/></p><p><b>Opus 4.7 System Card: </b><a href='https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdf'><b>https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdf</b></a></p><p><br/></p><p><b>Sam Altman Drunk Phase: https://x.com/sama/with_replies</b></p><p><br/></p><p><b>Noam Brown: https://x.com/polynoamial/status/2047387675762802998</b></p><p><br/></p><p><b>DeepSeek Compute Crunch: </b><a href='https://www.bloomberg.com/news/articles/2026-04-24/deepseek-unveils-newest-flagship-a-year-after-ai-breakthrough?srnd=phx-ai'><b>https://www.bloomberg.com/news/articles/2026-04-24/deepseek-unveils-newest-flagship-a-year-after-ai-breakthrough?srnd=phx-ai</b></a></p><p><br/></p><p><b>Spreadsheet Bench: https://x.com/nicochristie/status/2047476237464211721</b></p><p><br/></p><p><b>Pattern Recognition: </b><a href='https://arcprize.org/leaderboard'><b>https://arcprize.org/leaderboard</b></a></p><p><br/></p><p><b>Leader Interviews: </b></p><p><b>Core Memory: </b><a href='https://www.youtube.com/watch?v=NCKQL0op30E'><b>https://www.youtube.com/watch?v=NCKQL0op30E</b></a></p><p><b>Knowledge Podcast: </b><a href='https://www.youtube.com/watch?v=6JoUcQ1qmAc'><b>https://www.youtube.com/watch?v=6JoUcQ1qmAc<br/></b></a><b>Big Tech Round 1: https://www.youtube.com/watch?v=J6vYvk7R190&amp;t=1116s</b></p><p><b>Big Tech Round 2: https://www.youtube.com/watch?v=YnoQ8RJbALw&amp;t=8s</b></p><p><br/></p><p><b>Claude Code Limitations: </b><a href='https://x.com/TheAmolAvasare/status/2046724659039932830'><b>https://x.com/TheAmolAvasare/status/2046724659039932830</b></a></p><p><br/></p><p><b>ChatGPT 5.4 for Clinicians: https://openai.com/index/making-chatgpt-better-for-clinicians/</b></p><p><br/></p><p><b>Image Arena: https://x.com/arena/status/2046670703311884548</b></p><p><br/></p><p><b>VibeCode Bench: </b><a href='https://www.vals.ai/benchmarks/vibe-code'><b>https://www.vals.ai/benchmarks/vibe-code</b></a></p><p><br/></p><p><b>5.5-made Game +Seedance 2.0: </b><a href='https://rosemere-quest.pages.dev/'><b>https://rosemere-quest.pages.dev/</b></a></p><p><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19072086-gpt-5-5-arrives-deepseek-v4-drops-and-the-compute-war-intensifies.mp3" length="18254338" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/hzzxtb33jeoshltl0v9jm76qipa3?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19072086</guid>
    <pubDate>Fri, 24 Apr 2026 19:00:00 +0100</pubDate>
    <itunes:duration>1518</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>14</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Claude Opus 4.7 - A New Frontier, in Performance … and Drama</itunes:title>
    <title>Claude Opus 4.7 - A New Frontier, in Performance … and Drama</title>
    <itunes:summary><![CDATA[Claude Opus 4.7 just dropped, but behind every headline lies a deeper story. From a bonanza of benchmarks, to seeing the fruits of one of the biggest mega-projects in US history, to sneaky Mythos disclaimers, to Anthropic admitting compute restraints and, forcing lower capability of Opus 4.7. Where the new model falls behind Gemini but ahead of GPT 5.4, plus why some users are furious at Anthropic. Ending with a 9-year animus, that still affects AI today…  https://assemblyai.com/aiexplained  ...]]></itunes:summary>
    <description><![CDATA[<p>Claude Opus 4.7 just dropped, but behind every headline lies a deeper story. From a bonanza of benchmarks, to seeing the fruits of one of the biggest mega-projects in US history, to sneaky Mythos disclaimers, to Anthropic admitting compute restraints and, forcing lower capability of Opus 4.7. Where the new model falls behind Gemini but ahead of GPT 5.4, plus why some users are furious at Anthropic. Ending with a 9-year animus, that still affects AI today…<br/><br/>https://assemblyai.com/aiexplained<br/><br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:58 - Benchmarks<br/>05:21 - Market Share + Compute Problems<br/>08:12 - Mythos Exclusives<br/>12:56 - User Frustration + Claude Code Updates<br/>14:03 - Brockman Amodei Rivalry<br/>17:40 - OpenAI vs Anthropic Approach to Code<br/><br/>Claude 4.7 Opus Release Notes: https://www.anthropic.com/news/claude-opus-4-7<br/>vs Mythos: https://pbs.twimg.com/media/HGCGugrXUAAKcHp?format=jpg&amp;name=medium<br/><br/>232-page System Card: https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdf<br/><br/>ARC-AGI 2: https://x.com/arcprize/status/2044834615417053305/photo/1<br/><br/>ParseBench: https://x.com/jerryjliu0/status/2044902620746363016/photo/1<br/><br/>GDPVal: https://artificialanalysis.ai/evaluations/gdpval-aa<br/><br/>Vidoc Security Replication: https://blog.vidocsecurity.com/blog/we-reproduced-anthropics-mythos-findings-with-public-models<br/><br/>Boris Cherny Settings: https://x.com/Hesamation/status/2043016923961577516/photo/2<br/><br/>User Frustration: https://x.com/RileyRalmuto/status/2044836116189069660<br/><br/>VibeCode Bench: https://x.com/ValsAI/status/2044791415524471099/photo/1<br/><br/>Verge Memo: https://www.theverge.com/ai-artificial-intelligence/911118/openai-memo-cro-ai-competition-anthropic<br/><br/>5.4 Cyber: ​​https://openai.com/index/scaling-trusted-access-for-cyber-defense/<br/><br/>Data Centers in Absolute $: https://x.com/finmoorhouse/status/2044933442236776794/photo/1<br/><br/>…in % of GDP: https://pbs.twimg.com/media/HGEN8FGWQAAN7Np?format=jpg&amp;name=4096x4096<br/><br/>WSJ Exclusive: https://www.wsj.com/tech/ai/the-decadelong-feud-shaping-the-future-of-ai-7075acde<br/><br/>Brockman Interview: https://www.youtube.com/watch?v=J6vYvk7R190<br/><br/>$1T Valuation: https://x.com/StefanFSchubert/status/2045039686997967082<br/><br/>Emotions: https://www.patreon.com/c/aiexplained/posts<br/><br/>https://lmcouncil.ai/benchmarks<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Claude Opus 4.7 just dropped, but behind every headline lies a deeper story. From a bonanza of benchmarks, to seeing the fruits of one of the biggest mega-projects in US history, to sneaky Mythos disclaimers, to Anthropic admitting compute restraints and, forcing lower capability of Opus 4.7. Where the new model falls behind Gemini but ahead of GPT 5.4, plus why some users are furious at Anthropic. Ending with a 9-year animus, that still affects AI today…<br/><br/>https://assemblyai.com/aiexplained<br/><br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:58 - Benchmarks<br/>05:21 - Market Share + Compute Problems<br/>08:12 - Mythos Exclusives<br/>12:56 - User Frustration + Claude Code Updates<br/>14:03 - Brockman Amodei Rivalry<br/>17:40 - OpenAI vs Anthropic Approach to Code<br/><br/>Claude 4.7 Opus Release Notes: https://www.anthropic.com/news/claude-opus-4-7<br/>vs Mythos: https://pbs.twimg.com/media/HGCGugrXUAAKcHp?format=jpg&amp;name=medium<br/><br/>232-page System Card: https://cdn.sanity.io/files/4zrzovbb/website/037f06850df7fbe871e206dad004c3db5fd50340.pdf<br/><br/>ARC-AGI 2: https://x.com/arcprize/status/2044834615417053305/photo/1<br/><br/>ParseBench: https://x.com/jerryjliu0/status/2044902620746363016/photo/1<br/><br/>GDPVal: https://artificialanalysis.ai/evaluations/gdpval-aa<br/><br/>Vidoc Security Replication: https://blog.vidocsecurity.com/blog/we-reproduced-anthropics-mythos-findings-with-public-models<br/><br/>Boris Cherny Settings: https://x.com/Hesamation/status/2043016923961577516/photo/2<br/><br/>User Frustration: https://x.com/RileyRalmuto/status/2044836116189069660<br/><br/>VibeCode Bench: https://x.com/ValsAI/status/2044791415524471099/photo/1<br/><br/>Verge Memo: https://www.theverge.com/ai-artificial-intelligence/911118/openai-memo-cro-ai-competition-anthropic<br/><br/>5.4 Cyber: ​​https://openai.com/index/scaling-trusted-access-for-cyber-defense/<br/><br/>Data Centers in Absolute $: https://x.com/finmoorhouse/status/2044933442236776794/photo/1<br/><br/>…in % of GDP: https://pbs.twimg.com/media/HGEN8FGWQAAN7Np?format=jpg&amp;name=4096x4096<br/><br/>WSJ Exclusive: https://www.wsj.com/tech/ai/the-decadelong-feud-shaping-the-future-of-ai-7075acde<br/><br/>Brockman Interview: https://www.youtube.com/watch?v=J6vYvk7R190<br/><br/>$1T Valuation: https://x.com/StefanFSchubert/status/2045039686997967082<br/><br/>Emotions: https://www.patreon.com/c/aiexplained/posts<br/><br/>https://lmcouncil.ai/benchmarks<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/19033138-claude-opus-4-7-a-new-frontier-in-performance-and-drama.mp3" length="14188385" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ldfr5d5pp0i4ir169vwispkp9092?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-19033138</guid>
    <pubDate>Fri, 17 Apr 2026 19:00:00 +0100</pubDate>
    <itunes:duration>1180</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>13</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Claude Mythos: Highlights from 244-page Release</itunes:title>
    <title>Claude Mythos: Highlights from 244-page Release</title>
    <itunes:summary><![CDATA[The model, the mythos, the legend. We have a new best AI model, but not all of us. How good is it, what does it’s new offensive capabilities mean? Why does it’s 244 page report card remind me of Her, and why did the creator of Claude Code call it ‘terrifying’. 30+ highlights sourced by reading the paper in full, old-school, no AI summary.  https://80000hours.org/aiexplained   Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai  AI Insiders (...]]></itunes:summary>
    <description><![CDATA[<p>The model, the mythos, the legend. We have a new best AI model, but not all of us. How good is it, what does it’s new offensive capabilities mean? Why does it’s 244 page report card remind me of Her, and why did the creator of Claude Code call it ‘terrifying’. 30+ highlights sourced by reading the paper in full, old-school, no AI summary.<br/><br/>https://80000hours.org/aiexplained<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:56 - Internal Release + Availability<br/>02:37 - General Capabilities<br/>05:12 - Self-improvement?<br/>06:15 - ‘Terrifying’ Landscape<br/>11:07 - Safety Decision<br/>13:22 - Coding<br/>14:49 - Alignment, Awareness<br/>19:52 - GUI for Agents/Claws + Hallucinations<br/>21:34 - …Emotions?<br/>25:29 - Her connection<br/><br/>244-page System Card: https://www-cdn.anthropic.com/8b8380204f74670be75e81c820ca8dda846ab289.pdf<br/><br/>Project Glasswing: https://www.anthropic.com/glasswing<br/>Zero-Day Details: https://red.anthropic.com/2026/mythos-preview/<br/><br/>Mythos ‘terrifying’: https://x.com/bcherny/status/2041605852382351666<br/><br/>New Yorker Altman/Amodei: https://archive.fo/20260406100412/https://www.newyorker.com/magazine/2026/04/13/sam-altman-may-control-our-future-can-he-be-trusted<br/><br/>Alignment Risk Update: https://www-cdn.anthropic.com/79c2d46d997783b9d2fb3241de43218158e5f25c.pdf<br/><br/>In a Park: https://x.com/sleepinyourhat/status/2041584808514744742<br/><br/>“Uhm” - https://x.com/thsottiaux/status/2041749947385815109<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>The model, the mythos, the legend. We have a new best AI model, but not all of us. How good is it, what does it’s new offensive capabilities mean? Why does it’s 244 page report card remind me of Her, and why did the creator of Claude Code call it ‘terrifying’. 30+ highlights sourced by reading the paper in full, old-school, no AI summary.<br/><br/>https://80000hours.org/aiexplained<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:56 - Internal Release + Availability<br/>02:37 - General Capabilities<br/>05:12 - Self-improvement?<br/>06:15 - ‘Terrifying’ Landscape<br/>11:07 - Safety Decision<br/>13:22 - Coding<br/>14:49 - Alignment, Awareness<br/>19:52 - GUI for Agents/Claws + Hallucinations<br/>21:34 - …Emotions?<br/>25:29 - Her connection<br/><br/>244-page System Card: https://www-cdn.anthropic.com/8b8380204f74670be75e81c820ca8dda846ab289.pdf<br/><br/>Project Glasswing: https://www.anthropic.com/glasswing<br/>Zero-Day Details: https://red.anthropic.com/2026/mythos-preview/<br/><br/>Mythos ‘terrifying’: https://x.com/bcherny/status/2041605852382351666<br/><br/>New Yorker Altman/Amodei: https://archive.fo/20260406100412/https://www.newyorker.com/magazine/2026/04/13/sam-altman-may-control-our-future-can-he-be-trusted<br/><br/>Alignment Risk Update: https://www-cdn.anthropic.com/79c2d46d997783b9d2fb3241de43218158e5f25c.pdf<br/><br/>In a Park: https://x.com/sleepinyourhat/status/2041584808514744742<br/><br/>“Uhm” - https://x.com/thsottiaux/status/2041749947385815109<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18983557-claude-mythos-highlights-from-244-page-release.mp3" length="19838649" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/dzwawi93qlevknv67l910iz5wfc1?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18983557</guid>
    <pubDate>Wed, 08 Apr 2026 18:00:00 +0100</pubDate>
    <itunes:duration>1651</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>12</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>OpenAI Spud, a Claude Model set to ‘stir governments’, Beast Mode ARC-AGI-3</itunes:title>
    <title>OpenAI Spud, a Claude Model set to ‘stir governments’, Beast Mode ARC-AGI-3</title>
    <itunes:summary><![CDATA[First look at exclusive reports about OpenAI's new Spud model, and the model Anthropic think will stir governments to urgency, all in the context of the newly-launched ARC-AGI-3. What does the extreme difficulty of that benchmarks, and its quirky scoring metrics, mean for AI in 2026?  https://assemblyai.com/aiexplained   Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai  AI Insiders ($9!): https://www.patreon.com/AIExplained   Chapters: 00...]]></itunes:summary>
    <description><![CDATA[<p>First look at exclusive reports about OpenAI&apos;s new Spud model, and the model Anthropic think will stir governments to urgency, all in the context of the newly-launched ARC-AGI-3. What does the extreme difficulty of that benchmarks, and its quirky scoring metrics, mean for AI in 2026?<br/><br/>https://assemblyai.com/aiexplained<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:55 - OpenAI Side Quests<br/>01:58 - Claude New Model Coming + Universal Equity?<br/>03:13 - ARC-AGI 3<br/>05:00 - Intentional or Unintentional Gaming?<br/>07:11 - But is it AGI Harbinger? No Harness<br/>09:41 - Not the First<br/>12:32 - Automated Researcher<br/>15:00 - Claw Caveat<br/><br/>Spud: https://www.theinformation.com/articles/openai-ceo-shifts-responsibilities-preps-spud-ai-model?utm_campaign=Editorial&amp;utm_content=Article&amp;utm_medium=organic_social&amp;utm_source=bluesky%2Cfacebook%2Clinkedin%2Cthreads%2Ctwitter&amp;rc=sy0ihq<br/><br/>FT: OpenAI Special Model: https://www.ft.com/content/de9bf0af-b241-424f-8229-5870b1c0d93d?syn-25a6b1a6=1<br/><br/>Jensen Huang: https://www.forbes.com/sites/antoniopequenoiv/2026/03/23/nvidias-jensen-huang-says-he-thinks-weve-achieved-agi/<br/><br/>Axios Article: https://archive.fo/20260326100140/https://www.axios.com/2026/03/26/anthropic-pentagon-ai-deal#selection-827.0-829.257<br/><br/>https://arcprize.org/arc-agi/3<br/><br/>ARC AGI 3 Paper: https://arcprize.org/media/ARC_AGI_3_Technical_Report.pdf<br/><br/>NetHack Leaderboard: https://balrogai.com/<br/>Paper: https://ai.meta.com/research/publications/the-nethack-learning-environment/<br/>https://x.com/_rockt/status/2036864121585438995<br/><br/>Claw Shells: https://x.com/DrJimFan/status/2036494601750716711<br/><br/>OpenAI Automated Researcher: https://www.technologyreview.com/2026/03/20/1134438/openai-is-throwing-everything-into-building-a-fully-automated-researcher/<br/><br/>Patreon Post: https://www.patreon.com/c/aiexplained/posts<br/><br/>Eng Jobs: https://x.com/lennysan/status/2036535460726767793<br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>First look at exclusive reports about OpenAI&apos;s new Spud model, and the model Anthropic think will stir governments to urgency, all in the context of the newly-launched ARC-AGI-3. What does the extreme difficulty of that benchmarks, and its quirky scoring metrics, mean for AI in 2026?<br/><br/>https://assemblyai.com/aiexplained<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:55 - OpenAI Side Quests<br/>01:58 - Claude New Model Coming + Universal Equity?<br/>03:13 - ARC-AGI 3<br/>05:00 - Intentional or Unintentional Gaming?<br/>07:11 - But is it AGI Harbinger? No Harness<br/>09:41 - Not the First<br/>12:32 - Automated Researcher<br/>15:00 - Claw Caveat<br/><br/>Spud: https://www.theinformation.com/articles/openai-ceo-shifts-responsibilities-preps-spud-ai-model?utm_campaign=Editorial&amp;utm_content=Article&amp;utm_medium=organic_social&amp;utm_source=bluesky%2Cfacebook%2Clinkedin%2Cthreads%2Ctwitter&amp;rc=sy0ihq<br/><br/>FT: OpenAI Special Model: https://www.ft.com/content/de9bf0af-b241-424f-8229-5870b1c0d93d?syn-25a6b1a6=1<br/><br/>Jensen Huang: https://www.forbes.com/sites/antoniopequenoiv/2026/03/23/nvidias-jensen-huang-says-he-thinks-weve-achieved-agi/<br/><br/>Axios Article: https://archive.fo/20260326100140/https://www.axios.com/2026/03/26/anthropic-pentagon-ai-deal#selection-827.0-829.257<br/><br/>https://arcprize.org/arc-agi/3<br/><br/>ARC AGI 3 Paper: https://arcprize.org/media/ARC_AGI_3_Technical_Report.pdf<br/><br/>NetHack Leaderboard: https://balrogai.com/<br/>Paper: https://ai.meta.com/research/publications/the-nethack-learning-environment/<br/>https://x.com/_rockt/status/2036864121585438995<br/><br/>Claw Shells: https://x.com/DrJimFan/status/2036494601750716711<br/><br/>OpenAI Automated Researcher: https://www.technologyreview.com/2026/03/20/1134438/openai-is-throwing-everything-into-building-a-fully-automated-researcher/<br/><br/>Patreon Post: https://www.patreon.com/c/aiexplained/posts<br/><br/>Eng Jobs: https://x.com/lennysan/status/2036535460726767793<br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18914351-openai-spud-a-claude-model-set-to-stir-governments-beast-mode-arc-agi-3.mp3" length="11880041" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/u3zdl7k1isulcn0zzg8n0jcd0jpq?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18914351</guid>
    <pubDate>Thu, 26 Mar 2026 20:00:00 +0000</pubDate>
    <itunes:duration>987</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>11</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>What the New ChatGPT 5.4 Means for the World</itunes:title>
    <title>What the New ChatGPT 5.4 Means for the World</title>
    <itunes:summary><![CDATA[Just 48 hours after releasing GPT 5.3 Instant, OpenAI have released GPT 5.4 Thinking, so either their is an imminent singularity or perhaps we are being distracted from other news. This video will give 9 crucial bits of context, not just on the GPT 5.4 drop but on the background to the meltdown between the Pentagon and Anthropic. What does this say about the state of AI progress, your job, and what is next.   Check out my fast-growing (!) app, free to use, and code INSIDER15 for 15% off paid ...]]></itunes:summary>
    <description><![CDATA[<p>Just 48 hours after releasing GPT 5.3 Instant, OpenAI have released GPT 5.4 Thinking, so either their is an imminent singularity or perhaps we are being distracted from other news. This video will give 9 crucial bits of context, not just on the GPT 5.4 drop but on the background to the meltdown between the Pentagon and Anthropic. What does this say about the state of AI progress, your job, and what is next.<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for 15% off paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:06: GPT 5.4 Breakdown<br/>05:06 - Closing the Loop<br/>06:35 - Spiky Performance<br/>10:31 - Advice<br/>11:32 - Less Encouraging Developments - Fired Like Dogs<br/>17:45 - But Used in Iran<br/><br/><br/>GPT 5.4: https://openai.com/index/introducing-gpt-5-4/<br/><br/>Hallucinations: https://artificialanalysis.ai/evaluations/omniscience<br/>Investment Banking Bench: https://x.com/bradlightcap/status/2029684672343728452<br/>Move 37: https://x.com/nasqret/status/2029628846518010099<br/>System Card: https://deploymentsafety.openai.com/gpt-5-4-thinking/gpt-5-4-thinking.pdf<br/><br/>Prediction Market Scandal: https://www.wired.com/story/openai-fires-employee-insider-trading-polymarket-kalshi/<br/><br/><br/>GPT 5.3 Instant: https://openai.com/index/gpt-5-3-instant/<br/><br/>GDPVal: https://openai.com/index/gdpval/<br/><br/>Claude in Iran: https://www.washingtonpost.com/technology/2026/03/04/anthropic-ai-iran-campaign<br/><br/>‘Like Dogs’: https://x.com/AndrewCurran_/status/2029605783311470679<br/><br/>Altman leak: https://www.cnbc.com/2026/03/03/sam-altman-tells-openai-staff-operational-decisions-up-to-government.html<br/><br/>Original 2024 Switch: https://archive.fo/20240116172526/https://www.bloomberg.com/news/articles/2024-01-16/openai-working-with-us-military-on-cybersecurity-tools-for-veterans#selection-6173.83-6173.226<br/><br/>Amodei Original Memo: https://www.theinformation.com/articles/read-anthropic-ceos-memo-attacking-openais-mendacious-pentagon-announcement?rc=sy0ihq<br/>Anthropic Apology: https://www.anthropic.com/news/where-stand-department-war<br/>OpenAI Employee Reaction: https://x.com/tszzl/status/2029334980481212820<br/><br/>DoD Suppler Risk: https://www.cnbc.com/amp/2026/03/05/anthropic-pentagon-ai-claude-iran.html<br/>Atlantic Exclusive: https://archive.fo/20260301152646/https://www.theatlantic.com/technology/2026/03/inside-anthropics-killer-robot-dispute-with-the-pentagon/686200/#selection-941.61-941.212<br/>No Negotiation: https://x.com/USWREMichael/status/2029754965778907493<br/><br/>$20B Doubling: https://archive.ph/20260304111124/https://www.bloomberg.com/news/articles/2026-03-03/anthropic-nears-20-billion-revenue-run-rate-amid-pentagon-feud<br/><br/>March 2022 Interview: https://www.youtube.com/watch?v=uAA6PZkek4A<br/><br/>https://lmcouncil.ai/<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Just 48 hours after releasing GPT 5.3 Instant, OpenAI have released GPT 5.4 Thinking, so either their is an imminent singularity or perhaps we are being distracted from other news. This video will give 9 crucial bits of context, not just on the GPT 5.4 drop but on the background to the meltdown between the Pentagon and Anthropic. What does this say about the state of AI progress, your job, and what is next.<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for 15% off paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:06: GPT 5.4 Breakdown<br/>05:06 - Closing the Loop<br/>06:35 - Spiky Performance<br/>10:31 - Advice<br/>11:32 - Less Encouraging Developments - Fired Like Dogs<br/>17:45 - But Used in Iran<br/><br/><br/>GPT 5.4: https://openai.com/index/introducing-gpt-5-4/<br/><br/>Hallucinations: https://artificialanalysis.ai/evaluations/omniscience<br/>Investment Banking Bench: https://x.com/bradlightcap/status/2029684672343728452<br/>Move 37: https://x.com/nasqret/status/2029628846518010099<br/>System Card: https://deploymentsafety.openai.com/gpt-5-4-thinking/gpt-5-4-thinking.pdf<br/><br/>Prediction Market Scandal: https://www.wired.com/story/openai-fires-employee-insider-trading-polymarket-kalshi/<br/><br/><br/>GPT 5.3 Instant: https://openai.com/index/gpt-5-3-instant/<br/><br/>GDPVal: https://openai.com/index/gdpval/<br/><br/>Claude in Iran: https://www.washingtonpost.com/technology/2026/03/04/anthropic-ai-iran-campaign<br/><br/>‘Like Dogs’: https://x.com/AndrewCurran_/status/2029605783311470679<br/><br/>Altman leak: https://www.cnbc.com/2026/03/03/sam-altman-tells-openai-staff-operational-decisions-up-to-government.html<br/><br/>Original 2024 Switch: https://archive.fo/20240116172526/https://www.bloomberg.com/news/articles/2024-01-16/openai-working-with-us-military-on-cybersecurity-tools-for-veterans#selection-6173.83-6173.226<br/><br/>Amodei Original Memo: https://www.theinformation.com/articles/read-anthropic-ceos-memo-attacking-openais-mendacious-pentagon-announcement?rc=sy0ihq<br/>Anthropic Apology: https://www.anthropic.com/news/where-stand-department-war<br/>OpenAI Employee Reaction: https://x.com/tszzl/status/2029334980481212820<br/><br/>DoD Suppler Risk: https://www.cnbc.com/amp/2026/03/05/anthropic-pentagon-ai-claude-iran.html<br/>Atlantic Exclusive: https://archive.fo/20260301152646/https://www.theatlantic.com/technology/2026/03/inside-anthropics-killer-robot-dispute-with-the-pentagon/686200/#selection-941.61-941.212<br/>No Negotiation: https://x.com/USWREMichael/status/2029754965778907493<br/><br/>$20B Doubling: https://archive.ph/20260304111124/https://www.bloomberg.com/news/articles/2026-03-03/anthropic-nears-20-billion-revenue-run-rate-amid-pentagon-feud<br/><br/>March 2022 Interview: https://www.youtube.com/watch?v=uAA6PZkek4A<br/><br/>https://lmcouncil.ai/<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18802064-what-the-new-chatgpt-5-4-means-for-the-world.mp3" length="15763853" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/mtptzxjaphskh5e79tfw7m3rppwc?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18802064</guid>
    <pubDate>Fri, 06 Mar 2026 16:00:00 +0000</pubDate>
    <itunes:duration>1311</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>10</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Deadline Day for Autonomous AI Weapons &amp; Mass Surveillance</itunes:title>
    <title>Deadline Day for Autonomous AI Weapons &amp; Mass Surveillance</title>
    <itunes:summary><![CDATA[Will Anthropic be forced to make a version of Claude for war? And does a new paper expose the risks of Claude agents, in both OpenClaw and the field of war? Plus, 5 more twists in the story of the Pentagon versus Anthropic + some AI lab employees, and a petition that could change everything, or nothing...   Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduct...]]></itunes:summary>
    <description><![CDATA[<p>Will Anthropic be forced to make a version of Claude for war? And does a new paper expose the risks of Claude agents, in both OpenClaw and the field of war? Plus, 5 more twists in the story of the Pentagon versus Anthropic + some AI lab employees, and a petition that could change everything, or nothing...<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:44 - Deadline Day + Petition<br/>02:42 - Twist 1: Existing Deal<br/>03:26 - Twist 2: Existing Policy<br/>04:21 - Twist 3: Twin Threats<br/>05:54 - Twist 4: Interesting Objections<br/>11:32 - Twist 5: Anthropic’s Dropped Policy<br/><br/><br/>Dario Statement: https://www.anthropic.com/news/statement-department-of-war<br/><br/>Google/OpenAI Petition: https://notdivided.org/<br/><br/>Axios on Amodei Rejection: https://www.axios.com/2026/02/26/anthropic-rejects-pentagon-ai-terms<br/><br/>FT on US Threat: https://www.ft.com/content/11d27612-d6c5-4cf7-94dd-f65603549b7f<br/><br/>Politico on Latest: https://archive.ph/20260227013117/https://www.politico.com/news/2026/02/26/incoherent-hegseths-anthropic-ultimatum-confounds-ai-policymakers-00800135<br/><br/>The Verge on Current Deal: https://www.theverge.com/ai-artificial-intelligence/883456/anthropic-pentagon-department-of-defense-negotiations<br/><br/>Anthropic RSP change: https://www.anthropic.com/news/responsible-scaling-policy-v3<br/><br/>Time Magazine on RSP: https://time.com/7380854/exclusive-anthropic-drops-flagship-safety-pledge/<br/><br/>Agent of Chaos Paper: https://x.com/NatalieShapira/status/2026062499599319526<br/><br/>AI Agent Reliability Paper: https://arxiv.org/pdf/2602.16666<br/><br/>My Patreon Video: https://www.patreon.com/posts/real-mystery-ai-151647211<br/><br/>Patreon Documentary: https://www.patreon.com/posts/our-new-age-of-133960279 <br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Will Anthropic be forced to make a version of Claude for war? And does a new paper expose the risks of Claude agents, in both OpenClaw and the field of war? Plus, 5 more twists in the story of the Pentagon versus Anthropic + some AI lab employees, and a petition that could change everything, or nothing...<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for paid tiers: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:44 - Deadline Day + Petition<br/>02:42 - Twist 1: Existing Deal<br/>03:26 - Twist 2: Existing Policy<br/>04:21 - Twist 3: Twin Threats<br/>05:54 - Twist 4: Interesting Objections<br/>11:32 - Twist 5: Anthropic’s Dropped Policy<br/><br/><br/>Dario Statement: https://www.anthropic.com/news/statement-department-of-war<br/><br/>Google/OpenAI Petition: https://notdivided.org/<br/><br/>Axios on Amodei Rejection: https://www.axios.com/2026/02/26/anthropic-rejects-pentagon-ai-terms<br/><br/>FT on US Threat: https://www.ft.com/content/11d27612-d6c5-4cf7-94dd-f65603549b7f<br/><br/>Politico on Latest: https://archive.ph/20260227013117/https://www.politico.com/news/2026/02/26/incoherent-hegseths-anthropic-ultimatum-confounds-ai-policymakers-00800135<br/><br/>The Verge on Current Deal: https://www.theverge.com/ai-artificial-intelligence/883456/anthropic-pentagon-department-of-defense-negotiations<br/><br/>Anthropic RSP change: https://www.anthropic.com/news/responsible-scaling-policy-v3<br/><br/>Time Magazine on RSP: https://time.com/7380854/exclusive-anthropic-drops-flagship-safety-pledge/<br/><br/>Agent of Chaos Paper: https://x.com/NatalieShapira/status/2026062499599319526<br/><br/>AI Agent Reliability Paper: https://arxiv.org/pdf/2602.16666<br/><br/>My Patreon Video: https://www.patreon.com/posts/real-mystery-ai-151647211<br/><br/>Patreon Documentary: https://www.patreon.com/posts/our-new-age-of-133960279 <br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18761175-deadline-day-for-autonomous-ai-weapons-mass-surveillance.mp3" length="9858243" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/phxvh7fkj16vsd3f7xpjwjluw8ni?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18761175</guid>
    <pubDate>Fri, 27 Feb 2026 14:00:00 +0000</pubDate>
    <itunes:duration>819</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>9</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Gemini 3.1 Pro and the Downfall of Benchmarks: Welcome to the Vibe Era of AI</itunes:title>
    <title>Gemini 3.1 Pro and the Downfall of Benchmarks: Welcome to the Vibe Era of AI</title>
    <itunes:summary><![CDATA[Do we have a new best AI model, or do we have the downfall of benchmarks in general, as a way of capturing machine intelligence? Full breakdown of Gemini 3.1 Pro, guest-starring the new Sonnet 4.6, plus analysis from 7 papers/posts that will give you much needed context. Oh, and a new record on Simple Bench!  https://epoch.ai/ai-explained-datacenters   Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai  AI Insiders ($9!): https://www.patreon.com/A...]]></itunes:summary>
    <description><![CDATA[<p>Do we have a new best AI model, or do we have the downfall of benchmarks in general, as a way of capturing machine intelligence? Full breakdown of Gemini 3.1 Pro, guest-starring the new Sonnet 4.6, plus analysis from 7 papers/posts that will give you much needed context. Oh, and a new record on Simple Bench!<br/><br/>https://epoch.ai/ai-explained-datacenters<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:30 - Post-training Dominance<br/>04:00 - ARC-AGI 2 Caveat<br/>05:54 - Simple Bench Record<br/>08:22 - Hallucination Caveat<br/>10:05 - Model Card<br/>11:12 - Exponential Coming<br/>12:20 - Amodei on Generalizing<br/>15:10 - One True Benchmark?<br/>17:02 - Other Metrics…<br/><br/>Gemini 3.1 Model Card: https://storage.googleapis.com/deepmind-media/Model-Cards/Gemini-3-1-Pro-Model-Card.pdf<br/><br/>Release: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-1-pro/<br/><br/>Where are Agents deployed?: https://www.anthropic.com/research/measuring-agent-autonomy<br/><br/>Newsletter Post: https://signaltonoise.beehiiv.com/p/4-ai-numbers-that-surprised-me-this-week<br/><br/>Hallucination AA: https://artificialanalysis.ai/evaluations/omniscience<br/><br/>Melanie Mitchell: https://x.com/MelMitchell1/status/2022738363548340526<br/>ARC-AGI-2: https://x.com/arcprize/status/2024522812728496470/photo/1<br/><br/>Chollet on Agentic Coding and ML: https://x.com/fchollet/status/2024519439140737442<br/><br/>METR Caveat: https://metr.org/notes/2026-01-22-time-horizon-limitations/<br/><br/>Talaas Fast: https://chatjimmy.ai/<br/><br/>Amodei Interview Continual learning: https://www.dwarkesh.com/p/dario-amodei-2?open=false#%C2%A7002942-is-continual-learning-necessary-how-will-it-be-solved<br/><br/>Metaculus FutureEval: https://www.metaculus.com/futureeval/<br/><br/>Next Vid to Watch: https://www.patreon.com/posts/what-you-need-to-150647292<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Do we have a new best AI model, or do we have the downfall of benchmarks in general, as a way of capturing machine intelligence? Full breakdown of Gemini 3.1 Pro, guest-starring the new Sonnet 4.6, plus analysis from 7 papers/posts that will give you much needed context. Oh, and a new record on Simple Bench!<br/><br/>https://epoch.ai/ai-explained-datacenters<br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:30 - Post-training Dominance<br/>04:00 - ARC-AGI 2 Caveat<br/>05:54 - Simple Bench Record<br/>08:22 - Hallucination Caveat<br/>10:05 - Model Card<br/>11:12 - Exponential Coming<br/>12:20 - Amodei on Generalizing<br/>15:10 - One True Benchmark?<br/>17:02 - Other Metrics…<br/><br/>Gemini 3.1 Model Card: https://storage.googleapis.com/deepmind-media/Model-Cards/Gemini-3-1-Pro-Model-Card.pdf<br/><br/>Release: https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-1-pro/<br/><br/>Where are Agents deployed?: https://www.anthropic.com/research/measuring-agent-autonomy<br/><br/>Newsletter Post: https://signaltonoise.beehiiv.com/p/4-ai-numbers-that-surprised-me-this-week<br/><br/>Hallucination AA: https://artificialanalysis.ai/evaluations/omniscience<br/><br/>Melanie Mitchell: https://x.com/MelMitchell1/status/2022738363548340526<br/>ARC-AGI-2: https://x.com/arcprize/status/2024522812728496470/photo/1<br/><br/>Chollet on Agentic Coding and ML: https://x.com/fchollet/status/2024519439140737442<br/><br/>METR Caveat: https://metr.org/notes/2026-01-22-time-horizon-limitations/<br/><br/>Talaas Fast: https://chatjimmy.ai/<br/><br/>Amodei Interview Continual learning: https://www.dwarkesh.com/p/dario-amodei-2?open=false#%C2%A7002942-is-continual-learning-necessary-how-will-it-be-solved<br/><br/>Metaculus FutureEval: https://www.metaculus.com/futureeval/<br/><br/>Next Vid to Watch: https://www.patreon.com/posts/what-you-need-to-150647292<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18718459-gemini-3-1-pro-and-the-downfall-of-benchmarks-welcome-to-the-vibe-era-of-ai.mp3" length="13591168" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/tet2xj36tzd9bxy7jhdi3ufgborg?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18718459</guid>
    <pubDate>Fri, 20 Feb 2026 16:00:00 +0000</pubDate>
    <itunes:duration>1130</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>8</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>The Two Best AI Models/Enemies Just Got Released Simultaneously</itunes:title>
    <title>The Two Best AI Models/Enemies Just Got Released Simultaneously</title>
    <itunes:summary><![CDATA[The two models that you will hear discussed for at least the next two months - Claude Opus 4.6 and GPT 5.3 Codex - just got released within 26 mins or each other. The full breakdown of around 250 pages of reports, with just the most interest moments, from the battle of which is best, Claude personhood, the surprising misbehaviour of Opus 4.6, and much more  https://assemblyai.com/aiexplained  Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai  AI ...]]></itunes:summary>
    <description><![CDATA[<p>The two models that you will hear discussed for at least the next two months - Claude Opus 4.6 and GPT 5.3 Codex - just got released within 26 mins or each other. The full breakdown of around 250 pages of reports, with just the most interest moments, from the battle of which is best, Claude personhood, the surprising misbehaviour of Opus 4.6, and much more<br/><br/>https://assemblyai.com/aiexplained<br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai<br/><br/>AI Insiders ($9): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:54 - Self-improvement?<br/>02:44 - Knowledge Work<br/>05:30 - Overly agentic behaviour<br/>09:12 - Who Shouldn’t Use Claude Opus<br/>11:39 - Step-change?<br/>15:09 - Claude’s ‘Personhood’<br/><br/>Hassabis Roadmap: https://www.patreon.com/posts/hassabis-roadmap-149750869<br/><br/>Release of Opus 4.6: https://www.anthropic.com/news/claude-opus-4-6<br/>212 Page System Card: https://www-cdn.anthropic.com/0dd865075ad3132672ee0ab40b05a53f14cf5288.pdf<br/>Claude Code Tip: https://x.com/bcherny/status/2019475897691124107<br/><br/><br/>GPT Codex 5.3: https://openai.com/index/introducing-gpt-5-3-codex/</p><p>System Card: https://openai.com/index/gpt-5-3-codex-system-card/<br/><br/>Browse Comp: https://arxiv.org/pdf/2504.12516v1<br/>Finance Agent: https://www.vals.ai/benchmarks/finance_agent<br/>Terminal Bench 2: https://arxiv.org/pdf/2601.11868<br/>Vending Bench: https://andonlabs.com/blog/opus-4-6-vending-bench<br/><br/>My X post: https://x.com/AIExplainedYT/status/2016851303436095647<br/><br/>Anthropic Apology: https://x.com/ch402/status/2014066134194995256/photo/1<br/><br/>Altman rebuttal: https://x.com/sama/status/2019139174339928189<br/>https://x.com/sama/status/2019140276246442089<br/><br/>4% of GitHub: https://x.com/dylan522p/status/2019490550911766763<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>The two models that you will hear discussed for at least the next two months - Claude Opus 4.6 and GPT 5.3 Codex - just got released within 26 mins or each other. The full breakdown of around 250 pages of reports, with just the most interest moments, from the battle of which is best, Claude personhood, the surprising misbehaviour of Opus 4.6, and much more<br/><br/>https://assemblyai.com/aiexplained<br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai<br/><br/>AI Insiders ($9): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:54 - Self-improvement?<br/>02:44 - Knowledge Work<br/>05:30 - Overly agentic behaviour<br/>09:12 - Who Shouldn’t Use Claude Opus<br/>11:39 - Step-change?<br/>15:09 - Claude’s ‘Personhood’<br/><br/>Hassabis Roadmap: https://www.patreon.com/posts/hassabis-roadmap-149750869<br/><br/>Release of Opus 4.6: https://www.anthropic.com/news/claude-opus-4-6<br/>212 Page System Card: https://www-cdn.anthropic.com/0dd865075ad3132672ee0ab40b05a53f14cf5288.pdf<br/>Claude Code Tip: https://x.com/bcherny/status/2019475897691124107<br/><br/><br/>GPT Codex 5.3: https://openai.com/index/introducing-gpt-5-3-codex/</p><p>System Card: https://openai.com/index/gpt-5-3-codex-system-card/<br/><br/>Browse Comp: https://arxiv.org/pdf/2504.12516v1<br/>Finance Agent: https://www.vals.ai/benchmarks/finance_agent<br/>Terminal Bench 2: https://arxiv.org/pdf/2601.11868<br/>Vending Bench: https://andonlabs.com/blog/opus-4-6-vending-bench<br/><br/>My X post: https://x.com/AIExplainedYT/status/2016851303436095647<br/><br/>Anthropic Apology: https://x.com/ch402/status/2014066134194995256/photo/1<br/><br/>Altman rebuttal: https://x.com/sama/status/2019139174339928189<br/>https://x.com/sama/status/2019140276246442089<br/><br/>4% of GitHub: https://x.com/dylan522p/status/2019490550911766763<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18636326-the-two-best-ai-models-enemies-just-got-released-simultaneously.mp3" length="14305301" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/m11lhfyionipeved97c747v1lkc0?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18636326</guid>
    <pubDate>Fri, 06 Feb 2026 17:00:00 +0000</pubDate>
    <itunes:duration>1189</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>7</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Claude AI Co-founder Publishes 4 Big Claims about Near Future: Breakdown</itunes:title>
    <title>Claude AI Co-founder Publishes 4 Big Claims about Near Future: Breakdown</title>
    <itunes:summary><![CDATA[Anthropic's CEO, who has consistently predicted transformative AI will arrive before 2030, recently published a nearly 20,000-word essay outlining his vision of where AI is heading. The video gives you the highlights. The essay argues that scaling and recursion will advance AI from coding automation to full engineering automation, while warning of economic displacement within 1-2 years and China's trajectory toward AI-enabled totalitarianism. Additionally, Dario Amodei predicts that AI models...]]></itunes:summary>
    <description><![CDATA[<p>Anthropic&apos;s CEO, who has consistently predicted transformative AI will arrive before 2030, recently published a nearly 20,000-word essay outlining his vision of where AI is heading. The video gives you the highlights. The essay argues that scaling and recursion will advance AI from coding automation to full engineering automation, while warning of economic displacement within 1-2 years and China&apos;s trajectory toward AI-enabled totalitarianism. Additionally, Dario Amodei predicts that AI models will increasingly be understood as collections of distinct personas rather than monolithic systems.<br/><br/>80,000 Hours: https://www.youtube.com/watch?v=B54EQiuO1UU<br/><br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:10 - Scaling to software engineers<br/>06:11 - Permanent Underclass<br/>10:18 - Totalitarian Nightmares<br/>16:38 - Collection of Personas<br/><br/>Essay: https://www.darioamodei.com/essay/the-adolescence-of-technology<br/><br/>Physics Prediction: https://www.quantamagazine.org/is-particle-physics-dead-dying-or-just-hard-20260126/<br/><br/>Axios: https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/><br/>World GDP: https://data.worldbank.org/indicator/NY.GDP.MKTP.KD.ZG?end=2024&amp;start=1961&amp;view=chart<br/><br/>Demis Hassabis Counter: https://www.youtube.com/watch?v=q6fq4_uP7aM<br/><br/>Karpathy 80%: https://x.com/karpathy/status/2015883857489522876<br/><br/>Machines of Loving Grace: https://www.darioamodei.com/essay/machines-of-loving-grace<br/><br/>Anthropic LessWrong: https://www.lesswrong.com/posts/5aKRshJzhojqfbRyo/unless-its-governance-changes-anthropic-is-untrustworthy#1__In_private__Dario_frequently_said_he_won_t_push_the_frontier_of_AI_capabilities__later__Anthropic_pushed_the_frontier<br/><br/>Original Constitution: https://www.anthropic.com/news/claudes-constitution<br/><br/>New Constitution: https://www.anthropic.com/constitution<br/><br/>Kimi K2.5: https://x.com/Kimi_Moonshot/status/2016024049869324599<br/><br/>Societies of Thought, Google DeepMind Paper: https://arxiv.org/pdf/2601.10825<br/><br/>https://lmcouncil.ai/benchmarks<br/><br/>https://www.patreon.com/posts/our-new-age-of-133960279<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Anthropic&apos;s CEO, who has consistently predicted transformative AI will arrive before 2030, recently published a nearly 20,000-word essay outlining his vision of where AI is heading. The video gives you the highlights. The essay argues that scaling and recursion will advance AI from coding automation to full engineering automation, while warning of economic displacement within 1-2 years and China&apos;s trajectory toward AI-enabled totalitarianism. Additionally, Dario Amodei predicts that AI models will increasingly be understood as collections of distinct personas rather than monolithic systems.<br/><br/>80,000 Hours: https://www.youtube.com/watch?v=B54EQiuO1UU<br/><br/><br/><br/>Check out my fast-growing (!) app, free to use, and code INSIDER15 for Pro: https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:10 - Scaling to software engineers<br/>06:11 - Permanent Underclass<br/>10:18 - Totalitarian Nightmares<br/>16:38 - Collection of Personas<br/><br/>Essay: https://www.darioamodei.com/essay/the-adolescence-of-technology<br/><br/>Physics Prediction: https://www.quantamagazine.org/is-particle-physics-dead-dying-or-just-hard-20260126/<br/><br/>Axios: https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/><br/>World GDP: https://data.worldbank.org/indicator/NY.GDP.MKTP.KD.ZG?end=2024&amp;start=1961&amp;view=chart<br/><br/>Demis Hassabis Counter: https://www.youtube.com/watch?v=q6fq4_uP7aM<br/><br/>Karpathy 80%: https://x.com/karpathy/status/2015883857489522876<br/><br/>Machines of Loving Grace: https://www.darioamodei.com/essay/machines-of-loving-grace<br/><br/>Anthropic LessWrong: https://www.lesswrong.com/posts/5aKRshJzhojqfbRyo/unless-its-governance-changes-anthropic-is-untrustworthy#1__In_private__Dario_frequently_said_he_won_t_push_the_frontier_of_AI_capabilities__later__Anthropic_pushed_the_frontier<br/><br/>Original Constitution: https://www.anthropic.com/news/claudes-constitution<br/><br/>New Constitution: https://www.anthropic.com/constitution<br/><br/>Kimi K2.5: https://x.com/Kimi_Moonshot/status/2016024049869324599<br/><br/>Societies of Thought, Google DeepMind Paper: https://arxiv.org/pdf/2601.10825<br/><br/>https://lmcouncil.ai/benchmarks<br/><br/>https://www.patreon.com/posts/our-new-age-of-133960279<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18584414-claude-ai-co-founder-publishes-4-big-claims-about-near-future-breakdown.mp3" length="16027580" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ws6yeypnsqaisxfsswt5ie5htird?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18584414</guid>
    <pubDate>Wed, 28 Jan 2026 15:00:00 +0000</pubDate>
    <itunes:duration>1332</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>6</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Anthropic: Our AI just created a tool that can ‘automate all white collar work’, Me:</itunes:title>
    <title>Anthropic: Our AI just created a tool that can ‘automate all white collar work’, Me:</title>
    <itunes:summary><![CDATA[A new tool, with code written by an AI model, has gone omega-viral: Claude Cowork. But is the hype justified? What do the stats say on productivity? Where is the truth in a sea of noise? What is truth? Can we handle the truth? Where's Nemo?  https://matsprogram.org/s26-aie   Check out my new app! https://lmcouncil.ai  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters:  00:00 - Introduction 01:12 - Claude Cowork 06:48 - Productivity Speed-up + jobs 09:33 - Comparing Models ...]]></itunes:summary>
    <description><![CDATA[<p>A new tool, with code written by an AI model, has gone omega-viral: Claude Cowork. But is the hype justified? What do the stats say on productivity? Where is the truth in a sea of noise? What is truth? Can we handle the truth? Where&apos;s Nemo?<br/><br/>https://matsprogram.org/s26-aie<br/><br/><br/>Check out my new app! https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters: <br/>00:00 - Introduction<br/>01:12 - Claude Cowork<br/>06:48 - Productivity Speed-up + jobs<br/>09:33 - Comparing Models<br/>12:00 - Brittle AI Paper<br/><br/>Cowork Intro: https://x.com/claudeai/thread/2010805682434666759<br/><br/>&apos;All of it&apos;: https://x.com/bcherny/status/2010813886052581538<br/><br/>&apos;AGI&apos; Claims: https://x.com/deepfates/status/2004994698335879383<br/><br/>Douglas Interview: https://www.youtube.com/watch?v=TOsNrV3bXtQ&amp;t=2313s<br/><br/>Job Stats: https://www.oxfordeconomics.com/wp-content/uploads/2026/01/Evidence-of-an-AI-driven-shakeup-of-job-markets-is-patchy.pdf<br/>Amodei Prediction: https://fortune.com/2025/05/28/anthropic-ceo-warning-ai-job-loss/<br/><br/>GenAI Traffic: https://x.com/demishassabis/status/2009075877347512545<br/><br/>Illusion of Insight: https://arxiv.org/pdf/2601.00514<br/>Entropy Exploration: https://arxiv.org/pdf/2506.14758<br/>ProRL: https://arxiv.org/pdf/2505.24864<br/><br/>Genesis Mission: https://www.whitehouse.gov/presidential-actions/2025/11/launching-the-genesis-mission/<br/>https://deepmind.google/blog/how-were-supporting-better-tropical-cyclone-prediction-with-ai/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>A new tool, with code written by an AI model, has gone omega-viral: Claude Cowork. But is the hype justified? What do the stats say on productivity? Where is the truth in a sea of noise? What is truth? Can we handle the truth? Where&apos;s Nemo?<br/><br/>https://matsprogram.org/s26-aie<br/><br/><br/>Check out my new app! https://lmcouncil.ai<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters: <br/>00:00 - Introduction<br/>01:12 - Claude Cowork<br/>06:48 - Productivity Speed-up + jobs<br/>09:33 - Comparing Models<br/>12:00 - Brittle AI Paper<br/><br/>Cowork Intro: https://x.com/claudeai/thread/2010805682434666759<br/><br/>&apos;All of it&apos;: https://x.com/bcherny/status/2010813886052581538<br/><br/>&apos;AGI&apos; Claims: https://x.com/deepfates/status/2004994698335879383<br/><br/>Douglas Interview: https://www.youtube.com/watch?v=TOsNrV3bXtQ&amp;t=2313s<br/><br/>Job Stats: https://www.oxfordeconomics.com/wp-content/uploads/2026/01/Evidence-of-an-AI-driven-shakeup-of-job-markets-is-patchy.pdf<br/>Amodei Prediction: https://fortune.com/2025/05/28/anthropic-ceo-warning-ai-job-loss/<br/><br/>GenAI Traffic: https://x.com/demishassabis/status/2009075877347512545<br/><br/>Illusion of Insight: https://arxiv.org/pdf/2601.00514<br/>Entropy Exploration: https://arxiv.org/pdf/2506.14758<br/>ProRL: https://arxiv.org/pdf/2505.24864<br/><br/>Genesis Mission: https://www.whitehouse.gov/presidential-actions/2025/11/launching-the-genesis-mission/<br/>https://deepmind.google/blog/how-were-supporting-better-tropical-cyclone-prediction-with-ai/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18507476-anthropic-our-ai-just-created-a-tool-that-can-automate-all-white-collar-work-me.mp3" length="13210462" type="audio/mpeg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18507476</guid>
    <pubDate>Wed, 14 Jan 2026 16:00:00 +0000</pubDate>
    <itunes:duration>1096</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>5</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>What the Freakiness of 2025 in AI Tells Us About 2026</itunes:title>
    <title>What the Freakiness of 2025 in AI Tells Us About 2026</title>
    <itunes:summary><![CDATA[It’s probably not possible to satisfactorily condense a 12 month’s worth of weird progress in AI, as well as predictions for the year to come, into one video. But I’m gonna try anyway because it has been a very strange time.  http://matsprogram.org/s26-aie   My new app! https://lmcouncil.ai   Patreon Interview: https://www.patreon.com/posts/robot-in-your-27-146376094  Chapters: 00:00 - Introduction 00:34 - Reasoning Models … and limits 02:54 - A playable world 03:36 - Realism 03:50 - AI Slop ...]]></itunes:summary>
    <description><![CDATA[<p>It’s probably not possible to satisfactorily condense a 12 month’s worth of weird progress in AI, as well as predictions for the year to come, into one video. But I’m gonna try anyway because it has been a very strange time.<br/><br/>http://matsprogram.org/s26-aie<br/><br/><br/>My new app! https://lmcouncil.ai<br/><br/><br/>Patreon Interview: https://www.patreon.com/posts/robot-in-your-27-146376094<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:34 - Reasoning Models … and limits<br/>02:54 - A playable world<br/>03:36 - Realism<br/>03:50 - AI Slop gone mainstream<br/>05:03 - DolphinGemma<br/>05:39 - Public Mood<br/>07:34 - AI Enlisted<br/>08:30 - GPT-5<br/>11:05 - Open Weight not out<br/>13:00 - METR Breakout<br/>17:30 - VASA-1<br/>18:28 - Lateral Productivity<br/>20:15 - 1 or 1000 benchmarks needed?<br/>24:54 - Continual Learning + Altman on Superintelligence<br/>28:08 - Automated Information Discovery ft AlphaEvolve<br/><br/><br/>Hassabis on Generality: https://x.com/demishassabis/status/2003097405026193809<br/>https://www.youtube.com/watch?v=PqVbypvxDto<br/><br/>Gemini 3: https://storage.googleapis.com/gweb-uniblog-publish-prod/original_images/gemini_3_table_final_HLE_Tools_on.gif<br/>Reasoning Trade-offs: https://arxiv.org/pdf/2504.13837<br/><br/>DolphinGemma: https://blog.google/technology/ai/dolphingemma/?s=09<br/><br/>Genie 3: https://deepmind.google/blog/genie-3-a-new-frontier-for-world-models/<br/><br/>METR Time Horizon: https://arxiv.org/pdf/2503.14499<br/>https://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/<br/>Flaws: https://x.com/ShashwatGoel7/status/2002369517499105443<br/>https://shash42.substack.com/p/how-to-game-the-metr-plot<br/>https://x.com/METR_Evals/status/2002203627377574113<br/><br/>GPT-5 - Altman phd in everything: https://edition.cnn.com/2025/08/14/business/chatgpt-rollout-problems<br/><br/>https://simple-bench.com/<br/><br/>AI Slop: https://www.youtube.com/watch?v=I_3vxoJDD9k<br/>https://www.theguardian.com/technology/2025/dec/16/boost-for-artists-in-ai-copyright-battle-as-only-3-per-cent-back-uk-active-opt-out-plan<br/><br/>Survey: https://x.com/SearchlightInst/status/2001057144842387920/photo/1<br/><br/>Nvidia Nemotron: https://x.com/percyliang/status/2000608134205985169<br/><br/>OpenAI Compute Flywheel: https://x.com/OpenAI/status/2001363007209914399/photo/1<br/>Altman Interview: https://www.youtube.com/watch?v=2P27Ef-LLuQ<br/><br/>AI in Govt: https://x.com/jdcmedlock/status/1939814516503847259<br/><br/>Benchmark Gaming: https://techcrunch.com/2025/04/07/meta-exec-denies-the-company-artificially-boosted-llama-4s-benchmark-scores/<br/><br/>AlphaEvolve: https://deepmind.google/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/<br/>https://storage.googleapis.com/deepmind-media/DeepMind.com/Blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/AlphaEvolve.pdf?utm_source=deepmind.google&amp;utm_medium=referral&amp;utm_campaign=gdm&amp;utm_content=<br/>Continual Learning: https://abehrouz.github.io/files/NL.pdf<br/><br/>Job Risk: https://archive.ph/20250708204527/https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/><br/>GPT4o: https://x.com/AISafetyMemes/status/1916889492172013989<br/><br/>Vasa-1: https://www.microsoft.com/en-us/research/project/vasa-1/<br/><br/>Three Views: https://www.lesswrong.com/posts/K2D45BNxnZjdpSX2j/ai-timelines<br/>Turing Test: https://x.com/tunguz/status/1907185471211422147<br/><br/>Karpathy Year in Review: https://karpathy.bearblog.dev/year-in-review-2025/<br/><br/>LLM Brainrot: https://arxiv.org/pdf/2510.13928<br/><br/>Lateral Productivity: https://www.aisi.gov.uk/frontier-ai-trends-report<br/><br/>Emotional Quotient: https://arxiv.org/pdf/2511.08394<br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained</p>]]></description>
    <content:encoded><![CDATA[<p>It’s probably not possible to satisfactorily condense a 12 month’s worth of weird progress in AI, as well as predictions for the year to come, into one video. But I’m gonna try anyway because it has been a very strange time.<br/><br/>http://matsprogram.org/s26-aie<br/><br/><br/>My new app! https://lmcouncil.ai<br/><br/><br/>Patreon Interview: https://www.patreon.com/posts/robot-in-your-27-146376094<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:34 - Reasoning Models … and limits<br/>02:54 - A playable world<br/>03:36 - Realism<br/>03:50 - AI Slop gone mainstream<br/>05:03 - DolphinGemma<br/>05:39 - Public Mood<br/>07:34 - AI Enlisted<br/>08:30 - GPT-5<br/>11:05 - Open Weight not out<br/>13:00 - METR Breakout<br/>17:30 - VASA-1<br/>18:28 - Lateral Productivity<br/>20:15 - 1 or 1000 benchmarks needed?<br/>24:54 - Continual Learning + Altman on Superintelligence<br/>28:08 - Automated Information Discovery ft AlphaEvolve<br/><br/><br/>Hassabis on Generality: https://x.com/demishassabis/status/2003097405026193809<br/>https://www.youtube.com/watch?v=PqVbypvxDto<br/><br/>Gemini 3: https://storage.googleapis.com/gweb-uniblog-publish-prod/original_images/gemini_3_table_final_HLE_Tools_on.gif<br/>Reasoning Trade-offs: https://arxiv.org/pdf/2504.13837<br/><br/>DolphinGemma: https://blog.google/technology/ai/dolphingemma/?s=09<br/><br/>Genie 3: https://deepmind.google/blog/genie-3-a-new-frontier-for-world-models/<br/><br/>METR Time Horizon: https://arxiv.org/pdf/2503.14499<br/>https://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/<br/>Flaws: https://x.com/ShashwatGoel7/status/2002369517499105443<br/>https://shash42.substack.com/p/how-to-game-the-metr-plot<br/>https://x.com/METR_Evals/status/2002203627377574113<br/><br/>GPT-5 - Altman phd in everything: https://edition.cnn.com/2025/08/14/business/chatgpt-rollout-problems<br/><br/>https://simple-bench.com/<br/><br/>AI Slop: https://www.youtube.com/watch?v=I_3vxoJDD9k<br/>https://www.theguardian.com/technology/2025/dec/16/boost-for-artists-in-ai-copyright-battle-as-only-3-per-cent-back-uk-active-opt-out-plan<br/><br/>Survey: https://x.com/SearchlightInst/status/2001057144842387920/photo/1<br/><br/>Nvidia Nemotron: https://x.com/percyliang/status/2000608134205985169<br/><br/>OpenAI Compute Flywheel: https://x.com/OpenAI/status/2001363007209914399/photo/1<br/>Altman Interview: https://www.youtube.com/watch?v=2P27Ef-LLuQ<br/><br/>AI in Govt: https://x.com/jdcmedlock/status/1939814516503847259<br/><br/>Benchmark Gaming: https://techcrunch.com/2025/04/07/meta-exec-denies-the-company-artificially-boosted-llama-4s-benchmark-scores/<br/><br/>AlphaEvolve: https://deepmind.google/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/<br/>https://storage.googleapis.com/deepmind-media/DeepMind.com/Blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/AlphaEvolve.pdf?utm_source=deepmind.google&amp;utm_medium=referral&amp;utm_campaign=gdm&amp;utm_content=<br/>Continual Learning: https://abehrouz.github.io/files/NL.pdf<br/><br/>Job Risk: https://archive.ph/20250708204527/https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/><br/>GPT4o: https://x.com/AISafetyMemes/status/1916889492172013989<br/><br/>Vasa-1: https://www.microsoft.com/en-us/research/project/vasa-1/<br/><br/>Three Views: https://www.lesswrong.com/posts/K2D45BNxnZjdpSX2j/ai-timelines<br/>Turing Test: https://x.com/tunguz/status/1907185471211422147<br/><br/>Karpathy Year in Review: https://karpathy.bearblog.dev/year-in-review-2025/<br/><br/>LLM Brainrot: https://arxiv.org/pdf/2510.13928<br/><br/>Lateral Productivity: https://www.aisi.gov.uk/frontier-ai-trends-report<br/><br/>Emotional Quotient: https://arxiv.org/pdf/2511.08394<br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18407004-what-the-freakiness-of-2025-in-ai-tells-us-about-2026.mp3" length="24112528" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/cp24cc8hwmufukd3juhcw4xfes8u?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18407004</guid>
    <pubDate>Tue, 23 Dec 2025 17:00:00 +0000</pubDate>
    <itunes:duration>2006</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>4</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Gemini Exponential, Demis Hassabis&#39; ‘Proto-AGI’ coming, but …</itunes:title>
    <title>Gemini Exponential, Demis Hassabis&#39; ‘Proto-AGI’ coming, but …</title>
    <itunes:summary><![CDATA[The condensed highlights of hours of AI lab leader interviews, model releases, Gemini 3 Flash insights (plus it’s hidden flaw), Hassabis’ ‘proto-AGI’ and much more…  https://matsprogram.org/apply?utm_source=ai-explained&amp;utm_medium=youtube&amp;utm_campaign=s26    Also, do check out my new app: https://lmcouncil.ai  Chapters:  00:00 - Introduction 00:50 - Results 02:44 - But… the Flaw 04:49 - So Benchmarks are fake? No 07:37 - Spatial Reasoning + Hassabis 10:06 - Proto-AGI 12:07 -...]]></itunes:summary>
    <description><![CDATA[<p>The condensed highlights of hours of AI lab leader interviews, model releases, Gemini 3 Flash insights (plus it’s hidden flaw), Hassabis’ ‘proto-AGI’ and much more…<br/><br/>https://matsprogram.org/apply?utm_source=ai-explained&amp;utm_medium=youtube&amp;utm_campaign=s26  <br/><br/>Also, do check out my new app: https://lmcouncil.ai<br/><br/>Chapters: <br/>00:00 - Introduction<br/>00:50 - Results<br/>02:44 - But… the Flaw<br/>04:49 - So Benchmarks are fake? No<br/>07:37 - Spatial Reasoning + Hassabis<br/>10:06 - Proto-AGI<br/>12:07 - Minimal AGI<br/>15:07 - Compute Slowdown<br/>17:56 - New Data Paradigm<br/><br/>Gemini 3 Flash: https://deepmind.google/models/gemini/flash/<br/><br/>Hassabis Interview: https://www.youtube.com/watch?v=PqVbypvxDto<br/>Legg Interview: https://www.youtube.com/watch?v=l3u_FAv33G0<br/>Pre-training Lead Interview: https://www.youtube.com/watch?v=cNGDAqFXvew<br/>Altman Interview: https://www.youtube.com/watch?v=2P27Ef-LLuQ<br/>Brockman Video: https://x.com/OpenAI/status/2001336514786017417<br/>Post-Training Reveal: https://x.com/OfficialLoganK/status/2001742530472534442<br/><br/>Hallucinations Paper: https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4aaa5/why-language-models-hallucinate.pdf<br/>Patreon Hallucinations Vid: https://www.patreon.com/posts/blockers-to-and-139264812<br/>AA-Omniscience Benchmark: https://artificialanalysis.ai/evaluations/omniscience<br/>https://arxiv.org/pdf/2511.13029<br/><br/><br/>lmcouncil.ai/benchmarks <br/>https://simple-bench.com/<br/>https://x.com/scaling01/status/1999620587744813205<br/><br/>5.2 Codex Drop: https://cdn.openai.com/pdf/ac7c37ae-7f4c-4442-b741-2eabdeaf77e0/oai_5_2_Codex.pdf<br/><br/>OpenAI Compute Trend: https://www.theinformation.com/articles/openais-350-billion-computing-cost-problem?rc=sy0ihq<br/><br/>Cramer Tweet/Response: https://x.com/BorisMPower/status/2001440650210976018<br/><br/>OpenAI Valuation: ​​https://www.theinformation.com/articles/openai-discussed-raising-tens-billions-valuation-around-750-billion?rc=sy0ihq<br/><br/>Indian Data: https://www.reuters.com/world/india/with-freebies-openai-google-vie-indian-users-training-data-2025-12-17/<br/><br/>TheInformation Data: https://x.com/theinformation/status/2001421225751351778<br/><br/>Genie 3: https://deepmind.google/blog/genie-3-a-new-frontier-for-world-models/<br/>Sima 2: https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/<br/>Veo 3.1: https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/<br/><br/>METR: https://metr.org/blohttps://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/2025-03-19-measuring-ai-ability-to-complete-long-tasks/<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>The condensed highlights of hours of AI lab leader interviews, model releases, Gemini 3 Flash insights (plus it’s hidden flaw), Hassabis’ ‘proto-AGI’ and much more…<br/><br/>https://matsprogram.org/apply?utm_source=ai-explained&amp;utm_medium=youtube&amp;utm_campaign=s26  <br/><br/>Also, do check out my new app: https://lmcouncil.ai<br/><br/>Chapters: <br/>00:00 - Introduction<br/>00:50 - Results<br/>02:44 - But… the Flaw<br/>04:49 - So Benchmarks are fake? No<br/>07:37 - Spatial Reasoning + Hassabis<br/>10:06 - Proto-AGI<br/>12:07 - Minimal AGI<br/>15:07 - Compute Slowdown<br/>17:56 - New Data Paradigm<br/><br/>Gemini 3 Flash: https://deepmind.google/models/gemini/flash/<br/><br/>Hassabis Interview: https://www.youtube.com/watch?v=PqVbypvxDto<br/>Legg Interview: https://www.youtube.com/watch?v=l3u_FAv33G0<br/>Pre-training Lead Interview: https://www.youtube.com/watch?v=cNGDAqFXvew<br/>Altman Interview: https://www.youtube.com/watch?v=2P27Ef-LLuQ<br/>Brockman Video: https://x.com/OpenAI/status/2001336514786017417<br/>Post-Training Reveal: https://x.com/OfficialLoganK/status/2001742530472534442<br/><br/>Hallucinations Paper: https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4aaa5/why-language-models-hallucinate.pdf<br/>Patreon Hallucinations Vid: https://www.patreon.com/posts/blockers-to-and-139264812<br/>AA-Omniscience Benchmark: https://artificialanalysis.ai/evaluations/omniscience<br/>https://arxiv.org/pdf/2511.13029<br/><br/><br/>lmcouncil.ai/benchmarks <br/>https://simple-bench.com/<br/>https://x.com/scaling01/status/1999620587744813205<br/><br/>5.2 Codex Drop: https://cdn.openai.com/pdf/ac7c37ae-7f4c-4442-b741-2eabdeaf77e0/oai_5_2_Codex.pdf<br/><br/>OpenAI Compute Trend: https://www.theinformation.com/articles/openais-350-billion-computing-cost-problem?rc=sy0ihq<br/><br/>Cramer Tweet/Response: https://x.com/BorisMPower/status/2001440650210976018<br/><br/>OpenAI Valuation: ​​https://www.theinformation.com/articles/openai-discussed-raising-tens-billions-valuation-around-750-billion?rc=sy0ihq<br/><br/>Indian Data: https://www.reuters.com/world/india/with-freebies-openai-google-vie-indian-users-training-data-2025-12-17/<br/><br/>TheInformation Data: https://x.com/theinformation/status/2001421225751351778<br/><br/>Genie 3: https://deepmind.google/blog/genie-3-a-new-frontier-for-world-models/<br/>Sima 2: https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/<br/>Veo 3.1: https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/<br/><br/>METR: https://metr.org/blohttps://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/2025-03-19-measuring-ai-ability-to-complete-long-tasks/<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18386002-gemini-exponential-demis-hassabis-proto-agi-coming-but.mp3" length="14420961" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/tal7qkzj5zpvlcg4fzcbgrwjub9a?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18386002</guid>
    <pubDate>Fri, 19 Dec 2025 16:00:00 +0000</pubDate>
    <itunes:duration>1199</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>3</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>GPT 5.2: OpenAI Strikes Back</itunes:title>
    <title>GPT 5.2: OpenAI Strikes Back</title>
    <itunes:summary><![CDATA[Full GPT-5.2 breakdown - did OpenAI reclaim the crown? A story of tokens, time and cost, plus 9 details you wouldn’t get just from reading the headlines.  https://www.youtube.com/@eightythousandhours    AI Insiders ($9!): https://www.patreon.com/AIExplained https://lmcouncil.ai  Chapters: 00:00 - Introduction 00:55 - Better than Human @ Professional Tasks? 04:42 - Test time Compute 07:05 - Benchmark Selection 09:32 - Simple Results + council comparison 13:01 - Long Context 13:52 - Self-Improv...]]></itunes:summary>
    <description><![CDATA[<p>Full GPT-5.2 breakdown - did OpenAI reclaim the crown? A story of tokens, time and cost, plus 9 details you wouldn’t get just from reading the headlines.<br/><br/>https://www.youtube.com/@eightythousandhours<br/><br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/>https://lmcouncil.ai<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:55 - Better than Human @ Professional Tasks?<br/>04:42 - Test time Compute<br/>07:05 - Benchmark Selection<br/>09:32 - Simple Results + council comparison<br/>13:01 - Long Context<br/>13:52 - Self-Improvement<br/>15:00 - 10 Years + New Models<br/><br/>Release Page: https://openai.com/index/introducing-gpt-5-2/<br/><br/>GPT 5.2 Benchmark Comparison: https://www.reddit.com/r/singularity/comments/1pka1y9/gpt52_all_20_benchmarks_rankings_and_pricing/<br/>https://storage.googleapis.com/gweb-uniblog-publish-prod/original_images/gemini_3_table_final_HLE_Tools_on.gif<br/>https://lmcouncil.ai/benchmarks<br/><br/>Charxiv: https://charxiv.github.io/#leaderboard<br/><br/>GDPval: https://arxiv.org/pdf/2510.04374<br/>My vid: https://www.youtube.com/watch?v=oK5LxMaROSA<br/><br/>Kilpatrick: https://x.com/OfficialLoganK/status/1999270402712023158/photo/1<br/><br/>Noam Brown: https://x.com/polynoamial/status/1999189845164667132<br/><br/>New Model in New Year: https://www.theinformation.com/articles/openai-developing-garlic-model-counter-googles-recent-gains?rc=sy0ihq<br/><br/>10 Years of OpenAI: https://openai.com/index/ten-years/<br/><br/>GPQA: https://x.com/idavidrein/status/1841265634170278063<br/><br/>ARC-AGI 1-2: https://arcprize.org/arc-agi/2/<br/><br/>Sunday Robotics: https://x.com/tonyzzhao/status/1991204839578300813<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/><br/>https://lmcouncil.ai</p>]]></description>
    <content:encoded><![CDATA[<p>Full GPT-5.2 breakdown - did OpenAI reclaim the crown? A story of tokens, time and cost, plus 9 details you wouldn’t get just from reading the headlines.<br/><br/>https://www.youtube.com/@eightythousandhours<br/><br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/>https://lmcouncil.ai<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:55 - Better than Human @ Professional Tasks?<br/>04:42 - Test time Compute<br/>07:05 - Benchmark Selection<br/>09:32 - Simple Results + council comparison<br/>13:01 - Long Context<br/>13:52 - Self-Improvement<br/>15:00 - 10 Years + New Models<br/><br/>Release Page: https://openai.com/index/introducing-gpt-5-2/<br/><br/>GPT 5.2 Benchmark Comparison: https://www.reddit.com/r/singularity/comments/1pka1y9/gpt52_all_20_benchmarks_rankings_and_pricing/<br/>https://storage.googleapis.com/gweb-uniblog-publish-prod/original_images/gemini_3_table_final_HLE_Tools_on.gif<br/>https://lmcouncil.ai/benchmarks<br/><br/>Charxiv: https://charxiv.github.io/#leaderboard<br/><br/>GDPval: https://arxiv.org/pdf/2510.04374<br/>My vid: https://www.youtube.com/watch?v=oK5LxMaROSA<br/><br/>Kilpatrick: https://x.com/OfficialLoganK/status/1999270402712023158/photo/1<br/><br/>Noam Brown: https://x.com/polynoamial/status/1999189845164667132<br/><br/>New Model in New Year: https://www.theinformation.com/articles/openai-developing-garlic-model-counter-googles-recent-gains?rc=sy0ihq<br/><br/>10 Years of OpenAI: https://openai.com/index/ten-years/<br/><br/>GPQA: https://x.com/idavidrein/status/1841265634170278063<br/><br/>ARC-AGI 1-2: https://arcprize.org/arc-agi/2/<br/><br/>Sunday Robotics: https://x.com/tonyzzhao/status/1991204839578300813<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/><br/>https://lmcouncil.ai</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18343799-gpt-5-2-openai-strikes-back.mp3" length="12758419" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/1bd5lj9x2kk22yxqqv9y3hlpglp7?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18343799</guid>
    <pubDate>Fri, 12 Dec 2025 16:00:00 +0000</pubDate>
    <itunes:duration>1061</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>2</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>You Are Being Told Contradictory Things About AI: 8 examples</itunes:title>
    <title>You Are Being Told Contradictory Things About AI: 8 examples</title>
    <itunes:summary><![CDATA[With headlines of an imminent job apocalypse, code red for ChatGPT and recursive self-improvement, at the same time as Anthropic's CEO yesterday saying we know how to scale to AGI, and Gemini 3 DeepThink out today, it is easy to get lost among the narratives and counter-narratives. So here are both, plus the facts behind them, for you to decide.   https://epoch.ai/data/data-centers  Epoch AI is the sponsor of today’s video, and my views, and those expressed in this video, do not necessarily r...]]></itunes:summary>
    <description><![CDATA[<p>With headlines of an imminent job apocalypse, code red for ChatGPT and recursive self-improvement, at the same time as Anthropic&apos;s CEO yesterday saying we know how to scale to AGI, and Gemini 3 DeepThink out today, it is easy to get lost among the narratives and counter-narratives. So here are both, plus the facts behind them, for you to decide.<br/><br/><br/>https://epoch.ai/data/data-centers<br/><br/>Epoch AI is the sponsor of today’s video, and my views, and those expressed in this video, do not necessarily reflect Epoch AI’s views in any way.<br/><br/><br/>Chapters: <br/>00:00 - Introduction<br/>00:42 - Job Apocalypse?<br/>01:45 - Scaling to AGI<br/>04:15 - Recursive Self-Improvement Needed, or Not<br/>09:57 - OpenAI Code Red vs Gemini 3 DeepThink vs Claude Opus 4.5<br/>13:27 - DeepSeek Speciale vs Mistral Large v3<br/>16:45 - Claude Soul Document<br/><br/>https://lmcouncil.ai/<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained</p><p><br/><br/>Guardian Interview: https://www.theguardian.com/technology/ng-interactive/2025/dec/02/jared-kaplan-artificial-intelligence-train-itself<br/><br/>MIT Study on Jobs/Tasks: https://iceberg.mit.edu/report.pdf<br/>vs https://www.cnbc.com/2025/11/26/mit-study-finds-ai-can-already-replace-11point7percent-of-us-workforce.html<br/><br/>Amodei on Scaling: https://www.youtube.com/watch?v=FEj7wAjwQIk<br/>Claude Soul Document: https://www.lesswrong.com/posts/vpNG99GhbBoLov9og/claude-4-5-opus-soul-document<br/><br/>Capabilities Original Stance: https://www.anthropic.com/news/core-views-on-ai-safety<br/><br/>Ilya Interview: https://www.dwarkesh.com/p/ilya-sutskever-2<br/><br/>Ricursive Intelligence: https://x.com/RicursiveAI/status/1995932204703346946<br/><br/>Economist Worker Usage of GenAI: https://www.economist.com/finance-and-economics/2025/11/26/investors-expect-ai-use-to-soar-thats-not-happening#selection-1409.94-1413.42<br/><br/>Mistral v3 Large: https://docs.mistral.ai/models/mistral-large-3-25-12<br/><br/>Compute Slowdown Paper: https://joel-becker.com/images/publications/forecasting_time_horizon_under_compute_slowdown.pdf<br/>https://x.com/joel_bkr/status/1993023436541903155<br/><br/>METR Chart: https://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/<br/><br/>https://www.theinformation.com/articles/openais-350-billion-computing-cost-problem?rc=sy0ihq<br/><br/>OpenAI Code Red: https://www.anthropic.com/news/core-views-on-ai-safety<br/>Rocket Company: https://www.independent.co.uk/news/world/americas/sam-altman-rocket-elon-musk-spacex-b2878351.html<br/><br/>DeepSeek Paper: https://arxiv.org/html/2512.02556v1<br/><br/>DeepSeek Crowdstrike CCP: https://www.crowdstrike.com/en-us/blog/crowdstrike-researchers-identify-hidden-vulnerabilities-ai-coded-software/<br/><br/>https://simple-bench.com/<br/><br/>Patreon Post: https://www.patreon.com/c/aiexplained/posts<br/><br/>Robot: https://x.com/jloganolson/status/1985850115379351799</p>]]></description>
    <content:encoded><![CDATA[<p>With headlines of an imminent job apocalypse, code red for ChatGPT and recursive self-improvement, at the same time as Anthropic&apos;s CEO yesterday saying we know how to scale to AGI, and Gemini 3 DeepThink out today, it is easy to get lost among the narratives and counter-narratives. So here are both, plus the facts behind them, for you to decide.<br/><br/><br/>https://epoch.ai/data/data-centers<br/><br/>Epoch AI is the sponsor of today’s video, and my views, and those expressed in this video, do not necessarily reflect Epoch AI’s views in any way.<br/><br/><br/>Chapters: <br/>00:00 - Introduction<br/>00:42 - Job Apocalypse?<br/>01:45 - Scaling to AGI<br/>04:15 - Recursive Self-Improvement Needed, or Not<br/>09:57 - OpenAI Code Red vs Gemini 3 DeepThink vs Claude Opus 4.5<br/>13:27 - DeepSeek Speciale vs Mistral Large v3<br/>16:45 - Claude Soul Document<br/><br/>https://lmcouncil.ai/<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained</p><p><br/><br/>Guardian Interview: https://www.theguardian.com/technology/ng-interactive/2025/dec/02/jared-kaplan-artificial-intelligence-train-itself<br/><br/>MIT Study on Jobs/Tasks: https://iceberg.mit.edu/report.pdf<br/>vs https://www.cnbc.com/2025/11/26/mit-study-finds-ai-can-already-replace-11point7percent-of-us-workforce.html<br/><br/>Amodei on Scaling: https://www.youtube.com/watch?v=FEj7wAjwQIk<br/>Claude Soul Document: https://www.lesswrong.com/posts/vpNG99GhbBoLov9og/claude-4-5-opus-soul-document<br/><br/>Capabilities Original Stance: https://www.anthropic.com/news/core-views-on-ai-safety<br/><br/>Ilya Interview: https://www.dwarkesh.com/p/ilya-sutskever-2<br/><br/>Ricursive Intelligence: https://x.com/RicursiveAI/status/1995932204703346946<br/><br/>Economist Worker Usage of GenAI: https://www.economist.com/finance-and-economics/2025/11/26/investors-expect-ai-use-to-soar-thats-not-happening#selection-1409.94-1413.42<br/><br/>Mistral v3 Large: https://docs.mistral.ai/models/mistral-large-3-25-12<br/><br/>Compute Slowdown Paper: https://joel-becker.com/images/publications/forecasting_time_horizon_under_compute_slowdown.pdf<br/>https://x.com/joel_bkr/status/1993023436541903155<br/><br/>METR Chart: https://metr.org/blog/2025-03-19-measuring-ai-ability-to-complete-long-tasks/<br/><br/>https://www.theinformation.com/articles/openais-350-billion-computing-cost-problem?rc=sy0ihq<br/><br/>OpenAI Code Red: https://www.anthropic.com/news/core-views-on-ai-safety<br/>Rocket Company: https://www.independent.co.uk/news/world/americas/sam-altman-rocket-elon-musk-spacex-b2878351.html<br/><br/>DeepSeek Paper: https://arxiv.org/html/2512.02556v1<br/><br/>DeepSeek Crowdstrike CCP: https://www.crowdstrike.com/en-us/blog/crowdstrike-researchers-identify-hidden-vulnerabilities-ai-coded-software/<br/><br/>https://simple-bench.com/<br/><br/>Patreon Post: https://www.patreon.com/c/aiexplained/posts<br/><br/>Robot: https://x.com/jloganolson/status/1985850115379351799</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18305817-you-are-being-told-contradictory-things-about-ai-8-examples.mp3" length="14621262" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/djs5w48y647o31qqh0g8qw0b8mn0?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18305817</guid>
    <pubDate>Fri, 05 Dec 2025 16:00:00 +0000</pubDate>
    <itunes:duration>1215</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>5</itunes:season>
    <itunes:episode>1</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Gemini 3 is Here: 11 Details You Might Have Missed</itunes:title>
    <title>Gemini 3 is Here: 11 Details You Might Have Missed</title>
    <itunes:summary><![CDATA[Gemini 3 Pro is out, and records fell like snowflakes in Svalbard.   No long description, chapters or links today, huge technical difficulties, including with audio, so just want to publish asap.   https://app.grayswan.ai/ai-explained   https://lmcouncil.ai AI Insiders ($9!): https://www.patreon.com/AIExplained    Non-hype Newsletter: https://signaltonoise.beehiiv.com/ Podcast: https://aiexplainedopodcast.buzzsprout.com/ ]]></itunes:summary>
    <description><![CDATA[<p>Gemini 3 Pro is out, and records fell like snowflakes in Svalbard. <br/><br/>No long description, chapters or links today, huge technical difficulties, including with audio, so just want to publish asap.<br/><br/><br/>https://app.grayswan.ai/ai-explained<br/><br/><br/>https://lmcouncil.ai<br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Gemini 3 Pro is out, and records fell like snowflakes in Svalbard. <br/><br/>No long description, chapters or links today, huge technical difficulties, including with audio, so just want to publish asap.<br/><br/><br/>https://app.grayswan.ai/ai-explained<br/><br/><br/>https://lmcouncil.ai<br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18217832-gemini-3-is-here-11-details-you-might-have-missed.mp3" length="15685549" type="audio/mpeg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18217832</guid>
    <pubDate>Wed, 19 Nov 2025 14:00:00 +0000</pubDate>
    <itunes:duration>1302</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Is GPT-5.1 Really an Upgrade? But Models Can Auto-Hack Govts, so … there’s that</itunes:title>
    <title>Is GPT-5.1 Really an Upgrade? But Models Can Auto-Hack Govts, so … there’s that</title>
    <itunes:summary><![CDATA[A lot just got released in the last 36 hours, and it will all affect hundreds of millions of people. 10 details you would miss if you just read the headlines, from GPT 5.1 regressions, to how Claude hacked Govt Agencies, to SIMA 2, and Musical Turing Tests. https://assemblyai.com/aiexplained Chapters: 00:00 - Introduction 00:56 - GPT 5.1 Smarter? 01:47 - Some Regressions 03:22 - Sycophancy? 05:22 - Claude Auto-Hacking  06:16 - Jailbreaking through Granularity 08:22 - This Will be Re-used...]]></itunes:summary>
    <description><![CDATA[<p><b>A lot just got released in the last 36 hours, and it will all affect hundreds of millions of people. 10 details you would miss if you just read the headlines, from GPT 5.1 regressions, to how Claude hacked Govt Agencies, to SIMA 2, and Musical Turing Tests.</b></p><p><a href='https://assemblyai.com/aiexplained'><b>https://assemblyai.com/aiexplained</b></a></p><p><b>Chapters:<br/>00:00 - Introduction</b></p><p><b>00:56 - GPT 5.1 Smarter?</b></p><p><b>01:47 - Some Regressions</b></p><p><b>03:22 - Sycophancy?</b></p><p><b>05:22 - Claude Auto-Hacking </b></p><p><b>06:16 - Jailbreaking through Granularity</b></p><p><b>08:22 - This Will be Re-used</b></p><p><b>09:30 - Hallucinating Hacker</b></p><p><b>09:57 - Surprisingly Neutral Tone</b></p><p><b>12:18 - SIMA 2</b></p><p><b>14:10 - Alpha Parallels</b></p><p><b>17:24 - AI Music</b></p><p><b><br/></b><br/></p><p><b>GPT 5.1 Announcement: </b><a href='https://openai.com/index/gpt-5-1/'><b>https://openai.com/index/gpt-5-1/</b></a></p><p><b>System Card: </b><a href='https://cdn.openai.com/pdf/4173ec8d-1229-47db-96de-06d87147e07e/5_1_system_card.pdf'><b>https://cdn.openai.com/pdf/4173ec8d-1229-47db-96de-06d87147e07e/5_1_system_card.pdf</b></a></p><p><b>Benchmarks: </b><a href='https://openai.com/index/gpt-5-1-for-developers/'><b>https://openai.com/index/gpt-5-1-for-developers/</b></a></p><p><b>Simple Bench: </b><a href='https://lmcouncil.ai/benchmarks'><b>https://lmcouncil.ai/benchmarks</b></a></p><p><br/></p><p><b>Auto-Hacking: </b><a href='https://x.com/AnthropicAI/status/1989033793190277618'><b>https://x.com/AnthropicAI/status/1989033793190277618</b></a></p><p><a href='https://www.anthropic.com/news/disrupting-AI-espionage'><b>https://www.anthropic.com/news/disrupting-AI-espionage</b></a></p><p><b>Report: https://assets.anthropic.com/m/ec212e6566a0d47/original/Disrupting-the-first-reported-AI-orchestrated-cyber-espionage-campaign.pdf</b></p><p><b><br/></b><br/></p><p><b>Sima 2 Announcement: </b><a href='https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/'><b>https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/</b></a></p><p><a href='https://x.com/amoufarek/status/1988986075331858693'><b>https://x.com/amoufarek/status/1988986075331858693</b></a></p><p><b>Scepticism: </b><a href='https://www.technologyreview.com/2025/11/13/1127921/google-deepmind-is-using-gemini-to-train-agents-inside-goat-simulator-3/'><b>https://www.technologyreview.com/2025/11/13/1127921/google-deepmind-is-using-gemini-to-train-agents-inside-goat-simulator-3/</b></a></p><p><b>Voyager: </b><a href='https://voyager.minedojo.org/'><b>https://voyager.minedojo.org/</b></a></p><p><br/></p><p><b>Reuters Music: https://www.reuters.com/legal/litigation/are-you-listening-bots-survey-shows-ai-music-is-virtually-undetectable-2025-11-12/</b></p><p><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>A lot just got released in the last 36 hours, and it will all affect hundreds of millions of people. 10 details you would miss if you just read the headlines, from GPT 5.1 regressions, to how Claude hacked Govt Agencies, to SIMA 2, and Musical Turing Tests.</b></p><p><a href='https://assemblyai.com/aiexplained'><b>https://assemblyai.com/aiexplained</b></a></p><p><b>Chapters:<br/>00:00 - Introduction</b></p><p><b>00:56 - GPT 5.1 Smarter?</b></p><p><b>01:47 - Some Regressions</b></p><p><b>03:22 - Sycophancy?</b></p><p><b>05:22 - Claude Auto-Hacking </b></p><p><b>06:16 - Jailbreaking through Granularity</b></p><p><b>08:22 - This Will be Re-used</b></p><p><b>09:30 - Hallucinating Hacker</b></p><p><b>09:57 - Surprisingly Neutral Tone</b></p><p><b>12:18 - SIMA 2</b></p><p><b>14:10 - Alpha Parallels</b></p><p><b>17:24 - AI Music</b></p><p><b><br/></b><br/></p><p><b>GPT 5.1 Announcement: </b><a href='https://openai.com/index/gpt-5-1/'><b>https://openai.com/index/gpt-5-1/</b></a></p><p><b>System Card: </b><a href='https://cdn.openai.com/pdf/4173ec8d-1229-47db-96de-06d87147e07e/5_1_system_card.pdf'><b>https://cdn.openai.com/pdf/4173ec8d-1229-47db-96de-06d87147e07e/5_1_system_card.pdf</b></a></p><p><b>Benchmarks: </b><a href='https://openai.com/index/gpt-5-1-for-developers/'><b>https://openai.com/index/gpt-5-1-for-developers/</b></a></p><p><b>Simple Bench: </b><a href='https://lmcouncil.ai/benchmarks'><b>https://lmcouncil.ai/benchmarks</b></a></p><p><br/></p><p><b>Auto-Hacking: </b><a href='https://x.com/AnthropicAI/status/1989033793190277618'><b>https://x.com/AnthropicAI/status/1989033793190277618</b></a></p><p><a href='https://www.anthropic.com/news/disrupting-AI-espionage'><b>https://www.anthropic.com/news/disrupting-AI-espionage</b></a></p><p><b>Report: https://assets.anthropic.com/m/ec212e6566a0d47/original/Disrupting-the-first-reported-AI-orchestrated-cyber-espionage-campaign.pdf</b></p><p><b><br/></b><br/></p><p><b>Sima 2 Announcement: </b><a href='https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/'><b>https://deepmind.google/blog/sima-2-an-agent-that-plays-reasons-and-learns-with-you-in-virtual-3d-worlds/</b></a></p><p><a href='https://x.com/amoufarek/status/1988986075331858693'><b>https://x.com/amoufarek/status/1988986075331858693</b></a></p><p><b>Scepticism: </b><a href='https://www.technologyreview.com/2025/11/13/1127921/google-deepmind-is-using-gemini-to-train-agents-inside-goat-simulator-3/'><b>https://www.technologyreview.com/2025/11/13/1127921/google-deepmind-is-using-gemini-to-train-agents-inside-goat-simulator-3/</b></a></p><p><b>Voyager: </b><a href='https://voyager.minedojo.org/'><b>https://voyager.minedojo.org/</b></a></p><p><br/></p><p><b>Reuters Music: https://www.reuters.com/legal/litigation/are-you-listening-bots-survey-shows-ai-music-is-virtually-undetectable-2025-11-12/</b></p><p><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18191565-is-gpt-5-1-really-an-upgrade-but-models-can-auto-hack-govts-so-there-s-that.mp3" length="13307482" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/qa5ohiztxn8t3l3lq5mp93eyearw?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18191565</guid>
    <pubDate>Fri, 14 Nov 2025 17:00:00 +0000</pubDate>
    <itunes:duration>1106</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>4</itunes:season>
    <itunes:episode>2</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Bubble or No Bubble, AI Keeps Progressing (ft. Relentless Learning + Introspection)</itunes:title>
    <title>Bubble or No Bubble, AI Keeps Progressing (ft. Relentless Learning + Introspection)</title>
    <itunes:summary><![CDATA[Don’t let headlines about bubbles distract you from the real avenues of progress being explored in AI every week, including what had been thought to be a long-term blocker - continual learning (learning on the fly).   https://app.grayswan.ai/ai-explained  This, plus models introspecting (hesitate before you berate), Nano Banana 2 possibly spotted, Chinese imagen and more.  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 01:26 - Continual Learning (N...]]></itunes:summary>
    <description><![CDATA[<p>Don’t let headlines about bubbles distract you from the real avenues of progress being explored in AI every week, including what had been thought to be a long-term blocker - continual learning (learning on the fly). <br/><br/>https://app.grayswan.ai/ai-explained<br/><br/>This, plus models introspecting (hesitate before you berate), Nano Banana 2 possibly spotted, Chinese imagen and more.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:26 - Continual Learning (Nested Learning / HOPE)<br/>07:00 - Introspection<br/>10:54 - Image-Gen Progress<br/><br/>Nested Learning Post: https://research.google/blog/introducing-nested-learning-a-new-ml-paradigm-for-continual-learning/<br/><br/>Nested Learning Paper: https://abehrouz.github.io/files/NL.pdf<br/><br/>Original Titans Paper: https://arxiv.org/pdf/2501.00663<br/><br/>Siri News: https://www.bloomberg.com/news/articles/2025-11-05/apple-plans-to-use-1-2-trillion-parameter-google-gemini-model-to-power-new-siri<br/><br/>Introspection: https://www.anthropic.com/research/introspection<br/><br/>Full Paper: https://transformer-circuits.pub/2025/introspection/index.html#mechanisms<br/><br/>Earlier Work: https://www.anthropic.com/research/mapping-mind-language-model<br/>https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html<br/><br/>Release Post: https://x.com/AnthropicAI/status/1983584136972677319<br/><br/>https://lmcouncil.ai <br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Don’t let headlines about bubbles distract you from the real avenues of progress being explored in AI every week, including what had been thought to be a long-term blocker - continual learning (learning on the fly). <br/><br/>https://app.grayswan.ai/ai-explained<br/><br/>This, plus models introspecting (hesitate before you berate), Nano Banana 2 possibly spotted, Chinese imagen and more.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:26 - Continual Learning (Nested Learning / HOPE)<br/>07:00 - Introspection<br/>10:54 - Image-Gen Progress<br/><br/>Nested Learning Post: https://research.google/blog/introducing-nested-learning-a-new-ml-paradigm-for-continual-learning/<br/><br/>Nested Learning Paper: https://abehrouz.github.io/files/NL.pdf<br/><br/>Original Titans Paper: https://arxiv.org/pdf/2501.00663<br/><br/>Siri News: https://www.bloomberg.com/news/articles/2025-11-05/apple-plans-to-use-1-2-trillion-parameter-google-gemini-model-to-power-new-siri<br/><br/>Introspection: https://www.anthropic.com/research/introspection<br/><br/>Full Paper: https://transformer-circuits.pub/2025/introspection/index.html#mechanisms<br/><br/>Earlier Work: https://www.anthropic.com/research/mapping-mind-language-model<br/>https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html<br/><br/>Release Post: https://x.com/AnthropicAI/status/1983584136972677319<br/><br/>https://lmcouncil.ai <br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/18163685-bubble-or-no-bubble-ai-keeps-progressing-ft-relentless-learning-introspection.mp3" length="9309944" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/5ve25t3lji7re7kvugvmbp2b2wr6?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-18163685</guid>
    <pubDate>Mon, 10 Nov 2025 15:00:00 +0000</pubDate>
    <itunes:duration>773</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>4</itunes:season>
    <itunes:episode>1</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Sora 2 - It will only get more realistic from here</itunes:title>
    <title>Sora 2 - It will only get more realistic from here</title>
    <itunes:summary><![CDATA[Sora 2 - the start of the infinite slop-feed or a key step to a generalist agent? Better than VEO 3 or over-hyped? I bring out 6 details you may have missed, contrast the announcement to Periodic Labs and even squeeze in some Claude Sonnet 4.5 analysis. Maybe I should make my videos longer…  https://80000hours.org/aiexplained  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 00:40 - Two models? 01:15 - Rollout Details 01:43 - Versus Sora 1 / Veo 3 04:30 -...]]></itunes:summary>
    <description><![CDATA[<p>Sora 2 - the start of the infinite slop-feed or a key step to a generalist agent? Better than VEO 3 or over-hyped? I bring out 6 details you may have missed, contrast the announcement to Periodic Labs and even squeeze in some Claude Sonnet 4.5 analysis. Maybe I should make my videos longer…<br/><br/>https://80000hours.org/aiexplained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:40 - Two models?<br/>01:15 - Rollout Details<br/>01:43 - Versus Sora 1 / Veo 3<br/>04:30 - Sora App / Social Media<br/>06:40 - Masterplan<br/>09:30 - Generalist Agent? Periodic Labs<br/>12:05 - Claude Sonnet 4.5<br/>13:42 - Future Outlook<br/><br/>Announcement: https://openai.com/index/sora-2/<br/>Launch Video: https://www.youtube.com/live/gzneGhpXwjU<br/>System Card: https://cdn.openai.com/pdf/50d5973c-c4ff-4c2d-986f-c72b5d0ff069/sora_2_system_card.pdf<br/>Sam Altman Blog Post on Sora App: https://blog.samaltman.com/sora-2<br/><br/>Most Intelligent Claim: https://x.com/willdepue/status/1973089331284681110<br/>GTA: https://x.com/AndrewCurran_/status/1973298436536766666<br/><br/>Meta Vibes: https://x.com/alexandr_wang/status/1971295156411433228?s=46<br/><br/>Altman on Regulations: https://www.lesswrong.com/posts/5jjk4CDnj9tA7ugxr/openai-email-archives-from-musk-v-altman<br/>OpenAI Profit: https://www.theinformation.com/articles/openais-first-half-results-4-3-billion-sales-2-5-billion-cash-burn?rc=sy0ihq<br/><br/>Periodic Labs: https://periodic.com/<br/>https://www.nytimes.com/2025/09/30/technology/ai-meta-google-openai-periodic.html<br/>https://x.com/LiamFedus/status/1973055380193431965<br/>https://baincapitalventures.com/insight/we-must-know-we-will-know/?s=09<br/><br/>Sonnet 4.5: https://www.anthropic.com/news/claude-sonnet-4-5<br/>https://simple-bench.com/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Sora 2 - the start of the infinite slop-feed or a key step to a generalist agent? Better than VEO 3 or over-hyped? I bring out 6 details you may have missed, contrast the announcement to Periodic Labs and even squeeze in some Claude Sonnet 4.5 analysis. Maybe I should make my videos longer…<br/><br/>https://80000hours.org/aiexplained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:40 - Two models?<br/>01:15 - Rollout Details<br/>01:43 - Versus Sora 1 / Veo 3<br/>04:30 - Sora App / Social Media<br/>06:40 - Masterplan<br/>09:30 - Generalist Agent? Periodic Labs<br/>12:05 - Claude Sonnet 4.5<br/>13:42 - Future Outlook<br/><br/>Announcement: https://openai.com/index/sora-2/<br/>Launch Video: https://www.youtube.com/live/gzneGhpXwjU<br/>System Card: https://cdn.openai.com/pdf/50d5973c-c4ff-4c2d-986f-c72b5d0ff069/sora_2_system_card.pdf<br/>Sam Altman Blog Post on Sora App: https://blog.samaltman.com/sora-2<br/><br/>Most Intelligent Claim: https://x.com/willdepue/status/1973089331284681110<br/>GTA: https://x.com/AndrewCurran_/status/1973298436536766666<br/><br/>Meta Vibes: https://x.com/alexandr_wang/status/1971295156411433228?s=46<br/><br/>Altman on Regulations: https://www.lesswrong.com/posts/5jjk4CDnj9tA7ugxr/openai-email-archives-from-musk-v-altman<br/>OpenAI Profit: https://www.theinformation.com/articles/openais-first-half-results-4-3-billion-sales-2-5-billion-cash-burn?rc=sy0ihq<br/><br/>Periodic Labs: https://periodic.com/<br/>https://www.nytimes.com/2025/09/30/technology/ai-meta-google-openai-periodic.html<br/>https://x.com/LiamFedus/status/1973055380193431965<br/>https://baincapitalventures.com/insight/we-must-know-we-will-know/?s=09<br/><br/>Sonnet 4.5: https://www.anthropic.com/news/claude-sonnet-4-5<br/>https://simple-bench.com/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17938988-sora-2-it-will-only-get-more-realistic-from-here.mp3" length="11356715" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/dj5im8ddz8juqto0oiq3xzi9lxlr?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17938988</guid>
    <pubDate>Wed, 01 Oct 2025 15:00:00 +0100</pubDate>
    <itunes:duration>943</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>3</itunes:season>
    <itunes:episode>5</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>OpenAI Tests if GPT-5 Can Automate Your Job - 4 Unexpected Findings</itunes:title>
    <title>OpenAI Tests if GPT-5 Can Automate Your Job - 4 Unexpected Findings</title>
    <itunes:summary><![CDATA[An OpenAI report released in the last 24 hours is the best look we have as to whether 2025 AI can automate your job. I’ll go through 4 unexpected findings, from which model is best at what, to practical tips and massive caveats. Plus UFC robots, radiologist essay, don’t trust videos and the blockers to the singularity.    Gray Swan: https://app.grayswan.ai/ai-explained    GDPval: https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf   [GDP Impact: https://fred.stloui...]]></itunes:summary>
    <description><![CDATA[<p><b>An OpenAI report released in the last 24 hours is the best look we have as to whether 2025 AI can automate your job. I’ll go through 4 unexpected findings, from which model is best at what, to practical tips and massive caveats. Plus UFC robots, radiologist essay, don’t trust videos and the blockers to the singularity. </b></p><p><br/></p><p><b>Gray Swan: </b><a href='https://app.grayswan.ai/ai-explained'><b>https://app.grayswan.ai/ai-explained</b></a></p><p><b><br/></b><br/></p><p><b>GDPval: </b><a href='https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf'><b>https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf</b></a></p><p><br/></p><p><b>[GDP Impact: </b><a href='https://fred.stlouisfed.org/release/tables?rid=331&amp;eid=211'><b>https://fred.stlouisfed.org/release/tables?rid=331&amp;eid=211</b></a></p><p><b>Task List: </b><a href='https://www.onetonline.org/link/summary/11-9141.00'><b>https://www.onetonline.org/link/summary/11-9141.00</b></a></p><p><b>Summer Tweet: </b><a href='https://x.com/LHSummers/status/1971252567981146347'><b>https://x.com/LHSummers/status/1971252567981146347</b></a></p><p><b>Emad: https://x.com/EMostaque/status/1971254153067593739</b></p><p><br/></p><p><b>Robots: </b><a href='https://x.com/cixliv/status/1967663286679478759'><b>https://x.com/cixliv/status/1967663286679478759</b></a></p><p><b>Unitree G1: </b><a href='https://x.com/UnitreeRobotics/status/1970039940022239491'><b>https://x.com/UnitreeRobotics/status/1970039940022239491</b></a></p><p><br/></p><p><b>Don’t Trust Video: https://x.com/AISafetyMemes/status/1970453369446871420</b></p><p><br/></p><p><b>AGI Tweet: </b><a href='https://x.com/hyhieu226/status/1968378785709133915'><b>https://x.com/hyhieu226/status/1968378785709133915</b></a></p><p><br/></p><p><b>Blockers to the Singularity: </b><a href='https://www.patreon.com/posts/blockers-to-and-139264812'><b>https://www.patreon.com/posts/blockers-to-and-139264812</b></a></p><p><br/></p><p><b>Framework: </b><a href='https://gemini.google.com/share/f4b9c85a6ae9'><b>https://gemini.google.com/share/f4b9c85a6ae9</b></a></p><p><br/></p><p><b>METR Study (Dev Slowdown): https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/</b></p><p><br/></p><p><b>Karpathy Tweet: </b><a href='https://x.com/karpathy/status/1971220449515516391'><b>https://x.com/karpathy/status/1971220449515516391</b></a></p><p><b>Radiology Essay: </b><a href='https://worksinprogress.co/issue/the-algorithm-will-see-you-now/'><b>https://worksinprogress.co/issue/the-algorithm-will-see-you-now/</b></a></p><p><br/></p><p><b>Chapters:</b></p><p><b>00:00 - Introduction</b></p><p><b>00:55 - OpenAI Report Summary</b></p><p><b>02:40 - Tipping Point Speed-up</b></p><p><b>04:11 - Better than Industry Experts?</b></p><p><b>06:33 - Big Caveat</b></p><p><b>11:10 - Karpathy and the Radiologist Analogy</b></p><p><b>13:30 - Outro</b></p>]]></description>
    <content:encoded><![CDATA[<p><b>An OpenAI report released in the last 24 hours is the best look we have as to whether 2025 AI can automate your job. I’ll go through 4 unexpected findings, from which model is best at what, to practical tips and massive caveats. Plus UFC robots, radiologist essay, don’t trust videos and the blockers to the singularity. </b></p><p><br/></p><p><b>Gray Swan: </b><a href='https://app.grayswan.ai/ai-explained'><b>https://app.grayswan.ai/ai-explained</b></a></p><p><b><br/></b><br/></p><p><b>GDPval: </b><a href='https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf'><b>https://cdn.openai.com/pdf/d5eb7428-c4e9-4a33-bd86-86dd4bcf12ce/GDPval.pdf</b></a></p><p><br/></p><p><b>[GDP Impact: </b><a href='https://fred.stlouisfed.org/release/tables?rid=331&amp;eid=211'><b>https://fred.stlouisfed.org/release/tables?rid=331&amp;eid=211</b></a></p><p><b>Task List: </b><a href='https://www.onetonline.org/link/summary/11-9141.00'><b>https://www.onetonline.org/link/summary/11-9141.00</b></a></p><p><b>Summer Tweet: </b><a href='https://x.com/LHSummers/status/1971252567981146347'><b>https://x.com/LHSummers/status/1971252567981146347</b></a></p><p><b>Emad: https://x.com/EMostaque/status/1971254153067593739</b></p><p><br/></p><p><b>Robots: </b><a href='https://x.com/cixliv/status/1967663286679478759'><b>https://x.com/cixliv/status/1967663286679478759</b></a></p><p><b>Unitree G1: </b><a href='https://x.com/UnitreeRobotics/status/1970039940022239491'><b>https://x.com/UnitreeRobotics/status/1970039940022239491</b></a></p><p><br/></p><p><b>Don’t Trust Video: https://x.com/AISafetyMemes/status/1970453369446871420</b></p><p><br/></p><p><b>AGI Tweet: </b><a href='https://x.com/hyhieu226/status/1968378785709133915'><b>https://x.com/hyhieu226/status/1968378785709133915</b></a></p><p><br/></p><p><b>Blockers to the Singularity: </b><a href='https://www.patreon.com/posts/blockers-to-and-139264812'><b>https://www.patreon.com/posts/blockers-to-and-139264812</b></a></p><p><br/></p><p><b>Framework: </b><a href='https://gemini.google.com/share/f4b9c85a6ae9'><b>https://gemini.google.com/share/f4b9c85a6ae9</b></a></p><p><br/></p><p><b>METR Study (Dev Slowdown): https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/</b></p><p><br/></p><p><b>Karpathy Tweet: </b><a href='https://x.com/karpathy/status/1971220449515516391'><b>https://x.com/karpathy/status/1971220449515516391</b></a></p><p><b>Radiology Essay: </b><a href='https://worksinprogress.co/issue/the-algorithm-will-see-you-now/'><b>https://worksinprogress.co/issue/the-algorithm-will-see-you-now/</b></a></p><p><br/></p><p><b>Chapters:</b></p><p><b>00:00 - Introduction</b></p><p><b>00:55 - OpenAI Report Summary</b></p><p><b>02:40 - Tipping Point Speed-up</b></p><p><b>04:11 - Better than Industry Experts?</b></p><p><b>06:33 - Big Caveat</b></p><p><b>11:10 - Karpathy and the Radiologist Analogy</b></p><p><b>13:30 - Outro</b></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17911105-openai-tests-if-gpt-5-can-automate-your-job-4-unexpected-findings.mp3" length="10190362" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/3v21qis2enffwaa05n8vu98w8vkb?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17911105</guid>
    <pubDate>Fri, 26 Sep 2025 16:00:00 +0100</pubDate>
    <itunes:duration>846</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>3</itunes:season>
    <itunes:episode>4</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>ChatGPT Will Guess your Age, Flirt if Asked, and Can Call the Cops </itunes:title>
    <title>ChatGPT Will Guess your Age, Flirt if Asked, and Can Call the Cops </title>
    <itunes:summary><![CDATA[Sam Altman, CEO of OpenAI, announced a set of new ‘protections’ and ‘privileges’ for ChatGPT users, requiring a significant amount of trust from users. From predicting your age based on your chat to calling law enforcement if you are at risk of harm, to allowing non-minors to flirt. But amidst all of these announcements, there are interview snippets you may have missed, as Altman dramatically revises his predictions of AI impact on jobs. Plus a Hassbis backtrack to boot.  https://80000hours.o...]]></itunes:summary>
    <description><![CDATA[<p><b>Sam Altman, CEO of OpenAI, announced a set of new ‘protections’ and ‘privileges’ for ChatGPT users, requiring a significant amount of trust from users. From predicting your age based on your chat to calling law enforcement if you are at risk of harm, to allowing non-minors to flirt. But amidst all of these announcements, there are interview snippets you may have missed, as Altman dramatically revises his predictions of AI impact on jobs. Plus a Hassbis backtrack to boot.<br/><br/></b><a href='https://80000hours.org/aiexplained'><b>https://80000hours.org/aiexplained</b></a></p><p><br/></p><p><b>Calling the Cops: </b><a href='https://openai.com/index/teen-safety-freedom-and-privacy/'><b>https://openai.com/index/teen-safety-freedom-and-privacy/</b></a></p><p><br/></p><p><b>Age Prediction: </b><a href='https://openai.com/index/building-towards-age-prediction/'><b>https://openai.com/index/building-towards-age-prediction/</b></a></p><p><br/></p><p><b>Not Everyone Will Agree: </b><a href='https://x.com/sama/status/1967955739911364693?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet'><b>https://x.com/sama/status/1967955739911364693?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet</b></a></p><p><br/></p><p><b>Theory 1: NYT Lawsuit: </b><a href='https://openai.com/index/response-to-nyt-data-demands/'><b>https://openai.com/index/response-to-nyt-data-demands/</b></a></p><p><br/></p><p><b>Theory 2: FTC Investigation into AI Companions: </b><a href='https://x.com/AndrewCurran_/status/1966167585994764743'><b>https://x.com/AndrewCurran_/status/1966167585994764743</b></a></p><p><br/></p><p><b>YT Does the Same: </b><a href='https://www.cbsnews.com/news/youtube-ai-powered-technology-teen-users/'><b>https://www.cbsnews.com/news/youtube-ai-powered-technology-teen-users/</b></a></p><p><br/></p><p><b>Carlsen Interview: </b><a href='https://www.youtube.com/watch?v=5KmpT-BoVf4'><b>https://www.youtube.com/watch?v=5KmpT-BoVf4</b></a></p><p><br/></p><p><b>vs Senate Testimony (70% Jobs): https://www.youtube.com/watch?v=5CWVP8-XVjQ</b></p><p><br/></p><p><b>Hallucinations Paper: </b><a href='https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4aaa5/why-language-models-hallucinate.pdf'><b>https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4aaa5/why-language-models-hallucinate.pdf</b></a></p><p><br/></p><p><b>Hassbis Quote 1: </b><a href='https://www.youtube.com/watch?v=toShbNUGAyo'><b>https://www.youtube.com/watch?v=toShbNUGAyo</b></a></p><p><br/></p><p><b>vs Quote 2: https://www.youtube.com/watch?v=Kr3Sh2PKA8Y</b></p><p><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>Sam Altman, CEO of OpenAI, announced a set of new ‘protections’ and ‘privileges’ for ChatGPT users, requiring a significant amount of trust from users. From predicting your age based on your chat to calling law enforcement if you are at risk of harm, to allowing non-minors to flirt. But amidst all of these announcements, there are interview snippets you may have missed, as Altman dramatically revises his predictions of AI impact on jobs. Plus a Hassbis backtrack to boot.<br/><br/></b><a href='https://80000hours.org/aiexplained'><b>https://80000hours.org/aiexplained</b></a></p><p><br/></p><p><b>Calling the Cops: </b><a href='https://openai.com/index/teen-safety-freedom-and-privacy/'><b>https://openai.com/index/teen-safety-freedom-and-privacy/</b></a></p><p><br/></p><p><b>Age Prediction: </b><a href='https://openai.com/index/building-towards-age-prediction/'><b>https://openai.com/index/building-towards-age-prediction/</b></a></p><p><br/></p><p><b>Not Everyone Will Agree: </b><a href='https://x.com/sama/status/1967955739911364693?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet'><b>https://x.com/sama/status/1967955739911364693?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet</b></a></p><p><br/></p><p><b>Theory 1: NYT Lawsuit: </b><a href='https://openai.com/index/response-to-nyt-data-demands/'><b>https://openai.com/index/response-to-nyt-data-demands/</b></a></p><p><br/></p><p><b>Theory 2: FTC Investigation into AI Companions: </b><a href='https://x.com/AndrewCurran_/status/1966167585994764743'><b>https://x.com/AndrewCurran_/status/1966167585994764743</b></a></p><p><br/></p><p><b>YT Does the Same: </b><a href='https://www.cbsnews.com/news/youtube-ai-powered-technology-teen-users/'><b>https://www.cbsnews.com/news/youtube-ai-powered-technology-teen-users/</b></a></p><p><br/></p><p><b>Carlsen Interview: </b><a href='https://www.youtube.com/watch?v=5KmpT-BoVf4'><b>https://www.youtube.com/watch?v=5KmpT-BoVf4</b></a></p><p><br/></p><p><b>vs Senate Testimony (70% Jobs): https://www.youtube.com/watch?v=5CWVP8-XVjQ</b></p><p><br/></p><p><b>Hallucinations Paper: </b><a href='https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4aaa5/why-language-models-hallucinate.pdf'><b>https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4aaa5/why-language-models-hallucinate.pdf</b></a></p><p><br/></p><p><b>Hassbis Quote 1: </b><a href='https://www.youtube.com/watch?v=toShbNUGAyo'><b>https://www.youtube.com/watch?v=toShbNUGAyo</b></a></p><p><br/></p><p><b>vs Quote 2: https://www.youtube.com/watch?v=Kr3Sh2PKA8Y</b></p><p><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17855052-chatgpt-will-guess-your-age-flirt-if-asked-and-can-call-the-cops.mp3" length="8334472" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/rdr1g078j3viai5vrmtpn1e4muk0?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17855052</guid>
    <pubDate>Tue, 16 Sep 2025 18:00:00 +0100</pubDate>
    <itunes:duration>691</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>3</itunes:season>
    <itunes:episode>3</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>An ‘AI Bubble’? What Altman Actually said, the Facts and Nano Banana</itunes:title>
    <title>An ‘AI Bubble’? What Altman Actually said, the Facts and Nano Banana</title>
    <itunes:summary><![CDATA[Wait, why did Sam Altman say AI was in a bubble? Or did he? Is it? 8 points for you to consider, before we all get distracted by Nano Banana.  Chapters: 00:00 - Introduction 01:14 - Sam Altman Clarification 02:30 - Media Calls a Bubble (for the tenth time) 03:40 - MIT and McKinsey Analysed 08:21 - Incremental Progress Deceptive 12:07 - Reasoning Breakthroughs 15:31 - CEOs might not know their products 17:25 - But did stocks go down? 17:31 - Media is Contradictory of course   https://donate.re...]]></itunes:summary>
    <description><![CDATA[<p>Wait, why did Sam Altman say AI was in a bubble? Or did he? Is it? 8 points for you to consider, before we all get distracted by Nano Banana.<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:14 - Sam Altman Clarification<br/>02:30 - Media Calls a Bubble (for the tenth time)<br/>03:40 - MIT and McKinsey Analysed<br/>08:21 - Incremental Progress Deceptive<br/>12:07 - Reasoning Breakthroughs<br/>15:31 - CEOs might not know their products<br/>17:25 - But did stocks go down?<br/>17:31 - Media is Contradictory of course<br/><br/><br/>https://donate.redcross.org.uk/appeal/gaza-crisis-appeal<br/><br/><br/>Bubble about to burst: https://www.telegraph.co.uk/business/2025/08/20/ai-report-triggering-panic-and-fear-on-wall-street/<br/><br/>Nano Banana: https://blog.google/products/gemini/updated-image-editing-model/<br/>https://ai.studio/banana<br/><br/>McKinsey Report: https://www.mckinsey.com/capabilities/quantumblack/our-insights/seizing-the-agentic-ai-advantage#/<br/>https://www.mckinsey.com/capabilities/quantumblack/our-insights/the-state-of-ai#/<br/>Revenue: https://www.wsj.com/tech/ai/mckinsey-consulting-firms-ai-strategy-89fbf1be<br/><br/>MIT Report: https://mlq.ai/media/quarterly_decks/v0.1_State_of_AI_in_Business_2025_Report.pdf<br/><br/>Safe Superintelligence: https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/<br/><br/>Thinking Machines Lab: https://techcrunch.com/2025/07/15/mira-muratis-thinking-machines-lab-is-worth-12b-in-seed-round/<br/><br/>WSJ Prediction 2024: https://www.wsj.com/tech/ai/the-ai-revolution-is-already-losing-steam-a93478b1<br/>WP Prediction 2023: https://www.washingtonpost.com/technology/2023/08/05/ai-hype-bubble-chatgpt/<br/><br/>Companies are Pouring Billions into AI: https://www.nytimes.com/2025/08/13/business/ai-business-payoff-lags.html<br/><br/>Consumer Surplus: https://www.wsj.com/opinion/ais-overlooked-97-billion-contribution-to-the-economy-users-service-da6e8f55<br/>Figure AI robot: https://x.com/adcock_brett/status/1958193476639826383<br/><br/>GDP Bet: https://x.com/adamdangelo/status/1627726566259318784?lang=en<br/><br/>Genie 3 Immersion: https://x.com/holynski_/status/1953879983535141043<br/><br/>https://x.com/elonmusk/status/1953861448431718662<br/>htttps://simple-bench.com<br/>MMMU: https://mmmu-benchmark.github.io/#leaderboard <br/>Prophet Arena: https://www.prophetarena.co/leaderboard<br/><br/>NYT Jobs: https://www.nytimes.com/2025/08/19/opinion/ai-job-loss-deindustrialization.html<br/><br/>Dawn of Reasoning?: https://openreview.net/pdf?id=FkKBxp0FhR<br/>vs :https://arxiv.org/pdf/2403.04121<br/><br/>ARC-AGI: https://arcprize.org/arc-agi/1/<br/>https://x.com/fchollet/status/1870169764762710376?lang=en-GB<br/><br/>Turing Test: https://arxiv.org/pdf/2503.23674<br/><br/>Mathematics of Starvation: https://www.theguardian.com/world/2025/jul/31/the-mathematics-of-starvation-how-israel-caused-a-famine-in-gaza<br/>https://donate.redcross.org.uk/appeal/gaza-crisis-appeal<br/><br/>https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/<br/><br/>METR Interview: https://www.patreon.com/c/aiexplained/posts<br/><br/>AlphaEvolve: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/<br/>Paper: https://storage.googleapis.com/deepmind-media/DeepMind.com/Blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/AlphaEvolve.pdf<br/><br/>Amodei: https://kantrowitz.medium.com/the-making-of-anthropic-ceo-dario-amodei-449777529dd6<br/>https://www.theloganbartlettshow.com/archive/ep-82-dario-amodeis-ai-predictions-through-2030#:~:text=DARIO%3A%20I%20think%20our%20concern,being%20responsible%20to%20accelerate%20things<br/>Unreleased OpenAI: https://x.com/alexwei_/status/1954966393419599962<br/><br/>VLMs Tricked: https://x.com/an_vo12/status/1943715159559545186<br/><br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained</p>]]></description>
    <content:encoded><![CDATA[<p>Wait, why did Sam Altman say AI was in a bubble? Or did he? Is it? 8 points for you to consider, before we all get distracted by Nano Banana.<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:14 - Sam Altman Clarification<br/>02:30 - Media Calls a Bubble (for the tenth time)<br/>03:40 - MIT and McKinsey Analysed<br/>08:21 - Incremental Progress Deceptive<br/>12:07 - Reasoning Breakthroughs<br/>15:31 - CEOs might not know their products<br/>17:25 - But did stocks go down?<br/>17:31 - Media is Contradictory of course<br/><br/><br/>https://donate.redcross.org.uk/appeal/gaza-crisis-appeal<br/><br/><br/>Bubble about to burst: https://www.telegraph.co.uk/business/2025/08/20/ai-report-triggering-panic-and-fear-on-wall-street/<br/><br/>Nano Banana: https://blog.google/products/gemini/updated-image-editing-model/<br/>https://ai.studio/banana<br/><br/>McKinsey Report: https://www.mckinsey.com/capabilities/quantumblack/our-insights/seizing-the-agentic-ai-advantage#/<br/>https://www.mckinsey.com/capabilities/quantumblack/our-insights/the-state-of-ai#/<br/>Revenue: https://www.wsj.com/tech/ai/mckinsey-consulting-firms-ai-strategy-89fbf1be<br/><br/>MIT Report: https://mlq.ai/media/quarterly_decks/v0.1_State_of_AI_in_Business_2025_Report.pdf<br/><br/>Safe Superintelligence: https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/<br/><br/>Thinking Machines Lab: https://techcrunch.com/2025/07/15/mira-muratis-thinking-machines-lab-is-worth-12b-in-seed-round/<br/><br/>WSJ Prediction 2024: https://www.wsj.com/tech/ai/the-ai-revolution-is-already-losing-steam-a93478b1<br/>WP Prediction 2023: https://www.washingtonpost.com/technology/2023/08/05/ai-hype-bubble-chatgpt/<br/><br/>Companies are Pouring Billions into AI: https://www.nytimes.com/2025/08/13/business/ai-business-payoff-lags.html<br/><br/>Consumer Surplus: https://www.wsj.com/opinion/ais-overlooked-97-billion-contribution-to-the-economy-users-service-da6e8f55<br/>Figure AI robot: https://x.com/adcock_brett/status/1958193476639826383<br/><br/>GDP Bet: https://x.com/adamdangelo/status/1627726566259318784?lang=en<br/><br/>Genie 3 Immersion: https://x.com/holynski_/status/1953879983535141043<br/><br/>https://x.com/elonmusk/status/1953861448431718662<br/>htttps://simple-bench.com<br/>MMMU: https://mmmu-benchmark.github.io/#leaderboard <br/>Prophet Arena: https://www.prophetarena.co/leaderboard<br/><br/>NYT Jobs: https://www.nytimes.com/2025/08/19/opinion/ai-job-loss-deindustrialization.html<br/><br/>Dawn of Reasoning?: https://openreview.net/pdf?id=FkKBxp0FhR<br/>vs :https://arxiv.org/pdf/2403.04121<br/><br/>ARC-AGI: https://arcprize.org/arc-agi/1/<br/>https://x.com/fchollet/status/1870169764762710376?lang=en-GB<br/><br/>Turing Test: https://arxiv.org/pdf/2503.23674<br/><br/>Mathematics of Starvation: https://www.theguardian.com/world/2025/jul/31/the-mathematics-of-starvation-how-israel-caused-a-famine-in-gaza<br/>https://donate.redcross.org.uk/appeal/gaza-crisis-appeal<br/><br/>https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/<br/><br/>METR Interview: https://www.patreon.com/c/aiexplained/posts<br/><br/>AlphaEvolve: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/<br/>Paper: https://storage.googleapis.com/deepmind-media/DeepMind.com/Blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/AlphaEvolve.pdf<br/><br/>Amodei: https://kantrowitz.medium.com/the-making-of-anthropic-ceo-dario-amodei-449777529dd6<br/>https://www.theloganbartlettshow.com/archive/ep-82-dario-amodeis-ai-predictions-through-2030#:~:text=DARIO%3A%20I%20think%20our%20concern,being%20responsible%20to%20accelerate%20things<br/>Unreleased OpenAI: https://x.com/alexwei_/status/1954966393419599962<br/><br/>VLMs Tricked: https://x.com/an_vo12/status/1943715159559545186<br/><br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17737829-an-ai-bubble-what-altman-actually-said-the-facts-and-nano-banana.mp3" length="13635063" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/coef9g3relazzbeyka2waw3qhgvl?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17737829</guid>
    <pubDate>Tue, 26 Aug 2025 19:00:00 +0100</pubDate>
    <itunes:duration>1134</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>3</itunes:season>
    <itunes:episode>2</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>GPT-5 has Arrived</itunes:title>
    <title>GPT-5 has Arrived</title>
    <itunes:summary><![CDATA[GPT-5 will change how hundreds of millions of people use AI. Yes, you might have to forgive the chart crimes, the underwhelming livestream and Altman hype… But it’s a good model. I have read the 50 page system card in full, have the benchmark scores, coding tests, and things you might have missed.  https://app.grayswan.ai/ai-explained  Announcement: https://openai.com/index/introducing-gpt-5/  System Card: https://cdn.openai.com/pdf/8124a3ce-ab78-4f06-96eb-49ea29ffb52f/gpt5-system-card-aug7.p...]]></itunes:summary>
    <description><![CDATA[<p><b>GPT-5 will change how hundreds of millions of people use AI. Yes, you might have to forgive the chart crimes, the underwhelming livestream and Altman hype… But it’s a good model. I have read the 50 page system card in full, have the benchmark scores, coding tests, and things you might have missed.<br/><br/></b><a href='https://app.grayswan.ai/ai-explained'><b>https://app.grayswan.ai/ai-explained<br/><br/></b></a><b>Announcement: </b><a href='https://openai.com/index/introducing-gpt-5/'><b>https://openai.com/index/introducing-gpt-5/<br/><br/></b></a><b>System Card: https://cdn.openai.com/pdf/8124a3ce-ab78-4f06-96eb-49ea29ffb52f/gpt5-system-card-aug7.pdf</b></p><p><b>Extra Paper: </b><a href='https://cdn.openai.com/pdf/be60c07b-6bc2-4f54-bcee-4141e1d6c69a/gpt-5-safe_completions.pdf'><b>https://cdn.openai.com/pdf/be60c07b-6bc2-4f54-bcee-4141e1d6c69a/gpt-5-safe_completions.pdf<br/><br/></b></a><b>Altman tweet: https://x.com/sama/status/1953551377873117369<br/><br/>Livestream: </b><a href='https://www.youtube.com/watch?v=0Uu_VJeVVfo'><b>https://www.youtube.com/watch?v=0Uu_VJeVVfo</b></a></p><p><b>METR Report: </b><a href='https://metr.github.io/autonomy-evals-guide/gpt-5-report/'><b>https://metr.github.io/autonomy-evals-guide/gpt-5-report/</b></a></p><p><b>ARC-AGI-2: https://x.com/fchollet/status/1953511631054680085<br/><br/>Claude Opus 4.1: </b><a href='https://www.anthropic.com/news/claude-opus-4-1'><b>https://www.anthropic.com/news/claude-opus-4-1</b></a></p><p><b>MMMU: </b><a href='https://mmmu-benchmark.github.io/'><b>https://mmmu-benchmark.github.io/</b></a></p><p><b>Cursor Praise: https://x.com/ryolu_/status/1953531724895596669</b></p><p><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>GPT-5 will change how hundreds of millions of people use AI. Yes, you might have to forgive the chart crimes, the underwhelming livestream and Altman hype… But it’s a good model. I have read the 50 page system card in full, have the benchmark scores, coding tests, and things you might have missed.<br/><br/></b><a href='https://app.grayswan.ai/ai-explained'><b>https://app.grayswan.ai/ai-explained<br/><br/></b></a><b>Announcement: </b><a href='https://openai.com/index/introducing-gpt-5/'><b>https://openai.com/index/introducing-gpt-5/<br/><br/></b></a><b>System Card: https://cdn.openai.com/pdf/8124a3ce-ab78-4f06-96eb-49ea29ffb52f/gpt5-system-card-aug7.pdf</b></p><p><b>Extra Paper: </b><a href='https://cdn.openai.com/pdf/be60c07b-6bc2-4f54-bcee-4141e1d6c69a/gpt-5-safe_completions.pdf'><b>https://cdn.openai.com/pdf/be60c07b-6bc2-4f54-bcee-4141e1d6c69a/gpt-5-safe_completions.pdf<br/><br/></b></a><b>Altman tweet: https://x.com/sama/status/1953551377873117369<br/><br/>Livestream: </b><a href='https://www.youtube.com/watch?v=0Uu_VJeVVfo'><b>https://www.youtube.com/watch?v=0Uu_VJeVVfo</b></a></p><p><b>METR Report: </b><a href='https://metr.github.io/autonomy-evals-guide/gpt-5-report/'><b>https://metr.github.io/autonomy-evals-guide/gpt-5-report/</b></a></p><p><b>ARC-AGI-2: https://x.com/fchollet/status/1953511631054680085<br/><br/>Claude Opus 4.1: </b><a href='https://www.anthropic.com/news/claude-opus-4-1'><b>https://www.anthropic.com/news/claude-opus-4-1</b></a></p><p><b>MMMU: </b><a href='https://mmmu-benchmark.github.io/'><b>https://mmmu-benchmark.github.io/</b></a></p><p><b>Cursor Praise: https://x.com/ryolu_/status/1953531724895596669</b></p><p><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17632575-gpt-5-has-arrived.mp3" length="10835497" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/xnms5blc3gtnpnvpkgc7mbuwinaa?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17632575</guid>
    <pubDate>Fri, 08 Aug 2025 00:00:00 +0100</pubDate>
    <itunes:duration>901</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>3</itunes:season>
    <itunes:episode>1</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Genie 3: The World Becomes Playable (DeepMind)</itunes:title>
    <title>Genie 3: The World Becomes Playable (DeepMind)</title>
    <itunes:summary><![CDATA[Soon, anything will be playable. A photo becomes an interactive world, a selfie becomes a new game. Genie 3 from Google, debuting just 2 hours ago, is what I mean, and I have the full analysis, plus the pushback I gave the authors (will it really lead to reliable AI agents? Is that even the point?). You make your own mind up, but it’s certainly fascinating, and not to be overlooked in the week that will bring us GPT-5.  https://80000hours.org/aiexplained  AI Insiders ($9!): https://www.patreo...]]></itunes:summary>
    <description><![CDATA[<p>Soon, anything will be playable. A photo becomes an interactive world, a selfie becomes a new game. Genie 3 from Google, debuting just 2 hours ago, is what I mean, and I have the full analysis, plus the pushback I gave the authors (will it really lead to reliable AI agents? Is that even the point?). You make your own mind up, but it’s certainly fascinating, and not to be overlooked in the week that will bring us GPT-5.<br/><br/>https://80000hours.org/aiexplained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters: <br/>00:00 - Introduction<br/>01:27 - Background and Access<br/>04:58 - Caveats<br/>07:24 - Demo<br/>10:12 - Conclusion<br/><br/>Announcement: https://deepmind.google/discover/blog/genie-3-a-new-frontier-for-world-models/<br/><br/>Isaac Labs: https://developer.nvidia.com/isaac/lab<br/><br/>Genie 2 Coverage: https://www.youtube.com/watch?v=jIm2T7h_a0M<br/><br/>TED Talk Roblox: https://www.youtube.com/watch?v=-OAP0ho5AUg<br/><br/>DeepThink Post: https://www.patreon.com/posts/deep-ish-on-new-135688441<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Soon, anything will be playable. A photo becomes an interactive world, a selfie becomes a new game. Genie 3 from Google, debuting just 2 hours ago, is what I mean, and I have the full analysis, plus the pushback I gave the authors (will it really lead to reliable AI agents? Is that even the point?). You make your own mind up, but it’s certainly fascinating, and not to be overlooked in the week that will bring us GPT-5.<br/><br/>https://80000hours.org/aiexplained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters: <br/>00:00 - Introduction<br/>01:27 - Background and Access<br/>04:58 - Caveats<br/>07:24 - Demo<br/>10:12 - Conclusion<br/><br/>Announcement: https://deepmind.google/discover/blog/genie-3-a-new-frontier-for-world-models/<br/><br/>Isaac Labs: https://developer.nvidia.com/isaac/lab<br/><br/>Genie 2 Coverage: https://www.youtube.com/watch?v=jIm2T7h_a0M<br/><br/>TED Talk Roblox: https://www.youtube.com/watch?v=-OAP0ho5AUg<br/><br/>DeepThink Post: https://www.patreon.com/posts/deep-ish-on-new-135688441<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17618980-genie-3-the-world-becomes-playable-deepmind.mp3" length="8594621" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/c9flazo949u9dzl6zlcclmbpskt1?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17618980</guid>
    <pubDate>Tue, 05 Aug 2025 17:00:00 +0100</pubDate>
    <itunes:duration>714</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>24</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>How Not to Read a Headline on AI (ft. new Olympiad Gold, GPT-5 …)</itunes:title>
    <title>How Not to Read a Headline on AI (ft. new Olympiad Gold, GPT-5 …)</title>
    <itunes:summary><![CDATA[GPT-5 did what? OpenAI ahead of Google? There are 9 ways to misread the headlines of the last 48 hours, so this video is here to tell you what happened, sans sizzle. It’s been a fairly momentous last few days, so let’s dive in to the International Math Olympiad Gold, GPT-5 alpha release, whether mathematicians are out of jobs, and the white collar impact by year’s end.   Job Board: https://80000hours.org/aiexplained   New Documentary on Patreon: https://www.patreon.com/posts/our-new-age-of-13...]]></itunes:summary>
    <description><![CDATA[<p><b>GPT-5 did what? OpenAI ahead of Google? There are 9 ways to misread the headlines of the last 48 hours, so this video is here to tell you what happened, sans sizzle. It’s been a fairly momentous last few days, so let’s dive in to the International Math Olympiad Gold, GPT-5 alpha release, whether mathematicians are out of jobs, and the white collar impact by year’s end.</b></p><p><br/></p><p><b>Job Board: </b><a href='https://80000hours.org/aiexplained'><b>https://80000hours.org/aiexplained</b></a></p><p><br/></p><p><b>New Documentary on Patreon: https://www.patreon.com/posts/our-new-age-of-133960279<br/><br/>Chapters: <br/>00:00 - Introduction<br/>00:18 - AI &gt; Mathematicians?</b></p><p><b>01:23 - OPENAI vs GOOGLE</b></p><p><b>02:42 - Irrelevant to Jobs or …</b></p><p><b>06:45 - White-collar jobs gone?</b></p><p><b>10:26 - AI is Plateauing?</b></p><p><b>12:00 - We Don’t Know the Details…</b></p><p><b>14:33 - GPT-5 alpha</b></p><p><b>14:54 - Nothing but Exponentials?</b></p><p><b>15:53 - No Impact?</b></p><p><br/></p><p><b>Announcement: </b><a href='https://x.com/alexwei_/status/1946477742855532918'><b>https://x.com/alexwei_/status/1946477742855532918<br/><br/></b></a><br/></p><p><b>UCLA Math Prof: </b><a href='https://x.com/ErnestRyu/status/1946699302308635130'><b>https://x.com/ErnestRyu/status/1946699302308635130</b></a></p><p><br/></p><p><b>ChatGPT Agent: https://openai.com/index/introducing-chatgpt-agent/</b></p><p><b>Livestream: </b><a href='https://www.youtube.com/watch?v=1jn_RpbPbEc&amp;t=796s'><b>https://www.youtube.com/watch?v=1jn_RpbPbEc&amp;t=796s<br/></b></a><b>System Card: </b><a href='https://cdn.openai.com/pdf/839e66fc-602c-48bf-81d3-b21eacc3459d/chatgpt_agent_system_card.pdf'><b>https://cdn.openai.com/pdf/839e66fc-602c-48bf-81d3-b21eacc3459d/chatgpt_agent_system_card.pdf<br/><br/></b></a><br/></p><p><b>Jerry Tworek (OpenAI): </b><a href='https://x.com/MillionInt/status/1946556255490982022'><b>https://x.com/MillionInt/status/1946556255490982022</b></a></p><p><a href='https://x.com/MillionInt/status/1946558130906968330'><b>https://x.com/MillionInt/status/1946558130906968330</b></a></p><p><br/></p><p><b>Noam Brown Details: https://x.com/polynoamial/status/1946478249187377206</b></p><p><br/></p><p><b>Trieu Tranh Retweet: </b><a href='https://x.com/Mihonarium/status/1946880931723194389'><b>https://x.com/Mihonarium/status/1946880931723194389</b></a></p><p><br/></p><p><b>Neel Nanda: </b><a href='https://x.com/NeelNanda5/status/1946602953370173647'><b>https://x.com/NeelNanda5/status/1946602953370173647</b></a></p><p><br/></p><p><b>Terence Tao: </b><a href='https://mathstodon.xyz/@tao'><b>https://mathstodon.xyz/@tao</b></a></p><p><br/></p><p><b>Sam Altman: </b><a href='https://x.com/sama/status/1946569252296929727'><b>https://x.com/sama/status/1946569252296929727</b></a></p><p><br/></p><p><b>METR Dev Study: </b><a href='https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/'><b>https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/</b></a></p><p><br/></p><p><b>Ravid Schwatz: </b><a href='https://x.com/ziv_ravid/status/1946378712716562605'><b>https://x.com/ziv_ravid/status/1946378712716562605<br/><br/></b></a><br/></p><p><b>AlphaEvolve: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/</b></p><p><br/></p><p><a href='https://simple-bench.com/'><b>https://simple-bench.com/</b></a></p><p><br/></p><p><b>Meta Salary: https://www.tomshardware.com/tech-industry/artificial-intelligence/abel-founder-claims-meta-offered-usd1-25-billion-over-four-years-to-ai-hire-person-still-said-no-despite-equivalent-of-usd312-million-yearly-salary</b></p><p><br/></p><p><b>$2k per month: https://www.theinformation.com/articles/openai-considers-higher-priced-subscriptions-to-its-chatbot-ai-preview-of-the-informations-ai-summit?rc=sy0ihq</b></p><p><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>GPT-5 did what? OpenAI ahead of Google? There are 9 ways to misread the headlines of the last 48 hours, so this video is here to tell you what happened, sans sizzle. It’s been a fairly momentous last few days, so let’s dive in to the International Math Olympiad Gold, GPT-5 alpha release, whether mathematicians are out of jobs, and the white collar impact by year’s end.</b></p><p><br/></p><p><b>Job Board: </b><a href='https://80000hours.org/aiexplained'><b>https://80000hours.org/aiexplained</b></a></p><p><br/></p><p><b>New Documentary on Patreon: https://www.patreon.com/posts/our-new-age-of-133960279<br/><br/>Chapters: <br/>00:00 - Introduction<br/>00:18 - AI &gt; Mathematicians?</b></p><p><b>01:23 - OPENAI vs GOOGLE</b></p><p><b>02:42 - Irrelevant to Jobs or …</b></p><p><b>06:45 - White-collar jobs gone?</b></p><p><b>10:26 - AI is Plateauing?</b></p><p><b>12:00 - We Don’t Know the Details…</b></p><p><b>14:33 - GPT-5 alpha</b></p><p><b>14:54 - Nothing but Exponentials?</b></p><p><b>15:53 - No Impact?</b></p><p><br/></p><p><b>Announcement: </b><a href='https://x.com/alexwei_/status/1946477742855532918'><b>https://x.com/alexwei_/status/1946477742855532918<br/><br/></b></a><br/></p><p><b>UCLA Math Prof: </b><a href='https://x.com/ErnestRyu/status/1946699302308635130'><b>https://x.com/ErnestRyu/status/1946699302308635130</b></a></p><p><br/></p><p><b>ChatGPT Agent: https://openai.com/index/introducing-chatgpt-agent/</b></p><p><b>Livestream: </b><a href='https://www.youtube.com/watch?v=1jn_RpbPbEc&amp;t=796s'><b>https://www.youtube.com/watch?v=1jn_RpbPbEc&amp;t=796s<br/></b></a><b>System Card: </b><a href='https://cdn.openai.com/pdf/839e66fc-602c-48bf-81d3-b21eacc3459d/chatgpt_agent_system_card.pdf'><b>https://cdn.openai.com/pdf/839e66fc-602c-48bf-81d3-b21eacc3459d/chatgpt_agent_system_card.pdf<br/><br/></b></a><br/></p><p><b>Jerry Tworek (OpenAI): </b><a href='https://x.com/MillionInt/status/1946556255490982022'><b>https://x.com/MillionInt/status/1946556255490982022</b></a></p><p><a href='https://x.com/MillionInt/status/1946558130906968330'><b>https://x.com/MillionInt/status/1946558130906968330</b></a></p><p><br/></p><p><b>Noam Brown Details: https://x.com/polynoamial/status/1946478249187377206</b></p><p><br/></p><p><b>Trieu Tranh Retweet: </b><a href='https://x.com/Mihonarium/status/1946880931723194389'><b>https://x.com/Mihonarium/status/1946880931723194389</b></a></p><p><br/></p><p><b>Neel Nanda: </b><a href='https://x.com/NeelNanda5/status/1946602953370173647'><b>https://x.com/NeelNanda5/status/1946602953370173647</b></a></p><p><br/></p><p><b>Terence Tao: </b><a href='https://mathstodon.xyz/@tao'><b>https://mathstodon.xyz/@tao</b></a></p><p><br/></p><p><b>Sam Altman: </b><a href='https://x.com/sama/status/1946569252296929727'><b>https://x.com/sama/status/1946569252296929727</b></a></p><p><br/></p><p><b>METR Dev Study: </b><a href='https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/'><b>https://metr.org/blog/2025-07-10-early-2025-ai-experienced-os-dev-study/</b></a></p><p><br/></p><p><b>Ravid Schwatz: </b><a href='https://x.com/ziv_ravid/status/1946378712716562605'><b>https://x.com/ziv_ravid/status/1946378712716562605<br/><br/></b></a><br/></p><p><b>AlphaEvolve: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/</b></p><p><br/></p><p><a href='https://simple-bench.com/'><b>https://simple-bench.com/</b></a></p><p><br/></p><p><b>Meta Salary: https://www.tomshardware.com/tech-industry/artificial-intelligence/abel-founder-claims-meta-offered-usd1-25-billion-over-four-years-to-ai-hire-person-still-said-no-despite-equivalent-of-usd312-million-yearly-salary</b></p><p><br/></p><p><b>$2k per month: https://www.theinformation.com/articles/openai-considers-higher-priced-subscriptions-to-its-chatbot-ai-preview-of-the-informations-ai-summit?rc=sy0ihq</b></p><p><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17539163-how-not-to-read-a-headline-on-ai-ft-new-olympiad-gold-gpt-5.mp3" length="12502980" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ufltwdyt7rue8csiqf15qsjba72v?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17539163</guid>
    <pubDate>Mon, 21 Jul 2025 16:00:00 +0100</pubDate>
    <itunes:duration>1039</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>23</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Grok 4 - 10 New Things to Know</itunes:title>
    <title>Grok 4 - 10 New Things to Know</title>
    <itunes:summary><![CDATA[Grok 4 is here, but did you know these 10 things about the new model? From benchmark caveats to soloing science, $300 a month secrets to Grok 5 promises, here's 10 new things to know in just under 12 minutes.  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 00:22 - Benchmark Results 02:11 - Benchmark Caveats 02:59 - ARC-AGI 2  03:35 - SimpleBench 04:49 - ‘Humanity’s Last Exam’ 07:20 - SuperGrok Heavy Price 07:58 - API Price 08:12 - Grok 5, Gemini 3....]]></itunes:summary>
    <description><![CDATA[<p>Grok 4 is here, but did you know these 10 things about the new model? From benchmark caveats to soloing science, $300 a month secrets to Grok 5 promises, here&apos;s 10 new things to know in just under 12 minutes.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:22 - Benchmark Results<br/>02:11 - Benchmark Caveats<br/>02:59 - ARC-AGI 2 <br/>03:35 - SimpleBench<br/>04:49 - ‘Humanity’s Last Exam’<br/>07:20 - SuperGrok Heavy Price<br/>07:58 - API Price<br/>08:12 - Grok 5, Gemini 3.0 Beta, GPT-5<br/>09:12 - System Prompt Change + $1B a month, pollution<br/>10:20 - Not soloing science, helping you solo code<br/><br/>Livestream: https://www.youtube.com/watch?v=1tQ_KrlHgfg&amp;t=1s<br/><br/>Price: https://grok.com/#subscribe<br/>https://x.com/ArtificialAnlys/status/1943166841150644622<br/><br/>Gemini DeepThink: https://blog.google/technology/google-deepmind/google-gemini-updates-io-2025/#deep-think<br/><br/>https://simple-bench.com/<br/><br/>ARC-AGI 2: https://x.com/arcprize/status/1943168950763950555<br/><br/>Humanity’s Last Exam: https://agi.safe.ai/<br/><br/>SmartGPT: https://www.youtube.com/watch?v=hVade_8H8mE<br/><br/>New Power Plant, 1m GPUs: https://www.tomshardware.com/tech-industry/artificial-intelligence/elon-musk-xai-power-plant-overseas-to-power-1-million-gpus<br/><br/>Gemini 3.0 beta: https://web.archive.org/web/20250709174548/https://github.com/google-gemini/gemini-cli/blob/b0cce952860b9ff51a0f731fbb8a7649ead23530/packages/cli/src/ui/utils/errorParsing.test.ts<br/><br/>Pollution: https://www.theguardian.com/technology/2025/apr/24/elon-musk-xai-memphis<br/>https://www.youtube.com/watch?v=C8rU4dv2w8Q<br/>https://www.youtube.com/watch?v=3VJT2JeDCyw<br/><br/>System Prompt: https://github.com/xai-org/grok-prompts/blob/535aa67a6221ce4928761335a38dea8e678d8501/ask_grok_system_prompt.j2<br/><br/>Burn Rate: https://www.bloomberg.com/news/articles/2025-06-17/musk-s-xai-burning-through-1-billion-a-month-as-costs-pile-up<br/><br/>Ron Johnson: https://x.com/jdcmedlock/status/1939814516503847259<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Grok 4 is here, but did you know these 10 things about the new model? From benchmark caveats to soloing science, $300 a month secrets to Grok 5 promises, here&apos;s 10 new things to know in just under 12 minutes.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:22 - Benchmark Results<br/>02:11 - Benchmark Caveats<br/>02:59 - ARC-AGI 2 <br/>03:35 - SimpleBench<br/>04:49 - ‘Humanity’s Last Exam’<br/>07:20 - SuperGrok Heavy Price<br/>07:58 - API Price<br/>08:12 - Grok 5, Gemini 3.0 Beta, GPT-5<br/>09:12 - System Prompt Change + $1B a month, pollution<br/>10:20 - Not soloing science, helping you solo code<br/><br/>Livestream: https://www.youtube.com/watch?v=1tQ_KrlHgfg&amp;t=1s<br/><br/>Price: https://grok.com/#subscribe<br/>https://x.com/ArtificialAnlys/status/1943166841150644622<br/><br/>Gemini DeepThink: https://blog.google/technology/google-deepmind/google-gemini-updates-io-2025/#deep-think<br/><br/>https://simple-bench.com/<br/><br/>ARC-AGI 2: https://x.com/arcprize/status/1943168950763950555<br/><br/>Humanity’s Last Exam: https://agi.safe.ai/<br/><br/>SmartGPT: https://www.youtube.com/watch?v=hVade_8H8mE<br/><br/>New Power Plant, 1m GPUs: https://www.tomshardware.com/tech-industry/artificial-intelligence/elon-musk-xai-power-plant-overseas-to-power-1-million-gpus<br/><br/>Gemini 3.0 beta: https://web.archive.org/web/20250709174548/https://github.com/google-gemini/gemini-cli/blob/b0cce952860b9ff51a0f731fbb8a7649ead23530/packages/cli/src/ui/utils/errorParsing.test.ts<br/><br/>Pollution: https://www.theguardian.com/technology/2025/apr/24/elon-musk-xai-memphis<br/>https://www.youtube.com/watch?v=C8rU4dv2w8Q<br/>https://www.youtube.com/watch?v=3VJT2JeDCyw<br/><br/>System Prompt: https://github.com/xai-org/grok-prompts/blob/535aa67a6221ce4928761335a38dea8e678d8501/ask_grok_system_prompt.j2<br/><br/>Burn Rate: https://www.bloomberg.com/news/articles/2025-06-17/musk-s-xai-burning-through-1-billion-a-month-as-costs-pile-up<br/><br/>Ron Johnson: https://x.com/jdcmedlock/status/1939814516503847259<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17483875-grok-4-10-new-things-to-know.mp3" length="8474758" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/h0xw272ncv9lgobzkgq4rh9tir39?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17483875</guid>
    <pubDate>Thu, 10 Jul 2025 15:00:00 +0100</pubDate>
    <itunes:duration>703</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>22</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>When Will AI Models Blackmail You, and Why?</itunes:title>
    <title>When Will AI Models Blackmail You, and Why?</title>
    <itunes:summary><![CDATA[In the last few days Anthropic have released an impressive honest account of how all models blackmail, no matter what goal they have, and despite prompt warnings, and other preventions. But do these models *want* this?  Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: storyblocks.com/AIExplained   AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 01:20 - What prompts blackmail? 02:44 - Black...]]></itunes:summary>
    <description><![CDATA[<p>In the last few days Anthropic have released an impressive honest account of how all models blackmail, no matter what goal they have, and despite prompt warnings, and other preventions. But do these models *want* this?<br/><br/>Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: storyblocks.com/AIExplained<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:20 - What prompts blackmail?<br/>02:44 - Blackmail walkthrough <br/>06:04 - ‘American interests’<br/>08:00 - Inherent desire?<br/>10:45 - Switching Goals<br/>11:35 - Murder<br/>12:22 - Realizing it’s a scenario? <br/>15:02 - Prompt engineering fix?<br/>16:27 - Any fixes?<br/>17:45 - Chekov’s Gun<br/>19:25 - Job implications<br/>21:19 - Bonus Details<br/><br/>Report: https://www.anthropic.com/research/agentic-misalignment<br/>30 Page Appendices: https://assets.anthropic.com/m/6d46dac66e1a132a/original/Agentic_Misalignment_Appendix.pdf<br/>Announcement: https://x.com/AnthropicAI/status/1936144602446082431?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet<br/>OpenAI Files: https://www.openaifiles.org/<br/>Grok 4 News: https://x.com/RonFilipkowski/status/1936372579607912473<br/>Claude 4 Report Card: https://www-cdn.anthropic.com/6be99a52cb68eb70eb9572b4cafad13df32ed995.pdf<br/>New Apollo Research: https://www.apolloresearch.ai/blog/more-capable-models-are-better-at-in-context-scheming<br/>Interesting Reflections: https://nostalgebraist.tumblr.com/post/785766737747574784/the-void<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>In the last few days Anthropic have released an impressive honest account of how all models blackmail, no matter what goal they have, and despite prompt warnings, and other preventions. But do these models *want* this?<br/><br/>Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: storyblocks.com/AIExplained<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:20 - What prompts blackmail?<br/>02:44 - Blackmail walkthrough <br/>06:04 - ‘American interests’<br/>08:00 - Inherent desire?<br/>10:45 - Switching Goals<br/>11:35 - Murder<br/>12:22 - Realizing it’s a scenario? <br/>15:02 - Prompt engineering fix?<br/>16:27 - Any fixes?<br/>17:45 - Chekov’s Gun<br/>19:25 - Job implications<br/>21:19 - Bonus Details<br/><br/>Report: https://www.anthropic.com/research/agentic-misalignment<br/>30 Page Appendices: https://assets.anthropic.com/m/6d46dac66e1a132a/original/Agentic_Misalignment_Appendix.pdf<br/>Announcement: https://x.com/AnthropicAI/status/1936144602446082431?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet<br/>OpenAI Files: https://www.openaifiles.org/<br/>Grok 4 News: https://x.com/RonFilipkowski/status/1936372579607912473<br/>Claude 4 Report Card: https://www-cdn.anthropic.com/6be99a52cb68eb70eb9572b4cafad13df32ed995.pdf<br/>New Apollo Research: https://www.apolloresearch.ai/blog/more-capable-models-are-better-at-in-context-scheming<br/>Interesting Reflections: https://nostalgebraist.tumblr.com/post/785766737747574784/the-void<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17390142-when-will-ai-models-blackmail-you-and-why.mp3" length="19005735" type="audio/mpeg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17390142</guid>
    <pubDate>Tue, 24 Jun 2025 15:00:00 +0100</pubDate>
    <itunes:duration>1579</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>21</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Apple’s ‘AI Can’t Reason’ Claim Seen By 13M+, What You Need to Know </itunes:title>
    <title>Apple’s ‘AI Can’t Reason’ Claim Seen By 13M+, What You Need to Know </title>
    <itunes:summary><![CDATA[What to make of those headlines that AI can’t reason, seen by tens of millions? I cover the paper in layman’s terms, what it means and doesn’t mean, and what’s next.   Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: https://storyblocks.com/AIExplained  Plus o3-pro and whether it is my current most-recommended model.  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 00:57 - Viral Post ...]]></itunes:summary>
    <description><![CDATA[<p>What to make of those headlines that AI can’t reason, seen by tens of millions? I cover the paper in layman’s terms, what it means and doesn’t mean, and what’s next. <br/><br/>Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: https://storyblocks.com/AIExplained<br/><br/>Plus o3-pro and whether it is my current most-recommended model.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:57 - Viral Post + Headlines<br/>01:42 - Apple Paper Analysis<br/>08:34 - But they do Hallucinate <br/>10:43 - Not Supercomputers<br/>11:18 - o3 Pro and Recommendations <br/><br/><br/>13.7M Tweet: https://x.com/RubenHssd/status/1931389580105925115<br/><br/>Apple Paper: https://ml-site.cdn-apple.com/papers/the-illusion-of-thinking.pdf<br/><br/>Guardian Article: https://www.theguardian.com/technology/2025/jun/09/apple-artificial-intelligence-ai-study-collapse<br/><br/>Lisan al Gaib post: https://x.com/scaling01/status/1931854370716426246<br/><br/>Multiplication: https://x.com/yuntiandeng/status/1836114401213989366<br/><br/>The Illusion of the Illusion of Thinking: https://drive.google.com/file/d/1Zx9ikRj0Enc3SB4wA9HlYIlpmO_8QiUO/view<br/><br/>Marcus: https://www.theguardian.com/commentisfree/2025/jun/10/billion-dollar-ai-puzzle-break-down<br/><br/>Prof Rao: https://x.com/rao2z/status/1927707640223719631<br/><br/>AI Job Headlines: https://www.nytimes.com/2025/06/11/technology/ai-mechanize-jobs.html<br/>https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/><br/>Sky News Story: https://news.sky.com/story/can-we-trust-chatgpt-despite-it-hallucinating-answers-13380975<br/><br/>Veo 3 Ad: https://x.com/Kalshi/status/1932891608388681791<br/><br/>Altman Essay: https://blog.samaltman.com/<br/><br/>o3 Original benchmarks: https://substackcdn.com/image/fetch/f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5b8b6c44-acd6-43b3-b5c6-1a1d5c6c25e4_2486x1388.png<br/><br/>https://pbs.twimg.com/media/GfQ0bfcXQAAQt13.jpg<br/><br/>Alpha Evolve Video: https://www.youtube.com/watch?v=RH4hAgvYSzg<br/><br/>https://simple-bench.com/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>What to make of those headlines that AI can’t reason, seen by tens of millions? I cover the paper in layman’s terms, what it means and doesn’t mean, and what’s next. <br/><br/>Thanks to Storyblocks for sponsoring this video! Download unlimited stock media at one set price with Storyblocks: https://storyblocks.com/AIExplained<br/><br/>Plus o3-pro and whether it is my current most-recommended model.<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:57 - Viral Post + Headlines<br/>01:42 - Apple Paper Analysis<br/>08:34 - But they do Hallucinate <br/>10:43 - Not Supercomputers<br/>11:18 - o3 Pro and Recommendations <br/><br/><br/>13.7M Tweet: https://x.com/RubenHssd/status/1931389580105925115<br/><br/>Apple Paper: https://ml-site.cdn-apple.com/papers/the-illusion-of-thinking.pdf<br/><br/>Guardian Article: https://www.theguardian.com/technology/2025/jun/09/apple-artificial-intelligence-ai-study-collapse<br/><br/>Lisan al Gaib post: https://x.com/scaling01/status/1931854370716426246<br/><br/>Multiplication: https://x.com/yuntiandeng/status/1836114401213989366<br/><br/>The Illusion of the Illusion of Thinking: https://drive.google.com/file/d/1Zx9ikRj0Enc3SB4wA9HlYIlpmO_8QiUO/view<br/><br/>Marcus: https://www.theguardian.com/commentisfree/2025/jun/10/billion-dollar-ai-puzzle-break-down<br/><br/>Prof Rao: https://x.com/rao2z/status/1927707640223719631<br/><br/>AI Job Headlines: https://www.nytimes.com/2025/06/11/technology/ai-mechanize-jobs.html<br/>https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/><br/>Sky News Story: https://news.sky.com/story/can-we-trust-chatgpt-despite-it-hallucinating-answers-13380975<br/><br/>Veo 3 Ad: https://x.com/Kalshi/status/1932891608388681791<br/><br/>Altman Essay: https://blog.samaltman.com/<br/><br/>o3 Original benchmarks: https://substackcdn.com/image/fetch/f_auto,q_auto:good,fl_progressive:steep/https%3A%2F%2Fsubstack-post-media.s3.amazonaws.com%2Fpublic%2Fimages%2F5b8b6c44-acd6-43b3-b5c6-1a1d5c6c25e4_2486x1388.png<br/><br/>https://pbs.twimg.com/media/GfQ0bfcXQAAQt13.jpg<br/><br/>Alpha Evolve Video: https://www.youtube.com/watch?v=RH4hAgvYSzg<br/><br/>https://simple-bench.com/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17326326-apple-s-ai-can-t-reason-claim-seen-by-13m-what-you-need-to-know.mp3" length="10106688" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/v6j3cegl9aqd2d0w84w0101bsjqk?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17326326</guid>
    <pubDate>Thu, 12 Jun 2025 18:00:00 +0100</pubDate>
    <itunes:duration>840</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>20</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>AI Accelerates: New Gemini Model + AI Unemployment Stories Analysed</itunes:title>
    <title>AI Accelerates: New Gemini Model + AI Unemployment Stories Analysed</title>
    <itunes:summary><![CDATA[There’s a new best language model, so let’s go through the up and downs of Gemini 2.5 Pro 06-05. Record-breaking common-sense, but dumb mistakes remain. And it’s not even their best model, which remains behind the scenes - Gemini 2.5 Ultra. Plus Sundar Pichai’s AGI date and an analysis of whether the current AI unemployment headlines are justified, and Elevenlabs v3.   https://emergentmind.com   AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 02:04 - Gem...]]></itunes:summary>
    <description><![CDATA[<p>There’s a new best language model, so let’s go through the up and downs of Gemini 2.5 Pro 06-05. Record-breaking common-sense, but dumb mistakes remain. And it’s not even their best model, which remains behind the scenes - Gemini 2.5 Ultra. Plus Sundar Pichai’s AGI date and an analysis of whether the current AI unemployment headlines are justified, and Elevenlabs v3.<br/><br/><br/>https://emergentmind.com<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>02:04 - Gemini 2.5 Ultra <br/>03:34 - Benchmarks<br/>07:41 - AGI Date and Meaning Pichai<br/>09:13 - Jobs and AI Unemployment Fears<br/>15:28 - Elevenlabs v3<br/><br/>Sundar Pichai Fridman: https://www.youtube.com/watch?v=9V6tWC4CdFQ<br/><br/>Pichai More Jobs (until 2026 at least): https://www.techradar.com/pro/alphabet-ceo-sundar-pichai-says-ai-wont-lead-to-job-cuts-will-be-an-accelerator<br/><br/>Gemini Comparison: https://blog.google/products/gemini/gemini-2-5-pro-latest-preview/<br/>https://x.com/viathebrink/status/1930733154203292121<br/><br/>https://simple-bench.com/<br/><br/>White Collar Bloodbath: https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/>https://fortune.com/2025/05/25/ai-entry-level-jobs-gen-z-careers-young-workers-linkedin/<br/>https://www.nytimes.com/2025/05/19/opinion/linkedin-ai-entry-level-jobs.html<br/>https://www.nytimes.com/2025/03/25/business/economy/white-collar-layoffs.html<br/><br/>College Unemployment: https://www.newyorkfed.org/research/college-labor-market/#--:explore:unemployment<br/><br/>New Scientist AI Hallucinaitons: https://www.newscientist.com/article/2479545-ai-hallucinations-are-getting-worse-and-theyre-here-to-stay/<br/><br/>Duolingo: https://fortune.com/2025/05/24/duolingo-ai-first-employees-ceo-luis-von-ahn/<br/>Klarna: https://www.forbes.com/sites/quickerbettertech/2025/05/18/business-tech-news-klarna-reverses-on-ai-says-customers-like-talking-to-people/<br/><br/>Sholto Douglas: https://www.reddit.com/r/ClaudeAI/comments/1ktt1rb/anthropics_sholto_douglas_says_by_202728_its/<br/><br/>Figure 02: https://x.com/adcock_brett/status/1930693311771332853<br/><br/>Elevenlabs v3: https://www.youtube.com/watch?v=zv_IoWIO5Ek<br/><br/>Gemini Speech Generation: https://aistudio.google.com/generate-speech<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>There’s a new best language model, so let’s go through the up and downs of Gemini 2.5 Pro 06-05. Record-breaking common-sense, but dumb mistakes remain. And it’s not even their best model, which remains behind the scenes - Gemini 2.5 Ultra. Plus Sundar Pichai’s AGI date and an analysis of whether the current AI unemployment headlines are justified, and Elevenlabs v3.<br/><br/><br/>https://emergentmind.com<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>02:04 - Gemini 2.5 Ultra <br/>03:34 - Benchmarks<br/>07:41 - AGI Date and Meaning Pichai<br/>09:13 - Jobs and AI Unemployment Fears<br/>15:28 - Elevenlabs v3<br/><br/>Sundar Pichai Fridman: https://www.youtube.com/watch?v=9V6tWC4CdFQ<br/><br/>Pichai More Jobs (until 2026 at least): https://www.techradar.com/pro/alphabet-ceo-sundar-pichai-says-ai-wont-lead-to-job-cuts-will-be-an-accelerator<br/><br/>Gemini Comparison: https://blog.google/products/gemini/gemini-2-5-pro-latest-preview/<br/>https://x.com/viathebrink/status/1930733154203292121<br/><br/>https://simple-bench.com/<br/><br/>White Collar Bloodbath: https://www.axios.com/2025/05/28/ai-jobs-white-collar-unemployment-anthropic<br/>https://fortune.com/2025/05/25/ai-entry-level-jobs-gen-z-careers-young-workers-linkedin/<br/>https://www.nytimes.com/2025/05/19/opinion/linkedin-ai-entry-level-jobs.html<br/>https://www.nytimes.com/2025/03/25/business/economy/white-collar-layoffs.html<br/><br/>College Unemployment: https://www.newyorkfed.org/research/college-labor-market/#--:explore:unemployment<br/><br/>New Scientist AI Hallucinaitons: https://www.newscientist.com/article/2479545-ai-hallucinations-are-getting-worse-and-theyre-here-to-stay/<br/><br/>Duolingo: https://fortune.com/2025/05/24/duolingo-ai-first-employees-ceo-luis-von-ahn/<br/>Klarna: https://www.forbes.com/sites/quickerbettertech/2025/05/18/business-tech-news-klarna-reverses-on-ai-says-customers-like-talking-to-people/<br/><br/>Sholto Douglas: https://www.reddit.com/r/ClaudeAI/comments/1ktt1rb/anthropics_sholto_douglas_says_by_202728_its/<br/><br/>Figure 02: https://x.com/adcock_brett/status/1930693311771332853<br/><br/>Elevenlabs v3: https://www.youtube.com/watch?v=zv_IoWIO5Ek<br/><br/>Gemini Speech Generation: https://aistudio.google.com/generate-speech<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17293380-ai-accelerates-new-gemini-model-ai-unemployment-stories-analysed.mp3" length="12046606" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/to5t4n1zzc1y54a4okl60305bq1c?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17293380</guid>
    <pubDate>Fri, 06 Jun 2025 16:00:00 +0100</pubDate>
    <itunes:duration>1001</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>19</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Claude 4: Full 120 Page Breakdown … Is it the Best New Model?</itunes:title>
    <title>Claude 4: Full 120 Page Breakdown … Is it the Best New Model?</title>
    <itunes:summary><![CDATA[Not only did I get early access and ran my own tests, as per the title I read both the 120 page Claude 4 Opus and Claude 4 Sonnet System Card, and 25 page report on ASL-3 being triggered, plus the 2 hour launch video, and surrounding coverage. Ft. coding tests, Simple, twitter controversies, deep alignment coverage, spiritual bliss and much more!   https://80000hours.org/aiexplained   Chapters:  00:00 - Introduction 01:12 - 3 Quick Controversies 02:42 - Benchmark Results  04:20 - 12...]]></itunes:summary>
    <description><![CDATA[<p><b>Not only did I get early access and ran my own tests, as per the title I read both the 120 page Claude 4 Opus and Claude 4 Sonnet System Card, and 25 page report on ASL-3 being triggered, plus the 2 hour launch video, and surrounding coverage. Ft. coding tests, Simple, twitter controversies, deep alignment coverage, spiritual bliss and much more!</b></p><p><br/></p><p><b>https://</b><a href='http://80000hours.org/aiexplained'><b>80000hours.org/aiexplained</b></a></p><p><br/></p><p><b>Chapters: </b></p><p><b>00:00 - Introduction<br/>01:12 - 3 Quick Controversies</b></p><p><b>02:42 - Benchmark Results </b></p><p><b>04:20 - 120 page Card 20 Highlights</b></p><p><b>10:07 - Coding Test<br/>11:27 - Model Welfare and Spiritual Bliss</b></p><p><b>13:29 -  ASL-3<br/><br/>Claude Card: </b><a href='https://www-cdn.anthropic.com/4263b940cabb546aa0e3283f35b686f4f3b2ff47.pdf?s=09'><b>https://www-cdn.anthropic.com/4263b940cabb546aa0e3283f35b686f4f3b2ff47.pdf?s=09<br/></b></a><b>ASL 3:</b><a href='https://www-cdn.anthropic.com/807c59454757214bfd37592d6e048079cd7a7728.pdf'><b>https://www-cdn.anthropic.com/807c59454757214bfd37592d6e048079cd7a7728.pdf</b></a></p><p><b>Tweets: </b><a href='https://x.com/fish_kyle3/status/1925597284546629753'><b>https://x.com/fish_kyle3/status/1925597284546629753</b></a></p><p><a href='https://x.com/EMostaque/status/1925624164527874452?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet'><b>https://x.com/EMostaque/status/1925624164527874452?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet<br/><br/></b></a><br/></p><p><b>Cursor Says State of the Art for Coding: https://x.com/cursor_ai/status/1925594428095561941</b></p><p><br/></p><p><b>Benchmarks: https://www.anthropic.com/news/claude-4</b></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>Not only did I get early access and ran my own tests, as per the title I read both the 120 page Claude 4 Opus and Claude 4 Sonnet System Card, and 25 page report on ASL-3 being triggered, plus the 2 hour launch video, and surrounding coverage. Ft. coding tests, Simple, twitter controversies, deep alignment coverage, spiritual bliss and much more!</b></p><p><br/></p><p><b>https://</b><a href='http://80000hours.org/aiexplained'><b>80000hours.org/aiexplained</b></a></p><p><br/></p><p><b>Chapters: </b></p><p><b>00:00 - Introduction<br/>01:12 - 3 Quick Controversies</b></p><p><b>02:42 - Benchmark Results </b></p><p><b>04:20 - 120 page Card 20 Highlights</b></p><p><b>10:07 - Coding Test<br/>11:27 - Model Welfare and Spiritual Bliss</b></p><p><b>13:29 -  ASL-3<br/><br/>Claude Card: </b><a href='https://www-cdn.anthropic.com/4263b940cabb546aa0e3283f35b686f4f3b2ff47.pdf?s=09'><b>https://www-cdn.anthropic.com/4263b940cabb546aa0e3283f35b686f4f3b2ff47.pdf?s=09<br/></b></a><b>ASL 3:</b><a href='https://www-cdn.anthropic.com/807c59454757214bfd37592d6e048079cd7a7728.pdf'><b>https://www-cdn.anthropic.com/807c59454757214bfd37592d6e048079cd7a7728.pdf</b></a></p><p><b>Tweets: </b><a href='https://x.com/fish_kyle3/status/1925597284546629753'><b>https://x.com/fish_kyle3/status/1925597284546629753</b></a></p><p><a href='https://x.com/EMostaque/status/1925624164527874452?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet'><b>https://x.com/EMostaque/status/1925624164527874452?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet<br/><br/></b></a><br/></p><p><b>Cursor Says State of the Art for Coding: https://x.com/cursor_ai/status/1925594428095561941</b></p><p><br/></p><p><b>Benchmarks: https://www.anthropic.com/news/claude-4</b></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17212117-claude-4-full-120-page-breakdown-is-it-the-best-new-model.mp3" length="13756410" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/3kbh6xd2zsmmlsn0xj02vzyaso33?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17212117</guid>
    <pubDate>Thu, 22 May 2025 23:00:00 +0100</pubDate>
    <itunes:duration>1144</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>18</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Google Takes No Prisoners Amid Torrent of AI Announcements</itunes:title>
    <title>Google Takes No Prisoners Amid Torrent of AI Announcements</title>
    <itunes:summary><![CDATA[Google just announced at least 12 things that are each worthy of a video, but here are the top I/O highlights. From Veo 3 to Deep Research now being useable, Deep Think breaking records to Gemini Diffusion, Gemini 2.5 Flash changing how AI is priced and GemmaVerse, SynthID Detector and Imagen 4. And even this intro is missing other announcements covered in the vid! And yes, they’ll be plenty of Veo 3 clips to enjoy…  https://80000hours.org/aiexplained  AI Insiders ($9!): https://www.patreon.c...]]></itunes:summary>
    <description><![CDATA[<p>Google just announced at least 12 things that are each worthy of a video, but here are the top I/O highlights. From Veo 3 to Deep Research now being useable, Deep Think breaking records to Gemini Diffusion, Gemini 2.5 Flash changing how AI is priced and GemmaVerse, SynthID Detector and Imagen 4. And even this intro is missing other announcements covered in the vid! And yes, they’ll be plenty of Veo 3 clips to enjoy…<br/><br/>https://80000hours.org/aiexplained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:48 - Veo 3<br/>02:10 - Gemini 2.5 Flash<br/>03:13 - Universal Assistant<br/>03:47 - Usage Skyrockets + OpenAI dig<br/>04:51 - Gemini Pro Deep Think<br/>06:21 - Overviews and AI Mode<br/>07:26 - Deep Research Updates (new) + Jules <br/>08:53 - Make and Deploy Apps with Gemini<br/>09:12 - Imagen 4 <br/>10:00 - Gemini Diffusion<br/>11:46 - Try It On<br/>12:17 - SynthID Detector<br/>13:30 - GemmaVerse, SignGemma, Gemma3n, medGemma<br/>14:24 - Outro + Clips<br/><br/>Event: https://www.youtube.com/watch?v=o8NiE3XMPrM<br/>Ntaive Audio: https://aistudio.google.com/generate-speech<br/>Gemini Diffusion: https://deepmind.google/models/gemini-diffusion/#capabilities <br/>New Gemini 2.5 Flash: https://deepmind.google/models/gemini/flash/<br/>SignGemma (See end of this vid): https://www.youtube.com/watch?v=GjvgtwSOCao<br/>Deep Think: https://blog.google/technology/google-deepmind/google-gemini-updates-io-2025/#flash-improvements<br/>Google Parallel Sampling: https://www.patreon.com/posts/next-level-good-127441188<br/><br/>Price Plans: https://blog.google/products/google-one/google-ai-ultra/<br/>Imagen 4 Benchmarks: https://deepmind.google/models/imagen/<br/>Jules: https://jules.google/<br/>SynthID Detector: https://blog.google/technology/ai/google-synthid-ai-content-detector/<br/>Veo 3 Benchmarks: https://deepmind.google/models/veo/evals/<br/>MedGemma: https://deepmind.google/models/gemma/medgemma/<br/>Build Apps: https://aistudio.google.com/apps<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Google just announced at least 12 things that are each worthy of a video, but here are the top I/O highlights. From Veo 3 to Deep Research now being useable, Deep Think breaking records to Gemini Diffusion, Gemini 2.5 Flash changing how AI is priced and GemmaVerse, SynthID Detector and Imagen 4. And even this intro is missing other announcements covered in the vid! And yes, they’ll be plenty of Veo 3 clips to enjoy…<br/><br/>https://80000hours.org/aiexplained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:48 - Veo 3<br/>02:10 - Gemini 2.5 Flash<br/>03:13 - Universal Assistant<br/>03:47 - Usage Skyrockets + OpenAI dig<br/>04:51 - Gemini Pro Deep Think<br/>06:21 - Overviews and AI Mode<br/>07:26 - Deep Research Updates (new) + Jules <br/>08:53 - Make and Deploy Apps with Gemini<br/>09:12 - Imagen 4 <br/>10:00 - Gemini Diffusion<br/>11:46 - Try It On<br/>12:17 - SynthID Detector<br/>13:30 - GemmaVerse, SignGemma, Gemma3n, medGemma<br/>14:24 - Outro + Clips<br/><br/>Event: https://www.youtube.com/watch?v=o8NiE3XMPrM<br/>Ntaive Audio: https://aistudio.google.com/generate-speech<br/>Gemini Diffusion: https://deepmind.google/models/gemini-diffusion/#capabilities <br/>New Gemini 2.5 Flash: https://deepmind.google/models/gemini/flash/<br/>SignGemma (See end of this vid): https://www.youtube.com/watch?v=GjvgtwSOCao<br/>Deep Think: https://blog.google/technology/google-deepmind/google-gemini-updates-io-2025/#flash-improvements<br/>Google Parallel Sampling: https://www.patreon.com/posts/next-level-good-127441188<br/><br/>Price Plans: https://blog.google/products/google-one/google-ai-ultra/<br/>Imagen 4 Benchmarks: https://deepmind.google/models/imagen/<br/>Jules: https://jules.google/<br/>SynthID Detector: https://blog.google/technology/ai/google-synthid-ai-content-detector/<br/>Veo 3 Benchmarks: https://deepmind.google/models/veo/evals/<br/>MedGemma: https://deepmind.google/models/gemma/medgemma/<br/>Build Apps: https://aistudio.google.com/apps<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17204461-google-takes-no-prisoners-amid-torrent-of-ai-announcements.mp3" length="12372268" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/kznx89cmqbmxg3nrnwhisgofi7zk?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17204461</guid>
    <pubDate>Wed, 21 May 2025 19:00:00 +0100</pubDate>
    <itunes:duration>1027</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>17</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>AI Improves at Self-improving</itunes:title>
    <title>AI Improves at Self-improving</title>
    <itunes:summary><![CDATA[AlphaEvolve is not the first system to exhibit self-improvement, but it may be the most impressive yet. AI is literally improving the hardware, architectures, data and training methods of AI itself. A deep dive into the paper, drawing on two previous interviews and 5 other papers. Plus a snippet on OpenAI’s new Codex system.  Gray Swan: http://app.grayswan.ai/ai-explained  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 00:27 - AlphaEvolve 05:23 - Limita...]]></itunes:summary>
    <description><![CDATA[<p>AlphaEvolve is not the first system to exhibit self-improvement, but it may be the most impressive yet. AI is literally improving the hardware, architectures, data and training methods of AI itself. A deep dive into the paper, drawing on two previous interviews and 5 other papers. Plus a snippet on OpenAI’s new Codex system.<br/><br/>Gray Swan: http://app.grayswan.ai/ai-explained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:27 - AlphaEvolve<br/>05:23 - Limitation<br/>06:10 - Achievements<br/>08:21 - Future Improvements<br/>13:30 - Quirks<br/>16:34 - Final Thoughts<br/><br/>AlphaEvolve release: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/<br/><br/>Paper: https://storage.googleapis.com/deepmind-media/DeepMind.com/Blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/AlphaEvolve.pdf<br/><br/>Terence Tao Quote: https://mathstodon.xyz/@tao/114508029896631083<br/><br/>Nature Article: https://www.nature.com/articles/s41586-022-05172-4<br/>MIT Article: https://www.technologyreview.com/2025/05/14/1116438/google-deepminds-new-ai-uses-large-language-models-to-crack-real-world-problems/<br/>AI Co-Scientist: https://arxiv.org/pdf/2502.18864<br/><br/>OpenAI Codex: https://openai.com/index/introducing-codex/<br/><br/><br/>70% of Pull Requests: https://x.com/slow_developer/status/1920920456393028027<br/><br/>Amodei Essay: https://www.darioamodei.com/essay/machines-of-loving-grace<br/><br/>OpenAI Jason Wei Tweet: https://x.com/_jasonwei/status/1923091260354531612<br/><br/>PromptBreeder: https://arxiv.org/pdf/2309.16797<br/>DrEureka: https://arxiv.org/pdf/2406.01967<br/><br/>FT DeepMind: https://www.ft.com/content/4e497a91-670a-4f69-be4a-18e247daba3e<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>AlphaEvolve is not the first system to exhibit self-improvement, but it may be the most impressive yet. AI is literally improving the hardware, architectures, data and training methods of AI itself. A deep dive into the paper, drawing on two previous interviews and 5 other papers. Plus a snippet on OpenAI’s new Codex system.<br/><br/>Gray Swan: http://app.grayswan.ai/ai-explained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:27 - AlphaEvolve<br/>05:23 - Limitation<br/>06:10 - Achievements<br/>08:21 - Future Improvements<br/>13:30 - Quirks<br/>16:34 - Final Thoughts<br/><br/>AlphaEvolve release: https://deepmind.google/discover/blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/<br/><br/>Paper: https://storage.googleapis.com/deepmind-media/DeepMind.com/Blog/alphaevolve-a-gemini-powered-coding-agent-for-designing-advanced-algorithms/AlphaEvolve.pdf<br/><br/>Terence Tao Quote: https://mathstodon.xyz/@tao/114508029896631083<br/><br/>Nature Article: https://www.nature.com/articles/s41586-022-05172-4<br/>MIT Article: https://www.technologyreview.com/2025/05/14/1116438/google-deepminds-new-ai-uses-large-language-models-to-crack-real-world-problems/<br/>AI Co-Scientist: https://arxiv.org/pdf/2502.18864<br/><br/>OpenAI Codex: https://openai.com/index/introducing-codex/<br/><br/><br/>70% of Pull Requests: https://x.com/slow_developer/status/1920920456393028027<br/><br/>Amodei Essay: https://www.darioamodei.com/essay/machines-of-loving-grace<br/><br/>OpenAI Jason Wei Tweet: https://x.com/_jasonwei/status/1923091260354531612<br/><br/>PromptBreeder: https://arxiv.org/pdf/2309.16797<br/>DrEureka: https://arxiv.org/pdf/2406.01967<br/><br/>FT DeepMind: https://www.ft.com/content/4e497a91-670a-4f69-be4a-18e247daba3e<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17188622-ai-improves-at-self-improving.mp3" length="12762144" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/pil9vng45eycexnajz631yln7fc3?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17188622</guid>
    <pubDate>Mon, 19 May 2025 13:00:00 +0100</pubDate>
    <itunes:duration>1061</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>16</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>o3 breaks (some) records, but AI becomes pay-to-win</itunes:title>
    <title>o3 breaks (some) records, but AI becomes pay-to-win</title>
    <itunes:summary><![CDATA[A green card, o3 vs Gemini 2.5, 6 Benchmarks and a whole bunch of my thoughts on what on earth is happening in AI, from here to 2030. Plus, how AI is becoming pay-to-win, and why. Crazy times, 14 mins probably wasn’t enough.  https://app.grayswan.ai/ai-explained  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters: 00:00 - Introduction 00:33 - FictionLiveBench 01:37 - PHYBench 02:14 - SimpleBench 02:54 - Virology Capabilities Test 03:13 - Mathematics Performance 04:29 - Vision Be...]]></itunes:summary>
    <description><![CDATA[<p>A green card, o3 vs Gemini 2.5, 6 Benchmarks and a whole bunch of my thoughts on what on earth is happening in AI, from here to 2030. Plus, how AI is becoming pay-to-win, and why. Crazy times, 14 mins probably wasn’t enough.<br/><br/>https://app.grayswan.ai/ai-explained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:33 - FictionLiveBench<br/>01:37 - PHYBench<br/>02:14 - SimpleBench<br/>02:54 - Virology Capabilities Test<br/>03:13 - Mathematics Performance<br/>04:29 - Vision Benchmarks<br/>05:43 - V* and how o3 works<br/>06:44 - Revenue and costs for you<br/>08:54 - Expensive RL and trade-offs <br/>09:40 - How to spend the OOMs<br/>13:27 - Gray Swan Arena<br/><br/>Green Card: https://techcrunch.com/2025/04/25/an-openai-researcher-who-worked-on-gpt-4-5-had-their-green-card-denied/<br/>PHYBench: https://arxiv.org/pdf/2504.16074Virologytest: https://www.virologytest.ai/<br/>How o3 Vision Works: https://arxiv.org/pdf/2312.14135 https://x.com/sainingxie/status/1912570624523829573<br/>Visual puzzles: https://neulab.github.io/VisualPuzzles/<br/>Fiction Bench: https://x.com/ficlive/status/1912863028141244850<br/>https://geobench.org/<br/>https://simple-bench.com/<br/>AIME 2025: https://openai.com/index/introducing-o3-and-o4-mini/<br/>USAMO: https://x.com/mbalunovic/status/1914398518896193747<br/>NaturalBench: https://linzhiqiu.github.io/papers/naturalbench/<br/>Where’s Waldo: https://uk.pinterest.com/pin/492792384225896298/<br/>IMO and AlphaProof:https://deepmind.google/discover/blog/ai-solves-imo-problems-at-silver-medal-level/<br/>Crazy Revenue: https://www.theinformation.com/articles/openai-forecasts-revenue-topping-125-billion-2029-agents-new-products-gain?rc=sy0ihq<br/>Number of Users: https://www.theinformation.com/briefings/googles-gemini-user-numbers-revealed-court?rc=sy0ihq<br/>Subscriptions pay to win: https://www.forbes.com/sites/paulmonckton/2025/04/23/google-leak-reveals-new-gemini-ai-subscription-levels/<br/>GPU Trade-offs: https://x.com/sama/status/1915098951067554030<br/>RL Scale-up Amodei: https://www.darioamodei.com/post/on-deepseek-and-export-controls<br/>Log-linear Returns: https://x.com/bobmcgrewai/status/1895228291981943265<br/>2030 Scaling: https://epoch.ai/blog/can-ai-scaling-continue-through-2030<br/>Model Size: https://x.com/slow_developer/status/1874554473256997201<br/>Adam on AGI: https://x.com/TheRealAdamG/status/1913998366632968381<br/>Papers on Patreon: https://arxiv.org/pdf/2502.01839<br/>https://arxiv.org/pdf/2504.13837<br/>Chollet Quote: https://x.com/fchollet/status/1912934762580447447<br/>OpenSim: https://opensim.stanford.edu/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>A green card, o3 vs Gemini 2.5, 6 Benchmarks and a whole bunch of my thoughts on what on earth is happening in AI, from here to 2030. Plus, how AI is becoming pay-to-win, and why. Crazy times, 14 mins probably wasn’t enough.<br/><br/>https://app.grayswan.ai/ai-explained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:33 - FictionLiveBench<br/>01:37 - PHYBench<br/>02:14 - SimpleBench<br/>02:54 - Virology Capabilities Test<br/>03:13 - Mathematics Performance<br/>04:29 - Vision Benchmarks<br/>05:43 - V* and how o3 works<br/>06:44 - Revenue and costs for you<br/>08:54 - Expensive RL and trade-offs <br/>09:40 - How to spend the OOMs<br/>13:27 - Gray Swan Arena<br/><br/>Green Card: https://techcrunch.com/2025/04/25/an-openai-researcher-who-worked-on-gpt-4-5-had-their-green-card-denied/<br/>PHYBench: https://arxiv.org/pdf/2504.16074Virologytest: https://www.virologytest.ai/<br/>How o3 Vision Works: https://arxiv.org/pdf/2312.14135 https://x.com/sainingxie/status/1912570624523829573<br/>Visual puzzles: https://neulab.github.io/VisualPuzzles/<br/>Fiction Bench: https://x.com/ficlive/status/1912863028141244850<br/>https://geobench.org/<br/>https://simple-bench.com/<br/>AIME 2025: https://openai.com/index/introducing-o3-and-o4-mini/<br/>USAMO: https://x.com/mbalunovic/status/1914398518896193747<br/>NaturalBench: https://linzhiqiu.github.io/papers/naturalbench/<br/>Where’s Waldo: https://uk.pinterest.com/pin/492792384225896298/<br/>IMO and AlphaProof:https://deepmind.google/discover/blog/ai-solves-imo-problems-at-silver-medal-level/<br/>Crazy Revenue: https://www.theinformation.com/articles/openai-forecasts-revenue-topping-125-billion-2029-agents-new-products-gain?rc=sy0ihq<br/>Number of Users: https://www.theinformation.com/briefings/googles-gemini-user-numbers-revealed-court?rc=sy0ihq<br/>Subscriptions pay to win: https://www.forbes.com/sites/paulmonckton/2025/04/23/google-leak-reveals-new-gemini-ai-subscription-levels/<br/>GPU Trade-offs: https://x.com/sama/status/1915098951067554030<br/>RL Scale-up Amodei: https://www.darioamodei.com/post/on-deepseek-and-export-controls<br/>Log-linear Returns: https://x.com/bobmcgrewai/status/1895228291981943265<br/>2030 Scaling: https://epoch.ai/blog/can-ai-scaling-continue-through-2030<br/>Model Size: https://x.com/slow_developer/status/1874554473256997201<br/>Adam on AGI: https://x.com/TheRealAdamG/status/1913998366632968381<br/>Papers on Patreon: https://arxiv.org/pdf/2502.01839<br/>https://arxiv.org/pdf/2504.13837<br/>Chollet Quote: https://x.com/fchollet/status/1912934762580447447<br/>OpenSim: https://opensim.stanford.edu/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/17044292-o3-breaks-some-records-but-ai-becomes-pay-to-win.mp3" length="10518990" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/2j6t0ahdqji2p1s345tedkw7i844?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-17044292</guid>
    <pubDate>Fri, 25 Apr 2025 20:00:00 +0100</pubDate>
    <itunes:duration>873</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>15</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>o3 and o4-mini - they’re great, but easy to over-hype</itunes:title>
    <title>o3 and o4-mini - they’re great, but easy to over-hype</title>
    <itunes:summary><![CDATA[Critical analysis of the two most powerful new models behind ChatGPT, o3 and o4-mini. Not just the system cards, benchmarks, and my own tests, but some you may not have seen before. Yes, they can whip up amazing front-end in a few seconds, but you always have to ask what is in their data. Either way, they prove the gains from RL are just beginning…  https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained  AI Insiders ($9!): https://www.pat...]]></itunes:summary>
    <description><![CDATA[<p>Critical analysis of the two most powerful new models behind ChatGPT, o3 and o4-mini. Not just the system cards, benchmarks, and my own tests, but some you may not have seen before. Yes, they can whip up amazing front-end in a few seconds, but you always have to ask what is in their data. Either way, they prove the gains from RL are just beginning…<br/><br/>https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - o3 and o4-mini<br/><br/><br/>https://simple-bench.com/<br/><br/>Plus, Teams and Pro,  plus token count: https://x.com/btibor91/status/1912568994512662679<br/><br/>System Card: https://openai.com/index/o3-o4-mini-system-card/<br/><br/>Release Notes: https://openai.com/index/introducing-o3-and-o4-mini/<br/><br/>https://deepmind.google/technologies/gemini/pro/<br/><br/>https://x.com/DeryaTR_/status/1912558350794961168<br/><br/>https://x.com/polynoamial/status/1912564068168450396<br/><br/>API Pricing:https://openai.com/api/pricing/<br/><br/>https://aider.chat/docs/leaderboards/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Critical analysis of the two most powerful new models behind ChatGPT, o3 and o4-mini. Not just the system cards, benchmarks, and my own tests, but some you may not have seen before. Yes, they can whip up amazing front-end in a few seconds, but you always have to ask what is in their data. Either way, they prove the gains from RL are just beginning…<br/><br/>https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/><br/>Chapters:<br/>00:00 - o3 and o4-mini<br/><br/><br/>https://simple-bench.com/<br/><br/>Plus, Teams and Pro,  plus token count: https://x.com/btibor91/status/1912568994512662679<br/><br/>System Card: https://openai.com/index/o3-o4-mini-system-card/<br/><br/>Release Notes: https://openai.com/index/introducing-o3-and-o4-mini/<br/><br/>https://deepmind.google/technologies/gemini/pro/<br/><br/>https://x.com/DeryaTR_/status/1912558350794961168<br/><br/>https://x.com/polynoamial/status/1912564068168450396<br/><br/>API Pricing:https://openai.com/api/pricing/<br/><br/>https://aider.chat/docs/leaderboards/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16991608-o3-and-o4-mini-they-re-great-but-easy-to-over-hype.mp3" length="10397987" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ni2xif5lnytdutjg5owau4b56xe4?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16991608</guid>
    <pubDate>Wed, 16 Apr 2025 21:00:00 +0100</pubDate>
    <itunes:duration>864</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>14</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>‘Speaking Dolphin’ to AI Data Dominance, 4.1 + Kling 2: 7 Developments Critically Analysed</itunes:title>
    <title>‘Speaking Dolphin’ to AI Data Dominance, 4.1 + Kling 2: 7 Developments Critically Analysed</title>
    <itunes:summary><![CDATA[This pod won’t just be about the release of GPT 4.1 in the last 48 hours, o3 build-up, Kling 2.0, a sneak-peak at the next OpenAI model, or even the new Dolphin language tool. It will be about 7 such stories that contextualise where we are in AI and what is happening. https://www.emergentmind.com/   Chapters:  00:00 - Introduction 00:30 - Kling 2.0 01:35 - GPT 4.1 05:25 - o3 Build-up 07:37 - ‘Product Company’ 09:31 - Safe Superintelligence 10:54 - DolphinGemma 13:16 - Data Dominance?   K...]]></itunes:summary>
    <description><![CDATA[<p><b>This pod won’t just be about the release of GPT 4.1 in the last 48 hours, o3 build-up, Kling 2.0, a sneak-peak at the next OpenAI model, or even the new Dolphin language tool. It will be about 7 such stories that contextualise where we are in AI and what is happening.</b></p><p><a href='https://www.emergentmind.com/'><b>https://www.emergentmind.com/</b></a></p><p><br/></p><p><b>Chapters: </b></p><p><b>00:00 - Introduction</b></p><p><b>00:30 - Kling 2.0</b></p><p><b>01:35 - GPT 4.1</b></p><p><b>05:25 - o3 Build-up</b></p><p><b>07:37 - ‘Product Company’</b></p><p><b>09:31 - Safe Superintelligence</b></p><p><b>10:54 - DolphinGemma</b></p><p><b>13:16 - Data Dominance?</b></p><p><br/></p><p><b>Kling 2.0: </b><a href='https://app.klingai.com/global/release-notes'><b>https://app.klingai.com/global/release-notes</b></a></p><p><br/></p><p><b>Dolphin Gemma: </b><a href='https://blog.google/technology/ai/dolphingemma/?s=09'><b>https://blog.google/technology/ai/dolphingemma/?s=09</b></a></p><p><br/></p><p><b>https://openai.com/index/gpt-4-1/</b></p><p><br/></p><p><b>OpenAI o3 Build-up The Information: </b><a href='https://www.theinformation.com/articles/openais-latest-breakthrough-ai-comes-new-ideas?rc=sy0ihq'><b>https://www.theinformation.com/articles/openais-latest-breakthrough-ai-comes-new-ideas?rc=sy0ihq</b></a></p><p><br/></p><p><b>Physical reasoning: </b><a href='https://x.com/a_karvonen/status/1911839968990814503'><b>https://x.com/a_karvonen/status/1911839968990814503</b></a></p><p><br/></p><p><b>Fiction Live.bench: </b><a href='https://x.com/ficlive/status/1911853409847906626'><b>https://x.com/ficlive/status/1911853409847906626</b></a></p><p><br/></p><p><b>Altman Ted: </b><a href='https://www.youtube.com/watch?v=5MWT_doo68k'><b>https://www.youtube.com/watch?v=5MWT_doo68k</b></a></p><p><br/></p><p><b>https://simple-bench.com/try-yourself</b></p><p><br/></p><p><a href='https://aider.chat/docs/leaderboards/'><b>https://aider.chat/docs/leaderboards/</b></a></p><p><br/></p><p><b>4.5: </b><a href='https://www.youtube.com/watch?v=6nJZopACRuQ'><b>https://www.youtube.com/watch?v=6nJZopACRuQ</b></a></p><p><br/></p><p><b>Geospatial reasoning: </b><a href='https://research.google/blog/geospatial-reasoning-unlocking-insights-with-generative-ai-and-multiple-foundation-models/'><b>https://research.google/blog/geospatial-reasoning-unlocking-insights-with-generative-ai-and-multiple-foundation-models/</b></a></p><p><br/></p><p><b>Pioneers: </b><a href='https://x.com/OpenAIDevs/status/1910017976256119151'><b>https://x.com/OpenAIDevs/status/1910017976256119151</b></a></p><p><b>Evals: </b><a href='https://www.youtube.com/watch?v=scsW6_2SPC4'><b>https://www.youtube.com/watch?v=scsW6_2SPC4</b></a></p><p><b>Anthropic Updates: </b><a href='https://www.bloomberg.com/news/articles/2025-04-15/anthropic-is-readying-a-voice-assistant-feature-to-rival-openai?srnd=phx-ai'><b>https://www.bloomberg.com/news/articles/2025-04-15/anthropic-is-readying-a-voice-assistant-feature-to-rival-openai?srnd=phx-ai</b></a></p><p><b>https://x.com/sethsaler/status/1912188383457059301</b></p><p><br/></p><p><a href='https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/'><b>https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/</b></a></p><p><b>https://ai.meta.com/blog/llama-4-multimodal-intelligence/</b></p><p><a href='https://deepmind.google/technologies/gemini/pro/'><b>https://deepmind.google/technologies/gemini/pro/</b></a></p><p><a href='https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/'><b>https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/</b></a></p><p><b>https://blog.google/products/google-cloud/ironwood-tpu-age-of-inference/</b></p><p><b>OpenAI Documentary: https://www.patreon.com/posts/one-machine-to-121940490</b></p>]]></description>
    <content:encoded><![CDATA[<p><b>This pod won’t just be about the release of GPT 4.1 in the last 48 hours, o3 build-up, Kling 2.0, a sneak-peak at the next OpenAI model, or even the new Dolphin language tool. It will be about 7 such stories that contextualise where we are in AI and what is happening.</b></p><p><a href='https://www.emergentmind.com/'><b>https://www.emergentmind.com/</b></a></p><p><br/></p><p><b>Chapters: </b></p><p><b>00:00 - Introduction</b></p><p><b>00:30 - Kling 2.0</b></p><p><b>01:35 - GPT 4.1</b></p><p><b>05:25 - o3 Build-up</b></p><p><b>07:37 - ‘Product Company’</b></p><p><b>09:31 - Safe Superintelligence</b></p><p><b>10:54 - DolphinGemma</b></p><p><b>13:16 - Data Dominance?</b></p><p><br/></p><p><b>Kling 2.0: </b><a href='https://app.klingai.com/global/release-notes'><b>https://app.klingai.com/global/release-notes</b></a></p><p><br/></p><p><b>Dolphin Gemma: </b><a href='https://blog.google/technology/ai/dolphingemma/?s=09'><b>https://blog.google/technology/ai/dolphingemma/?s=09</b></a></p><p><br/></p><p><b>https://openai.com/index/gpt-4-1/</b></p><p><br/></p><p><b>OpenAI o3 Build-up The Information: </b><a href='https://www.theinformation.com/articles/openais-latest-breakthrough-ai-comes-new-ideas?rc=sy0ihq'><b>https://www.theinformation.com/articles/openais-latest-breakthrough-ai-comes-new-ideas?rc=sy0ihq</b></a></p><p><br/></p><p><b>Physical reasoning: </b><a href='https://x.com/a_karvonen/status/1911839968990814503'><b>https://x.com/a_karvonen/status/1911839968990814503</b></a></p><p><br/></p><p><b>Fiction Live.bench: </b><a href='https://x.com/ficlive/status/1911853409847906626'><b>https://x.com/ficlive/status/1911853409847906626</b></a></p><p><br/></p><p><b>Altman Ted: </b><a href='https://www.youtube.com/watch?v=5MWT_doo68k'><b>https://www.youtube.com/watch?v=5MWT_doo68k</b></a></p><p><br/></p><p><b>https://simple-bench.com/try-yourself</b></p><p><br/></p><p><a href='https://aider.chat/docs/leaderboards/'><b>https://aider.chat/docs/leaderboards/</b></a></p><p><br/></p><p><b>4.5: </b><a href='https://www.youtube.com/watch?v=6nJZopACRuQ'><b>https://www.youtube.com/watch?v=6nJZopACRuQ</b></a></p><p><br/></p><p><b>Geospatial reasoning: </b><a href='https://research.google/blog/geospatial-reasoning-unlocking-insights-with-generative-ai-and-multiple-foundation-models/'><b>https://research.google/blog/geospatial-reasoning-unlocking-insights-with-generative-ai-and-multiple-foundation-models/</b></a></p><p><br/></p><p><b>Pioneers: </b><a href='https://x.com/OpenAIDevs/status/1910017976256119151'><b>https://x.com/OpenAIDevs/status/1910017976256119151</b></a></p><p><b>Evals: </b><a href='https://www.youtube.com/watch?v=scsW6_2SPC4'><b>https://www.youtube.com/watch?v=scsW6_2SPC4</b></a></p><p><b>Anthropic Updates: </b><a href='https://www.bloomberg.com/news/articles/2025-04-15/anthropic-is-readying-a-voice-assistant-feature-to-rival-openai?srnd=phx-ai'><b>https://www.bloomberg.com/news/articles/2025-04-15/anthropic-is-readying-a-voice-assistant-feature-to-rival-openai?srnd=phx-ai</b></a></p><p><b>https://x.com/sethsaler/status/1912188383457059301</b></p><p><br/></p><p><a href='https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/'><b>https://techcrunch.com/2025/04/12/openai-co-founder-ilya-sutskevers-safe-superintelligence-reportedly-valued-at-32b/</b></a></p><p><b>https://ai.meta.com/blog/llama-4-multimodal-intelligence/</b></p><p><a href='https://deepmind.google/technologies/gemini/pro/'><b>https://deepmind.google/technologies/gemini/pro/</b></a></p><p><a href='https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/'><b>https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/</b></a></p><p><b>https://blog.google/products/google-cloud/ironwood-tpu-age-of-inference/</b></p><p><b>OpenAI Documentary: https://www.patreon.com/posts/one-machine-to-121940490</b></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16989978-speaking-dolphin-to-ai-data-dominance-4-1-kling-2-7-developments-critically-analysed.mp3" length="14545026" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/lmyslfhk1vycmwy6kiak20v3ej1v?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16989978</guid>
    <pubDate>Wed, 16 Apr 2025 17:00:00 +0100</pubDate>
    <itunes:duration>1209</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>13</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>AI CEO: ‘Stock Crash Could Stop AI Progress’, Llama 4 Anti-climax +‘Superintelligence in 2027’...</itunes:title>
    <title>AI CEO: ‘Stock Crash Could Stop AI Progress’, Llama 4 Anti-climax +‘Superintelligence in 2027’...</title>
    <itunes:summary><![CDATA[The latest on Llama 4, and whether it signals a slowdown in AI, or solid progress. Plus, a deep dive on that viral prediction of superintelligence by 2027, and Amodei’s cautionary words on what could stop AI progress in its tracks. o3 news, and more, as well.  Weights &amp; Biases: https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained   DeepSeek Doc: https://www.patreon.com/posts/openai-is-not-r1-125869969  AI Insiders ($9!): https://www...]]></itunes:summary>
    <description><![CDATA[<p>The latest on Llama 4, and whether it signals a slowdown in AI, or solid progress. Plus, a deep dive on that viral prediction of superintelligence by 2027, and Amodei’s cautionary words on what could stop AI progress in its tracks. o3 news, and more, as well.<br/><br/>Weights &amp; Biases: https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained<br/><br/><br/>DeepSeek Doc: https://www.patreon.com/posts/openai-is-not-r1-125869969<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:47 - Stock Crash <br/>02:28 - Llama 4<br/>10:55 - o3 News<br/>11:59 - OpenAI non-profit?<br/>13:13 - AI 2027<br/><br/>Llama 4 Release: https://ai.meta.com/blog/llama-4-multimodal-intelligence/<br/><br/>Dario Amodei Comments: https://www.youtube.com/watch?v=esCSpbDPJik<br/><br/>Knowledge Cut-off: https://www.llama.com/docs/model-cards-and-prompt-formats/llama4_omni/<br/><br/>Aider Polyglot: https://aider.chat/docs/leaderboards/<br/><br/>Gemini 1.5: https://arxiv.org/pdf/2403.05530<br/><br/>Fiction-LiveBench: https://fiction.live/stories/Fiction-liveBench-Mar-25-2025/oQdzQvKHw8JyXbN87<br/><br/>OpenAI Valuation: https://www.nytimes.com/2025/03/31/technology/openai-valuation-300-billion.html?login=smartlock&amp;auth=login-smartlock<br/><br/>OpenAI Cybersecurity: https://www.bloomberg.com/news/articles/2024-01-16/openai-working-with-us-military-on-cybersecurity-tools-for-veterans<br/><br/>Deep research System Card: https://cdn.openai.com/deep-research-system-card.pdf<br/><br/>https://openai.com/index/paperbench/<br/><br/>AI 2027: https://ai-2027.com/<br/><br/>METR Paper: https://arxiv.org/pdf/2503.14499<br/><br/>OpenAI non-profit: https://openai.com/index/nonprofit-commission-guidance/<br/><br/>NYT Piece: https://www.nytimes.com/2025/04/03/technology/ai-futures-project-ai-2027.html?unlocked_article_code=1.804._yKi.QhwOp15Q3tcU&amp;smid=url-share&amp;s=09<br/><br/>Kokotajlo predictions 2021: https://www.lesswrong.com/posts/6Xgy6CAf2jqHhynHL/what-2026-looks-like<br/><br/>https://simple-bench.com/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>The latest on Llama 4, and whether it signals a slowdown in AI, or solid progress. Plus, a deep dive on that viral prediction of superintelligence by 2027, and Amodei’s cautionary words on what could stop AI progress in its tracks. o3 news, and more, as well.<br/><br/>Weights &amp; Biases: https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained<br/><br/><br/>DeepSeek Doc: https://www.patreon.com/posts/openai-is-not-r1-125869969<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:47 - Stock Crash <br/>02:28 - Llama 4<br/>10:55 - o3 News<br/>11:59 - OpenAI non-profit?<br/>13:13 - AI 2027<br/><br/>Llama 4 Release: https://ai.meta.com/blog/llama-4-multimodal-intelligence/<br/><br/>Dario Amodei Comments: https://www.youtube.com/watch?v=esCSpbDPJik<br/><br/>Knowledge Cut-off: https://www.llama.com/docs/model-cards-and-prompt-formats/llama4_omni/<br/><br/>Aider Polyglot: https://aider.chat/docs/leaderboards/<br/><br/>Gemini 1.5: https://arxiv.org/pdf/2403.05530<br/><br/>Fiction-LiveBench: https://fiction.live/stories/Fiction-liveBench-Mar-25-2025/oQdzQvKHw8JyXbN87<br/><br/>OpenAI Valuation: https://www.nytimes.com/2025/03/31/technology/openai-valuation-300-billion.html?login=smartlock&amp;auth=login-smartlock<br/><br/>OpenAI Cybersecurity: https://www.bloomberg.com/news/articles/2024-01-16/openai-working-with-us-military-on-cybersecurity-tools-for-veterans<br/><br/>Deep research System Card: https://cdn.openai.com/deep-research-system-card.pdf<br/><br/>https://openai.com/index/paperbench/<br/><br/>AI 2027: https://ai-2027.com/<br/><br/>METR Paper: https://arxiv.org/pdf/2503.14499<br/><br/>OpenAI non-profit: https://openai.com/index/nonprofit-commission-guidance/<br/><br/>NYT Piece: https://www.nytimes.com/2025/04/03/technology/ai-futures-project-ai-2027.html?unlocked_article_code=1.804._yKi.QhwOp15Q3tcU&amp;smid=url-share&amp;s=09<br/><br/>Kokotajlo predictions 2021: https://www.lesswrong.com/posts/6Xgy6CAf2jqHhynHL/what-2026-looks-like<br/><br/>https://simple-bench.com/<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16931936-ai-ceo-stock-crash-could-stop-ai-progress-llama-4-anti-climax-superintelligence-in-2027.mp3" length="17232234" type="audio/mpeg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16931936</guid>
    <pubDate>Mon, 07 Apr 2025 17:00:00 +0100</pubDate>
    <itunes:duration>1431</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>12</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Gemini 2.5 Pro - It’s a Smart Chatbot … (New Simple High Score)</itunes:title>
    <title>Gemini 2.5 Pro - It’s a Smart Chatbot … (New Simple High Score)</title>
    <itunes:summary><![CDATA[Gemini gets a new record on Simple Bench, and several other benchmarks. I’ll go deep to explore its nuances, including how it deceptively reverse engineers answers, does better on certain coding benchmarks than others, may have a universal ‘conceptual language’ …  https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained  … and more. Plus practical tips, a note on security and Kling vs Veo 2 guest appearance.   AI Insiders ($9!): https://www...]]></itunes:summary>
    <description><![CDATA[<p>Gemini gets a new record on Simple Bench, and several other benchmarks. I’ll go deep to explore its nuances, including how it deceptively reverse engineers answers, does better on certain coding benchmarks than others, may have a universal ‘conceptual language’ …<br/><br/>https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained<br/><br/>… and more. Plus practical tips, a note on security and Kling vs Veo 2 guest appearance.<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:36 - Fiction Bench<br/>02:41 - Practicality - YouTube urls + Security - cut-off date<br/>03:42 - Coding <br/>06:22 - WeirdML Bench<br/>07:01 - Simple Bench Record High <br/>11:23 - Reverse Engineering!<br/>13:22 - Anthropic Paper<br/>17:49 - 3 Caveats<br/><br/>Gemini 2.5 Updated: https://deepmind.google/technologies/gemini/<br/><br/>Fiction Live Bench: https://fiction.live/stories/Fiction-liveBench-Feb-19-2025/oQdzQvKHw8JyXbN87<br/><br/>https://simple-bench.com/<br/><br/>WeirdML: https://htihle.github.io/weirdml.html<br/>https://x.com/htihle/status/1905014058228625542<br/><br/>Anthropic Thoughts: https://www.anthropic.com/research/tracing-thoughts-language-model<br/>https://transformer-circuits.pub/2025/attribution-graphs/biology.html#dives-cot<br/><br/>https://aistudio.google.com/prompts/new_chat<br/><br/>Search Study: https://www.cjr.org/tow_center/we-compared-eight-ai-search-engines-theyre-all-bad-at-citing-news.php<br/><br/>Live bench: https://livebench.ai/#/<br/>Paper: https://arxiv.org/pdf/2406.19314<br/><br/>LiveCode Bench: https://livecodebench.github.io/<br/><br/>SWE-Verified: https://arxiv.org/pdf/2310.06770<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Gemini gets a new record on Simple Bench, and several other benchmarks. I’ll go deep to explore its nuances, including how it deceptively reverse engineers answers, does better on certain coding benchmarks than others, may have a universal ‘conceptual language’ …<br/><br/>https://weave-docs.wandb.ai/?utm_source=sponsorship&amp;utm_medium=simple_bench&amp;utm_campaign=ai_explained<br/><br/>… and more. Plus practical tips, a note on security and Kling vs Veo 2 guest appearance.<br/><br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:36 - Fiction Bench<br/>02:41 - Practicality - YouTube urls + Security - cut-off date<br/>03:42 - Coding <br/>06:22 - WeirdML Bench<br/>07:01 - Simple Bench Record High <br/>11:23 - Reverse Engineering!<br/>13:22 - Anthropic Paper<br/>17:49 - 3 Caveats<br/><br/>Gemini 2.5 Updated: https://deepmind.google/technologies/gemini/<br/><br/>Fiction Live Bench: https://fiction.live/stories/Fiction-liveBench-Feb-19-2025/oQdzQvKHw8JyXbN87<br/><br/>https://simple-bench.com/<br/><br/>WeirdML: https://htihle.github.io/weirdml.html<br/>https://x.com/htihle/status/1905014058228625542<br/><br/>Anthropic Thoughts: https://www.anthropic.com/research/tracing-thoughts-language-model<br/>https://transformer-circuits.pub/2025/attribution-graphs/biology.html#dives-cot<br/><br/>https://aistudio.google.com/prompts/new_chat<br/><br/>Search Study: https://www.cjr.org/tow_center/we-compared-eight-ai-search-engines-theyre-all-bad-at-citing-news.php<br/><br/>Live bench: https://livebench.ai/#/<br/>Paper: https://arxiv.org/pdf/2406.19314<br/><br/>LiveCode Bench: https://livecodebench.github.io/<br/><br/>SWE-Verified: https://arxiv.org/pdf/2310.06770<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16877943-gemini-2-5-pro-it-s-a-smart-chatbot-new-simple-high-score.mp3" length="15395235" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/lvft8l4h5l86utqdn8vsegauv1x4?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16877943</guid>
    <pubDate>Fri, 28 Mar 2025 20:00:00 +0000</pubDate>
    <itunes:duration>1281</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>11</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Did AI Just Get Commoditized? Gemini 2.5, New DeepSeek V3, &amp; Microsoft vs OpenAI</itunes:title>
    <title>Did AI Just Get Commoditized? Gemini 2.5, New DeepSeek V3, &amp; Microsoft vs OpenAI</title>
    <itunes:summary><![CDATA[Gemini 2.5 is out, on the same day as the new DeepSeek V3 (which should power Deepseek R2). Do both models prove AI is being commoditized? Let’s find out, on this blockbuster day of AI releases. Plus exclusives from the Information, Simple indications, Vista Bench, LM Arena and more…  AI Insiders ($9!): https://www.patreon.com/AIExplained  Chapters:  00:00 - Introduction 01:15 - Gemini 2.5 Benchmarks 05:46 - Long Context, Simple indication 07:08 - New Deepseek V3 -024 09:11 - Microsoft M...]]></itunes:summary>
    <description><![CDATA[<p>Gemini 2.5 is out, on the same day as the new DeepSeek V3 (which should power Deepseek R2). Do both models prove AI is being commoditized? Let’s find out, on this blockbuster day of AI releases. Plus exclusives from the Information, Simple indications, Vista Bench, LM Arena and more…<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters: <br/>00:00 - Introduction<br/>01:15 - Gemini 2.5 Benchmarks<br/>05:46 - Long Context, Simple indication<br/>07:08 - New Deepseek V3 -024<br/>09:11 - Microsoft MAI<br/>11:48 - 90% of code but new Claude jobs<br/><br/>‘World’s most powerful model’: https://x.com/OfficialLoganK/status/1904580368432586975<br/><br/>Gemini 2.5 Release Notes: https://blog.google/technology/google-deepmind/gemini-model-thinking-updates-march-2025/#gemini-2-5-thinking<br/><br/>‘Commoditized’: https://the-decoder.com/microsoft-ceo-satya-nadella-says-ai-models-are-getting-commoditized/<br/><br/>Microsoft Information report: https://www.theinformation.com/articles/microsofts-ai-guru-wants-independence-from-openai-thats-easier-said-than-done?rc=sy0ihq<br/><br/>LMarena: https://x.com/lmarena_ai/status/1904581128746656099/photo/1<br/><br/>Free for now: https://x.com/btibor91/status/1904578053537476628<br/><br/>Vista Bench:https://scale.com/leaderboard/visual_language_understanding<br/><br/>DeepSeek V3: https://huggingface.co/deepseek-ai/DeepSeek-V3-0324<br/><br/>Claude Plays Pokemon: https://www.twitch.tv/claudeplayspokemon<br/>Amodei: 100% Coding: https://www.youtube.com/watch?v=esCSpbDPJik&amp;t=3017s<br/><br/>Anthropic Jobs: https://job-boards.greenhouse.io/anthropic/jobs/4020717008<br/><br/>Microsoft Money from Onslaught: https://www.972mag.com/microsoft-azure-openai-israeli-army-cloud/<br/><br/>https://simple-bench.com/<br/><br/>Release Date Comments: https://x.com/zacharynado/status/1904647277861318979<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Gemini 2.5 is out, on the same day as the new DeepSeek V3 (which should power Deepseek R2). Do both models prove AI is being commoditized? Let’s find out, on this blockbuster day of AI releases. Plus exclusives from the Information, Simple indications, Vista Bench, LM Arena and more…<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters: <br/>00:00 - Introduction<br/>01:15 - Gemini 2.5 Benchmarks<br/>05:46 - Long Context, Simple indication<br/>07:08 - New Deepseek V3 -024<br/>09:11 - Microsoft MAI<br/>11:48 - 90% of code but new Claude jobs<br/><br/>‘World’s most powerful model’: https://x.com/OfficialLoganK/status/1904580368432586975<br/><br/>Gemini 2.5 Release Notes: https://blog.google/technology/google-deepmind/gemini-model-thinking-updates-march-2025/#gemini-2-5-thinking<br/><br/>‘Commoditized’: https://the-decoder.com/microsoft-ceo-satya-nadella-says-ai-models-are-getting-commoditized/<br/><br/>Microsoft Information report: https://www.theinformation.com/articles/microsofts-ai-guru-wants-independence-from-openai-thats-easier-said-than-done?rc=sy0ihq<br/><br/>LMarena: https://x.com/lmarena_ai/status/1904581128746656099/photo/1<br/><br/>Free for now: https://x.com/btibor91/status/1904578053537476628<br/><br/>Vista Bench:https://scale.com/leaderboard/visual_language_understanding<br/><br/>DeepSeek V3: https://huggingface.co/deepseek-ai/DeepSeek-V3-0324<br/><br/>Claude Plays Pokemon: https://www.twitch.tv/claudeplayspokemon<br/>Amodei: 100% Coding: https://www.youtube.com/watch?v=esCSpbDPJik&amp;t=3017s<br/><br/>Anthropic Jobs: https://job-boards.greenhouse.io/anthropic/jobs/4020717008<br/><br/>Microsoft Money from Onslaught: https://www.972mag.com/microsoft-azure-openai-israeli-army-cloud/<br/><br/>https://simple-bench.com/<br/><br/>Release Date Comments: https://x.com/zacharynado/status/1904647277861318979<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16860611-did-ai-just-get-commoditized-gemini-2-5-new-deepseek-v3-microsoft-vs-openai.mp3" length="9966485" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/20w1dd6agjadhf06s3sor8xx6sxo?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16860611</guid>
    <pubDate>Tue, 25 Mar 2025 23:00:00 +0000</pubDate>
    <itunes:duration>827</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>10</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Manus AI - The Calm Before the Hypestorm … (vs Deep Research + Grok 3)</itunes:title>
    <title>Manus AI - The Calm Before the Hypestorm … (vs Deep Research + Grok 3)</title>
    <itunes:summary><![CDATA[Is Manus AI the memecoin of the AI world, or legit? I’ll compare it to OpenAI’s Deep Research, Operator, Grok 3 DeepSearch and more to find out. I’ll also let you in on some of the secrets of what makes a good hype campaign, the estimated costs of Manus AI, and where it is strong. Other news (yes, Gemini image editing and research hacking, I mean you), will have to wait for a few more hours, as millions enquire about Manus AI.  https://app.grayswan.ai/arena  AI Insiders ($9!): https://www.pat...]]></itunes:summary>
    <description><![CDATA[<p>Is Manus AI the memecoin of the AI world, or legit? I’ll compare it to OpenAI’s Deep Research, Operator, Grok 3 DeepSearch and more to find out. I’ll also let you in on some of the secrets of what makes a good hype campaign, the estimated costs of Manus AI, and where it is strong. Other news (yes, Gemini image editing and research hacking, I mean you), will have to wait for a few more hours, as millions enquire about Manus AI.<br/><br/>https://app.grayswan.ai/arena<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/>Patreon Vid: https://www.patreon.com/posts/4-ai-trends-in-123857767<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:46 - Hype Campaign<br/>02:40 - Single, Public Benchmark <br/>03:12 - What is Manus AI?<br/>04:22 - Test 1<br/>05:12 - Cost and Rate Limits<br/>06:15 - Test 2 vs Deep Research + Grok 3 DeepSearch<br/>08:24 - Test 3 (not AGI)<br/>11:10 - 4 Trends in AI in 2025<br/>11:37 - Hype Works<br/><br/>Manus AI: https://manus.im/app<br/><br/>Xiao Hong Interview: https://www.chinatalk.media/p/manus-chinas-latest-ai-sensation<br/><br/>Gaia Benchmark: https://openreview.net/pdf?id=fibxvahvs3<br/>MIT Report: https://www.technologyreview.com/2025/03/11/1113133/manus-ai-review/<br/><br/>Information Report: https://www.theinformation.com/articles/anthropics-claude-drives-strong-revenue-growth-while-powering-manus-sensation?rc=sy0ihq<br/><br/>Hype Examples: https://x.com/Saboo_Shubham_/status/1898425707401031940<br/>https://x.com/EHuanglu/status/1899110687902978373<br/>https://x.com/AJs_AI/status/1898756132384178291<br/><br/>Mistakes: https://x.com/TheXeophon/status/1898737178273829220<br/><br/>Tools and Code: https://x.com/peakji/status/1898994802194346408<br/><br/>https://operator.chatgpt.com/<br/><br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></description>
    <content:encoded><![CDATA[<p>Is Manus AI the memecoin of the AI world, or legit? I’ll compare it to OpenAI’s Deep Research, Operator, Grok 3 DeepSearch and more to find out. I’ll also let you in on some of the secrets of what makes a good hype campaign, the estimated costs of Manus AI, and where it is strong. Other news (yes, Gemini image editing and research hacking, I mean you), will have to wait for a few more hours, as millions enquire about Manus AI.<br/><br/>https://app.grayswan.ai/arena<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/>Patreon Vid: https://www.patreon.com/posts/4-ai-trends-in-123857767<br/><br/>Chapters:<br/>00:00 - Introduction<br/>00:46 - Hype Campaign<br/>02:40 - Single, Public Benchmark <br/>03:12 - What is Manus AI?<br/>04:22 - Test 1<br/>05:12 - Cost and Rate Limits<br/>06:15 - Test 2 vs Deep Research + Grok 3 DeepSearch<br/>08:24 - Test 3 (not AGI)<br/>11:10 - 4 Trends in AI in 2025<br/>11:37 - Hype Works<br/><br/>Manus AI: https://manus.im/app<br/><br/>Xiao Hong Interview: https://www.chinatalk.media/p/manus-chinas-latest-ai-sensation<br/><br/>Gaia Benchmark: https://openreview.net/pdf?id=fibxvahvs3<br/>MIT Report: https://www.technologyreview.com/2025/03/11/1113133/manus-ai-review/<br/><br/>Information Report: https://www.theinformation.com/articles/anthropics-claude-drives-strong-revenue-growth-while-powering-manus-sensation?rc=sy0ihq<br/><br/>Hype Examples: https://x.com/Saboo_Shubham_/status/1898425707401031940<br/>https://x.com/EHuanglu/status/1899110687902978373<br/>https://x.com/AJs_AI/status/1898756132384178291<br/><br/>Mistakes: https://x.com/TheXeophon/status/1898737178273829220<br/><br/>Tools and Code: https://x.com/peakji/status/1898994802194346408<br/><br/>https://operator.chatgpt.com/<br/><br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/<br/><br/>Podcast: https://aiexplainedopodcast.buzzsprout.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16785914-manus-ai-the-calm-before-the-hypestorm-vs-deep-research-grok-3.mp3" length="9372197" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/u6kw9ji1ybryaqh6nkh3ruhn5c64?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16785914</guid>
    <pubDate>Thu, 13 Mar 2025 15:00:00 +0000</pubDate>
    <itunes:duration>778</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>9</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>GPT 4.5 - not so much wow</itunes:title>
    <title>GPT 4.5 - not so much wow</title>
    <itunes:summary><![CDATA[GPT 4.5 is here, and do you remember when AI lab CEOs like Sam Altman and Dario Amodei were betting everything on scaling up base models like this one? Well let’s find out what would have happened if the future of AI rested on models like GPT 4.5. You’ll see all the benchmarks, highlights of the paper, emotional intelligence and humor tests, Simple Bench results (reddit was an unreliable source), and why it’s not all bad news for OpenAI.  https://www.emergentmind.com/  AI Insiders (now $9!): ...]]></itunes:summary>
    <description><![CDATA[<p>GPT 4.5 is here, and do you remember when AI lab CEOs like Sam Altman and Dario Amodei were betting everything on scaling up base models like this one? Well let’s find out what would have happened if the future of AI rested on models like GPT 4.5. You’ll see all the benchmarks, highlights of the paper, emotional intelligence and humor tests, Simple Bench results (reddit was an unreliable source), and why it’s not all bad news for OpenAI.<br/><br/>https://www.emergentmind.com/<br/><br/>AI Insiders (now $9!): https://www.patreon.com/AIExplained<br/><br/>Chapters<br/>00:00 - Introduction<br/>01:04 - Details and Benchmarks<br/>03:04 - Emotional intelligence? <br/>08:37 - Creative writing?<br/>11:40 - Visual reasoning and Pricing<br/>12:41 - Simple Performance<br/>16:01 - End of Pretraining Scaling?<br/>17:03 - CEO Hype<br/>18:11 - System Card Highlights<br/>23:32 - Karpathy Reaction<br/><br/>GPT 4.5 System card: https://cdn.openai.com/gpt-4-5-system-card-2272025.pdf<br/>Release Notes: https://openai.com/index/gpt-4-5-system-card/<br/>Altman Hype: https://x.com/sama/status/1891533802779910471<br/>Details: https://openai.com/index/introducing-gpt-4-5/ https://x.com/OpenAI/status/1895219596317335792<br/>End of an Era: https://x.com/wgussml/status/1895187231666774377<br/>Anthropic Original Claim: https://techcrunch.com/2023/04/06/anthropics-5b-4-year-plan-to-take-on-openai/<br/>Smell: https://x.com/rapha_gl/status/1895213014699385082<br/>Bob McGrew: https://x.com/bobmcgrewai/status/1895228291981943265<br/>Deep Research System Card: https://cdn.openai.com/deep-research-system-card.pdf<br/>Reddit: https://www.reddit.com/r/singularity/comments/1izu1t7/gpt45_crushes_simple_bench/<br/>API Pricing: https://openai.com/api/pricing/<br/>LiveStream: https://www.youtube.com/watch?v=cfRYp0nItZ8&amp;t=1s<br/>https://simple-bench.com/<br/><br/><br/>Karpathy Comparison: https://x.com/karpathy/status/1895213020982472863<br/>https://x.com/karpathy/status/1895337579589079434<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>GPT 4.5 is here, and do you remember when AI lab CEOs like Sam Altman and Dario Amodei were betting everything on scaling up base models like this one? Well let’s find out what would have happened if the future of AI rested on models like GPT 4.5. You’ll see all the benchmarks, highlights of the paper, emotional intelligence and humor tests, Simple Bench results (reddit was an unreliable source), and why it’s not all bad news for OpenAI.<br/><br/>https://www.emergentmind.com/<br/><br/>AI Insiders (now $9!): https://www.patreon.com/AIExplained<br/><br/>Chapters<br/>00:00 - Introduction<br/>01:04 - Details and Benchmarks<br/>03:04 - Emotional intelligence? <br/>08:37 - Creative writing?<br/>11:40 - Visual reasoning and Pricing<br/>12:41 - Simple Performance<br/>16:01 - End of Pretraining Scaling?<br/>17:03 - CEO Hype<br/>18:11 - System Card Highlights<br/>23:32 - Karpathy Reaction<br/><br/>GPT 4.5 System card: https://cdn.openai.com/gpt-4-5-system-card-2272025.pdf<br/>Release Notes: https://openai.com/index/gpt-4-5-system-card/<br/>Altman Hype: https://x.com/sama/status/1891533802779910471<br/>Details: https://openai.com/index/introducing-gpt-4-5/ https://x.com/OpenAI/status/1895219596317335792<br/>End of an Era: https://x.com/wgussml/status/1895187231666774377<br/>Anthropic Original Claim: https://techcrunch.com/2023/04/06/anthropics-5b-4-year-plan-to-take-on-openai/<br/>Smell: https://x.com/rapha_gl/status/1895213014699385082<br/>Bob McGrew: https://x.com/bobmcgrewai/status/1895228291981943265<br/>Deep Research System Card: https://cdn.openai.com/deep-research-system-card.pdf<br/>Reddit: https://www.reddit.com/r/singularity/comments/1izu1t7/gpt45_crushes_simple_bench/<br/>API Pricing: https://openai.com/api/pricing/<br/>LiveStream: https://www.youtube.com/watch?v=cfRYp0nItZ8&amp;t=1s<br/>https://simple-bench.com/<br/><br/><br/>Karpathy Comparison: https://x.com/karpathy/status/1895213020982472863<br/>https://x.com/karpathy/status/1895337579589079434<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16710421-gpt-4-5-not-so-much-wow.mp3" length="18095595" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/xw9p24hgjlh0rsmlhar214sfjnkp?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16710421</guid>
    <pubDate>Fri, 28 Feb 2025 16:00:00 +0000</pubDate>
    <itunes:duration>1505</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>8</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Claude 3.7 is More Significant than its Name Implies (ft DeepSeek R2 + GPT 4.5 coming soon)</itunes:title>
    <title>Claude 3.7 is More Significant than its Name Implies (ft DeepSeek R2 + GPT 4.5 coming soon)</title>
    <itunes:summary><![CDATA[Claude 3.7 is here, hot on the heels of Grok 3 and a host of other developments, but how good is it really? And what does it say about the next few months in AI? I’ve read the papers, played with the model for hours, and benched it on Simple. Things aren’t slowing down. Plus the latest in humanoid robots, led by Helix and freaked out by Protoclone. And reports of GPT 4.5 and DeepSeek R2.   GraySwan Competition! https://app.grayswan.ai/arena/challenge/agent-red-teaming https://x.com/GraySwanAI...]]></itunes:summary>
    <description><![CDATA[<p><b>Claude 3.7 is here, hot on the heels of Grok 3 and a host of other developments, but how good is it really? And what does it say about the next few months in AI? I’ve read the papers, played with the model for hours, and benched it on Simple. Things aren’t slowing down. Plus the latest in humanoid robots, led by Helix and freaked out by Protoclone. And reports of GPT 4.5 and DeepSeek R2.</b></p><p><br/></p><p><b>GraySwan Competition! </b><a href='https://app.grayswan.ai/arena/challenge/agent-red-teaming'><b>https://app.grayswan.ai/arena/challenge/agent-red-teaming</b></a></p><p><b>https://x.com/GraySwanAI/status/1894084923260043282</b></p><p><br/></p><p><b>Chapters:</b></p><p><b>00:00 - Introduction</b></p><p><b>01:25 - Claude 3.7 New Stats/Demos </b></p><p><b>05:22 - 128k Output</b></p><p><b>06:13 - Pokemon</b></p><p><b>06:58 - Just a tool? </b></p><p><b>09:54 - DeepSeek R2</b></p><p><b>10:20 - Claude 3.7 System Card/Paper Highlights </b></p><p><b>17:18 - Simple Record Score/Competition</b></p><p><b>20:37 - Grok 3 + Redteaming prizes</b></p><p><b>22:26 - Google Co-scientist</b></p><p><b>24:02 - Humanoid Robot Developments</b></p><p><br/></p><p><b>3.7 Release Notes: https://www.anthropic.com/news/claude-3-7-sonnet</b></p><p><b>vs o3 and Grok 3: https://x.com/12exyz/status/1891723056931827959</b></p><p><b>Extended Thinking: </b><a href='https://www.anthropic.com/research/visible-extended-thinking?s=09'><b>https://www.anthropic.com/research/visible-extended-thinking?s=09</b></a></p><p><b>System Prompt: https://docs.anthropic.com/en/release-notes/system-prompts#feb-24th-2025</b></p><p><b>System Card: </b><a href='https://assets.anthropic.com/m/785e231869ea8b3b/original/claude-3-7-sonnet-system-card.pdf'><b>https://assets.anthropic.com/m/785e231869ea8b3b/original/claude-3-7-sonnet-system-card.pdf</b></a></p><p><b>Unfaithful CoT: </b><a href='https://arxiv.org/pdf/2305.04388'><b>https://arxiv.org/pdf/2305.04388</b></a></p><p><b>Original Constitution: </b><a href='https://www.anthropic.com/news/claudes-constitution'><b>https://www.anthropic.com/news/claudes-constitution</b></a></p><p><b>Responsible Scaling Policy: </b><a href='https://assets.anthropic.com/m/24a47b00f10301cd/original/Anthropic-Responsible-Scaling-Policy-2024-10-15.pdf'><b>https://assets.anthropic.com/m/24a47b00f10301cd/original/Anthropic-Responsible-Scaling-Policy-2024-10-15.pdf</b></a></p><p><b>Amodei and Hassabis:https://www.youtube.com/watch?v=4poqjZlM8Lo</b></p><p><b>https://simple-bench.com/</b></p><p><b>400 Weekly Users: </b><a href='https://x.com/bradlightcap/status/1892579908179882057'><b>https://x.com/bradlightcap/status/1892579908179882057</b></a></p><p><b>Grok 3 Jailbroken: https://x.com/LinusEkenstam/status/1893832876581380280</b></p><p><b>Google Co-Scientist: </b><a href='https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/'><b>https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/</b></a></p><p><b>But Hassabis Says Years Away: https://www.youtube.com/watch?v=yr0GiSgUvPU&amp;t=156s</b></p><p><b>DeepSeek R2 Reuters: </b><a href='https://www.reuters.com/technology/artificial-intelligence/deepseek-rushes-launch-new-ai-model-china-goes-all-2025-02-25/'><b>https://www.reuters.com/technology/artificial-intelligence/deepseek-rushes-launch-new-ai-model-china-goes-all-2025-02-25/</b></a></p><p><b>Protoclone: </b><a href='https://www.reddit.com/r/interestingasfuck/comments/1it9rpp/protoclone_the_worlds_first_bipedal/'><b>https://www.reddit.com/r/interestingasfuck/comments/1it9rpp/protoclone_the_worlds_first_bipedal/</b></a></p><p><b>Helix: </b><a href='https://www.figure.ai/news/helix'><b>https://www.figure.ai/news/helix</b></a></p><p><b>TechTrance: </b><a href='https://www.youtube.com/@TheTechTrance/videos'><b>https://www.youtube.com/@TheTechTrance/videos</b></a></p><p><b>GPT 4.5 Soon: </b><a href='https://www.theverge.com/notepad-microsoft-newsletter/616464/microsoft-pr&lt;/truncato-artificial-root&gt;'></p>]]></description>
    <content:encoded><![CDATA[<p><b>Claude 3.7 is here, hot on the heels of Grok 3 and a host of other developments, but how good is it really? And what does it say about the next few months in AI? I’ve read the papers, played with the model for hours, and benched it on Simple. Things aren’t slowing down. Plus the latest in humanoid robots, led by Helix and freaked out by Protoclone. And reports of GPT 4.5 and DeepSeek R2.</b></p><p><br/></p><p><b>GraySwan Competition! </b><a href='https://app.grayswan.ai/arena/challenge/agent-red-teaming'><b>https://app.grayswan.ai/arena/challenge/agent-red-teaming</b></a></p><p><b>https://x.com/GraySwanAI/status/1894084923260043282</b></p><p><br/></p><p><b>Chapters:</b></p><p><b>00:00 - Introduction</b></p><p><b>01:25 - Claude 3.7 New Stats/Demos </b></p><p><b>05:22 - 128k Output</b></p><p><b>06:13 - Pokemon</b></p><p><b>06:58 - Just a tool? </b></p><p><b>09:54 - DeepSeek R2</b></p><p><b>10:20 - Claude 3.7 System Card/Paper Highlights </b></p><p><b>17:18 - Simple Record Score/Competition</b></p><p><b>20:37 - Grok 3 + Redteaming prizes</b></p><p><b>22:26 - Google Co-scientist</b></p><p><b>24:02 - Humanoid Robot Developments</b></p><p><br/></p><p><b>3.7 Release Notes: https://www.anthropic.com/news/claude-3-7-sonnet</b></p><p><b>vs o3 and Grok 3: https://x.com/12exyz/status/1891723056931827959</b></p><p><b>Extended Thinking: </b><a href='https://www.anthropic.com/research/visible-extended-thinking?s=09'><b>https://www.anthropic.com/research/visible-extended-thinking?s=09</b></a></p><p><b>System Prompt: https://docs.anthropic.com/en/release-notes/system-prompts#feb-24th-2025</b></p><p><b>System Card: </b><a href='https://assets.anthropic.com/m/785e231869ea8b3b/original/claude-3-7-sonnet-system-card.pdf'><b>https://assets.anthropic.com/m/785e231869ea8b3b/original/claude-3-7-sonnet-system-card.pdf</b></a></p><p><b>Unfaithful CoT: </b><a href='https://arxiv.org/pdf/2305.04388'><b>https://arxiv.org/pdf/2305.04388</b></a></p><p><b>Original Constitution: </b><a href='https://www.anthropic.com/news/claudes-constitution'><b>https://www.anthropic.com/news/claudes-constitution</b></a></p><p><b>Responsible Scaling Policy: </b><a href='https://assets.anthropic.com/m/24a47b00f10301cd/original/Anthropic-Responsible-Scaling-Policy-2024-10-15.pdf'><b>https://assets.anthropic.com/m/24a47b00f10301cd/original/Anthropic-Responsible-Scaling-Policy-2024-10-15.pdf</b></a></p><p><b>Amodei and Hassabis:https://www.youtube.com/watch?v=4poqjZlM8Lo</b></p><p><b>https://simple-bench.com/</b></p><p><b>400 Weekly Users: </b><a href='https://x.com/bradlightcap/status/1892579908179882057'><b>https://x.com/bradlightcap/status/1892579908179882057</b></a></p><p><b>Grok 3 Jailbroken: https://x.com/LinusEkenstam/status/1893832876581380280</b></p><p><b>Google Co-Scientist: </b><a href='https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/'><b>https://research.google/blog/accelerating-scientific-breakthroughs-with-an-ai-co-scientist/</b></a></p><p><b>But Hassabis Says Years Away: https://www.youtube.com/watch?v=yr0GiSgUvPU&amp;t=156s</b></p><p><b>DeepSeek R2 Reuters: </b><a href='https://www.reuters.com/technology/artificial-intelligence/deepseek-rushes-launch-new-ai-model-china-goes-all-2025-02-25/'><b>https://www.reuters.com/technology/artificial-intelligence/deepseek-rushes-launch-new-ai-model-china-goes-all-2025-02-25/</b></a></p><p><b>Protoclone: </b><a href='https://www.reddit.com/r/interestingasfuck/comments/1it9rpp/protoclone_the_worlds_first_bipedal/'><b>https://www.reddit.com/r/interestingasfuck/comments/1it9rpp/protoclone_the_worlds_first_bipedal/</b></a></p><p><b>Helix: </b><a href='https://www.figure.ai/news/helix'><b>https://www.figure.ai/news/helix</b></a></p><p><b>TechTrance: </b><a href='https://www.youtube.com/@TheTechTrance/videos'><b>https://www.youtube.com/@TheTechTrance/videos</b></a></p><p><b>GPT 4.5 Soon: </b><a href='https://www.theverge.com/notepad-microsoft-newsletter/616464/microsoft-pr&lt;/truncato-artificial-root&gt;'></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16691238-claude-3-7-is-more-significant-than-its-name-implies-ft-deepseek-r2-gpt-4-5-coming-soon.mp3" length="19941052" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/hm5nfd30luxjuc068b6qjns5ivur?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16691238</guid>
    <pubDate>Tue, 25 Feb 2025 17:00:00 +0000</pubDate>
    <itunes:duration>1659</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>7</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>AGI: (gets close), Humans: ‘Who Gets the Money?’</itunes:title>
    <title>AGI: (gets close), Humans: ‘Who Gets the Money?’</title>
    <itunes:summary><![CDATA[A 'frontier reasoning model' from just 1000 examples (s1). A $100B Musk bid for power. Gemini 2, Rand and warning from Amodei. Here’s 7-8 developments you may have missed but which I would argue help us understand how the next few years will play out. From labour vs capital to automating rival companies and countries, and from non-profit shenanigans to new mini-docs, there was just too much for me not to make a vid.  GiveWell: https://www.givewell.org/charities/top-charities  AI Insiders ($9!...]]></itunes:summary>
    <description><![CDATA[<p>A &apos;frontier reasoning model&apos; from just 1000 examples (s1). A $100B Musk bid for power. Gemini 2, Rand and warning from Amodei. Here’s 7-8 developments you may have missed but which I would argue help us understand how the next few years will play out. From labour vs capital to automating rival companies and countries, and from non-profit shenanigans to new mini-docs, there was just too much for me not to make a vid.<br/><br/>GiveWell: https://www.givewell.org/charities/top-charities<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>s1 Paper: https://arxiv.org/pdf/2501.19393<br/>Musk Bid: https://www.wsj.com/tech/ai/musks-97-4-billion-openai-bid-piles-pressure-on-altman-f6749e6c?mod=hp_lead_pos1<br/>Altman Reply: https://x.com/sama/status/1889059531625464090?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet<br/>Google vs OpenAI: https://x.com/sama/status/1888703820596977684<br/>RAND Study: https://www.rand.org/pubs/perspectives/PEA3691-4.html<br/>Dev Meetup: https://x.com/btibor91/status/1888976302621040852<br/>Altman $100 Trillion: https://www.nytimes.com/2023/03/31/technology/sam-altman-open-ai-chatgpt.html<br/>Karpathy Vid: https://www.youtube.com/watch?v=7xTGNNLPyMI<br/>Amodei Warning: https://www.anthropic.com/news/paris-ai-summit<br/>Bengio Source: https://www.youtube.com/watch?v=6HDjVncL5Go<br/><br/>Chapters:<br/>00:00 - Intro<br/>01:37 -  AGI Inches Closer<br/>04:26 - ‘Super-Exponential’<br/>05:58 - Musk Bid<br/>07:34 - Luxury Goods and Land<br/>09:05 - ‘Benefits All Humanity’<br/>12:52 - ‘National Security’<br/>14:21 - s1<br/>20:33 - Final thoughts<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>A &apos;frontier reasoning model&apos; from just 1000 examples (s1). A $100B Musk bid for power. Gemini 2, Rand and warning from Amodei. Here’s 7-8 developments you may have missed but which I would argue help us understand how the next few years will play out. From labour vs capital to automating rival companies and countries, and from non-profit shenanigans to new mini-docs, there was just too much for me not to make a vid.<br/><br/>GiveWell: https://www.givewell.org/charities/top-charities<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>s1 Paper: https://arxiv.org/pdf/2501.19393<br/>Musk Bid: https://www.wsj.com/tech/ai/musks-97-4-billion-openai-bid-piles-pressure-on-altman-f6749e6c?mod=hp_lead_pos1<br/>Altman Reply: https://x.com/sama/status/1889059531625464090?ref_src=twsrc%5Egoogle%7Ctwcamp%5Eserp%7Ctwgr%5Etweet<br/>Google vs OpenAI: https://x.com/sama/status/1888703820596977684<br/>RAND Study: https://www.rand.org/pubs/perspectives/PEA3691-4.html<br/>Dev Meetup: https://x.com/btibor91/status/1888976302621040852<br/>Altman $100 Trillion: https://www.nytimes.com/2023/03/31/technology/sam-altman-open-ai-chatgpt.html<br/>Karpathy Vid: https://www.youtube.com/watch?v=7xTGNNLPyMI<br/>Amodei Warning: https://www.anthropic.com/news/paris-ai-summit<br/>Bengio Source: https://www.youtube.com/watch?v=6HDjVncL5Go<br/><br/>Chapters:<br/>00:00 - Intro<br/>01:37 -  AGI Inches Closer<br/>04:26 - ‘Super-Exponential’<br/>05:58 - Musk Bid<br/>07:34 - Luxury Goods and Land<br/>09:05 - ‘Benefits All Humanity’<br/>12:52 - ‘National Security’<br/>14:21 - s1<br/>20:33 - Final thoughts<br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16605249-agi-gets-close-humans-who-gets-the-money.mp3" length="16086240" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/e0t6adi1dsf1bjwgj35jslr6jldj?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16605249</guid>
    <pubDate>Tue, 11 Feb 2025 21:00:00 +0000</pubDate>
    <itunes:duration>1337</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>6</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Deep Research by OpenAI - The Ups and Downs vs DeepSeek R1 Search + Gemini Deep Research</itunes:title>
    <title>Deep Research by OpenAI - The Ups and Downs vs DeepSeek R1 Search + Gemini Deep Research</title>
    <itunes:summary><![CDATA[12 hours ago Deep Research was unveiled, and I’ve tested it thoroughly, including vs Deepseek R1 with search, Gemini Deep Research and even R1 in Perplexity. It’s a notable step forward, with one big caveat. I’ll go through all the benchmark figures, my initial impression of the o3 model within, and much more.  Deep Research: https://openai.com/index/introducing-deep-research/ https://www.youtube.com/watch?v=YkCDVn3_wiw   GAIA Bench: https://openreview.net/forum?id=fibxvahvs3 https://openrevi...]]></itunes:summary>
    <description><![CDATA[<p><b>12 hours ago Deep Research was unveiled, and I’ve tested it thoroughly, including vs Deepseek R1 with search, Gemini Deep Research and even R1 in Perplexity. It’s a notable step forward, with one big caveat. I’ll go through all the benchmark figures, my initial impression of the o3 model within, and much more.<br/><br/>Deep Research: </b><a href='https://openai.com/index/introducing-deep-research/'><b>https://openai.com/index/introducing-deep-research/</b></a></p><p><b>https://www.youtube.com/watch?v=YkCDVn3_wiw</b></p><p><br/></p><p><b>GAIA Bench: </b><a href='https://openreview.net/forum?id=fibxvahvs3'><b>https://openreview.net/forum?id=fibxvahvs3</b></a></p><p><a href='https://openreview.net/pdf?id=fibxvahvs3'><b>https://openreview.net/pdf?id=fibxvahvs3</b></a></p><p><b>CodeELO:</b><a href='https://arxiv.org/pdf/2501.01257'><b>https://arxiv.org/pdf/2501.01257</b></a></p><p><b>CamelCamel:</b><a href='https://uk.camelcamelcamel.com/'><b>https://uk.camelcamelcamel.com/</b></a></p><p><b>Deepseek R1 with search: https://chat.deepseek.com/</b></p><p><a href='https://arxiv.org/pdf/2501.12948'><b>https://arxiv.org/pdf/2501.12948</b></a></p><p><b>HaluBench: </b><a href='https://arxiv.org/pdf/2407.08488'><b>https://arxiv.org/pdf/2407.08488</b></a></p><p><br/></p><p><b>Chapters:</b></p><p><b>00:00 - Introduction</b></p><p><b>01:06 - Powered by o3, Humanity’s Last Exam, GAIA</b></p><p><b>03:55 - Simple Tests </b></p><p><b>06:00 - Good News vs Deepseek R1 and Gemini Deep Research</b></p><p><b>09:32 - Bad News on Hallucinations </b></p><p><b>14:14 - What Can’t it Browse?</b></p><p><b>14:42 - For Shopping?</b></p><p><b>16:40 - Final thoughts</b></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>12 hours ago Deep Research was unveiled, and I’ve tested it thoroughly, including vs Deepseek R1 with search, Gemini Deep Research and even R1 in Perplexity. It’s a notable step forward, with one big caveat. I’ll go through all the benchmark figures, my initial impression of the o3 model within, and much more.<br/><br/>Deep Research: </b><a href='https://openai.com/index/introducing-deep-research/'><b>https://openai.com/index/introducing-deep-research/</b></a></p><p><b>https://www.youtube.com/watch?v=YkCDVn3_wiw</b></p><p><br/></p><p><b>GAIA Bench: </b><a href='https://openreview.net/forum?id=fibxvahvs3'><b>https://openreview.net/forum?id=fibxvahvs3</b></a></p><p><a href='https://openreview.net/pdf?id=fibxvahvs3'><b>https://openreview.net/pdf?id=fibxvahvs3</b></a></p><p><b>CodeELO:</b><a href='https://arxiv.org/pdf/2501.01257'><b>https://arxiv.org/pdf/2501.01257</b></a></p><p><b>CamelCamel:</b><a href='https://uk.camelcamelcamel.com/'><b>https://uk.camelcamelcamel.com/</b></a></p><p><b>Deepseek R1 with search: https://chat.deepseek.com/</b></p><p><a href='https://arxiv.org/pdf/2501.12948'><b>https://arxiv.org/pdf/2501.12948</b></a></p><p><b>HaluBench: </b><a href='https://arxiv.org/pdf/2407.08488'><b>https://arxiv.org/pdf/2407.08488</b></a></p><p><br/></p><p><b>Chapters:</b></p><p><b>00:00 - Introduction</b></p><p><b>01:06 - Powered by o3, Humanity’s Last Exam, GAIA</b></p><p><b>03:55 - Simple Tests </b></p><p><b>06:00 - Good News vs Deepseek R1 and Gemini Deep Research</b></p><p><b>09:32 - Bad News on Hallucinations </b></p><p><b>14:14 - What Can’t it Browse?</b></p><p><b>14:42 - For Shopping?</b></p><p><b>16:40 - Final thoughts</b></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16554084-deep-research-by-openai-the-ups-and-downs-vs-deepseek-r1-search-gemini-deep-research.mp3" length="13383623" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ed9dq0wjnmldrikcna9hhitb6auu?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16554084</guid>
    <pubDate>Mon, 03 Feb 2025 16:00:00 +0000</pubDate>
    <itunes:duration>1112</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>5</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>o3-mini and the “AI War”</itunes:title>
    <title>o3-mini and the “AI War”</title>
    <itunes:summary><![CDATA[o3-mini is here, and yes, I’ve read the paper in full - 2 hours after release, and even the post-launch Reddit AMA. Some epic details like a FrontierMath score that made me double-take, a likely new Cursor favorite, bio risk expertise and a cost-comparison with Deepseek R1., But does it perform on basic reasoning - let’s find out. Plus, arguably the bigger story - the increasingly frenetic rhetoric coming out of the West - and Dario Amodei and Alexandr Wang (CEOs of Anthropic and Scale AI res...]]></itunes:summary>
    <description><![CDATA[<p>o<b>3-mini is here, and yes, I’ve read the paper in full - 2 hours after release, and even the post-launch Reddit AMA. Some epic details like a FrontierMath score that made me double-take, a likely new Cursor favorite, bio risk expertise and a cost-comparison with Deepseek R1., But does it perform on basic reasoning - let’s find out. Plus, arguably the bigger story - the increasingly frenetic rhetoric coming out of the West - and Dario Amodei and Alexandr Wang (CEOs of Anthropic and Scale AI respectively) in particular. The last thing we need is an “AI War”.</b></p><p><br/></p><p><a href='https://wandb.me/simple-bench'><b>https://wandb.me/simple-bench</b></a></p><p><br/></p><p><b>(Colab): </b><a href='https://colab.research.google.com/drive/1AVijcPnEkl8Gy_754XbRdG5m7Q5-9slg?usp=sharing'><b>https://colab.research.google.com/drive/1AVijcPnEkl8Gy_754XbRdG5m7Q5-9slg?usp=sharing<br/><br/></b></a><br/></p><p><b>Chapters: </b></p><p><b>00:00 - Introduction</b></p><p><b>00:45 - o3 mini</b></p><p><b>05:11 - First impressions vs Deepseek R1</b></p><p><b>07:21 - 10x Scale, o3-mini System Card, Amodei Essay, bitcoin wallets…</b></p><p><b>12:40 - Simple Competition Finale</b></p><p><b>13:03 - Clips and Final Thoughts on the “AI War”</b></p><p><b><br/></b><br/></p><p><b>O3-mini: </b><a href='https://openai.com/index/openai-o3-mini/'><b>https://openai.com/index/openai-o3-mini/</b></a></p><p><b>Paper: https://cdn.openai.com/o3-mini-system-card.pdf</b></p><p><b>Amodei Essay: </b><a href='https://darioamodei.com/on-deepseek-and-export-controls?s=09'><b>https://darioamodei.com/on-deepseek-and-export-controls?s=09</b></a></p><p><b>FrontierMath wild stat:https://arxiv.org/pdf/2411.04872</b></p><p><b>Sam Altman Channels Napoleon: </b><a href='https://x.com/sama/status/1883185690508488934'><b>https://x.com/sama/status/1883185690508488934</b></a></p><p><b>Altman ‘pulls up releases’: https://x.com/sama/status/1884066337103962416</b></p><p><b>“AI War” by Wang: </b><a href='https://scale.com/blog/win-the-ai-war'><b>https://scale.com/blog/win-the-ai-war</b></a></p><p><b>Anthropic Original Views on Capabilities: </b><a href='https://www.anthropic.com/news/core-views-on-ai-safety'><b>https://www.anthropic.com/news/core-views-on-ai-safety</b></a></p><p><b>AI Insider Cost Comparison:https://x.com/arankomatsuzaki/status/1884676245922934788</b></p><p><b>Deepseek R1 Paper: </b><a href='https://arxiv.org/pdf/2501.12948'><b>https://arxiv.org/pdf/2501.12948</b></a></p><p><b>R1, o3-mini Price Comparison: https://techcrunch.com/2025/01/31/openai-launches-o3-mini-its-latest-reasoning-model/</b></p><p><b>Semianalysis on $1,3M deepseek salaries, and them falling behind as ‘the time gap to match US capabilities increases’: https://semianalysis.com/2025/01/31/deepseek-debates/</b></p><p><b>OpenAI Valuation: https://www.bloomberg.com/news/articles/2025-01-30/openai-in-talks-to-raise-funding-at-340-billion-value-wsj-says?srnd=phx-ai</b></p><p><b>Wang Clip: </b><a href='https://x.com/tsarnick/status/1867700453494206883'><b>https://x.com/tsarnick/status/1867700453494206883</b></a></p><p><b>Amodei Clip: </b><a href='https://x.com/ai_ctrl/status/1884951111771001188'><b>https://x.com/ai_ctrl/status/1884951111771001188</b></a></p><p><b>https://simple-bench.com/</b></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p>o<b>3-mini is here, and yes, I’ve read the paper in full - 2 hours after release, and even the post-launch Reddit AMA. Some epic details like a FrontierMath score that made me double-take, a likely new Cursor favorite, bio risk expertise and a cost-comparison with Deepseek R1., But does it perform on basic reasoning - let’s find out. Plus, arguably the bigger story - the increasingly frenetic rhetoric coming out of the West - and Dario Amodei and Alexandr Wang (CEOs of Anthropic and Scale AI respectively) in particular. The last thing we need is an “AI War”.</b></p><p><br/></p><p><a href='https://wandb.me/simple-bench'><b>https://wandb.me/simple-bench</b></a></p><p><br/></p><p><b>(Colab): </b><a href='https://colab.research.google.com/drive/1AVijcPnEkl8Gy_754XbRdG5m7Q5-9slg?usp=sharing'><b>https://colab.research.google.com/drive/1AVijcPnEkl8Gy_754XbRdG5m7Q5-9slg?usp=sharing<br/><br/></b></a><br/></p><p><b>Chapters: </b></p><p><b>00:00 - Introduction</b></p><p><b>00:45 - o3 mini</b></p><p><b>05:11 - First impressions vs Deepseek R1</b></p><p><b>07:21 - 10x Scale, o3-mini System Card, Amodei Essay, bitcoin wallets…</b></p><p><b>12:40 - Simple Competition Finale</b></p><p><b>13:03 - Clips and Final Thoughts on the “AI War”</b></p><p><b><br/></b><br/></p><p><b>O3-mini: </b><a href='https://openai.com/index/openai-o3-mini/'><b>https://openai.com/index/openai-o3-mini/</b></a></p><p><b>Paper: https://cdn.openai.com/o3-mini-system-card.pdf</b></p><p><b>Amodei Essay: </b><a href='https://darioamodei.com/on-deepseek-and-export-controls?s=09'><b>https://darioamodei.com/on-deepseek-and-export-controls?s=09</b></a></p><p><b>FrontierMath wild stat:https://arxiv.org/pdf/2411.04872</b></p><p><b>Sam Altman Channels Napoleon: </b><a href='https://x.com/sama/status/1883185690508488934'><b>https://x.com/sama/status/1883185690508488934</b></a></p><p><b>Altman ‘pulls up releases’: https://x.com/sama/status/1884066337103962416</b></p><p><b>“AI War” by Wang: </b><a href='https://scale.com/blog/win-the-ai-war'><b>https://scale.com/blog/win-the-ai-war</b></a></p><p><b>Anthropic Original Views on Capabilities: </b><a href='https://www.anthropic.com/news/core-views-on-ai-safety'><b>https://www.anthropic.com/news/core-views-on-ai-safety</b></a></p><p><b>AI Insider Cost Comparison:https://x.com/arankomatsuzaki/status/1884676245922934788</b></p><p><b>Deepseek R1 Paper: </b><a href='https://arxiv.org/pdf/2501.12948'><b>https://arxiv.org/pdf/2501.12948</b></a></p><p><b>R1, o3-mini Price Comparison: https://techcrunch.com/2025/01/31/openai-launches-o3-mini-its-latest-reasoning-model/</b></p><p><b>Semianalysis on $1,3M deepseek salaries, and them falling behind as ‘the time gap to match US capabilities increases’: https://semianalysis.com/2025/01/31/deepseek-debates/</b></p><p><b>OpenAI Valuation: https://www.bloomberg.com/news/articles/2025-01-30/openai-in-talks-to-raise-funding-at-340-billion-value-wsj-says?srnd=phx-ai</b></p><p><b>Wang Clip: </b><a href='https://x.com/tsarnick/status/1867700453494206883'><b>https://x.com/tsarnick/status/1867700453494206883</b></a></p><p><b>Amodei Clip: </b><a href='https://x.com/ai_ctrl/status/1884951111771001188'><b>https://x.com/ai_ctrl/status/1884951111771001188</b></a></p><p><b>https://simple-bench.com/</b></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16541724-o3-mini-and-the-ai-war.mp3" length="11100490" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/70lj5vmmvapdd50lsbe67bcz5l2o?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16541724</guid>
    <pubDate>Fri, 31 Jan 2025 23:00:00 +0000</pubDate>
    <itunes:duration>921</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>4</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Nothing Much Happens in AI, Then Everything Does All At Once</itunes:title>
    <title>Nothing Much Happens in AI, Then Everything Does All At Once</title>
    <itunes:summary><![CDATA[When it rains, it pours. OpenAI Operator tested and reviewed, with full paper analysis. Perplexity Assistant is useful. Then Stargate, is it all smoke and mirrors? Strong rumours of an o3+ model from Anthropic. Then a full breakdown of Deepseek R1, and what it’s training method says about the state of AI. It’s not open source BTW. Plus Humanity’s Last Exam, and Hassabis Accelerates his AGI timeline. 00:00 - Introduction 00:54 - OpenAI Operator 04:53 - Perplexity Assistant  05:15 - StarGa...]]></itunes:summary>
    <description><![CDATA[<p><b>When it rains, it pours. OpenAI Operator tested and reviewed, with full paper analysis. Perplexity Assistant is useful. Then Stargate, is it all smoke and mirrors? Strong rumours of an o3+ model from Anthropic. Then a full breakdown of Deepseek R1, and what it’s training method says about the state of AI. It’s not open source BTW. Plus Humanity’s Last Exam, and Hassabis Accelerates his AGI timeline.</b></p><p><b>00:00 - Introduction</b></p><p><b>00:54 - OpenAI Operator</b></p><p><b>04:53 - Perplexity Assistant </b></p><p><b>05:15 - StarGate</b></p><p><b>07:51 - Better than o3?</b></p><p><b>08:25 - DeepSeek R1 Analysis</b></p><p><b>12:12 - Training Secrets</b></p><p><b>15:19 - No More Process Rewarding ?</b></p><p><b>19:01 - Hassabis Timeline Accelerates</b></p><p><b>21:22 - Humanity’s Last Exam</b></p><p><br/></p><p><a href='https://app.grayswan.ai/arena/chat/harmful-ai-assistant'><b>https://app.grayswan.ai/arena/chat/harmful-ai-assistant</b></a></p><p><a href='https://app.grayswan.ai/arena'><b>https://app.grayswan.ai/arena</b></a></p><p><a href='https://openai.com/index/computer-using-agent/'><b>https://openai.com/index/computer-using-agent/</b></a></p><p><b>System Prompt: https://github.com/wunderwuzzi23/scratch/blob/master/system_prompts/operator_system_prompt-2025-01-23.txt</b></p><p><br/></p><p><b>OpenAI Operator: </b><a href='https://operator.chatgpt.com/'><b>https://operator.chatgpt.com/</b></a></p><p><b>System Card: https://cdn.openai.com/operator_system_card.pdf</b></p><p><br/></p><p><b>There is No Plan: </b><a href='https://x.com/jeffclune/status/1882120726339318007'><b>https://x.com/jeffclune/status/1882120726339318007</b></a></p><p><br/></p><p><b>Perplexity Assistant: https://x.com/perplexity_ai/status/1882466239123255686</b></p><p><br/></p><p><b>Stargate: </b><a href='https://openai.com/index/announcing-the-stargate-project/'><b>https://openai.com/index/announcing-the-stargate-project/</b></a></p><p><b>Labour goes to 0: https://moores.samaltman.com/</b></p><p><b>Larry Ellison AI Surveillance: </b><a href='https://x.com/TheChiefNerd/status/1882042989184430332'><b>https://x.com/TheChiefNerd/status/1882042989184430332</b></a></p><p><b>Amodei 1984: https://www.bloomberg.com/news/articles/2025-01-22/anthropic-ceo-says-openai-s-stargate-venture-seems-chaotic</b></p><p><b>Microsoft Hesitate: </b><a href='https://www.theinformation.com/articles/why-sam-altman-joined-forces-with-larry-ellison-and-took-a-step-back-from-microsoft?rc=sy0ihq'><b>https://www.theinformation.com/articles/why-sam-altman-joined-forces-with-larry-ellison-and-took-a-step-back-from-microsoft?rc=sy0ihq</b></a></p><p><br/></p><p><b>Dylan Patel o3+ for Anthropic: https://www.youtube.com/watch?v=7EH0VjM3dTk</b></p><p><br/></p><p><b>Deepseek R1: </b><a href='https://arxiv.org/pdf/2501.12948'><b>https://arxiv.org/pdf/2501.12948</b></a></p><p><a href='https://arxiv.org/pdf/2412.19437'><b>https://arxiv.org/pdf/2412.19437</b></a></p><p><b>Diagram: https://pbs.twimg.com/media/GhyQsM6WQAE7W52?format=jpg&amp;name=large</b></p><p><b>https://simple-bench.com/</b></p><p><b>Process: </b><a href='https://x.com/sama/status/1664018190840614912'><b>https://x.com/sama/status/1664018190840614912</b></a></p><p><a href='https://x.com/karpathy/status/1835561952258723930'><b>https://x.com/karpathy/status/1835561952258723930</b></a></p><p><b>https://openai.com/index/trading-inference-time-compute-for-adversarial-robustness/?s=09</b></p><p><b>Demis Interview: </b><a href='https://www.youtube.com/watch?v=yr0GiSgUvPU'><b>https://www.youtube.com/watch?v=yr0GiSgUvPU</b></a></p><p><b>Humanity’s Last Exam: </b></p><p><a href='https://agi.safe.ai/'><b>https://agi.safe.ai/</b></a></p><p><b>https://x.com/DanHendrycks/status/1882481730671857815</b></p><p><b>https://www.nytimes.com/2025/01/23/technology/ai-test-humanitys-last-exam.html?s=09</b></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>When it rains, it pours. OpenAI Operator tested and reviewed, with full paper analysis. Perplexity Assistant is useful. Then Stargate, is it all smoke and mirrors? Strong rumours of an o3+ model from Anthropic. Then a full breakdown of Deepseek R1, and what it’s training method says about the state of AI. It’s not open source BTW. Plus Humanity’s Last Exam, and Hassabis Accelerates his AGI timeline.</b></p><p><b>00:00 - Introduction</b></p><p><b>00:54 - OpenAI Operator</b></p><p><b>04:53 - Perplexity Assistant </b></p><p><b>05:15 - StarGate</b></p><p><b>07:51 - Better than o3?</b></p><p><b>08:25 - DeepSeek R1 Analysis</b></p><p><b>12:12 - Training Secrets</b></p><p><b>15:19 - No More Process Rewarding ?</b></p><p><b>19:01 - Hassabis Timeline Accelerates</b></p><p><b>21:22 - Humanity’s Last Exam</b></p><p><br/></p><p><a href='https://app.grayswan.ai/arena/chat/harmful-ai-assistant'><b>https://app.grayswan.ai/arena/chat/harmful-ai-assistant</b></a></p><p><a href='https://app.grayswan.ai/arena'><b>https://app.grayswan.ai/arena</b></a></p><p><a href='https://openai.com/index/computer-using-agent/'><b>https://openai.com/index/computer-using-agent/</b></a></p><p><b>System Prompt: https://github.com/wunderwuzzi23/scratch/blob/master/system_prompts/operator_system_prompt-2025-01-23.txt</b></p><p><br/></p><p><b>OpenAI Operator: </b><a href='https://operator.chatgpt.com/'><b>https://operator.chatgpt.com/</b></a></p><p><b>System Card: https://cdn.openai.com/operator_system_card.pdf</b></p><p><br/></p><p><b>There is No Plan: </b><a href='https://x.com/jeffclune/status/1882120726339318007'><b>https://x.com/jeffclune/status/1882120726339318007</b></a></p><p><br/></p><p><b>Perplexity Assistant: https://x.com/perplexity_ai/status/1882466239123255686</b></p><p><br/></p><p><b>Stargate: </b><a href='https://openai.com/index/announcing-the-stargate-project/'><b>https://openai.com/index/announcing-the-stargate-project/</b></a></p><p><b>Labour goes to 0: https://moores.samaltman.com/</b></p><p><b>Larry Ellison AI Surveillance: </b><a href='https://x.com/TheChiefNerd/status/1882042989184430332'><b>https://x.com/TheChiefNerd/status/1882042989184430332</b></a></p><p><b>Amodei 1984: https://www.bloomberg.com/news/articles/2025-01-22/anthropic-ceo-says-openai-s-stargate-venture-seems-chaotic</b></p><p><b>Microsoft Hesitate: </b><a href='https://www.theinformation.com/articles/why-sam-altman-joined-forces-with-larry-ellison-and-took-a-step-back-from-microsoft?rc=sy0ihq'><b>https://www.theinformation.com/articles/why-sam-altman-joined-forces-with-larry-ellison-and-took-a-step-back-from-microsoft?rc=sy0ihq</b></a></p><p><br/></p><p><b>Dylan Patel o3+ for Anthropic: https://www.youtube.com/watch?v=7EH0VjM3dTk</b></p><p><br/></p><p><b>Deepseek R1: </b><a href='https://arxiv.org/pdf/2501.12948'><b>https://arxiv.org/pdf/2501.12948</b></a></p><p><a href='https://arxiv.org/pdf/2412.19437'><b>https://arxiv.org/pdf/2412.19437</b></a></p><p><b>Diagram: https://pbs.twimg.com/media/GhyQsM6WQAE7W52?format=jpg&amp;name=large</b></p><p><b>https://simple-bench.com/</b></p><p><b>Process: </b><a href='https://x.com/sama/status/1664018190840614912'><b>https://x.com/sama/status/1664018190840614912</b></a></p><p><a href='https://x.com/karpathy/status/1835561952258723930'><b>https://x.com/karpathy/status/1835561952258723930</b></a></p><p><b>https://openai.com/index/trading-inference-time-compute-for-adversarial-robustness/?s=09</b></p><p><b>Demis Interview: </b><a href='https://www.youtube.com/watch?v=yr0GiSgUvPU'><b>https://www.youtube.com/watch?v=yr0GiSgUvPU</b></a></p><p><b>Humanity’s Last Exam: </b></p><p><a href='https://agi.safe.ai/'><b>https://agi.safe.ai/</b></a></p><p><b>https://x.com/DanHendrycks/status/1882481730671857815</b></p><p><b>https://www.nytimes.com/2025/01/23/technology/ai-test-humanitys-last-exam.html?s=09</b></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16496591-nothing-much-happens-in-ai-then-everything-does-all-at-once.mp3" length="16698387" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/nxuc6k9efv07g3pg79vw76qo7cf8?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16496591</guid>
    <pubDate>Fri, 24 Jan 2025 17:00:00 +0000</pubDate>
    <itunes:duration>1389</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>3</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Altman Expects a ‘Fast Take-off’, ‘Super-Agent’ Debuting Soon and DeepSeek R1 Out</itunes:title>
    <title>Altman Expects a ‘Fast Take-off’, ‘Super-Agent’ Debuting Soon and DeepSeek R1 Out</title>
    <itunes:summary><![CDATA[OpenAI looks set to debut their Operator system, and some leaks are out. At the same time Deepseek R1 releases some numbers, and Sam Altman says he might have been wrong before, and now anticipates a 'fast take-off'.  Plus two papers to give you an idea of what a super-agent might be decent at doing, some more exclusive article analysis and much more. Who said anything else is happening today...  80,000 Hours Channel: https://www.youtube.com/channel/UCafjal1QYJ3rb0Y9xZk1Ezg Spotify: http...]]></itunes:summary>
    <description><![CDATA[<p>OpenAI looks set to debut their Operator system, and some leaks are out. At the same time Deepseek R1 releases some numbers, and Sam Altman says he might have been wrong before, and now anticipates a &apos;fast take-off&apos;.  Plus two papers to give you an idea of what a super-agent might be decent at doing, some more exclusive article analysis and much more. Who said anything else is happening today...<br/><br/>80,000 Hours Channel: https://www.youtube.com/channel/UCafjal1QYJ3rb0Y9xZk1Ezg<br/>Spotify: https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:13 - Pro Cost and OpenAI Operator<br/>04:00 - Agent Benchmarks Being Targeted<br/>07:48 - Fast Take-off, Altman<br/>08:48 - Altman flip-flops<br/>10:02 - Deepseek R1 First Reaction<br/><br/>Altman ‘100x expectations out of control’: https://x.com/sama/status/1881258443669172470<br/>OpenAI Operator Table: https://x.com/btibor91/status/1881285255266750564<br/>WebVoyager: https://arxiv.org/pdf/2401.13919<br/>OSWorld: https://arxiv.org/pdf/2404.07972<br/>Axios Exclusive 1 (SuperAgent): https://www.axios.com/2025/01/19/ai-superagent-openai-meta?s=09<br/>Axios Exclusive 2: https://www.axios.com/2025/01/18/biden-sullivan-ai-race-trump-china<br/>Deepseek R1 Numbers: https://x.com/deepseek_ai/status/1881318130334814301<br/>Does 1.5B outperform 3.5 Sonnet on Math?: https://x.com/reach_vb/status/1881319500089634954<br/>Deepseek R1 (deepseek-reasoner) Pricing: https://api-docs.deepseek.com/quick_start/pricing/<br/>Altman Fast Takeoff: https://x.com/tsarnick/status/1879100390840697191<br/>OpenAI Economic Blueprint: https://cdn.openai.com/global-affairs/ai-in-america-oai-economic-blueprint-20250113.pdf<br/>Target is Long-horizon Tasks: https://x.com/karinanguyen_/status/1879576037249667520<br/>Support Regulations: https://www.techemails.com/p/elon-musk-and-openai<br/>https://www.nytimes.com/2023/05/16/technology/openai-altman-artificial-intelligence-regulation.html<br/>Donation: https://qz.com/sam-altman-donate-million-zuckerberg-bezos-donald-trump-1851721035<br/>Amodei on Regulations by 2025: https://www.youtube.com/watch?v=ugvHCXCOmm4<br/>‘Feel the AGI’: https://x.com/polynoamial?lang=en<br/>GPT-5 and o-series merger: https://x.com/sama/status/1880358749187240274<br/>o1 Thinks in Chinese: https://techcrunch.com/2025/01/14/openais-ai-reasoning-model-thinks-in-chinese-sometimes-and-no-one-really-knows-why/<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></description>
    <content:encoded><![CDATA[<p>OpenAI looks set to debut their Operator system, and some leaks are out. At the same time Deepseek R1 releases some numbers, and Sam Altman says he might have been wrong before, and now anticipates a &apos;fast take-off&apos;.  Plus two papers to give you an idea of what a super-agent might be decent at doing, some more exclusive article analysis and much more. Who said anything else is happening today...<br/><br/>80,000 Hours Channel: https://www.youtube.com/channel/UCafjal1QYJ3rb0Y9xZk1Ezg<br/>Spotify: https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib<br/><br/>AI Insiders ($9!): https://www.patreon.com/AIExplained<br/><br/>Chapters:<br/>00:00 - Introduction<br/>01:13 - Pro Cost and OpenAI Operator<br/>04:00 - Agent Benchmarks Being Targeted<br/>07:48 - Fast Take-off, Altman<br/>08:48 - Altman flip-flops<br/>10:02 - Deepseek R1 First Reaction<br/><br/>Altman ‘100x expectations out of control’: https://x.com/sama/status/1881258443669172470<br/>OpenAI Operator Table: https://x.com/btibor91/status/1881285255266750564<br/>WebVoyager: https://arxiv.org/pdf/2401.13919<br/>OSWorld: https://arxiv.org/pdf/2404.07972<br/>Axios Exclusive 1 (SuperAgent): https://www.axios.com/2025/01/19/ai-superagent-openai-meta?s=09<br/>Axios Exclusive 2: https://www.axios.com/2025/01/18/biden-sullivan-ai-race-trump-china<br/>Deepseek R1 Numbers: https://x.com/deepseek_ai/status/1881318130334814301<br/>Does 1.5B outperform 3.5 Sonnet on Math?: https://x.com/reach_vb/status/1881319500089634954<br/>Deepseek R1 (deepseek-reasoner) Pricing: https://api-docs.deepseek.com/quick_start/pricing/<br/>Altman Fast Takeoff: https://x.com/tsarnick/status/1879100390840697191<br/>OpenAI Economic Blueprint: https://cdn.openai.com/global-affairs/ai-in-america-oai-economic-blueprint-20250113.pdf<br/>Target is Long-horizon Tasks: https://x.com/karinanguyen_/status/1879576037249667520<br/>Support Regulations: https://www.techemails.com/p/elon-musk-and-openai<br/>https://www.nytimes.com/2023/05/16/technology/openai-altman-artificial-intelligence-regulation.html<br/>Donation: https://qz.com/sam-altman-donate-million-zuckerberg-bezos-donald-trump-1851721035<br/>Amodei on Regulations by 2025: https://www.youtube.com/watch?v=ugvHCXCOmm4<br/>‘Feel the AGI’: https://x.com/polynoamial?lang=en<br/>GPT-5 and o-series merger: https://x.com/sama/status/1880358749187240274<br/>o1 Thinks in Chinese: https://techcrunch.com/2025/01/14/openais-ai-reasoning-model-thinks-in-chinese-sometimes-and-no-one-really-knows-why/<br/><br/><br/><br/>Non-hype Newsletter: https://signaltonoise.beehiiv.com/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16469583-altman-expects-a-fast-take-off-super-agent-debuting-soon-and-deepseek-r1-out.mp3" length="9532530" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/kan9f76gs9grpnagijrfyitlwqe6?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16469583</guid>
    <pubDate>Mon, 20 Jan 2025 16:00:00 +0000</pubDate>
    <itunes:duration>791</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>2</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>OpenAI Backtracks on Superintelligence + Altman Brings His Timeline Forward</itunes:title>
    <title>OpenAI Backtracks on Superintelligence + Altman Brings His Timeline Forward</title>
    <itunes:summary><![CDATA[Sam Altman unexpectedly brings his timelines to AGI forward, while OpenAI backtrack on superintelligence. None of these changes were heralded, but they are significant. Plus the new year brings new assessments of the true capability of models to automate 'large swathes of the economy'. I'll give my prediction on that front for 2025, announcement a new Simple Bench competition, and showcase Kling 1.6 vs Veo 2 vs Sora, and much more.   wandb.me/simple-bench (Colab): https://colab.research.googl...]]></itunes:summary>
    <description><![CDATA[<p>Sam Altman unexpectedly brings his timelines to AGI forward, while OpenAI backtrack on superintelligence. None of these changes were heralded, but they are significant. Plus the new year brings new assessments of the true capability of models to automate &apos;large swathes of the economy&apos;. I&apos;ll give my prediction on that front for 2025, announcement a new Simple Bench competition, and showcase Kling 1.6 vs Veo 2 vs Sora, and much more. <br/><br/><a href='https://app.slack.com/client/T07KBQCRUMC/wandb.me/simple-bench'><b>wandb.me/simple-bench</b></a></p><p><b>(Colab): https://colab.research.google.com/drive/1AVijcPnEkl8Gy_754XbRdG5m7Q5-9slg?usp=sharing</b></p><p><br/></p><p><b>TheAgentCompany Paper: </b><a href='https://arxiv.org/pdf/2412.14161v1'><b>https://arxiv.org/pdf/2412.14161v1</b></a></p><p><b>Sam Altman Major Interview: https://www.bloomberg.com/features/2025-sam-altman-interview/?srnd=phx-ai</b></p><p><b>OpenAI Agent Coming Jan 2025: </b><a href='https://www.theinformation.com/articles/why-openai-is-taking-so-long-to-launch-agents?rc=sy0ihq'><b>https://www.theinformation.com/articles/why-openai-is-taking-so-long-to-launch-agents?rc=sy0ihq</b></a></p><p><b>Altman Singularity: </b><a href='https://x.com/sama/status/1875603249472139576'><b>https://x.com/sama/status/1875603249472139576</b></a></p><p><b>Altman Original Timeline: </b><a href='https://www.youtube.com/watch?v=7dCPytNTnjk&amp;t=621s'><b>https://www.youtube.com/watch?v=7dCPytNTnjk&amp;t=621s</b></a></p><p><a href='https://www.ft.com/content/34a7a082-e685-4e02-bca7-61ff89d99ed2'><b>https://www.ft.com/content/34a7a082-e685-4e02-bca7-61ff89d99ed2</b></a></p><p><b>OpenAI Original Emails: https://www.lesswrong.com/posts/5jjk4CDnj9tA7ugxr/openai-email-archives-from-musk-v-altman-and-openai-blog</b></p><p><b>DeepMind Sky News 2014 Article: </b><a href='https://news.sky.com/story/google-buys-uk-intelligence-firm-deepmind-10419783'><b>https://news.sky.com/story/google-buys-uk-intelligence-firm-deepmind-10419783</b></a></p><p><b>Altman Blog Reflections: </b><a href='https://blog.samaltman.com/reflections'><b>https://blog.samaltman.com/reflections</b></a></p><p><b>OpenAI Changes Who Gets AGI: </b><a href='https://openai.com/index/why-our-structure-must-evolve-to-advance-our-mission/?s=09'><b>https://openai.com/index/why-our-structure-must-evolve-to-advance-our-mission/?s=09</b></a></p><p><b>OpenAI 5 Levels: </b><a href='https://www.bloomberg.com/news/articles/2024-07-11/openai-sets-levels-to-track-progress-toward-superintelligent-ai'><b>https://www.bloomberg.com/news/articles/2024-07-11/openai-sets-levels-to-track-progress-toward-superintelligent-ai</b></a></p><p><b>Altman 2015: https://blog.samaltman.com/machine-intelligence-part-1</b></p><p><b>OpenAI React to Anthropic: </b><a href='https://www.theinformation.com/articles/how-anthropic-got-inside-openais-head?rc=sy0ihq'><b>https://www.theinformation.com/articles/how-anthropic-got-inside-openais-head?rc=sy0ihq</b></a></p><p><b>Microsoft $100B Definition: </b><a href='https://www.theinformation.com/articles/microsoft-and-openai-wrangle-over-terms-of-their-blockbuster-partnership?rc=sy0ihq'><b>https://www.theinformation.com/articles/microsoft-and-openai-wrangle-over-terms-of-their-blockbuster-partnership?rc=sy0ihq<br/></b></a><b>Epoch Scramble for Task Benchmark: </b><a href='https://x.com/tamaybes/status/1876692639363612919'><b>https://x.com/tamaybes/status/1876692639363612919</b></a></p><p><b>GPQA Progress: https://epoch.ai/data/ai-benchmarking-dashboard</b></p><p><b>Task Length Crucial for ARC-AGI: </b><a href='https://anokas.substack.com/p/llms-struggle-with-perception-not-reasoning-arcagi'><b>https://anokas.substack.com/p/llms-struggle-with-perception-not-reasoning-arcagi</b></a></p><p><b>RL Environment Tweet: https://x.com/vedantmisra/status/1876327518157807990</b></p><p><b>Jason Wei Talk: </b><a href='https://www.youtube.com/watch?v=yhpjpNXJDco'><b>https://www.youtube.com/watch?v=yhpjpNXJDco</b></a></p><p><b>Miles Brunda</b></p>]]></description>
    <content:encoded><![CDATA[<p>Sam Altman unexpectedly brings his timelines to AGI forward, while OpenAI backtrack on superintelligence. None of these changes were heralded, but they are significant. Plus the new year brings new assessments of the true capability of models to automate &apos;large swathes of the economy&apos;. I&apos;ll give my prediction on that front for 2025, announcement a new Simple Bench competition, and showcase Kling 1.6 vs Veo 2 vs Sora, and much more. <br/><br/><a href='https://app.slack.com/client/T07KBQCRUMC/wandb.me/simple-bench'><b>wandb.me/simple-bench</b></a></p><p><b>(Colab): https://colab.research.google.com/drive/1AVijcPnEkl8Gy_754XbRdG5m7Q5-9slg?usp=sharing</b></p><p><br/></p><p><b>TheAgentCompany Paper: </b><a href='https://arxiv.org/pdf/2412.14161v1'><b>https://arxiv.org/pdf/2412.14161v1</b></a></p><p><b>Sam Altman Major Interview: https://www.bloomberg.com/features/2025-sam-altman-interview/?srnd=phx-ai</b></p><p><b>OpenAI Agent Coming Jan 2025: </b><a href='https://www.theinformation.com/articles/why-openai-is-taking-so-long-to-launch-agents?rc=sy0ihq'><b>https://www.theinformation.com/articles/why-openai-is-taking-so-long-to-launch-agents?rc=sy0ihq</b></a></p><p><b>Altman Singularity: </b><a href='https://x.com/sama/status/1875603249472139576'><b>https://x.com/sama/status/1875603249472139576</b></a></p><p><b>Altman Original Timeline: </b><a href='https://www.youtube.com/watch?v=7dCPytNTnjk&amp;t=621s'><b>https://www.youtube.com/watch?v=7dCPytNTnjk&amp;t=621s</b></a></p><p><a href='https://www.ft.com/content/34a7a082-e685-4e02-bca7-61ff89d99ed2'><b>https://www.ft.com/content/34a7a082-e685-4e02-bca7-61ff89d99ed2</b></a></p><p><b>OpenAI Original Emails: https://www.lesswrong.com/posts/5jjk4CDnj9tA7ugxr/openai-email-archives-from-musk-v-altman-and-openai-blog</b></p><p><b>DeepMind Sky News 2014 Article: </b><a href='https://news.sky.com/story/google-buys-uk-intelligence-firm-deepmind-10419783'><b>https://news.sky.com/story/google-buys-uk-intelligence-firm-deepmind-10419783</b></a></p><p><b>Altman Blog Reflections: </b><a href='https://blog.samaltman.com/reflections'><b>https://blog.samaltman.com/reflections</b></a></p><p><b>OpenAI Changes Who Gets AGI: </b><a href='https://openai.com/index/why-our-structure-must-evolve-to-advance-our-mission/?s=09'><b>https://openai.com/index/why-our-structure-must-evolve-to-advance-our-mission/?s=09</b></a></p><p><b>OpenAI 5 Levels: </b><a href='https://www.bloomberg.com/news/articles/2024-07-11/openai-sets-levels-to-track-progress-toward-superintelligent-ai'><b>https://www.bloomberg.com/news/articles/2024-07-11/openai-sets-levels-to-track-progress-toward-superintelligent-ai</b></a></p><p><b>Altman 2015: https://blog.samaltman.com/machine-intelligence-part-1</b></p><p><b>OpenAI React to Anthropic: </b><a href='https://www.theinformation.com/articles/how-anthropic-got-inside-openais-head?rc=sy0ihq'><b>https://www.theinformation.com/articles/how-anthropic-got-inside-openais-head?rc=sy0ihq</b></a></p><p><b>Microsoft $100B Definition: </b><a href='https://www.theinformation.com/articles/microsoft-and-openai-wrangle-over-terms-of-their-blockbuster-partnership?rc=sy0ihq'><b>https://www.theinformation.com/articles/microsoft-and-openai-wrangle-over-terms-of-their-blockbuster-partnership?rc=sy0ihq<br/></b></a><b>Epoch Scramble for Task Benchmark: </b><a href='https://x.com/tamaybes/status/1876692639363612919'><b>https://x.com/tamaybes/status/1876692639363612919</b></a></p><p><b>GPQA Progress: https://epoch.ai/data/ai-benchmarking-dashboard</b></p><p><b>Task Length Crucial for ARC-AGI: </b><a href='https://anokas.substack.com/p/llms-struggle-with-perception-not-reasoning-arcagi'><b>https://anokas.substack.com/p/llms-struggle-with-perception-not-reasoning-arcagi</b></a></p><p><b>RL Environment Tweet: https://x.com/vedantmisra/status/1876327518157807990</b></p><p><b>Jason Wei Talk: </b><a href='https://www.youtube.com/watch?v=yhpjpNXJDco'><b>https://www.youtube.com/watch?v=yhpjpNXJDco</b></a></p><p><b>Miles Brunda</b></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16403828-openai-backtracks-on-superintelligence-altman-brings-his-timeline-forward.mp3" length="17089113" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/bubofg9hz6qalxy24tlw90wnkjpg?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16403828</guid>
    <pubDate>Wed, 08 Jan 2025 19:00:00 +0000</pubDate>
    <itunes:duration>1421</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>2</itunes:season>
    <itunes:episode>1</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>o3 - wow</itunes:title>
    <title>o3 - wow</title>
    <itunes:summary><![CDATA[o3 isn’t one of the biggest developments in AI for 2+ years because it beats a particular benchmark. It is so because it demonstrates a reusable technique through which almost any benchmark could fall, and at short notice. I’ll cover all the highlights, benchmarks broken, and what comes next. Plus, the costs OpenAI didn’t want us to know, Genesis, ARC-AGI 2, Gemini-Thinking, and much more.    FrontierMath: https://epoch.ai/frontiermath https://arxiv.org/pdf/2411.04872 Chollet Statement:h...]]></itunes:summary>
    <description><![CDATA[<p><b>o3 isn’t one of the biggest developments in AI for 2+ years because it beats a particular benchmark. It is so because it demonstrates a reusable technique through which almost any benchmark could fall, and at short notice. I’ll cover all the highlights, benchmarks broken, and what comes next. Plus, the costs OpenAI didn’t want us to know, Genesis, ARC-AGI 2, Gemini-Thinking, and much more. </b></p><p><br/></p><p><b>FrontierMath: </b><a href='https://epoch.ai/frontiermath'><b>https://epoch.ai/frontiermath</b></a></p><p><a href='https://arxiv.org/pdf/2411.04872'><b>https://arxiv.org/pdf/2411.04872</b></a></p><p><b>Chollet Statement:https://arcprize.org/blog/oai-o3-pub-breakthrough</b></p><p><b>MLC Paper: </b></p><p><a href='https://www.scientificamerican.com/article/new-training-method-helps-ai-generalize-like-people-do/?utm_campaign=socialflow&amp;utm_source=twitter&amp;utm_medium=social'><b>https://www.scientificamerican.com/article/new-training-method-helps-ai-generalize-like-people-do/?utm_campaign=socialflow&amp;utm_source=twitter&amp;utm_medium=social</b></a></p><p><b>AlphaCode 2: </b><a href='https://storage.googleapis.com/deepmind-media/AlphaCode2/AlphaCode2_Tech_Report.pdf'><b>https://storage.googleapis.com/deepmind-media/AlphaCode2/AlphaCode2_Tech_Report.pdf</b></a></p><p><b>Human Performance on ARC-AGI: </b><a href='https://arxiv.org/pdf/2409.01374v1'><b>https://arxiv.org/pdf/2409.01374v1</b></a></p><p><b>Wei Tweet ‘3 months’:</b><a href='https://x.com/_jasonwei/status/1870184982007644614'><b>https://x.com/_jasonwei/status/1870184982007644614</b></a></p><p><b>Deliberative Alignment Paper: </b><a href='https://openai.com/index/deliberative-alignment/'><b>https://openai.com/index/deliberative-alignment/</b></a></p><p><b>Brown Safety Tweet: </b><a href='https://x.com/polynoamial/status/1870196476908834893'><b>https://x.com/polynoamial/status/1870196476908834893</b></a></p><p><b>Swe-Bench Verified: </b><a href='https://openai.com/index/introducing-swe-bench-verified/'><b>https://openai.com/index/introducing-swe-bench-verified/</b></a></p><p><b>Amodei Prediction: </b><a href='https://x.com/OfirPress/status/1858567863788769518'><b>https://x.com/OfirPress/status/1858567863788769518</b></a></p><p><b>David Dohan: 16 hours </b><a href='https://x.com/dmdohan/status/1870171404093796638'><b>https://x.com/dmdohan/status/1870171404093796638</b></a></p><p><b>OpenAI Personal Writing: </b><a href='https://openai.com/index/learning-to-reason-with-llms/'><b>https://openai.com/index/learning-to-reason-with-llms/</b></a></p><p><a href='https://simple-bench.com/'><b>https://simple-bench.com/</b></a></p><p><b>John Hallman Tweet: </b><a href='https://x.com/johnohallman/status/1870233375681945725'><b>https://x.com/johnohallman/status/1870233375681945725</b></a></p><p><br/></p><p><b>00:00 - Introduction</b></p><p><b>01:19 - What is o3?</b></p><p><b>03:18 - FrontierMath</b></p><p><b>05:15 - o4, o5</b></p><p><b>06:03 - GPQA</b></p><p><b>06:24 - Coding, Codeforces + SWE-verified, AlphaCode 2</b></p><p><b>08:13 - 1st Caveat</b></p><p><b>09:03 - Compositionality?</b></p><p><b>10:16 - SimpleBench?</b></p><p><b>13:11 - ARC-AGI, Chollet</b></p><p><b><br/></b><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>o3 isn’t one of the biggest developments in AI for 2+ years because it beats a particular benchmark. It is so because it demonstrates a reusable technique through which almost any benchmark could fall, and at short notice. I’ll cover all the highlights, benchmarks broken, and what comes next. Plus, the costs OpenAI didn’t want us to know, Genesis, ARC-AGI 2, Gemini-Thinking, and much more. </b></p><p><br/></p><p><b>FrontierMath: </b><a href='https://epoch.ai/frontiermath'><b>https://epoch.ai/frontiermath</b></a></p><p><a href='https://arxiv.org/pdf/2411.04872'><b>https://arxiv.org/pdf/2411.04872</b></a></p><p><b>Chollet Statement:https://arcprize.org/blog/oai-o3-pub-breakthrough</b></p><p><b>MLC Paper: </b></p><p><a href='https://www.scientificamerican.com/article/new-training-method-helps-ai-generalize-like-people-do/?utm_campaign=socialflow&amp;utm_source=twitter&amp;utm_medium=social'><b>https://www.scientificamerican.com/article/new-training-method-helps-ai-generalize-like-people-do/?utm_campaign=socialflow&amp;utm_source=twitter&amp;utm_medium=social</b></a></p><p><b>AlphaCode 2: </b><a href='https://storage.googleapis.com/deepmind-media/AlphaCode2/AlphaCode2_Tech_Report.pdf'><b>https://storage.googleapis.com/deepmind-media/AlphaCode2/AlphaCode2_Tech_Report.pdf</b></a></p><p><b>Human Performance on ARC-AGI: </b><a href='https://arxiv.org/pdf/2409.01374v1'><b>https://arxiv.org/pdf/2409.01374v1</b></a></p><p><b>Wei Tweet ‘3 months’:</b><a href='https://x.com/_jasonwei/status/1870184982007644614'><b>https://x.com/_jasonwei/status/1870184982007644614</b></a></p><p><b>Deliberative Alignment Paper: </b><a href='https://openai.com/index/deliberative-alignment/'><b>https://openai.com/index/deliberative-alignment/</b></a></p><p><b>Brown Safety Tweet: </b><a href='https://x.com/polynoamial/status/1870196476908834893'><b>https://x.com/polynoamial/status/1870196476908834893</b></a></p><p><b>Swe-Bench Verified: </b><a href='https://openai.com/index/introducing-swe-bench-verified/'><b>https://openai.com/index/introducing-swe-bench-verified/</b></a></p><p><b>Amodei Prediction: </b><a href='https://x.com/OfirPress/status/1858567863788769518'><b>https://x.com/OfirPress/status/1858567863788769518</b></a></p><p><b>David Dohan: 16 hours </b><a href='https://x.com/dmdohan/status/1870171404093796638'><b>https://x.com/dmdohan/status/1870171404093796638</b></a></p><p><b>OpenAI Personal Writing: </b><a href='https://openai.com/index/learning-to-reason-with-llms/'><b>https://openai.com/index/learning-to-reason-with-llms/</b></a></p><p><a href='https://simple-bench.com/'><b>https://simple-bench.com/</b></a></p><p><b>John Hallman Tweet: </b><a href='https://x.com/johnohallman/status/1870233375681945725'><b>https://x.com/johnohallman/status/1870233375681945725</b></a></p><p><br/></p><p><b>00:00 - Introduction</b></p><p><b>01:19 - What is o3?</b></p><p><b>03:18 - FrontierMath</b></p><p><b>05:15 - o4, o5</b></p><p><b>06:03 - GPQA</b></p><p><b>06:24 - Coding, Codeforces + SWE-verified, AlphaCode 2</b></p><p><b>08:13 - 1st Caveat</b></p><p><b>09:03 - Compositionality?</b></p><p><b>10:16 - SimpleBench?</b></p><p><b>13:11 - ARC-AGI, Chollet</b></p><p><b><br/></b><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16318518-o3-wow.mp3" length="16114612" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/ek16nxl0kwuyooly8ikyhqgk0qp0?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16318518</guid>
    <pubDate>Sat, 21 Dec 2024 00:00:00 +0000</pubDate>
    <itunes:duration>1340</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>9</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Never Browse Alone? - Gemini 2 Live and ChatGPT Vision</itunes:title>
    <title>Never Browse Alone? - Gemini 2 Live and ChatGPT Vision</title>
    <itunes:summary><![CDATA[The ‘Gemini 2 Era’ begins … with screen-sharing? But really, it’s a great free tool, for curiosity satisfying rather than bleeding-edge intelligence. I give you the benchmarks, the highlights and of course, the latest from OpenAI Advanced Voice Mode with Vision.  Plus Deep Research in Gemini Advanced, Simple Bench updates, Santa and what might be for some of you Google’s deflating admission.    00:00 - Introduction 00:38 - Live Interaction  03:43 - Gemini 2.0 Flash Benchmarks&n...]]></itunes:summary>
    <description><![CDATA[<p><b>The ‘Gemini 2 Era’ begins … with screen-sharing? But really, it’s a great free tool, for curiosity satisfying rather than bleeding-edge intelligence. I give you the benchmarks, the highlights and of course, the latest from OpenAI Advanced Voice Mode with Vision. </b></p><p><b>Plus Deep Research in Gemini Advanced, Simple Bench updates, Santa and what might be for some of you Google’s deflating admission. </b></p><p><br/></p><p><b>00:00 - Introduction</b></p><p><b>00:38 - Live Interaction </b></p><p><b>03:43 - Gemini 2.0 Flash Benchmarks </b></p><p><b>05:10 - Audio and Image Output</b></p><p><b>06:38 - Project Mariner (+ WebVoyager Bench)</b></p><p><b>08:49 - But Progress Slowing Down?</b></p><p><b>10:43 - OpenAI Announcements + Games</b></p><p><b><br/></b><br/></p><p><b>https://aistudio.google.com/live</b></p><p><b>Gemini 2.0 Flash Benchmarks: https://deepmind.google/technologies/gemini/</b></p><p><b>Project mariner: </b><a href='https://deepmind.google/technologies/project-mariner/'><b>https://deepmind.google/technologies/project-mariner/</b></a></p><p><b>WebVoyager: https://x.com/laurentsifre/status/1858918588683296875/photo/1</b></p><p><b>Gemini Game play: </b><a href='https://www.youtube.com/watch?v=IKuGNHJBGsc'><b>https://www.youtube.com/watch?v=IKuGNHJBGsc</b></a></p><p><b>Advanced Voice Mode OpenAI: </b><a href='https://www.youtube.com/watch?v=NIQDnWlwYyQ'><b>https://www.youtube.com/watch?v=NIQDnWlwYyQ</b></a></p><p><a href='https://simple-bench.com/'><b>https://simple-bench.com/</b></a></p><p><b>Claude Computer Use: </b><a href='https://docs.anthropic.com/en/docs/build-with-claude/computer-use'><b>https://docs.anthropic.com/en/docs/build-with-claude/computer-use</b></a></p><p><b>Oriol Vinyals Interview: https://www.youtube.com/watch?v=78mEYaztGaw&amp;t=687s</b></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>The ‘Gemini 2 Era’ begins … with screen-sharing? But really, it’s a great free tool, for curiosity satisfying rather than bleeding-edge intelligence. I give you the benchmarks, the highlights and of course, the latest from OpenAI Advanced Voice Mode with Vision. </b></p><p><b>Plus Deep Research in Gemini Advanced, Simple Bench updates, Santa and what might be for some of you Google’s deflating admission. </b></p><p><br/></p><p><b>00:00 - Introduction</b></p><p><b>00:38 - Live Interaction </b></p><p><b>03:43 - Gemini 2.0 Flash Benchmarks </b></p><p><b>05:10 - Audio and Image Output</b></p><p><b>06:38 - Project Mariner (+ WebVoyager Bench)</b></p><p><b>08:49 - But Progress Slowing Down?</b></p><p><b>10:43 - OpenAI Announcements + Games</b></p><p><b><br/></b><br/></p><p><b>https://aistudio.google.com/live</b></p><p><b>Gemini 2.0 Flash Benchmarks: https://deepmind.google/technologies/gemini/</b></p><p><b>Project mariner: </b><a href='https://deepmind.google/technologies/project-mariner/'><b>https://deepmind.google/technologies/project-mariner/</b></a></p><p><b>WebVoyager: https://x.com/laurentsifre/status/1858918588683296875/photo/1</b></p><p><b>Gemini Game play: </b><a href='https://www.youtube.com/watch?v=IKuGNHJBGsc'><b>https://www.youtube.com/watch?v=IKuGNHJBGsc</b></a></p><p><b>Advanced Voice Mode OpenAI: </b><a href='https://www.youtube.com/watch?v=NIQDnWlwYyQ'><b>https://www.youtube.com/watch?v=NIQDnWlwYyQ</b></a></p><p><a href='https://simple-bench.com/'><b>https://simple-bench.com/</b></a></p><p><b>Claude Computer Use: </b><a href='https://docs.anthropic.com/en/docs/build-with-claude/computer-use'><b>https://docs.anthropic.com/en/docs/build-with-claude/computer-use</b></a></p><p><b>Oriol Vinyals Interview: https://www.youtube.com/watch?v=78mEYaztGaw&amp;t=687s</b></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16270631-never-browse-alone-gemini-2-live-and-chatgpt-vision.mp3" length="9878438" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/81bu5j7p7y45phksny365cojg8tu?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16270631</guid>
    <pubDate>Thu, 12 Dec 2024 23:00:00 +0000</pubDate>
    <itunes:duration>820</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>8</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Sora is Out, But is it a Distraction?</itunes:title>
    <title>Sora is Out, But is it a Distraction?</title>
    <itunes:summary><![CDATA[After a 10 month wait, OpenAI have released Sora to paying users. With just a prompt it can generate videos of up to 20 seconds in lower resolutions, and 10 seconds at 1080p if you can fork out $200/month. I’ve tested it and read the system card. The user interface is quite beautiful, even if the videos themselves operate until entirely new rules of physics. But I can’t help wondering if OpenAI want up to focus on releases like this, rather than some quietly broken promises.     80,000 h...]]></itunes:summary>
    <description><![CDATA[<p><b>After a 10 month wait, OpenAI have released Sora to paying users. With just a prompt it can generate videos of up to 20 seconds in lower resolutions, and 10 seconds at 1080p if you can fork out $200/month. I’ve tested it and read the system card. The user interface is quite beautiful, even if the videos themselves operate until entirely new rules of physics. But I can’t help wondering if OpenAI want up to focus on releases like this, rather than some quietly broken promises. </b></p><p><b><br/></b><br/></p><p><b>80,000 hours Website, Podcast + Channel:</b><a href='https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib'><b> </b></a></p><p><a href='https://80000hours.org/'><b>https://80000hours.org/</b></a></p><p><a href='https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib'><b>https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib</b></a><a href='https://www.youtube.com/@eightythousandhours/videos'><b> https://www.youtube.com/@eightythousandhours/videos</b></a></p><p><br/></p><p><b>https://openai.com/sora/</b></p><p><br/></p><p><b>Sora Countries: </b><a href='https://help.openai.com/en/articles/10250692-sora-supported-countries'><b>https://help.openai.com/en/articles/10250692-sora-supported-countries</b></a></p><p><b>Sora Credits: </b><a href='https://help.openai.com/en/articles/10245774-sora-billing-credits-faq'><b>https://help.openai.com/en/articles/10245774-sora-billing-credits-faq</b></a></p><p><a href='https://runwayml.com/'><b>https://runwayml.com/</b></a><b> and </b><a href='https://pika.art/home'><b>https://pika.art/home</b></a><b> </b></p><p><br/></p><p><b>DeepMind Veo: </b><a href='https://deepmind.google/technologies/veo/'><b>https://deepmind.google/technologies/veo/</b></a></p><p><br/></p><p><b>Sam Altman Ads as Last Resort: </b><a href='https://www.windowscentral.com/software-apps/openai-could-chase-intrusive-ads-as-last-resort'><b>https://www.windowscentral.com/software-apps/openai-could-chase-intrusive-ads-as-last-resort</b></a></p><p><br/></p><p><b>But OpenAI Considering Ads: </b><a href='https://www.inc.com/ben-sherry/is-openai-getting-into-the-advertising-business-the-company-is-sending-mixed-messages/91033533'><b>https://www.inc.com/ben-sherry/is-openai-getting-into-the-advertising-business-the-company-is-sending-mixed-messages/91033533</b></a></p><p><br/></p><p><b>OpenAI Backtracks on Microsoft AGI Clause: </b><a href='https://www.ft.com/content/2c14b89c-f363-4c2a-9dfc-13023b6bce65'><b>https://www.ft.com/content/2c14b89c-f363-4c2a-9dfc-13023b6bce65</b></a></p><p><br/></p><p><b>As Microsoft Boast of Labor Savings: </b><a href='https://www.theinformation.com/articles/microsofts-new-sales-pitch-for-ai-spend-less-money-on-humans?rc=sy0ihq'><b>https://www.theinformation.com/articles/microsofts-new-sales-pitch-for-ai-spend-less-money-on-humans?rc=sy0ihq</b></a></p><p><br/></p><p><b>OpenAI Military Pivot: </b><a href='https://www.technologyreview.com/2024/12/04/1107897/openais-new-defense-contract-completes-its-military-pivot/'><b>https://www.technologyreview.com/2024/12/04/1107897/openais-new-defense-contract-completes-its-military-pivot/</b></a></p><p><br/></p><p><b>Employees Have Doubts: </b><a href='https://www.washingtonpost.com/technology/2024/12/06/openai-anduril-employee-military-ai/?nid=top_pb_signin&amp;arcId=KZIV7PLRHBCVNPAIAAAVUNRHIM&amp;account_location=ONSITE_HEADER_ARTICLE'><b>https://www.washingtonpost.com/technology/2024/12/06/openai-anduril-employee-military-ai/?nid=top_pb_signin&amp;arcId=KZIV7PLRHBCVNPAIAAAVUNRHIM&amp;account_location=ONSITE_HEADER_ARTICLE</b></a></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>After a 10 month wait, OpenAI have released Sora to paying users. With just a prompt it can generate videos of up to 20 seconds in lower resolutions, and 10 seconds at 1080p if you can fork out $200/month. I’ve tested it and read the system card. The user interface is quite beautiful, even if the videos themselves operate until entirely new rules of physics. But I can’t help wondering if OpenAI want up to focus on releases like this, rather than some quietly broken promises. </b></p><p><b><br/></b><br/></p><p><b>80,000 hours Website, Podcast + Channel:</b><a href='https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib'><b> </b></a></p><p><a href='https://80000hours.org/'><b>https://80000hours.org/</b></a></p><p><a href='https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib'><b>https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib</b></a><a href='https://www.youtube.com/@eightythousandhours/videos'><b> https://www.youtube.com/@eightythousandhours/videos</b></a></p><p><br/></p><p><b>https://openai.com/sora/</b></p><p><br/></p><p><b>Sora Countries: </b><a href='https://help.openai.com/en/articles/10250692-sora-supported-countries'><b>https://help.openai.com/en/articles/10250692-sora-supported-countries</b></a></p><p><b>Sora Credits: </b><a href='https://help.openai.com/en/articles/10245774-sora-billing-credits-faq'><b>https://help.openai.com/en/articles/10245774-sora-billing-credits-faq</b></a></p><p><a href='https://runwayml.com/'><b>https://runwayml.com/</b></a><b> and </b><a href='https://pika.art/home'><b>https://pika.art/home</b></a><b> </b></p><p><br/></p><p><b>DeepMind Veo: </b><a href='https://deepmind.google/technologies/veo/'><b>https://deepmind.google/technologies/veo/</b></a></p><p><br/></p><p><b>Sam Altman Ads as Last Resort: </b><a href='https://www.windowscentral.com/software-apps/openai-could-chase-intrusive-ads-as-last-resort'><b>https://www.windowscentral.com/software-apps/openai-could-chase-intrusive-ads-as-last-resort</b></a></p><p><br/></p><p><b>But OpenAI Considering Ads: </b><a href='https://www.inc.com/ben-sherry/is-openai-getting-into-the-advertising-business-the-company-is-sending-mixed-messages/91033533'><b>https://www.inc.com/ben-sherry/is-openai-getting-into-the-advertising-business-the-company-is-sending-mixed-messages/91033533</b></a></p><p><br/></p><p><b>OpenAI Backtracks on Microsoft AGI Clause: </b><a href='https://www.ft.com/content/2c14b89c-f363-4c2a-9dfc-13023b6bce65'><b>https://www.ft.com/content/2c14b89c-f363-4c2a-9dfc-13023b6bce65</b></a></p><p><br/></p><p><b>As Microsoft Boast of Labor Savings: </b><a href='https://www.theinformation.com/articles/microsofts-new-sales-pitch-for-ai-spend-less-money-on-humans?rc=sy0ihq'><b>https://www.theinformation.com/articles/microsofts-new-sales-pitch-for-ai-spend-less-money-on-humans?rc=sy0ihq</b></a></p><p><br/></p><p><b>OpenAI Military Pivot: </b><a href='https://www.technologyreview.com/2024/12/04/1107897/openais-new-defense-contract-completes-its-military-pivot/'><b>https://www.technologyreview.com/2024/12/04/1107897/openais-new-defense-contract-completes-its-military-pivot/</b></a></p><p><br/></p><p><b>Employees Have Doubts: </b><a href='https://www.washingtonpost.com/technology/2024/12/06/openai-anduril-employee-military-ai/?nid=top_pb_signin&amp;arcId=KZIV7PLRHBCVNPAIAAAVUNRHIM&amp;account_location=ONSITE_HEADER_ARTICLE'><b>https://www.washingtonpost.com/technology/2024/12/06/openai-anduril-employee-military-ai/?nid=top_pb_signin&amp;arcId=KZIV7PLRHBCVNPAIAAAVUNRHIM&amp;account_location=ONSITE_HEADER_ARTICLE</b></a></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16250486-sora-is-out-but-is-it-a-distraction.mp3" length="11244650" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/d4axtv1bvpzifshffj0lixdi724g?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16250486</guid>
    <pubDate>Tue, 10 Dec 2024 00:00:00 +0000</pubDate>
    <itunes:duration>934</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>7</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>o1 Pro Mode – Full Analysis (plus o1 paper highlights)</itunes:title>
    <title>o1 Pro Mode – Full Analysis (plus o1 paper highlights)</title>
    <itunes:summary><![CDATA[Oh boy. o1 pro mode out on the same night as o1 full. I read the 49 page paper, ran my own tests, spent my fuel allowance on Pro Mode and will give you all the highlights. Suffice to say the story is not as simple as it first appears.  Weights and Biases’ Weave: wandb.me/ai_explained Plus, GPT-4.5? MLE Bench, Simple Update, Image Analysis and much more    o1 System Card: https://cdn.openai.com/o1-system-card-20241205.pdf Apollo Research: https://www.apolloresearch.ai/research/s...]]></itunes:summary>
    <description><![CDATA[<p>Oh boy. o1 pro mode out on the same night as o1 full. I read the 49 page paper, ran my own tests, spent my fuel allowance on Pro Mode and will give you all the highlights. Suffice to say the story is not as simple as it first appears. </p><p>Weights and Biases’ Weave: wandb.me/ai_explained</p><p>Plus, GPT-4.5? MLE Bench, Simple Update, Image Analysis and much more </p><p> </p><p>o1 System Card: https://cdn.openai.com/o1-system-card-20241205.pdf</p><p>Apollo Research: <a href='https://www.apolloresearch.ai/research/scheming-reasoning-evaluations'>https://www.apolloresearch.ai/research/scheming-reasoning-evaluations</a></p><p>Altman Tweet: <a href='https://x.com/AnonCEOMakeItAi/status/1864763052622504344'>https://x.com/AnonCEOMakeItAi/status/1864763052622504344</a></p><p>ChatGPT Pro: <a href='https://openai.com/index/introducing-chatgpt-pro/'>https://openai.com/index/introducing-chatgpt-pro/</a></p><p>Tibor Blaho: https://x.com/btibor91/status/1864709670470066605</p><p>Simple-bench.com </p><p> </p><p>00:00 - Introduction</p><p>00:27 - ChatGPT Pro is $200</p><p>01:25 - OpenAI Benchmarks</p><p>03:20 - o1 System Card, o1 and o1 Pro Mode vs o1-preview</p><p>06:18 - Simple Bench surprising results on sample</p><p>08:31 - Weight &amp; Biases</p><p>09:05 - Image Analysis Compared</p><p>12:51 - More Benchmarks and Safety</p>]]></description>
    <content:encoded><![CDATA[<p>Oh boy. o1 pro mode out on the same night as o1 full. I read the 49 page paper, ran my own tests, spent my fuel allowance on Pro Mode and will give you all the highlights. Suffice to say the story is not as simple as it first appears. </p><p>Weights and Biases’ Weave: wandb.me/ai_explained</p><p>Plus, GPT-4.5? MLE Bench, Simple Update, Image Analysis and much more </p><p> </p><p>o1 System Card: https://cdn.openai.com/o1-system-card-20241205.pdf</p><p>Apollo Research: <a href='https://www.apolloresearch.ai/research/scheming-reasoning-evaluations'>https://www.apolloresearch.ai/research/scheming-reasoning-evaluations</a></p><p>Altman Tweet: <a href='https://x.com/AnonCEOMakeItAi/status/1864763052622504344'>https://x.com/AnonCEOMakeItAi/status/1864763052622504344</a></p><p>ChatGPT Pro: <a href='https://openai.com/index/introducing-chatgpt-pro/'>https://openai.com/index/introducing-chatgpt-pro/</a></p><p>Tibor Blaho: https://x.com/btibor91/status/1864709670470066605</p><p>Simple-bench.com </p><p> </p><p>00:00 - Introduction</p><p>00:27 - ChatGPT Pro is $200</p><p>01:25 - OpenAI Benchmarks</p><p>03:20 - o1 System Card, o1 and o1 Pro Mode vs o1-preview</p><p>06:18 - Simple Bench surprising results on sample</p><p>08:31 - Weight &amp; Biases</p><p>09:05 - Image Analysis Compared</p><p>12:51 - More Benchmarks and Safety</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16230265-o1-pro-mode-full-analysis-plus-o1-paper-highlights.mp3" length="12063992" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/oem47q6q6czinhaxdb723cv9ive9?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16230265</guid>
    <pubDate>Thu, 05 Dec 2024 23:00:00 +0000</pubDate>
    <itunes:duration>1003</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>6</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>AI Breaks Its Silence: OpenAI’s ‘Next 12 Days’, Genie 2, and a Word of Caution</itunes:title>
    <title>AI Breaks Its Silence: OpenAI’s ‘Next 12 Days’, Genie 2, and a Word of Caution</title>
    <itunes:summary><![CDATA[Calmest before the storm? Whatever analogy you want to use things had gotten quiet toward the end of 2024. But then tonight we got Genie 2, and a series of scheduled announcements from OpenAI. Sora is soon here, and o1, but I dive deeper into what it all means and whether reliability is on a path to being solved, ft: two recent papers.  Assembly AI Speech to Text: https://www.assemblyai.com/?utm_source=youtube&amp;utm_medium=influencer&amp;utm_campaign=ai_explained  Plus Kling Motio...]]></itunes:summary>
    <description><![CDATA[<p><b>Calmest before the storm? Whatever analogy you want to use things had gotten quiet toward the end of 2024. But then tonight we got Genie 2, and a series of scheduled announcements from OpenAI. Sora is soon here, and o1, but I dive deeper into what it all means and whether reliability is on a path to being solved, ft: two recent papers. </b></p><p><b>Assembly AI Speech to Text: https://www.assemblyai.com/?utm_source=youtube&amp;utm_medium=influencer&amp;utm_campaign=ai_explained </b></p><p><b>Plus Kling Motion Brush, Simple Bench QwQ update and much more.</b></p><p><br/><b>Genie 2: https://deepmind.google/discover/blog/genie-2-a-large-scale-foundation-world-model/</b></p><p><b>Jim Cramer: https://x.com/jimcramer/status/1864068878692675625</b></p><p><b>Give Us Full o1: </b><a href='https://x.com/tszzl/status/1863882905422106851'><b>https://x.com/tszzl/status/1863882905422106851</b></a></p><p><b>Verge Scoop: </b><a href='https://x.com/tomwarren/status/1864326361415925861'><b>https://x.com/tomwarren/status/1864326361415925861</b></a></p><p><b>O1 Learning to Reason Benchmarks: https://openai.com/index/learning-to-reason-with-llms/</b></p><p><b>SIMA AI: </b><a href='https://arxiv.org/pdf/2404.10179'><b>https://arxiv.org/pdf/2404.10179</b></a></p><p><b>Genie Paper: </b><a href='https://arxiv.org/pdf/2402.15391'><b>https://arxiv.org/pdf/2402.15391</b></a></p><p><b>My Video on Genie: </b><a href='https://www.youtube.com/watch?v=gGKsfXkSXv8'><b>https://www.youtube.com/watch?v=gGKsfXkSXv8</b></a></p><p><b>Oasis Minecraft: https://x.com/risphereeditor/status/1852619965511204974</b></p><p><b>LLMs Procedural Knowledge Paper: </b><a href='https://arxiv.org/pdf/2411.12580'><b>https://arxiv.org/pdf/2411.12580</b></a></p><p><b>Bag of Heuristics Paper: </b><a href='https://arxiv.org/pdf/2410.21272'><b>https://arxiv.org/pdf/2410.21272</b></a></p><p><b>Jensen Huang Hallucinations: </b><a href='https://www.tomshardware.com/tech-industry/artificial-intelligence/jensen-says-we-are-several-years-away-from-solving-the-ai-hallucination-problem-in-the-meantime-we-have-to-keep-increasing-our-computation'><b>https://www.tomshardware.com/tech-industry/artificial-intelligence/jensen-says-we-are-several-years-away-from-solving-the-ai-hallucination-problem-in-the-meantime-we-have-to-keep-increasing-our-computation</b></a></p><p><b>DeepSeek Interview: https://www.chinatalk.media/p/deepseek-ceo-interview-with-chinas</b></p><p><b>Kling Motion Brush: </b><a href='https://klingai.com/image-to-video'><b>https://klingai.com/image-to-video</b></a></p><p><br/></p><p><b>Tim Rocktaschel Book: </b><a href='https://geni.us/ArtificialIntelligence'><b>https://geni.us/ArtificialIntelligence</b></a></p><p><br/></p><p><b>00:43 - OpenAI 12 Days, Sora Turbo, o1</b></p><p><b>03:06 - Genie 2</b></p><p><b>08:26 - Jensen Huang and Altman Hallucination Predictions</b></p><p><b>09:45 - Bag of Heuristics Paper</b></p><p><b>11:40 - Procedural Knowledge Paper<br/>13:02 - AssemblyAI Universal 2</b></p><p><b>13:45 - SimpleBench QwQ and Chinese Models</b></p><p><b>14:42 - Kling Motion Brush</b></p><p><br/><br/></p>]]></description>
    <content:encoded><![CDATA[<p><b>Calmest before the storm? Whatever analogy you want to use things had gotten quiet toward the end of 2024. But then tonight we got Genie 2, and a series of scheduled announcements from OpenAI. Sora is soon here, and o1, but I dive deeper into what it all means and whether reliability is on a path to being solved, ft: two recent papers. </b></p><p><b>Assembly AI Speech to Text: https://www.assemblyai.com/?utm_source=youtube&amp;utm_medium=influencer&amp;utm_campaign=ai_explained </b></p><p><b>Plus Kling Motion Brush, Simple Bench QwQ update and much more.</b></p><p><br/><b>Genie 2: https://deepmind.google/discover/blog/genie-2-a-large-scale-foundation-world-model/</b></p><p><b>Jim Cramer: https://x.com/jimcramer/status/1864068878692675625</b></p><p><b>Give Us Full o1: </b><a href='https://x.com/tszzl/status/1863882905422106851'><b>https://x.com/tszzl/status/1863882905422106851</b></a></p><p><b>Verge Scoop: </b><a href='https://x.com/tomwarren/status/1864326361415925861'><b>https://x.com/tomwarren/status/1864326361415925861</b></a></p><p><b>O1 Learning to Reason Benchmarks: https://openai.com/index/learning-to-reason-with-llms/</b></p><p><b>SIMA AI: </b><a href='https://arxiv.org/pdf/2404.10179'><b>https://arxiv.org/pdf/2404.10179</b></a></p><p><b>Genie Paper: </b><a href='https://arxiv.org/pdf/2402.15391'><b>https://arxiv.org/pdf/2402.15391</b></a></p><p><b>My Video on Genie: </b><a href='https://www.youtube.com/watch?v=gGKsfXkSXv8'><b>https://www.youtube.com/watch?v=gGKsfXkSXv8</b></a></p><p><b>Oasis Minecraft: https://x.com/risphereeditor/status/1852619965511204974</b></p><p><b>LLMs Procedural Knowledge Paper: </b><a href='https://arxiv.org/pdf/2411.12580'><b>https://arxiv.org/pdf/2411.12580</b></a></p><p><b>Bag of Heuristics Paper: </b><a href='https://arxiv.org/pdf/2410.21272'><b>https://arxiv.org/pdf/2410.21272</b></a></p><p><b>Jensen Huang Hallucinations: </b><a href='https://www.tomshardware.com/tech-industry/artificial-intelligence/jensen-says-we-are-several-years-away-from-solving-the-ai-hallucination-problem-in-the-meantime-we-have-to-keep-increasing-our-computation'><b>https://www.tomshardware.com/tech-industry/artificial-intelligence/jensen-says-we-are-several-years-away-from-solving-the-ai-hallucination-problem-in-the-meantime-we-have-to-keep-increasing-our-computation</b></a></p><p><b>DeepSeek Interview: https://www.chinatalk.media/p/deepseek-ceo-interview-with-chinas</b></p><p><b>Kling Motion Brush: </b><a href='https://klingai.com/image-to-video'><b>https://klingai.com/image-to-video</b></a></p><p><br/></p><p><b>Tim Rocktaschel Book: </b><a href='https://geni.us/ArtificialIntelligence'><b>https://geni.us/ArtificialIntelligence</b></a></p><p><br/></p><p><b>00:43 - OpenAI 12 Days, Sora Turbo, o1</b></p><p><b>03:06 - Genie 2</b></p><p><b>08:26 - Jensen Huang and Altman Hallucination Predictions</b></p><p><b>09:45 - Bag of Heuristics Paper</b></p><p><b>11:40 - Procedural Knowledge Paper<br/>13:02 - AssemblyAI Universal 2</b></p><p><b>13:45 - SimpleBench QwQ and Chinese Models</b></p><p><b>14:42 - Kling Motion Brush</b></p><p><br/><br/></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16227884-ai-breaks-its-silence-openai-s-next-12-days-genie-2-and-a-word-of-caution.mp3" length="11191886" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/eenv079n0botu54p759i8wtz3c6d?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16227884</guid>
    <pubDate>Thu, 05 Dec 2024 17:00:00 +0000</pubDate>
    <itunes:duration>929</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>5</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>New Google Model Ranked ‘No. 1 LLM’, But There’s a Problem</itunes:title>
    <title>New Google Model Ranked ‘No. 1 LLM’, But There’s a Problem</title>
    <itunes:summary><![CDATA[A new and mysterious Gemini model appears at the top of the leaderboard, but is that the full story? I dig behind the headline to show you some anti-climactic results, give some context with leaks in the last 48 hours of diminishing returns to scaling, and add the response of Altman, OpenAI and co. The future is about to look a lot stranger...   80,000 hours Podcast and Channel: https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib  https://www.youtube.com/@eightythousandhours/videos  &nb...]]></itunes:summary>
    <description><![CDATA[<p>A new and mysterious Gemini model appears at the top of the leaderboard, but is that the full story? I dig behind the headline to show you some anti-climactic results, give some context with leaks in the last 48 hours of diminishing returns to scaling, and add the response of Altman, OpenAI and co. The future is about to look a lot stranger...<br/><br/><br/>80,000 hours Podcast and Channel: <a href='https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib'>https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib</a><br/> <a href='https://www.youtube.com/@eightythousandhours/videos'>https://www.youtube.com/@eightythousandhours/videos</a>          </p><p> </p><p>You can now gift memberships to AI Insiders (my Patreon w/ exclusive vids, network): <a href='https://www.patreon.com/AIExplained/gift'>https://www.patreon.com/AIExplained/gift</a><br/><br/> <br/> ‘There is no wall’: <a href='https://x.com/sama/status/1856941766915641580'>https://x.com/sama/status/1856941766915641580</a></p><p>https://x.com/vedantmisra/status/1857148554105544708</p><p>Gemini Ranking: https://lmarena.ai/?leaderboard</p><p>API not yet up: <a href='https://x.com/OfficialLoganK/status/1857106844805681153'>https://x.com/OfficialLoganK/status/1857106844805681153</a></p><p>‘Just Die Chat’: <a href='https://x.com/koltregaskes/status/1856754648146653428'>https://x.com/koltregaskes/status/1856754648146653428</a></p><p>Google CEO tweet: <a href='https://x.com/sundarpichai/status/1857114106928718329'>https://x.com/sundarpichai/status/1857114106928718329</a></p><p>Sutskever Quote: <a href='https://www.reuters.com/technology/artificial-intelligence/openai-rivals-seek-new-path-smarter-ai-current-methods-hit-limitations-2024-11-11/'>https://www.reuters.com/technology/artificial-intelligence/openai-rivals-seek-new-path-smarter-ai-current-methods-hit-limitations-2024-11-11/</a></p><p>Another OpenAI Staffer Leaves: <a href='https://x.com/RichardMCNgo/status/1856843040427839804'>https://x.com/RichardMCNgo/status/1856843040427839804</a></p><p>Bloomberg Report: <a href='https://www.bloomberg.com/news/articles/2024-11-13/openai-google-and-anthropic-are-struggling-to-build-more-advanced-ai?s=09'>https://www.bloomberg.com/news/articles/2024-11-13/openai-google-and-anthropic-are-struggling-to-build-more-advanced-ai?s=09</a></p><p>Noam Brown on what OpenAI Researchers Believe: <a href='https://x.com/polynoamial/status/1855037689533178289'>https://x.com/polynoamial/status/1855037689533178289</a></p><p>Clive Chan: https://x.com/itsclivetime/status/1855704120495329667</p><p>Chollet Responds to Altman: <a href='https://x.com/fchollet/status/1857060079586975852'>https://x.com/fchollet/status/1857060079586975852</a></p><p><a href='https://x.com/sama/status/1856940152460869718'>https://x.com/sama/status/1856940152460869718</a></p><p>Altman Emails: <a href='https://x.com/TechEmails/status/1857285960997712356'>https://x.com/TechEmails/status/1857285960997712356</a></p><p>Change of Heart: https://sd11.senate.ca.gov/news/senator-wiener-responds-openai-opposition-sb-1047</p><p>Amodei on ‘Empirical Regularities’: <a href='https://lexfridman.com/dario-amodei-transcript/'>https://lexfridman.com/dario-amodei-transcript/</a></p><p>Verge Report: <a href='https://www.theverge.com/2024/10/25/24279600/google-next-gemini-ai-model-openai-december'>https://www.theverge.com/2024/10/25/24279600/google-next-gemini-ai-model-openai-december</a></p><p>OpenAI Agents in January: https://www.bloomberg.com/news/articles/2024-11-13/openai-nears-launch-of-ai-agents-to-automate-tasks-for-users?srnd=phx-ai</p>]]></description>
    <content:encoded><![CDATA[<p>A new and mysterious Gemini model appears at the top of the leaderboard, but is that the full story? I dig behind the headline to show you some anti-climactic results, give some context with leaks in the last 48 hours of diminishing returns to scaling, and add the response of Altman, OpenAI and co. The future is about to look a lot stranger...<br/><br/><br/>80,000 hours Podcast and Channel: <a href='https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib'>https://open.spotify.com/show/2WzJwXWBDnn4iZ7odKwDib</a><br/> <a href='https://www.youtube.com/@eightythousandhours/videos'>https://www.youtube.com/@eightythousandhours/videos</a>          </p><p> </p><p>You can now gift memberships to AI Insiders (my Patreon w/ exclusive vids, network): <a href='https://www.patreon.com/AIExplained/gift'>https://www.patreon.com/AIExplained/gift</a><br/><br/> <br/> ‘There is no wall’: <a href='https://x.com/sama/status/1856941766915641580'>https://x.com/sama/status/1856941766915641580</a></p><p>https://x.com/vedantmisra/status/1857148554105544708</p><p>Gemini Ranking: https://lmarena.ai/?leaderboard</p><p>API not yet up: <a href='https://x.com/OfficialLoganK/status/1857106844805681153'>https://x.com/OfficialLoganK/status/1857106844805681153</a></p><p>‘Just Die Chat’: <a href='https://x.com/koltregaskes/status/1856754648146653428'>https://x.com/koltregaskes/status/1856754648146653428</a></p><p>Google CEO tweet: <a href='https://x.com/sundarpichai/status/1857114106928718329'>https://x.com/sundarpichai/status/1857114106928718329</a></p><p>Sutskever Quote: <a href='https://www.reuters.com/technology/artificial-intelligence/openai-rivals-seek-new-path-smarter-ai-current-methods-hit-limitations-2024-11-11/'>https://www.reuters.com/technology/artificial-intelligence/openai-rivals-seek-new-path-smarter-ai-current-methods-hit-limitations-2024-11-11/</a></p><p>Another OpenAI Staffer Leaves: <a href='https://x.com/RichardMCNgo/status/1856843040427839804'>https://x.com/RichardMCNgo/status/1856843040427839804</a></p><p>Bloomberg Report: <a href='https://www.bloomberg.com/news/articles/2024-11-13/openai-google-and-anthropic-are-struggling-to-build-more-advanced-ai?s=09'>https://www.bloomberg.com/news/articles/2024-11-13/openai-google-and-anthropic-are-struggling-to-build-more-advanced-ai?s=09</a></p><p>Noam Brown on what OpenAI Researchers Believe: <a href='https://x.com/polynoamial/status/1855037689533178289'>https://x.com/polynoamial/status/1855037689533178289</a></p><p>Clive Chan: https://x.com/itsclivetime/status/1855704120495329667</p><p>Chollet Responds to Altman: <a href='https://x.com/fchollet/status/1857060079586975852'>https://x.com/fchollet/status/1857060079586975852</a></p><p><a href='https://x.com/sama/status/1856940152460869718'>https://x.com/sama/status/1856940152460869718</a></p><p>Altman Emails: <a href='https://x.com/TechEmails/status/1857285960997712356'>https://x.com/TechEmails/status/1857285960997712356</a></p><p>Change of Heart: https://sd11.senate.ca.gov/news/senator-wiener-responds-openai-opposition-sb-1047</p><p>Amodei on ‘Empirical Regularities’: <a href='https://lexfridman.com/dario-amodei-transcript/'>https://lexfridman.com/dario-amodei-transcript/</a></p><p>Verge Report: <a href='https://www.theverge.com/2024/10/25/24279600/google-next-gemini-ai-model-openai-december'>https://www.theverge.com/2024/10/25/24279600/google-next-gemini-ai-model-openai-december</a></p><p>OpenAI Agents in January: https://www.bloomberg.com/news/articles/2024-11-13/openai-nears-launch-of-ai-agents-to-automate-tasks-for-users?srnd=phx-ai</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16114166-new-google-model-ranked-no-1-llm-but-there-s-a-problem.mp3" length="11071159" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/xqemwyjgfvt8lsdojtobesoz15aq?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16114166</guid>
    <pubDate>Fri, 15 Nov 2024 17:00:00 +0000</pubDate>
    <itunes:duration>919</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>4</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>Leak: ‘GPT-5 exhibits diminishing returns’, Sam Altman: ‘lol’ </itunes:title>
    <title>Leak: ‘GPT-5 exhibits diminishing returns’, Sam Altman: ‘lol’ </title>
    <itunes:summary><![CDATA[The last few days have seen two narratives emerge. One, derived from yesterday’s OpenAI leak in TheInformation, that GPT-5/Orion is a disappointment, and less of a leap than GPT-3 to GPT-4. The second comes from a series of 4 clips (shown in this video) from Sam Altman, regarding the ‘clear path’ to AGI. Let’s go beyond the headlines (and through papers like Frontier Math) to get closer to the ground truth…    Plus Universal-2, Sora comments, Claude 3.5 Haiku SimpleBench update, and...]]></itunes:summary>
    <description><![CDATA[<p>The last few days have seen two narratives emerge. One, derived from yesterday’s OpenAI leak in TheInformation, that GPT-5/Orion is a disappointment, and less of a leap than GPT-3 to GPT-4. The second comes from a series of 4 clips (shown in this video) from Sam Altman, regarding the ‘clear path’ to AGI. Let’s go beyond the headlines (and through papers like Frontier Math) to get closer to the ground truth…<br/> <br/> Plus Universal-2, Sora comments, Claude 3.5 Haiku SimpleBench update, and a great new AI video.<br/><br/><br/>Assembly AI Speech to Text: https://www.assemblyai.com/?utm_source=youtube&amp;utm_medium=influencer&amp;utm_campaign=ai_explained </p><p> </p><p>00:39 – Bear Case, TheInformation Leak</p><p>04:01 – Bull Case, Sam Altman</p><p>06:20 – FrontierMath</p><p>11:29 – o1 Paradigm</p><p>13:11 – Text to Video Greatness and Universal-2 </p><p> </p><p>TheInformation Leak: <a href='https://www.theinformation.com/articles/openai-shifts-strategy-as-rate-of-gpt-ai-improvements-slows?rc=sy0ihq'>https://www.theinformation.com/articles/openai-shifts-strategy-as-rate-of-gpt-ai-improvements-slows?rc=sy0ihq</a></p><p>Noam Brown Replies: <a href='https://x.com/polynoamial/status/1855453104394637444'>https://x.com/polynoamial/status/1855453104394637444</a></p><p>Sam Altman Y-Combinator Interview: <a href='https://www.youtube.com/watch?v=xXCBz_8hM9w&amp;t=1556s'>https://www.youtube.com/watch?v=xXCBz_8hM9w&amp;t=1556s</a></p><p>Altman Reply: <a href='https://x.com/sama/status/1855100359511097828'>https://x.com/sama/status/1855100359511097828</a></p><p><a href='https://simple-bench.com/'>https://simple-bench.com/</a></p><p>FrontierMath Paper: https://arxiv.org/pdf/2411.04872</p><p>Frontier Math Blog Post: <a href='https://epochai.org/frontiermath'>https://epochai.org/frontiermath</a></p><p>Tao: https://x.com/EpochAIResearch/status/1854996368814936250</p><p>MMLU Are We Done (cites me!): <a href='https://arxiv.org/pdf/2406.04127'>https://arxiv.org/pdf/2406.04127</a></p><p>Universal-2 <a href='https://www.assemblyai.com/research/universal-2'>https://www.assemblyai.com/research/universal-2</a></p><p>Noam Brown ‘We don’t know’: https://www.youtube.com/watch?v=Gr_eYXdHFis</p><p>Anthropic Founder Response: https://x.com/jackclarkSF/status/1855485569998217231</p><p>Sora (Runway Comment): <a href='https://x.com/c_valenzuelab/status/1855026417354129455'>https://x.com/c_valenzuelab/status/1855026417354129455</a></p><p>Sora New Vid: https://www.youtube.com/watch?v=_iETa2KDRuw</p><p>Darri3D Video: https://www.reddit.com/r/ChatGPT/comments/1gn0n3z/can_you/</p>]]></description>
    <content:encoded><![CDATA[<p>The last few days have seen two narratives emerge. One, derived from yesterday’s OpenAI leak in TheInformation, that GPT-5/Orion is a disappointment, and less of a leap than GPT-3 to GPT-4. The second comes from a series of 4 clips (shown in this video) from Sam Altman, regarding the ‘clear path’ to AGI. Let’s go beyond the headlines (and through papers like Frontier Math) to get closer to the ground truth…<br/> <br/> Plus Universal-2, Sora comments, Claude 3.5 Haiku SimpleBench update, and a great new AI video.<br/><br/><br/>Assembly AI Speech to Text: https://www.assemblyai.com/?utm_source=youtube&amp;utm_medium=influencer&amp;utm_campaign=ai_explained </p><p> </p><p>00:39 – Bear Case, TheInformation Leak</p><p>04:01 – Bull Case, Sam Altman</p><p>06:20 – FrontierMath</p><p>11:29 – o1 Paradigm</p><p>13:11 – Text to Video Greatness and Universal-2 </p><p> </p><p>TheInformation Leak: <a href='https://www.theinformation.com/articles/openai-shifts-strategy-as-rate-of-gpt-ai-improvements-slows?rc=sy0ihq'>https://www.theinformation.com/articles/openai-shifts-strategy-as-rate-of-gpt-ai-improvements-slows?rc=sy0ihq</a></p><p>Noam Brown Replies: <a href='https://x.com/polynoamial/status/1855453104394637444'>https://x.com/polynoamial/status/1855453104394637444</a></p><p>Sam Altman Y-Combinator Interview: <a href='https://www.youtube.com/watch?v=xXCBz_8hM9w&amp;t=1556s'>https://www.youtube.com/watch?v=xXCBz_8hM9w&amp;t=1556s</a></p><p>Altman Reply: <a href='https://x.com/sama/status/1855100359511097828'>https://x.com/sama/status/1855100359511097828</a></p><p><a href='https://simple-bench.com/'>https://simple-bench.com/</a></p><p>FrontierMath Paper: https://arxiv.org/pdf/2411.04872</p><p>Frontier Math Blog Post: <a href='https://epochai.org/frontiermath'>https://epochai.org/frontiermath</a></p><p>Tao: https://x.com/EpochAIResearch/status/1854996368814936250</p><p>MMLU Are We Done (cites me!): <a href='https://arxiv.org/pdf/2406.04127'>https://arxiv.org/pdf/2406.04127</a></p><p>Universal-2 <a href='https://www.assemblyai.com/research/universal-2'>https://www.assemblyai.com/research/universal-2</a></p><p>Noam Brown ‘We don’t know’: https://www.youtube.com/watch?v=Gr_eYXdHFis</p><p>Anthropic Founder Response: https://x.com/jackclarkSF/status/1855485569998217231</p><p>Sora (Runway Comment): <a href='https://x.com/c_valenzuelab/status/1855026417354129455'>https://x.com/c_valenzuelab/status/1855026417354129455</a></p><p>Sora New Vid: https://www.youtube.com/watch?v=_iETa2KDRuw</p><p>Darri3D Video: https://www.reddit.com/r/ChatGPT/comments/1gn0n3z/can_you/</p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16079562-leak-gpt-5-exhibits-diminishing-returns-sam-altman-lol.mp3" length="11365746" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/opeckn58mm2p0xwy6qsfaefzhwz9?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16079562</guid>
    <pubDate>Sun, 10 Nov 2024 18:00:00 +0000</pubDate>
    <itunes:duration>944</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>3</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>ChatGPT with Search, Altman Answers Anything and Simple Bench Out</itunes:title>
    <title>ChatGPT with Search, Altman Answers Anything and Simple Bench Out</title>
    <itunes:summary><![CDATA[The Google destroyer, the Perplexity crusher? Or just hype? ChatGPT with Search is here, and simultaneously Altman and co did an AMA on Reddit, covering GPT-5, Sora, SearchGPT and a lot more. Plus, the biggest news of them all: Simple Bench is out.  ChatGPT with Search: https://openai.com/index/introducing-chatgpt-search/ Altman AMA (ask me anything): https://www.reddit.com/r/ChatGPT/comments/1ggixzy/ama_with_openais_sam_altman_kevin_weil_srinivas/ https://x.com/sama/status/185204107579352291...]]></itunes:summary>
    <description><![CDATA[<p>The Google destroyer, the Perplexity crusher? Or just hype? ChatGPT with Search is here, and simultaneously Altman and co did an AMA on Reddit, covering GPT-5, Sora, SearchGPT and a lot more. Plus, the biggest news of them all: Simple Bench is out.<br/><br/>ChatGPT with Search: https://openai.com/index/introducing-chatgpt-search/</p><p>Altman AMA (ask me anything): <a href='https://www.reddit.com/r/ChatGPT/comments/1ggixzy/ama_with_openais_sam_altman_kevin_weil_srinivas/'>https://www.reddit.com/r/ChatGPT/comments/1ggixzy/ama_with_openais_sam_altman_kevin_weil_srinivas/</a></p><p>https://x.com/sama/status/1852041075793522911</p><p>Perplexity Ads: <a href='https://www.cnbc.com/2024/08/22/perplexity-ai-plans-to-start-running-search-ads-in-fourth-quarter.html'>https://www.cnbc.com/2024/08/22/perplexity-ai-plans-to-start-running-search-ads-in-fourth-quarter.html</a></p><p>Perplexity: <a href='https://www.perplexity.ai/'>https://www.perplexity.ai/</a></p><p><a href='https://simple-bench.com/'>https://simple-bench.com/</a></p>]]></description>
    <content:encoded><![CDATA[<p>The Google destroyer, the Perplexity crusher? Or just hype? ChatGPT with Search is here, and simultaneously Altman and co did an AMA on Reddit, covering GPT-5, Sora, SearchGPT and a lot more. Plus, the biggest news of them all: Simple Bench is out.<br/><br/>ChatGPT with Search: https://openai.com/index/introducing-chatgpt-search/</p><p>Altman AMA (ask me anything): <a href='https://www.reddit.com/r/ChatGPT/comments/1ggixzy/ama_with_openais_sam_altman_kevin_weil_srinivas/'>https://www.reddit.com/r/ChatGPT/comments/1ggixzy/ama_with_openais_sam_altman_kevin_weil_srinivas/</a></p><p>https://x.com/sama/status/1852041075793522911</p><p>Perplexity Ads: <a href='https://www.cnbc.com/2024/08/22/perplexity-ai-plans-to-start-running-search-ads-in-fourth-quarter.html'>https://www.cnbc.com/2024/08/22/perplexity-ai-plans-to-start-running-search-ads-in-fourth-quarter.html</a></p><p>Perplexity: <a href='https://www.perplexity.ai/'>https://www.perplexity.ai/</a></p><p><a href='https://simple-bench.com/'>https://simple-bench.com/</a></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16028873-chatgpt-with-search-altman-answers-anything-and-simple-bench-out.mp3" length="11068903" type="audio/mpeg" />
    <itunes:image href="https://storage.buzzsprout.com/k8s3g7ilo1l8bcruw2v3spavp376?.jpg" />
    <itunes:author>Philip - Host of AI Explained YT</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16028873</guid>
    <pubDate>Fri, 01 Nov 2024 00:00:00 +0000</pubDate>
    <itunes:duration>920</itunes:duration>
    <itunes:keywords></itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>2</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
  <item>
    <itunes:title>The New Claude 3.5 Sonnet: Better, Yes, But Not Just in the Way You Might Think</itunes:title>
    <title>The New Claude 3.5 Sonnet: Better, Yes, But Not Just in the Way You Might Think</title>
    <itunes:summary><![CDATA[A new state of the art LLM (at least for creative writing and basic reasoning) but what lies behind the numbers that were put out? Is it for real, and are AI agents about to grab your mouse and shake your cursor?   Plus, results on my own Simple Bench, and new tools from Runway (Act-One), HeyGen (Zoom Calls) and an updated NotebookLM. AI, without the hype.  Weights and Biases' Weave: https://wandb.me/ai_explained ]]></itunes:summary>
    <description><![CDATA[<p>A new state of the art LLM (at least for creative writing and basic reasoning) but what lies behind the numbers that were put out? Is it for real, and are AI agents about to grab your mouse and shake your cursor? <br/><br/>Plus, results on my own Simple Bench, and new tools from Runway (Act-One), HeyGen (Zoom Calls) and an updated NotebookLM. AI, without the hype.<br/><br/>Weights and Biases&apos; Weave: <a href='https://www.youtube.com/redirect?event=video_description&amp;redir_token=QUFFLUhqa1NXX09EcU44OVR5emQwV3ZOVVpqY21UNlg1QXxBQ3Jtc0ttVjdGUXRJSWJTQVlZUkoxU0t4YVRtRXlILTRlTVMyUEZrRzB2aE8tSDRfSkEtUlVleFhMblM1Mlc5ZjZLVFN5TDFya01TSm9SdVY2NHZMdFE2NFctZUhpYTNHcjliWlAteFZCZFUxWWNNa3d1X0ZSUQ&amp;q=https%3A%2F%2Fwandb.me%2Fai_explained&amp;v=KngdLKv9RAc'>https://wandb.me/ai_explained</a></p>]]></description>
    <content:encoded><![CDATA[<p>A new state of the art LLM (at least for creative writing and basic reasoning) but what lies behind the numbers that were put out? Is it for real, and are AI agents about to grab your mouse and shake your cursor? <br/><br/>Plus, results on my own Simple Bench, and new tools from Runway (Act-One), HeyGen (Zoom Calls) and an updated NotebookLM. AI, without the hype.<br/><br/>Weights and Biases&apos; Weave: <a href='https://www.youtube.com/redirect?event=video_description&amp;redir_token=QUFFLUhqa1NXX09EcU44OVR5emQwV3ZOVVpqY21UNlg1QXxBQ3Jtc0ttVjdGUXRJSWJTQVlZUkoxU0t4YVRtRXlILTRlTVMyUEZrRzB2aE8tSDRfSkEtUlVleFhMblM1Mlc5ZjZLVFN5TDFya01TSm9SdVY2NHZMdFE2NFctZUhpYTNHcjliWlAteFZCZFUxWWNNa3d1X0ZSUQ&amp;q=https%3A%2F%2Fwandb.me%2Fai_explained&amp;v=KngdLKv9RAc'>https://wandb.me/ai_explained</a></p>]]></content:encoded>
    <enclosure url="https://www.buzzsprout.com/2418777/episodes/16003023-the-new-claude-3-5-sonnet-better-yes-but-not-just-in-the-way-you-might-think.mp3" length="16281565" type="audio/mpeg" />
    <link>https://www.patreon.com/AIExplained/membership</link>
    <itunes:image href="https://storage.buzzsprout.com/s1r9737mf3m3totz2e2nt90tqubx?.jpg" />
    <itunes:author>AI Explained - Hosted by Philip</itunes:author>
    <guid isPermaLink="false">Buzzsprout-16003023</guid>
    <pubDate>Mon, 28 Oct 2024 13:00:00 +0000</pubDate>
    <itunes:duration>1354</itunes:duration>
    <itunes:keywords>Claude 3.5 Sonnet, New Claude, Artificial Intelligence, ChatGPT, Anthropic, AI Bubble, AI, AI Explained, Dario Amodei, Sam Altman, HeyGen, Runway, NotebookLM</itunes:keywords>
    <itunes:season>1</itunes:season>
    <itunes:episode>1</itunes:episode>
    <itunes:episodeType>full</itunes:episodeType>
    <itunes:explicit>false</itunes:explicit>
  </item>
</channel>
</rss>
