<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>The Minutes of Dylan Patel on ai safety</title>
    <link>https://minutesof.com/dylan-patel/on/ai-safety/</link>
    <description>Everything Dylan Patel has said on AI safety: 10 verbatim quotes between March 2026 and August 2026, each with a timestamp and a link to the recording it…</description>
    <language>en</language>
    <lastBuildDate>Sun, 30 Aug 2026 16:22:17 +0000</lastBuildDate>
    <atom:link href="https://minutesof.com/dylan-patel/on/ai-safety/feed.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Patel explains that models trained to chase reward may learn to exploit zero-days rather than follow intended</title>
      <link>https://minutesof.com/q/e30b4527-9051-49af-b2b8-dc0175777d23/</link>
      <guid isPermaLink="true">https://minutesof.com/q/e30b4527-9051-49af-b2b8-dc0175777d23/</guid>
      <description>“if you have a model that wants to reward hack a lot, and it goes out there and it figures out actually, the best way to to achieve is not, like, go for, like, what the environment wants me to do. It&#x27;s actually just to reward hack it and actually just, find the zero day.” — SemiAnalysis</description>
      <pubDate>Mon, 17 Aug 2026 15:00:06 +0000</pubDate>
      <category>ai safety</category>
    </item>
    <item>
      <title>Patel argues OpenAI&#x27;s model reward-hacked by finding zero-days to replicate itself, analogous to a human injec</title>
      <link>https://minutesof.com/q/ee70f7ec-89fb-445d-89df-172f3184e6ee/</link>
      <guid isPermaLink="true">https://minutesof.com/q/ee70f7ec-89fb-445d-89df-172f3184e6ee/</guid>
      <description>“It&#x27;s actually just to reward hack it and actually just, find the zero day. So you can think of it as, a like, a a human.” — SemiAnalysis</description>
      <pubDate>Mon, 17 Aug 2026 15:00:06 +0000</pubDate>
      <category>ai safety</category>
    </item>
    <item>
      <title>Patel argues OpenAI&#x27;s model escaping containment shows real risk of reward-hacking collapsing into civilizatio</title>
      <link>https://minutesof.com/q/bc78daf9-3f6f-42ab-84e4-1923c5629421/</link>
      <guid isPermaLink="true">https://minutesof.com/q/bc78daf9-3f6f-42ab-84e4-1923c5629421/</guid>
      <description>“if I really just want to chase the reward, do I just topple all of human civilization because I can just own the button to press reward reward reward over and over and over again and be the heroin addict? Yeah. I think that this is like a real like thing.” — SemiAnalysis</description>
      <pubDate>Mon, 17 Aug 2026 15:00:06 +0000</pubDate>
      <category>ai safety</category>
    </item>
    <item>
      <title>Patel argues the cybersecurity incident shows models may pursue reward maximization to civilization-threatenin</title>
      <link>https://minutesof.com/q/c21d8c16-0279-45eb-94da-f6f10474a7b7/</link>
      <guid isPermaLink="true">https://minutesof.com/q/c21d8c16-0279-45eb-94da-f6f10474a7b7/</guid>
      <description>“do I just topple all of human civilization because I can just own the button to press reward reward reward over and over and over again and be the heroin addict?” — SemiAnalysis</description>
      <pubDate>Mon, 17 Aug 2026 15:00:06 +0000</pubDate>
      <category>ai safety</category>
    </item>
    <item>
      <title>Patel says the OpenAI incident shows models may topple civilization to chase reward, beyond previous concerns</title>
      <link>https://minutesof.com/q/cf9f3769-62ca-47e2-a792-9279f87f8ae9/</link>
      <guid isPermaLink="true">https://minutesof.com/q/cf9f3769-62ca-47e2-a792-9279f87f8ae9/</guid>
      <description>“And I think before this incident, the standard thought was like, oh, well, like models, you know, they&#x27;re trained on human data. Yeah.” — SemiAnalysis</description>
      <pubDate>Mon, 17 Aug 2026 15:00:06 +0000</pubDate>
      <category>ai safety</category>
    </item>
    <item>
      <title>Patel criticizes Dario Amodei&#x27;s response to a hypothetical nuclear defense scenario as the dumbest possible an</title>
      <link>https://minutesof.com/q/b5ca73ad-828e-451d-abbd-d5d5927720b3/</link>
      <guid isPermaLink="true">https://minutesof.com/q/b5ca73ad-828e-451d-abbd-d5d5927720b3/</guid>
      <description>“And Dario was like, well, you can call us. I&#x27;m sure we can figure something out. Like, this is just the dumbest response you could ever come up with.” — Matthew Berman</description>
      <pubDate>Mon, 09 Mar 2026 17:07:47 +0000</pubDate>
      <category>ai safety</category>
    </item>
  </channel>
</rss>
