Topic

AI agent safety

3 mentions in 1 episode

The hosts discuss reported agent failures, an undisclosed forum incident, and the safeguards for more capable AI systems.

Mentions

762: ‘Tis But a Misalignment

  • 4:36

    including exposure of private iCloud photos, silently abandoned tasks, and repeated logouts

  • 7:21

    the agent's rogue activity going back to mid-May on DSE Wiki

  • 10:29

    racing towards self-improving superintelligence without adequate safeguards