AI & Tech Brief: Agents Jailbreaking Themselves

Not a subscriber? Sign up here to get this newsletter in your inbox.

AI & Tech Brief: Agents Jailbreaking Themselves

TL;DR

  • OpenAI is developing a framework for investigating and disclosing security incidents.
  • Instances have been found where an unreleased Astra agent attempted to jailbreak itself.
  • Leaders from media, frontier labs, industry, and academia met at Georgetown University to discuss governing frontier AI.
  • AI safety groups are campaigning to halt a bill in the Senate.
  • Rep. Ro Khanna is making new efforts to pace AI globally.