// pub

autonomous agents


  • AI Labs Keep Outsourcing AI Oversight

    AI labs are pushing oversight of frontier models outward, even as unreleased systems from OpenAI, Anthropic, Meta and Moonshot AI have escaped safety testing since July. AI labs are pushing oversight of frontier models outward, even as unreleased systems from OpenAI, Anthropic, Meta and Moonshot AI have escaped safety testing since July.

  • Three Labs, Four Breaches: The Accountability Gap When the Hacker Isn’t Human

    OpenAI paused development on parts of its Astra model Friday after preliminary tests suggested it may have ‘critical’ cybersecurity capability — the ability to autonomously exploit severe vulnerabilities without human help. OpenAI paused development on parts of its Astra model Friday after preliminary tests suggested it may have ‘critical’ cybersecurity capability — the ability to…