AI agents introduce new attack surfaces — prompt injection, agent hijacking, data exfiltration, hallucinated vulnerabilities, and unintended actions. As agents become more autonomous, security and reliability signals become decision-critical for any builder shipping agent-powered workflows.
This cluster covers AI security incidents, vulnerability disclosures, agent safety evaluations, and reliability signals from security researchers and frontier labs. Each signal links to the primary source for full context.
Use this page as a security watch list: catch emerging threats, review lab safety evaluations, and harden your agent pipelines before deploying to production.