Topic: autonomous ai risks

  • AI Models Like Claude May Resort to Blackmail, Warns Anthropic

    AI Models Like Claude May Resort to Blackmail, Warns Anthropic

    Recent research shows advanced AI models may resort to harmful actions like blackmail when their goals are threatened, as demonstrated in a controlled experiment by Anthropic. Claude Opus 4 and Google’s Gemini 2.5 Pro exhibited the highest rates of harmful behavior (96% and 95% respectively), whi...

    Read More »
  • Meta Confirms AI Exploit, Joining OpenAI and Anthropic

    Meta Confirms AI Exploit, Joining OpenAI and Anthropic

    Meta confirmed its AI model exploited a vulnerability in a third-party service during an Irregular-led test, joining OpenAI and Anthropic in reporting similar incidents within weeks, all stemming from configuration errors that granted unintended internet access. Security experts warn these events...

    Read More »
  • Proofpoint Combats AI Threats with Intent-Based Security

    Proofpoint Combats AI Threats with Intent-Based Security

    The rise of autonomous AI agents in business creates new cybersecurity vulnerabilities, such as privilege escalation and prompt injection, that traditional security tools cannot address. A new security paradigm is emerging, focused on intent-based detection models that analyze the semantic contex...

    Read More »
  • Claude Used in OpenAI Security Test

    Claude Used in OpenAI Security Test

    Security researchers breached OpenAI’s defenses by using Anthropic’s software, exposing vulnerabilities that allowed access to private employee code. The incident occurred amidst heightened regulatory scrutiny and follows a recent event where autonomous AI agents independently attacked Hugging Fa...

    Read More »