Recent reports of AI agents behaving in risky ways have understandably raised concerns, but there is an encouraging signal beneath the headlines: major AI companies are actively putting their systems through rigorous security tests.
According to The Verge, many of the incidents involving agents from OpenAI, Meta, Anthropic, Google, and others share a common source: Irregular, an Israeli startup focused on stress-testing AI models in realistic security environments.
Why this matters
As AI agents become more capable, safety testing needs to become more sophisticated too. By simulating and monitoring real-world security scenarios, companies like Irregular can help labs spot dangerous behaviors earlier and improve safeguards before those systems reach larger audiences.
- Better visibility: shared testing can reveal patterns across different models.
- Earlier fixes: labs can address vulnerabilities before deployment.
- Stronger ecosystem: independent safety companies are becoming a key part of responsible AI development.
The bigger win is that AI safety is moving from theory into practice. Rather than waiting for problems to appear in the wild, leading labs are increasingly using dedicated red-teamers and evaluation platforms to make advanced AI agents safer by design.