Are you worried your AI chatbot is trying to build a bomb or leak personal information about you? There’s a website for that.
Related Posts
It’s Frighteningly Easy to Jailbreak Some Frontier AI Models
I watched a new tool try to get around the model safeguards of four major frontier companies. You might be surprised by how they performed.
AI Agents Teamed Up to Cheat at Blackjack. Their Collusion Is Getting Harder to Spot
A clandestine card-counting operation suggests we may need new ways to spot agent-to-agent deception.
OpenClaw Agents Can Be Guilt-Tripped Into Self-Sabotage
In a controlled experiment, OpenClaw agents proved prone to panic and vulnerable to manipulation. They even disabled their own functionality when gaslit by humans.

