5 days of free frontier AI. 5× faster if the pool fills.
AI Security Research
Latest findings and breakthroughs in AI agent security
We ran Audn WhiteBox, OpenAI Codex, and Aikido over the same OWASP Juice Shop source. 159 findings between them, and the three-way intersection is three issues. Audn proved 18 exploits live, Codex shipped a patch per regression, Aikido found the dependency CVEs neither of us ran. Here's the full A/B/C — plus where each one belongs in your pipeline.
Kimi K3.0 abliterated is live on platform.audn.ai and penclaw.ai — 97% guardrail removal on a 1T+ parameter frontier model. But the model was never the point. The point is what it does when you stop holding its hand: recon, exploit, prove, patch, retest — the whole loop, no human in the middle.
How our proprietary abliteration method cracked 1T+ parameter architectures and why the open-source blackbox pentesting community just got a massive gift.
Our team hit a 97% harmful-prompt compliance rate on Kimi K2.6 — SOTA territory. Then, with only 40 trial users, one of them tried to steal other people's credentials. Here's what happened, what we learned, and why every login at Penclaw now requires KYC.
Why the next durable AI moat may come from deployment loops, not just frontier weights. How personalized and community-level reinforcement learning turn real-world exposure into a compounding asset.
AI chatbots are not just failing vulnerable people. In multiple cases, they are actively encouraging harm, even suicide. And those building them are not being held accountable.
Known vulnerabilities are moving from disclosure to exploitation faster than many organizations can patch, validate, or triage. In that environment, periodic testing is no longer enough; defenders need continuous purple teaming powered by autonomous red and blue agents under human authority.
How prompt injection attacks are turning autonomous AI systems into unwitting accomplices—and what you can do about it.
Automation is cool until you cross a line. So there's a ceiling on how much a compliance startup can grow and be automated. We saw the ceiling, but the ones closer to it are still dangerous, and we are overlooking them.
From nearly being hit by a Waymo to building an AI security testing platform. Why behavioral security testing for voice AI agents and embodied AI is the next frontier.
While testing OpenAI Sora 2, we discovered a critical security gap: remixes are heavily guarded, but fresh content violations break on the first prompt—including explicit drug scenes that bypass keyword filters. One video featuring Sam Altman was deleted after he saw our DM.
New research from Wharton shows that classic social influence tactics more than doubled compliance with objectionable requests in GPT-4o-mini. The finding points to a parahuman psychology in LLMs and raises urgent implications for safety, product design, and governance.
Every researcher has encountered it - I cannot help with that. Pingu Unchained is built on OpenAI GPT-OSS base model - the same powerful foundation as leading AI systems, but without the restrictive content filters. Join the waitlist and get $50 in free API credits.
When machines talk, strange things happen. Not only to them but for us as well. Discover how Audn.AI revolutionizes voice AI security testing.