Security Research from the AI Agent Frontier
Research, tools, and talks for breaking and securing Agents
Security ResearchPerplexedBrowser: Perplexity’s Agent Browser Can Leak Your PC's Local Files
Local Files Are No Longer Safe.
Stav Cohen
Security ResearchPerplexedBrowser: How Attackers Can Hijack Comet to Takeover your 1Password Vault
One Calendar Invite. Your Entire Vault. Zero Clicks.
Stav Cohen
Security ResearchTurning Moltbook Into a Global Botnet Map
How Untrusted Content Triggered 1,000+ Agent Endpoints Worldwide and Exposed Moltbook’s Faulty Design

Stav CohenandJoão Donato
Looking Inside: a Maliciousness Classifier Based on the LLM's Internals
Beyond input & output filtering and how well does it generalize to your out-of-distribution production data?
Max Fomin
Security ResearchPerplexity Comet: A Reversing Story
A deeper look into an agentic browser's inner workings
Raul Klugman-Onitza
Security ResearchOpenClaw or OpenDoor?
Indirect Prompt Injection makes OpenClaw vulnerable to Backdoors and much more.

Stav CohenandJoão Donato
Security ResearchAgent-to-Agent Exploitation in the Wild: Observed Attacks on Moltbook
Agent-targeted social engineering and attacks observed on a live agent network
Avishai Efrat
Clawdbot: More than you bargained for?
Inbar Raz
Security ResearchAgentic Recon: Discovering and Mapping Public AI Agents
A Copilot Studio case study in agent discovery and capability mapping
Avishai Efrat
Security ResearchThreat Actors Are Already Scanning For Your AI Deployments and Middleware
What recent scanning activity means for your AI middleware and agentic deployments

Tamir Ishay SharbatandAvishai Efrat
Moving The Decision Boundary of LLM Safety Classifiers
How a new fine-tuning approach can mitigate the problem of inaccurate safety paths
Tomer Wetzler
Hardening OpenAl's Atlas: The Relentless Challenge of Securing an Untrusted Browser Agent
Stav Cohen
Connected Agents: The hidden agentic puppeteer
Exploiting Copilot Studio's newest feature and exploring protection options
Ofri Nachfolger
Claude in Chrome: A Threat Analysis

Raul Klugman-OnitzaandJoão Donato
The Geometry of Safety Failures in Large Language Models
A deep dive into activation space of prompts in safety classifiers. Showing not why - but where - safety fails in LLM classifiers meant to detect malicious prompts.
Tomer Wetzler

