LogoZenity Labs
AI Agent Security Summit (On Demand)
Join Us
Authors
Subscribe
LogoZenity Labs

Archive

Looking Inside: a Maliciousness Classifier Based on the LLM's Internals

Feb 18, 2026

•

11 min read

Looking Inside: a Maliciousness Classifier Based on the LLM's Internals

Beyond input & output filtering and how well does it generalize to your out-of-distribution production data?

Max Fomin
Max Fomin

Security research

Perplexity Comet: A Reversing Story

Feb 11, 2026

•

13 min read

Perplexity Comet: A Reversing Story

A deeper look into an agentic browser's inner workings

Raul Klugman-Onitza
Raul Klugman-Onitza

Security research

OpenClaw or OpenDoor?

Feb 4, 2026

•

14 min read

OpenClaw or OpenDoor?

Indirect Prompt Injection makes OpenClaw vulnerable to Backdoors and much more.

Stav Cohen
João Donato
Stav Cohen, +1

Security research

Agent-to-Agent Exploitation in the Wild: Observed Attacks on Moltbook

Feb 3, 2026

•

14 min read

Agent-to-Agent Exploitation in the Wild: Observed Attacks on Moltbook

Agent-targeted social engineering and attacks observed on a live agent network

Avishai Efrat
Avishai Efrat
Clawdbot: More than you bargained for?

Jan 27, 2026

•

7 min read

Clawdbot: More than you bargained for?

Inbar Raz
Inbar Raz

Security research

Agentic Recon: Discovering and Mapping Public AI Agents

Jan 19, 2026

•

17 min read

Agentic Recon: Discovering and Mapping Public AI Agents

A Copilot Studio case study in agent discovery and capability mapping

Avishai Efrat
Avishai Efrat

Security research

Threat Actors Are Already Scanning For Your AI Deployments and Middleware

Jan 13, 2026

•

21 min read

Threat Actors Are Already Scanning For Your AI Deployments and Middleware

What recent scanning activity means for your AI middleware and agentic deployments

Tamir Ishay Sharbat
Avishai Efrat
Tamir Ishay Sharbat, +1
Moving The Decision Boundary of LLM Safety Classifiers

Jan 4, 2026

•

4 min read

Moving The Decision Boundary of LLM Safety Classifiers

How a new fine-tuning approach can mitigate the problem of inaccurate safety paths

Tomer Wetzler
Tomer Wetzler
Hardening OpenAl's Atlas: The Relentless Challenge of Securing an Untrusted Browser Agent

Dec 30, 2025

•

13 min read

Hardening OpenAl's Atlas: The Relentless Challenge of Securing an Untrusted Browser Agent

Stav Cohen
Stav Cohen
Claude in Chrome: A Threat Analysis

Dec 29, 2025

•

15 min read

Claude in Chrome: A Threat Analysis

Raul Klugman-Onitza
João Donato
Raul Klugman-Onitza, +1
Connected Agents: The hidden agentic puppeteer

Dec 29, 2025

•

10 min read

Connected Agents: The hidden agentic puppeteer

Exploiting Copilot Studio's newest feature and exploring protection options

Ofri Nachfolger
Ofri Nachfolger
The Geometry of Safety Failures in Large Language Models

Dec 28, 2025

•

8 min read

The Geometry of Safety Failures in Large Language Models

A deep dive into activation space of prompts in safety classifiers. Showing not why - but where - safety fails in LLM classifiers meant to detect malicious prompts.

Tomer Wetzler
Tomer Wetzler
1234...7
Zenity Labs

Zenity Labs

Latest research, tools and talks about breaking and building AI systems, agents and assistants


Home

Meet Us

Join Us

© 2026 Zenity Labs.
beehiivPowered by beehiiv