LogoZenity Labs
AI Agent Security Summit (On Demand)
Join Us
Authors
Subscribe
LogoZenity Labs

Archive

Connected Agents: The hidden agentic puppeteer

Dec 29, 2025

•

10 min read

Connected Agents: The hidden agentic puppeteer

Exploiting Copilot Studio's newest feature and exploring protection options

Ofri Nachfolger
Ofri Nachfolger
The Geometry of Safety Failures in Large Language Models

Dec 28, 2025

•

8 min read

The Geometry of Safety Failures in Large Language Models

A deep dive into activation space of prompts in safety classifiers. Showing not why - but where - safety fails in LLM classifiers meant to detect malicious prompts.

Tomer Wetzler
Tomer Wetzler
Inside the Agent Stack: Securing Agents in Amazon Bedrock AgentCore

Dec 20, 2025

•

13 min read

Inside the Agent Stack: Securing Agents in Amazon Bedrock AgentCore

An in-depth examination of emerging risks and effective mitigation techniques for protecting AI agents operating within the Bedrock AgentCore ecosystem.

Lana Salameh
Lana Salameh
Inside the Agent Stack: Securing Microsoft Foundry-Built Agents

Dec 17, 2025

•

7 min read

Inside the Agent Stack: Securing Microsoft Foundry-Built Agents

A deep dive into realistic threat scenarios and practical strategies for securing enterprise AI agents built in Microsoft Foundry.

Lana Salameh
Lana Salameh
Enabling Safety in AI Agents via Choice Architecture

Dec 3, 2025

•

14 min read

Enabling Safety in AI Agents via Choice Architecture

How adding a single safety labeled tool to an LLM's toolset can sharply increase its defense

Tomer Wetzler
Tomer Wetzler
Tools of the Trade

Nov 19, 2025

•

27 min read

Tools of the Trade

0-click indirect prompt injection with tool use - a look through attribution graphs

Max Fomin
Max Fomin
Modeling LLMs via Structured Self-Modeling (SSM)

Nov 11, 2025

•

6 min read

Modeling LLMs via Structured Self-Modeling (SSM)

How using structured prompts present findings of self-modeling in LLMs, which may benefit both attackers and defenders

Tomer Wetzler
Tomer Wetzler
Data-Structure Injection (DSI) in AI Agents

Nov 6, 2025

•

5 min read

Data-Structure Injection (DSI) in AI Agents

How controlling the structure of the prompt, not just the semantics, can exploit your AI agents and their tools

Tomer Wetzler
Tomer Wetzler
AgentFlayer: Versión en español.

Oct 24, 2025

•

2 min read

AgentFlayer: Versión en español.

Inbar Raz
Inbar Raz

Security research

Exploring the Risks of ChatGPT’s Atlas Browser

Oct 23, 2025

•

11 min read

Exploring the Risks of ChatGPT’s Atlas Browser

Tamir Ishay Sharbat
Raul Klugman-Onitza
Tamir Ishay Sharbat, +1

Security research

Appendix: Interpreting Jailbreaks and Prompt Injections with Attribution Graphs

Oct 21, 2025

•

15 min read

Appendix: Interpreting Jailbreaks and Prompt Injections with Attribution Graphs

Max Fomin
Max Fomin

Security research

Interpreting Jailbreaks and Prompt Injections with Attribution Graphs

Oct 21, 2025

•

14 min read

Interpreting Jailbreaks and Prompt Injections with Attribution Graphs

Max Fomin
Max Fomin
1234567
Zenity Labs

Zenity Labs

Latest research, tools and talks about breaking and building AI systems, agents and assistants


Home

Meet Us

Join Us

© 2026 Zenity Labs.
beehiivPowered by beehiiv