Feb 18, 2026
•
11 min read
Beyond input & output filtering and how well does it generalize to your out-of-distribution production data?
Security research
Feb 11, 2026
13 min read
A deeper look into an agentic browser's inner workings
Feb 4, 2026
14 min read
Indirect Prompt Injection makes OpenClaw vulnerable to Backdoors and much more.
Feb 3, 2026
Agent-targeted social engineering and attacks observed on a live agent network
Jan 27, 2026
7 min read
Jan 19, 2026
17 min read
A Copilot Studio case study in agent discovery and capability mapping
Jan 13, 2026
21 min read
What recent scanning activity means for your AI middleware and agentic deployments
Jan 4, 2026
4 min read
How a new fine-tuning approach can mitigate the problem of inaccurate safety paths
Dec 30, 2025
Dec 29, 2025
15 min read
10 min read
Exploiting Copilot Studio's newest feature and exploring protection options
Dec 28, 2025
8 min read
A deep dive into activation space of prompts in safety classifiers. Showing not why - but where - safety fails in LLM classifiers meant to detect malicious prompts.