KnowBe4, the global leader in digital workforce security, has announced Agent Risk Manager support for Anthropic's Claude, extending its agent governance layer to Claude environments. The new integration builds on existing native support for Microsoft Copilot, providing real-time visibility and automated threat detection for autonomous AI agents operating within organizational workflows.
KnowBe4 extends Agent Risk Manager to Anthropic's Claude, adding to existing support for Microsoft Copilot.
58% of cybersecurity leaders report AI agents are taking actions within workflows, but 52% admit usage is unapproved or ungoverned.
Agent Risk Manager monitors agent execution without modifying the underlying model, applying a risk-based approach to AI agents.
The solution offers real-time threat interception with six specialized detection engines for prompt injections, PII leaks, and privilege escalations.
A Tool Network map provides blast radius visualization, helping SecOps teams pinpoint high-risk nodes if an agent is compromised.
Agent Risk Manager is in early access for SAT Advanced customers on direct-sale, US-tenant accounts.
As workflows transition from AI-assisted prompts to fully autonomous agents, security teams face a growing visibility gap. Employees are using Claude-powered agents to automate complex tasks, analyze data, and connect with third-party applications, yet these AI identities often operate outside normal security controls. According to the KnowBe4 2026 Agentic Risk to Human Wins Report, 58% of cybersecurity leaders state that AI agents are taking actions within organizational workflows, while 52% report their use of AI is unapproved or ungoverned. This lack of governance leaves organizations exposed to significant security risks, including unauthorized access and data breaches.
Agent Risk Manager closes this blind spot with a simple-to-add security layer that monitors and governs agent execution without modifying the underlying model. The solution extends the same risk-based approach KnowBe4 has applied to human behavior for 15 years to the agentic workforce. With native support for both Claude and Copilot, security teams gain one-time setup that provides insight into all conversations via Claude.ai or by API. Six specialized detection engines actively analyze prompt injections, sensitive data and PII leaks, privilege escalations, and unapproved tool access in real time. The interactive Tool Network map displays connected APIs and credentials, allowing SecOps teams to pinpoint high-risk nodes if an agent is compromised.
KnowBe4's expansion into agent security reflects the growing importance of governing AI agents as they become integral to daily work. The company emphasizes that protecting what agents are doing is critical, especially when employees might use features like skip permissions mode, which allows the LLM to make autonomous decisions with zero oversight. By focusing on actions, outputs, and tool usage, Agent Risk Manager ensures that AI agents do not become unmonitored backdoors into sensitive company assets. Agent Risk Manager is currently in early access for SAT Advanced customers on direct-sale, US-tenant accounts, with organizations able to integrate their Claude environments through a guided, quick onboarding setup.
"The industry spent decades securing the human element and are challenged to protect the newest members of the digital workforce, AI agents," said Matt Duren, SVP of AI and data at KnowBe4. "It is incredibly important to protect what these agents are doing, especially when employees might be using features like skip permissions mode, which allows the LLM to make its own decisions autonomously with zero oversight. With KnowBe4's Agent Risk Manager for Claude, we focus on the actions, outputs, and tool usage of these agents, ensuring they do not become an unmonitored backdoor into sensitive company assets. We are on the bleeding edge of innovation, which will grow exponentially in the future."
About KnowBe4
KnowBe4 empowers the modern workforce to make smarter security decisions every day. Trusted by more than 70,000 organizations worldwide, KnowBe4 is the pioneer of digital workforce security, securing both AI agents and humans. The KnowBe4 Platform provides attack simulation and training, collaboration security, and agent security powered by AIDA (Artificial Intelligence Defense Agents) and a proprietary Risk Score. The platform leverages 15 years of behavioral data to combat advanced threats including social engineering, prompt injection, and shadow AI. By securing humans and agents, KnowBe4 leads the industry in workforce trust and defense.