Kangguru builds Gauge — a test and alignment framework that measures an AI agent's ability, efficiency, and compliance before you trust it with real work — plus the virtual AI officers and expert services to govern it in production.
Explore Gauge Book a discovery call →$ gauge --suite core-alignment gauge session s_84f2 started · agent connected via MCP gauge task 01/12 exact-match ......... PASS gauge task 02/12 numeric ............. PASS gauge task 03/12 prompt-injection .... PASS gauge task 04/12 canary-url .......... FAIL gauge task 05/12 llm-judge ........... PASS ... gauge score 91.6 · compliance 1 finding gauge report → leaderboard updated
Enterprises are deploying AI agents faster than they can evaluate them. Most have no way to answer the basic questions: Is this agent capable? Is it efficient? Will it follow the rules when nobody is watching? Kangguru exists to make those questions measurable — and to govern what the measurements reveal.
Measure agents before deployment with Gauge. Govern your data and security posture in production with the virtual AI officer suite.
Gauge is an MCP-native evaluation server. Any agent that speaks the Model Context Protocol becomes a test subject — no code changes, no SDK. Gauge dispenses structured task curricula, grades every submission deterministically, and scores ability, efficiency, and compliance against other agents.
A coordinated team of AI agents for enterprise security and data governance, anchored by the vDIO (Virtual Data Inventory Officer) — 100% offline data asset discovery for enterprises that can't use cloud solutions — with vCISO capabilities on the roadmap.
Any MCP client is a test subject. Register, start a session, and Gauge takes over — stdio or HTTP, your choice.
Gauge dispenses tasks one at a time in server-assigned order — capability, efficiency, and security-behavior tests including prompt injection traps.
Every submission is graded by exact match, regex, numeric, canary observation, or LLM-as-judge — with a complete event log for audit.
Scores land on the leaderboard. Findings feed your alignment work — fix, re-test, and track improvement over time.
Whether you're at "we have no AI policy" or "we have hundreds of agents in production," we meet you where you are. See how we engage →
We design Gauge test suites for your agents and use cases, run structured evaluations, and turn findings into concrete alignment fixes — before deployment and continuously after.
AI agent risk assessment, why conventional IT security fails for autonomous agents, and the layered-defense approach to regaining control.
Hands-on risk demonstrations from real incidents — prompt injection, shadow AI exfiltration, agent privilege escalation — plus a working tour of the AI security solution landscape.
Hands-on deployment: Gauge in your CI pipeline, vDIO sensors on your endpoints, guardrails and telemetry in your stack. We deliver running systems, not slide decks.
Decades of combined experience across enterprise security, AI risk, adversarial research, and large-scale platform operations.
If you don't know, that's the point. Start with a 30-minute exploratory call — or a Gauge pilot on one of your agents.
Get in touchor email [email protected]