Test before you deploy. Protect at runtime. Catch regressions as prompts, tools and models evolve. And prove it at every release.
Apache-2.0 · No login required · Self-hostable
Firewall tiers from sanitization to an LLM judge.
OWASP-aligned attack categories.
Not a one-time scan. An always-on security layer for your AI estate that adapts as your agents, models, and data sources evolve.
Automated adversarial & behavioral testing. OWASP-aligned attack scenarios cover the full threat surface — prompt injection, jailbreaks, data exfiltration, tool abuse. Attacks adapt across single-turn, multi-turn, and agentic modes to find what static tests miss.
The Humanbound Firewall. Sits between users and your agent. Blocks prompt injections and policy violations before they reach the model. Four defence tiers, an agent-specific classifier trained on your own test data, and an LLM judge for deep contextual analysis.
Continuous assurance campaigns. When models update, prompts change, or new data sources connect, testing adapts automatically and prioritises the areas where coverage gaps are widest. A living posture score that reflects reality, not a point-in-time PDF.
The testing engine, SDK, and firewall are all Apache-2.0. Run a full security test from your terminal using your own API keys — or go fully air-gapped with Ollama. Same engine that powers the platform. Nothing held back, nothing artificially limited.
When your needs grow beyond a single agent, the platform adds continuous monitoring, finding lifecycle management, cross-session intelligence, and managed infrastructure.
Loading...
Humanbound does not produce a summary and leave you to figure out what it means. Every vulnerability gets a severity rating, an OWASP classification, and a reproducible evidence trail.
What security leaders, engineers, and developers ask before they deploy.
Humanbound turns every test, attack and runtime signal into evidence you can act on, share and defend.
Free plan available, no card required