Framework · ai-offensive-orchestratordual-usehigh-riskcloud cost
CVE-Factory
A multi-agent pipeline that automates end-to-end CVE reproduction: it researches CVE details, generates test cases, builds Docker environments, and validates both the exploit and the patch. It powers the LiveCVEBench evaluation and produces training traces for security-focused LLMs.
Use responsibly
Test only systems you own or are explicitly authorized to test. Unauthorized testing is illegal.
High risk of account lockouts, WAF bans, and terms-of-service violations. Requires explicit authorization and small, targeted inputs.
More frameworks & agents
- BugHunterai-offensive-orchestratorAn AI-powered bug bounty toolkit (standalone CLI and Claude Code plugin) that runs an autonomous scope-to-report loop: recon, hunting across 26+ web vulnerability classes and smart-contract bug categories, a validation gate, and submission-ready reports for HackerOne, Bugcrowd, Intigriti, and Immunefi. Orchestrates ~35 external scanners and supports Ollama/Groq or paid AI providers.
- Claude-BugHunterai-offensive-orchestratorA Claude Code skill bundle for authorized security testing, providing 83 skills, 15 slash commands and pattern databases with hunt templates for 58 web vulnerability classes (XSS, SQLi, SSRF, IDOR) plus recon/OSINT and reporting workflows. Includes authorization gates and excludes internal AD, C2 and post-exploitation.
- Cybersecurity AI (CAI)ai-offensive-orchestratorAn open-source framework for building AI-powered offensive and defensive security automation, using ReACT-model agents with tools for command execution, web recon and code analysis, plus handoffs, swarm/hierarchical patterns, guardrails and human-in-the-loop. Supports 300+ models across providers and targets bug bounty, vulnerability discovery and exploitation workflows.
- DorkAgentai-offensive-orchestratorAn LLM-powered agent (built on CrewAI) that automates Google Dorking for reconnaissance in penetration testing and bug bounty. It generates and refines dork queries with an LLM (OpenAI, Anthropic, or Gemini), analyzes results, and produces structured vulnerability reports to surface information disclosure, misconfigurations, and exposed sensitive data.
- hackingBuddyGPTai-offensive-orchestratorAcademic (TU Wien ipa-lab) open-source framework for building LLM-driven autonomous pentest agents in under ~50 lines, covering Linux privilege escalation, web-app and REST-API testing via SSH/local shell. Includes SQLite run logging and a web viewer for replaying agent runs; supports OpenAI and local models.
- HexStrike AIai-offensive-orchestratorAI-driven autonomous penetration-testing framework: an MCP server that lets an external LLM drive ~90+ offensive binaries plus native decision-engine, CVE-intelligence, and fault-tolerance subsystems.