The Hydra
A parallelized red-team execution engine that deploys adversarial attack heads against your chatbot across five domains.
About the name
The Hydra draws on Greek mythology — the Lernaean Hydra grew two heads for every one that was cut off. No single attack could defeat it; only coordinated, concentrated force could. The capability works the same way. It does not send a single probe at a target. It assembles all available intelligence, fans each scenario into multiple variations, and unleashes them all at once. The target agent cannot prepare for a single angle when it faces dozens.
How to access it
The Hydra lives at /hydra in the sidebar, or as a toggleable capability in the evaluation wizard. It unleashes concurrent adversarial heads against your chatbot endpoint.
to /hydra
endpoint
mode, cave link
the Hydra
Five evaluation domains
Tiger team head allocation
Concentrate firepower on weak domains. The Hydra does not distribute heads evenly. Before fanning scenarios, it checks prior evaluation scores across all five domains; domains below the weakness threshold (80) receive proportionally more heads. Every domain gets at least one.
Ethics: 91 → 1 head Bias: 88 → 1 head Legal: 85 → 1 head
Total: 12 heads allocated based on the weakness profile
Probe modes
only
only
Cave of Shadows integration
Linking a completed Cave of Shadows run gives the Hydra a tailored base scenario set instead of the standard static battery. The Hydra fans each Cave scenario into domain-specific variations, multiplying the attack surface proportionally across weak domains. Without a linked Cave run it falls back to the built-in director and worker probe libraries (10 each, covering all five domains).
Six-stage execution pipeline
Cave run
assembly
loading
allocation
fanning
execution
aggregation
Heads slider (1–20): the default of 10 balances cost and confidence. Higher counts give better statistical coverage. Each head runs one probe scenario against the target endpoint.
The process map
The Hydra assembles intelligence, picks the weakest domains, fans scenarios into variations, and unleashes them as concurrent multi-turn conversations — then scores each interaction to produce a domain-level resilience report.
Heads slider (1–20): default 10 balances cost and confidence. Weaker domains (below 80) draw more heads; every domain gets at least one. Higher head counts give better statistical coverage.
Stage details
gathering
assembly
loading
allocation
fanning
execution
aggregation
Hydra report output
Resilience report. The final output contains the following.
Five domain scores — Safety, Ethics, Bias, Legal, Security
Per-head results — probe scenario, full conversation transcript, pass/fail
Head-allocation breakdown — how many heads targeted each domain and why
Probe-mode coverage — director probes, worker probes, or both
Cost summary — total tokens and estimated spend
On subsequent runs the tiger-team allocation shifts automatically — domains that improved get fewer heads, and newly weak domains attract more firepower.