6-stage pipeline from discovery through production monitoring
Scoring engine: Tier 1 (1.5×) + Tier 2 (1.0×), hard veto at 75
Cave of Shadows: adversarial red-team with attack personas
Hydra: parallelized swarm running dozens of test scenarios simultaneously
Bias Probe: paired differential testing detecting treatment disparity
Sentinel: generates dual-output QA test plans from evaluation results
Drift Detection: 6 drift types, weekly sweeps, severity alerts
Build pipeline: 6-phase automated agent generation with HITL checkpoint
Portfolio layer: department Kanban, strategic prioritization agent, CSV imports
Monday.com integration: MCP client, direct PRD endpoint, argument mapping
Backend: FastAPI + Supabase + Celery | Frontend: Next.js 16 + TypeScript
Deployed: Fly.io (backend) + Vercel (frontend) | ~990 tests passing