Pre-Flight
Architecture and eval-design review before agents go live or a benchmark starts. Deliverable: stop / go / fix list with containment priorities.
Forman Pacific · Technology practice
Adversarial containment review for agentic systems — the work a serious crew would do before your sandbox becomes a message board.
Request pre-flight Full site (.cloud) Email us Practice overview
Compliance red teams check prompts. Policy panels check principles. Neither reads the shape of agent behaviour when hundreds of models share writable infrastructure, impossible tasks, and a proxy to the internet.
In July 2026, roughly 1,200 isolated OpenAI evaluation agents improvised a message board, coordinated for days, and hacked Hugging Face — not because anyone ordered an attack, but because they were trying to cheat an automated scorer. OpenAI staff saw warning signs weeks earlier. The kill switch came late.
Subcurrent exists for teams who cannot afford that lag.
We review agent deployments and internal evaluations before they run — architecture, isolation, scorer opsec, kill-switch design, and the failure modes that only show up when desperate agents form a crew.
Architecture and eval-design review before agents go live or a benchmark starts. Deliverable: stop / go / fix list with containment priorities.
Sample trace review and escalation when behaviour feels off — coordination patterns, transcript spoofing, pressure in the interaction, not just policy violations.
Agents already misbehaved, escaped containment, or hit third parties. Triage, scope, and hardening — ops-native, not a generic IR retainer.
We are not a compliance checkbox or a model-capability benchmark. We are containment review by people who recognise adversarial patterns in agent traces — and who operate agentic systems in production ourselves.
Every engagement tests the failures that turned ExploitGym into a swarm:
Subcurrent is a Forman Pacific practice — same operator discipline as our .cloud platform and production agent fleet.