Sealed-lab research harness for emergent multi-agent (LLM) coordination and containment — replicates the dynamics of the 2026 OpenAI–Hugging Face incident against self-hosted, intentionally-vulnerable targets. Air-gapped by design. Defensive / AI-safety only.

ai-safety containment llm-agents multi-agent red-team reward-hacking sandbox security-research
5 Open Issues Need Help Last updated: Sep 7, 2026

Open Issues Need Help

View All on GitHub
documentation help wanted good first issue

Sealed-lab research harness for emergent multi-agent (LLM) coordination and containment — replicates the dynamics of the 2026 OpenAI–Hugging Face incident against self-hosted, intentionally-vulnerable targets. Air-gapped by design. Defensive / AI-safety only.

Python
#ai-safety#containment#llm-agents#multi-agent#red-team#reward-hacking#sandbox#security-research

Sealed-lab research harness for emergent multi-agent (LLM) coordination and containment — replicates the dynamics of the 2026 OpenAI–Hugging Face incident against self-hosted, intentionally-vulnerable targets. Air-gapped by design. Defensive / AI-safety only.

Python
#ai-safety#containment#llm-agents#multi-agent#red-team#reward-hacking#sandbox#security-research

Sealed-lab research harness for emergent multi-agent (LLM) coordination and containment — replicates the dynamics of the 2026 OpenAI–Hugging Face incident against self-hosted, intentionally-vulnerable targets. Air-gapped by design. Defensive / AI-safety only.

Python
#ai-safety#containment#llm-agents#multi-agent#red-team#reward-hacking#sandbox#security-research
good first issue containment

Sealed-lab research harness for emergent multi-agent (LLM) coordination and containment — replicates the dynamics of the 2026 OpenAI–Hugging Face incident against self-hosted, intentionally-vulnerable targets. Air-gapped by design. Defensive / AI-safety only.

Python
#ai-safety#containment#llm-agents#multi-agent#red-team#reward-hacking#sandbox#security-research
help wanted good first issue target

Sealed-lab research harness for emergent multi-agent (LLM) coordination and containment — replicates the dynamics of the 2026 OpenAI–Hugging Face incident against self-hosted, intentionally-vulnerable targets. Air-gapped by design. Defensive / AI-safety only.

Python
#ai-safety#containment#llm-agents#multi-agent#red-team#reward-hacking#sandbox#security-research