A counter-espionage RL environment for scalable AI oversight. Train LLM agents to detect deception, trace leaks, and neutralize sleeper agents in a 160-turn corporate simulation. Built for the Meta PyTorch OpenEnv Hackathon.

1 stars 1 forks 1 watchers Python Apache License 2.0
ai-safety deception-detection hackathon llm-agents multi-agent openenv pytorch reinforcement-learning
1 Open Issue Need Help Last updated: Aug 8, 2026

Open Issues Need Help

View All on GitHub

A counter-espionage RL environment for scalable AI oversight. Train LLM agents to detect deception, trace leaks, and neutralize sleeper agents in a 160-turn corporate simulation. Built for the Meta PyTorch OpenEnv Hackathon.

Python
#ai-safety#deception-detection#hackathon#llm-agents#multi-agent#openenv#pytorch#reinforcement-learning