Open Issues Need Help
View All on GitHub security: prototype containerized evaluator isolation about 3 hours ago
help wanted security
Rust systems lab and reproducible RL-style environment for evaluating coding-agent reliability, refactoring, and performance work.
Python
#ai-evaluation#benchmarking#coding-agents#concurrency#python#reinforcement-learning#rust#software-engineering#systems-programming#testing
docs: add an end-to-end task-author walkthrough about 3 hours ago
documentation good first issue
Rust systems lab and reproducible RL-style environment for evaluating coding-agent reliability, refactoring, and performance work.
Python
#ai-evaluation#benchmarking#coding-agents#concurrency#python#reinforcement-learning#rust#software-engineering#systems-programming#testing
help wanted testing
Rust systems lab and reproducible RL-style environment for evaluating coding-agent reliability, refactoring, and performance work.
Python
#ai-evaluation#benchmarking#coding-agents#concurrency#python#reinforcement-learning#rust#software-engineering#systems-programming#testing
docs: add Windows and WSL verification guide about 3 hours ago
documentation good first issue
Rust systems lab and reproducible RL-style environment for evaluating coding-agent reliability, refactoring, and performance work.
Python
#ai-evaluation#benchmarking#coding-agents#concurrency#python#reinforcement-learning#rust#software-engineering#systems-programming#testing