Open-source agent evaluation framework - score any AI agent across providers with standardized benchmarks

agents ai ai-agents benchmark evaluation llm llmops open-source python testing typescript
3 Open Issues Need Help Last updated: Aug 8, 2026

Open Issues Need Help

View All on GitHub
good first issue provider

Open-source agent evaluation framework - score any AI agent across providers with standardized benchmarks

Python
#agents#ai#ai-agents#benchmark#evaluation#llm#llmops#open-source#python#testing#typescript
enhancement good first issue

Open-source agent evaluation framework - score any AI agent across providers with standardized benchmarks

Python
#agents#ai#ai-agents#benchmark#evaluation#llm#llmops#open-source#python#testing#typescript
good first issue provider

Open-source agent evaluation framework - score any AI agent across providers with standardized benchmarks

Python
#agents#ai#ai-agents#benchmark#evaluation#llm#llmops#open-source#python#testing#typescript