Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.

1 stars 0 forks 1 watchers Python Apache License 2.0
ai-evaluation ai-safety benchmarking llm-evaluation performance-testing python
5 Open Issues Need Help Last updated: Aug 6, 2026

Open Issues Need Help

View All on GitHub
documentation good first issue

Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.

Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
enhancement good first issue

Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.

Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
good first issue github_actions

Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.

Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
enhancement help wanted

Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.

Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python
enhancement help wanted python

Local-first, auditable evaluation framework for AI quality, safety, compliance evidence, and serving performance.

Python
#ai-evaluation#ai-safety#benchmarking#llm-evaluation#performance-testing#python