Open Issues Need Help
View All on GitHub Evaluate LangFair as a tool for evaluations and ideas for the taxonomy about 1 month ago
help wanted taxonomy evaluations
The AI Alliance project to define a reference stack for AI model and system evaluation, with evaluations, benchmarks, and leaderboards.
Makefile
#ai#safety#trust
POC candidate: benchmark cards about 1 month ago
help wanted evaluations
The AI Alliance project to define a reference stack for AI model and system evaluation, with evaluations, benchmarks, and leaderboards.
Makefile
#ai#safety#trust