Open Issues Need Help
View All on GitHub Reject empty result arrays so an empty eval run cannot pass CI about 8 hours ago
bug good first issue
CI for AI agents. Record agent runs, replay them against prompt and model changes, and catch regressions and unsafe tool calls before they ship. Diffs each suite against its previous run and returns shouldFail to gate the merge.
TypeScript
#ai-agents#ci#evaluation#nextjs#regression-testing#supabase#typescript