Evaluation for voice agents. Catches what text evals cannot see: mis-hearing, missing confirmation, latency, barge-in. Everyone can demo a voice agent; this tells you if yours is getting worse.

ai-agents evaluation llm python speech-to-text voice-ai
2 Open Issues Need Help Last updated: Aug 5, 2026

Open Issues Need Help

View All on GitHub
good first issue

Evaluation for voice agents. Catches what text evals cannot see: mis-hearing, missing confirmation, latency, barge-in. Everyone can demo a voice agent; this tells you if yours is getting worse.

Python
#ai-agents#evaluation#llm#python#speech-to-text#voice-ai
good first issue

Evaluation for voice agents. Catches what text evals cannot see: mis-hearing, missing confirmation, latency, barge-in. Everyone can demo a voice agent; this tells you if yours is getting worse.

Python
#ai-agents#evaluation#llm#python#speech-to-text#voice-ai