Evaluation for voice agents. Catches what text evals cannot see: mis-hearing, missing confirmation, latency, barge-in. Everyone can demo a voice agent; this tells you if yours is getting worse.

ai-agents evaluation llm python speech-to-text voice-ai
0 Open Issues Need Help Last updated: Aug 5, 2026

Open Issues Need Help

View All on GitHub

No open issues

This project doesn't have any open help-wanted issues at the moment.