Open source repositories tagged with #ai-agent-evaluation, ranked by health score.
Core engine behind Calibrate, a framework for evaluating AI agents: speech-to-text, text-to-speech, LLM evaluation, end-to-end simulations