Open source repositories tagged with #agent-evaluation-tools, ranked by health score.
Core engine behind Calibrate, a framework for evaluating AI agents: speech-to-text, text-to-speech, LLM evaluation, end-to-end simulations