AgentCompass introduces a unified, open‑source evaluation infrastructure for LLM‑based agents, designed to be lightweight and extensible. It structures the evaluation process around three independent components to improve reproducibility and reduce redundant engineering. The framework aims to address the fragmentation of current agent evaluation pipelines.

→ View original source