Trustra agent-eval: open-core evaluation and tracing harness for AI agents
TLDR
Trustra's agent-eval is an open-core harness for evaluating and tracing AI agents. It runs test datasets against agents, records every tool call and LLM call with latency and cost, and outputs structured evaluation reports. The core runner is free; a hosted dashboard is offered as a paid tier.