A
licenseNot graded
qualityC
maintenanceEnables reproducible evaluation of AI agents on multi-source evidence investigation tasks, providing Docker-based scenarios, MCP tools, and automated scoring for payout reconciliation, performance-claim verification, and incident reconstruction.
MIT