verify_agent
Checks agent final responses for hallucinated claims and evaluates execution traces across six dimensions: factuality, sourcing, instruction compliance, tool-claim consistency, task completion, and reflection.
Instructions
Agent 输出与轨迹自检。对最终回答跑逐声明幻觉核验,并对执行轨迹做六维结构化评估(事实性/来源/指令合规/工具声明一致/任务完成/反思)。返回 main.claims、steps[].checks、quota。耗 detect 额度。
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| task | No | 原始任务指令(task_completion 裁判用) | |
| speed | No | 速度模式:fast|standard|deep,默认 standard | |
| steps | No | agent 执行轨迹,逐步六维评估 | |
| domain | No | 领域模式:general|medical|legal|finance|education|government | |
| agent_output | Yes | agent 最终输出文本(≥1 字,≤50000 字) | |
| tool_schemas | No | 工具 Schema(与 trajectory 端点对齐) |