cognitive.evaluate_generalization_benchmarks
Evaluate broad generalization across spatial commons, multi-agent arenas, and sequential causal puzzles.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| benchmark_filter | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||