rd2d_bw
Selects the appropriate bandwidth for 2D boundary regression discontinuity, using treatment, running variables, and outcome to define the neighborhood for causal analysis near the boundary.
Instructions
Bandwidth selection for 2D boundary RD.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| p | No | Polynomial order. | |
| y | Yes | Outcome variable name. | |
| x1 | Yes | Running variable names. | |
| x2 | Yes | Running variable names. | |
| detail | No | Payload depth: 'minimal' (~150 tokens) for sub-step calls where only the point estimate is needed; 'standard' (~1K tokens) for diagnostics + coefficient table; 'agent' (~2K tokens, default) adds violations / next_steps / suggested_functions so the LLM can plan its next call without another round-trip. | agent |
| kernel | No | Kernel function. | triangular |
| approach | No | 'distance' or 'location'. | distance |
| boundary | No | Boundary function f(x1) -> x2. None implies x1 = 0. | |
| as_handle | No | If true, cache the fitted result on the server and return result_id + result_uri alongside the JSON payload so a subsequent tools/call can chain without re-running. | |
| data_path | Yes | Absolute path or URL to a data file. Supported: .csv / .tsv / .txt (delimited), .parquet / .pq, .feather / .arrow, .xlsx / .xls, .dta (Stata), .json / .jsonl. Schemes: file://, s3://, gs://, https://. | |
| result_id | No | Optional handle to a previously-fitted result (returned by an earlier call when as_handle=true). Tools that operate on a fitted object accept this in place of re-supplying data_path + columns. | |
| treatment | Yes | Binary treatment indicator. | |
| data_columns | No | Optional column projection. Parquet/Feather/Stata loaders honour this for fast partial reads. | |
| data_sample_n | No | Optional uniform random subsample size (seed=0, deterministic) — useful on huge panels. |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
No arguments | |||