Run read-only SQL queries over Indian legislation and court judgment parquet files to filter and retrieve legal records by act, case, year, or text pattern.
Enables inspection and management of Apache Airflow DAGs, runs, and logs across multiple instances. It provides tools for monitoring workflows and performing gated write operations like triggering DAGs or clearing task instances.
Enables querying and managing Apache Druid datasources through natural language, including SQL queries, datasource exploration, and cluster connectivity testing.
List installed Apache Airflow provider packages with their versions, descriptions, and included operators, hooks, and sensors. Responds to queries about provider availability and integration details.
Retrieve real column names and data types for a parquet file using its hf:// URI. Use it to verify field names before running SQL queries when no schema is documented.
Check dataset build status and retrieve download links for Parquet files when ready. Track progress through 10 steps and get file metadata including size and expiration.
Catalog every data set under given roots into a Parquet index with metadata, integrity checks, and personal-data flags for openreadout_search; resume interrupted crawls by calling again.
Retrieve the Python source code of an Apache Airflow DAG by providing its DAG ID. Returns the DAG file content and file token for inspection and debugging.
Convert raw lab-instrument files to open formats such as OME-TIFF, OME-Zarr, mzML, Parquet, NWB, JCAMP, or RDML, leaving the source untouched and verifying the output.