evaluate_claim
Assess how rigorously a scientific or AI claim was derived and how well evidence supports it. Grades methodology, statistics, reproducibility with cited sources.
Instructions
Evaluate how rigorously a scientific, biomedical, clinical, or AI/ML claim was derived, and how well the evidence supports it.
Use this whenever someone wants to assess, screen, sanity-check, or do due diligence on a research claim, a study, a paper, an abstract, a preprint, a grant or a pitch: whether the methodology is sound, whether the data was harmonized and controlled properly, whether the statistics hold, whether the result reproduces, and how well the public literature, clinical trials, and filings back it. It grades each evidence dimension and returns an overall reading with sources cited by hard id (PMID, DOI, NCT, NIH grant, SEC filing). Works for drug, omics, target-validation, diagnostic, and AI-model claims.
A screen is free and needs no token. A full attested dossier (deep) is the signed, independently re-checkable audit and needs access.
Args: claim: the claim to evaluate, in one or two sentences. context: optional background (stage, field, the decision at hand). documents: optional source text (a deck, abstract, or paper). depth: "screen" for a fast, free pass (no token needed), "deep" for the full attested dossier (needs access). mode: "research" (science only) or "diligence" (adds the capital reading); deep only.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| mode | No | research | |
| claim | Yes | ||
| depth | No | screen | |
| context | No | ||
| documents | No |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| result | Yes |