review_paper
Run a full pre-submission review of a paper PDF: verify claims, prior work, citations, methodology, and writing. Returns evidence-backed findings with quotes, no scores.
Instructions
Run the full Reviewer Zero review of a paper PDF (claims, prior work, citations with support checks, methodology, writing, format) with our prompts, using the user's own ANTHROPIC_API_KEY. It costs real money (usually about $1.50 to $2.00) and takes 7 to 12 minutes.
ALWAYS call it first with confirm=false: that spends nothing and returns the estimate and the hard limit. Show both to the user and call again with confirm=true only after they agree. If a call times out, call again with the same PDF: finished steps are served from the cache and only what is left is paid for.
novelty="full" (default) sends the 150 best candidates per claim to the reranker; novelty="lite" sends 60 (quicker and cheaper, lower recall; see the README for the measured recall and time of each).
Report findings evidence-first, each with its quote. Never give a score, rating or accept/reject verdict, and never rewrite the user's text. Without ANTHROPIC_API_KEY or a local GROBID container it explains what is missing; check_format, check_citations, find_prior_work, paper and citations work without a key.
Sent off your machine by this tool: the paper's text to Anthropic under your own key; search queries and cited works' titles/DOIs/arXiv ids to the index, OpenAlex, Crossref and arXiv.
Privacy: runs on your machine. Your PDF and its text never leave it, except to Anthropic under your own API key when you call review_paper. What is sent: search queries and paper keys to the Reviewer Zero index (it counts requests per API key and stores nothing else), and, for check_citations, the titles, DOIs and arXiv ids of the works the paper cites, to the index and to OpenAlex, Crossref and arXiv. No telemetry.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| venue | Yes | ||
| confirm | No | ||
| novelty | No | full | |
| pdf_path | Yes |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| notes | No | ||
| phase | Yes | 'plan' (confirm=false): nothing spent; show the estimate and ask the user. 'review': the full run. | |
| steps | No | ||
| format | Yes | ||
| findings | No | ||
| limit_usd | Yes | Hard limit: the run stops calling the model before spending more than this. | |
| spent_usd | No | ||
| meta_review | No | The allowlisted meta-review (no number anywhere; every sentence passed the D12 leak filter). None when the meta step did not run or failed. | |
| estimate_usd | Yes | About this much (sum of per-step medians for the steps not yet done; D5). | |
| novelty_mode | No | full: 150 candidates per claim go to the LLM reranker; lite: 60 (quicker, cheaper). | full |
| steps_to_run | Yes | ||
| steps_already_done | Yes | Finished in an earlier call on the same PDF; rerunning them is served from the cache. |