download_ddb_text
Download OCR text of a single page or entire issue from the German newspaper collection. Provide a page or issue ID to get cached text files.
Instructions
Download the OCR text of a page, or of every page of an issue.
Args: identifier: A page id ('ITEMID-pagename') or an issue item id ('ITEMID') refresh: Ignore any cached copy and fetch again
Returns: Path to the cached text file. Files run to tens of kilobytes per page, so read slices of them rather than the whole thing.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| refresh | No | ||
| identifier | Yes |
Output Schema
| Name | Required | Description | Default |
|---|---|---|---|
| result | Yes |