google_news_scraper
Fetch Google News headlines by keyword, site: or when: query or topic, and resolve each article's real publisher URL—one row per article for RAG feeds.
Instructions
Google News Scraper returns headlines from Google News RSS search and topic feeds for any keyword, site: or when: query, and resolves each article's real publisher URL — one row per article. Billed to your own Apify account: ~$0.0005 per article (Apify free-plan price, lower on paid plans).
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| topics | No | Section topics — Enter Google News section topics to fetch instead of (or in addition to) a search, e.g. TECHNOLOGY or BUSINESS. Accepted values: WORLD, NATION, BUSINESS, TECHNOLOGY, ENTERTAINMENT, SCIENCE, SPORTS, HEALTH. Unknown values are ignored. | |
| country | No | Country — Enter the 2-letter country/edition code, e.g. US for the United States or GB for the United Kingdom. Combined with Language to build the hl/gl/ceid feed parameters. | US |
| queries | Yes | Search queries — Enter Google News search terms, one row is returned per matching article. Supports Google's search operators, e.g. web scraping, "exact phrase", site:reuters.com AI, or -unwanted. Leave empty and use Topics below instead if you only want section feeds. Example: ["web scraping"]. | |
| language | No | Language — Enter the 2-letter interface language code Google News should use, e.g. en for English or fr for French. Combined with Country to build the hl/gl/ceid feed parameters. | en |
| sinceDays | No | Only articles from the last N days — Enter how many days back to search, e.g. 7 for the last week. Appends Google's when:Nd search operator to every query (topic feeds ignore this — they are always "latest"). Enter 0 to disable and return whatever Google's default ranking gives. | |
| maxItemsPerQuery | No | Max items per query/topic — Enter the maximum number of articles to keep per query or topic, e.g. 50. Google's own feed rarely returns more than ~100 items for a single request no matter how high this is set. | |
| resolvePublisherUrls | No | Resolve publisher URLs — Keep this on to follow each article's news.google.com redirect link and fetch the publisher's real URL (2 extra requests per article, using an undocumented Google endpoint — best-effort, see README). Turn it off for a much faster run that only returns the Google News link. |