# Google News Scraper — Articles by Keyword or Topic (`hichemdev/google-news-scraper`) Actor

Monitor Google News: scrape articles by keyword, topic, or publisher — title, source, publish date, link and snippet. Great for brand and competitor monitoring.

- **URL**: https://apify.com/hichemdev/google-news-scraper.md
- **Developed by:** [Hichem Ben Moussa](https://apify.com/hichemdev) (community)
- **Categories:** News
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.00 / 1,000 articles

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## 📰 Google News Scraper — Articles by Keyword or Topic

**Monitor the news without an API key — scrape Google News for any keyword, brand, or topic and get headlines, sources, publish dates, links, and snippets as clean structured data.**

Fast, reliable, and cheap: it reads Google News' own public feeds, so there's nothing to break and no proxy required.

***

### ✨ What you get for every article

| Field | Example |
|---|---|
| **title** | "Xi pitches China as leader of new global AI order" |
| **source** | Reuters |
| **publishedAt** | 2026-07-17T02:34:00.000Z |
| **snippet** | short description |
| **url** | link to the article |
| **query / language / country** | what produced the result |

***

### 🎯 Who it's for

- **Brand & PR monitoring** — track every mention of your company or product.
- **Competitor intelligence** — watch what's being written about rivals.
- **Market & trend research** — follow a topic across hundreds of publishers.
- **Newsletters & dashboards** — power a daily digest from structured data.
- **AI / sentiment analysis** — feed clean headlines into an LLM.

***

### 🚀 How to use it

1. Add **search queries** (e.g. `artificial intelligence`, `"your brand name"`), and/or pick **topics** (Business, Technology, Sports…).
2. Set **max articles**, plus **language** and **country** for the edition you want.
3. Click **Start** and export to CSV, Excel, JSON, or Google Sheets.

**Tip:** schedule it daily and pipe results into Slack or a spreadsheet for an automatic media-monitoring feed.

***

### 📥 Example input

```json
{
  "queries": ["artificial intelligence", "\"Acme Corp\""],
  "topics": ["TECHNOLOGY"],
  "maxItems": 50,
  "language": "en",
  "country": "US"
}
```

Google News search operators work too — e.g. `tesla when:7d`, `site:reuters.com AI`.

***

### 📤 Example output

```json
{
  "title": "Xi pitches China as leader of new global AI order",
  "source": "Reuters",
  "url": "https://news.google.com/rss/articles/...",
  "publishedAt": "2026-07-17T02:34:00.000Z",
  "publishedAtText": "Fri, 17 Jul 2026 02:34:00 GMT",
  "snippet": "China's president used the summit to...",
  "query": "artificial intelligence",
  "language": "en",
  "country": "US"
}
```

***

### ❓ FAQ

**Do I need an API key or proxy?**
Neither. It reads Google News' public feeds — no key, no quota, and no proxy needed for normal use.

**Can I monitor a specific publisher?**
Yes — use a search operator like `site:reuters.com climate`.

**How fresh are the results?**
As fresh as Google News itself — typically minutes old.

**Which languages and countries work?**
Any Google News edition — set `language` (e.g. `en`, `fr`, `de`) and `country` (e.g. `US`, `GB`, `FR`).

**Is this legal?**
It reads publicly available news feeds. You're responsible for how you use and republish the content.

***

### 🗺️ Roadmap

- Resolve publisher URLs behind Google's redirect links
- Full article text extraction (pair with the Website Content Crawler)
- Date-range filtering
- Deduplication across queries

***

*Built and maintained by [hichemdev](https://apify.com/hichemdev). Found a bug or want a feature? Open an issue on the Actor's **Issues** tab.*

# Actor input Schema

## `queries` (type: `array`):

Keywords, brands, or phrases to monitor, e.g. "electric vehicles" or "Apify". Google News search operators work too.

## `topics` (type: `array`):

Google News topic sections to pull headlines from instead of a keyword search.

## `maxItems` (type: `integer`):

Maximum number of articles to collect per query or topic.

## `language` (type: `string`):

Language code for results, e.g. en, fr, de, es.

## `country` (type: `string`):

Country code for the news edition, e.g. US, GB, FR, DE.

## `proxyConfiguration` (type: `object`):

Optional. Google News RSS is public, so no proxy is needed for normal use.

## Actor input object example

```json
{
  "queries": [
    "artificial intelligence"
  ],
  "topics": [],
  "maxItems": 15,
  "language": "en",
  "country": "US",
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "queries": [
        "artificial intelligence"
    ],
    "topics": [],
    "maxItems": 15,
    "language": "en",
    "country": "US"
};

// Run the Actor and wait for it to finish
const run = await client.actor("hichemdev/google-news-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "queries": ["artificial intelligence"],
    "topics": [],
    "maxItems": 15,
    "language": "en",
    "country": "US",
}

# Run the Actor and wait for it to finish
run = client.actor("hichemdev/google-news-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "queries": [
    "artificial intelligence"
  ],
  "topics": [],
  "maxItems": 15,
  "language": "en",
  "country": "US"
}' |
apify call hichemdev/google-news-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=hichemdev/google-news-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/Bp6FPrbeeYTanNILN/builds/jb6GiGuJSm15SEv46/openapi.json
