# Hacker News Scraper: Stories, Comments & Search (`scrapemint/hacker-news-scraper`) Actor

Search and scrape Hacker News: stories and comments by keyword, author, points, date or category (front page, Ask HN, Show HN). One clean row per story or comment with points, comment count, author and links. No API key.

- **URL**: https://apify.com/scrapemint/hacker-news-scraper.md
- **Developed by:** [Ken M](https://apify.com/scrapemint) (community)
- **Categories:** News, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: 5.00 out of 5 stars

## Pricing

$2.00 / 1,000 item rows

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Hacker News Scraper: Stories, Comments & Search

Search and scrape [Hacker News](https://news.ycombinator.com) - stories and comments - by keyword, author, points, date or category. No API key, no login, no browser. Built on the official Algolia HN Search API.

### What you get

**Stories** - one row each with title, external link, points, comment count, author, post text (for Ask/Show HN), tags, and the HN discussion link.

**Comments** - one row each with the comment text, author, the story it belongs to (title and URL), and links.

### Ways to use it

```json
{
    "queries": ["kubernetes"],
    "contentType": "stories",
    "minPoints": 50,
    "sortBy": "relevance"
}
```

- **queries** - keywords (one search per line); leave empty to browse a category
- **contentType** - stories, comments, or both
- **category** - front page, Ask HN, Show HN, or polls
- **author** - only a specific user's posts
- **minPoints / minComments** - only items that got traction
- **sinceDays** - recent items only
- **sortBy** - best match or newest

Examples: the current front page, the top "Show HN" launches over 100 points, every comment mentioning your product, or one author's whole history.

### Who uses this

- **Founders and marketers**: monitor mentions of your product or competitors and find launches gaining traction.
- **Developers and researchers**: track what the tech community is discussing, and pull discussions for analysis.
- **Trend and market analysts**: measure attention on a technology or company over time by points and comments.
- **Recruiters and community teams**: follow active authors and threads in a niche.

Complements our Hacker News lead actors (which focus on hiring posts and lead alerts) with general story and comment scraping.

### Pricing

A small fee per row (one per story or comment). Searches that match nothing are free note rows, and the first 2 rows of every run are free.

### Notes

- Source: Algolia Hacker News Search API, the same search that powers HN's own search box. Points and comment counts reflect the last time Algolia indexed the item.
- Relevance search reaches the first ~1,000 results per query; narrow with filters (points, date, author) or use "newest" to page deeper.

# Actor input Schema

## `queries` (type: `array`):

Keywords to search for, one per line ("kubernetes", "rust async"). Leave empty to just browse a category (e.g. the front page) using the filters below.

## `contentType` (type: `string`):

Return stories, comments, or both.

## `category` (type: `string`):

Limit to a Hacker News category (stories only).

## `author` (type: `string`):

Optional: only items by this Hacker News username (e.g. "pg").

## `minPoints` (type: `integer`):

Only items with at least this many points. 0 = no minimum.

## `minComments` (type: `integer`):

Stories only: at least this many comments. 0 = no minimum.

## `sinceDays` (type: `integer`):

Only items from the last N days. 0 = no date limit.

## `sortBy` (type: `string`):

Best match (relevance) or newest first.

## `maxPerQuery` (type: `integer`):

How many items to return per search term (or per category browse).

## `maxRows` (type: `integer`):

Stop after this many rows in total.

## Actor input object example

```json
{
  "queries": [
    "kubernetes"
  ],
  "contentType": "stories",
  "category": "any",
  "minPoints": 0,
  "minComments": 0,
  "sinceDays": 0,
  "sortBy": "relevance",
  "maxPerQuery": 50,
  "maxRows": 1000
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "queries": [
        "kubernetes"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("scrapemint/hacker-news-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "queries": ["kubernetes"] }

# Run the Actor and wait for it to finish
run = client.actor("scrapemint/hacker-news-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "queries": [
    "kubernetes"
  ]
}' |
apify call scrapemint/hacker-news-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=scrapemint/hacker-news-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/acts/1Pcqgg65I5pT7K7Au/builds/9YzHYvgOHMBMZ5FTT/openapi.json
