# 📕 RedNote (Xiaohongshu) Comments Scraper (`ethereal_wool/xiaohongshu-comments-scraper`) Actor

Scrape comments from RedNote / Xiaohongshu (小红书) notes by ID or URL — comment text, author, likes, time, IP location and note links. No login, structured JSON/CSV/Excel, pay per result.

- **URL**: https://apify.com/ethereal\_wool/xiaohongshu-comments-scraper.md
- **Developed by:** [Jackie Chen](https://apify.com/ethereal_wool) (community)
- **Categories:** Social media, Videos, Automation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 1 bookmarks
- **User rating**: No ratings yet

## Pricing

$10.00 / 1,000 comments

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## RedNote (Xiaohongshu) Comments Scraper

Collect the current top-level discussion under specific Xiaohongshu notes. Provide note IDs and receive clean comment rows with text, author identity, likes, time, location and a link back to the parent note.

> **Unofficial / independent tool.** This Actor is not affiliated with, authorized, sponsored, or endorsed by Xiaohongshu. It retrieves publicly available data through a third-party API. You are responsible for using the output in compliance with Xiaohongshu's terms and all applicable laws.

### What this Actor does

This Actor focuses on one job: fetch top-level comments for one or more xiaohongshu notes by note id on **xiaohongshu.com**. This Actor prices each delivered comment directly instead of hiding an extra paid API call inside a general note-search run.

- Fetches the current top-level comment page for each note ID.
- Returns comment text and author identity.
- Includes likes, publication time, IP location and reply count.
- Links every comment back to its parent Xiaohongshu note.

### Input

| Field | Type | Description |
| --- | --- | --- |
| `noteIds` | array | Note IDs from Xiaohongshu URLs or search output. Returns the current top-level comment page for each note. |
| `maxItems` | integer | Maximum records to return (caps your spend). |
| `proxyConfiguration` | object | Optional Apify Proxy settings. |

#### Example input

```json
{
  "noteIds": [
    "69d8ab67000000022200b884"
  ],
  "maxItems": 10
}
```

### Output

The Actor returns one dataset item per comment. Each item is a flat, analysis-ready JSON record. Example of a real returned item:

```json
{
  "commentId": "comment-sample-1",
  "noteId": "69d8ab67000000022200b884",
  "content": "这个防晒会搓泥吗？",
  "author": "小夏",
  "authorId": "user-sample-1",
  "likeCount": 18,
  "publishedAt": 1783728600,
  "ipLocation": "广东",
  "subCommentCount": 3,
  "id": "comment-sample-1",
  "url": "https://www.xiaohongshu.com/explore/69d8ab67000000022200b884",
  "source": "xiaohongshu-comments"
}
```

#### Output fields

| Field | Description |
| --- | --- |
| `commentId` | Comment ID |
| `noteId` | Parent note ID |
| `content` | Comment text |
| `author` | Comment author |
| `authorId` | Comment author user ID |
| `likeCount` | Comment like count |
| `publishedAt` | Comment timestamp |
| `ipLocation` | Comment IP location when available |
| `subCommentCount` | Reply count |
| `url` | Canonical link to the item on the source site |
| `id` | Stable identifier for the item (when available) |
| `source` | Which list / query the item came from |

### How it works

- **Direct API, no browser.** Data is fetched over HTTP — no headless browser, no login, no cookies to manage.
- **Honest failure.** Transient upstream blocks (rate limits, edge protection) are retried with exponential backoff. If the source stays unavailable, the run fails loudly instead of returning a misleading empty dataset.
- **De-duplicated.** Items are de-duplicated by their identifier within a run.
- **Pay per result.** Each delivered row charges one `result` event ($0.002); `maxItems` is a hard cap on both volume and spend.

### Use cases

- Analyze customer language and sentiment under product notes.
- Collect objections and questions for market research.
- Monitor discussion around campaigns or competitor posts.
- Feed comment evidence into social-listening agents.

### Integration

Run it from the Apify Console, on a schedule, or call it programmatically via the Apify API, the JavaScript / Python clients, or MCP. Output can be exported as JSON, CSV, or Excel, or pushed to your own storage.

### FAQ

**Do I need a Xiaohongshu account, cookies, or to log in?** No. The Actor only reads publicly available data.

**How am I billed?** $0.002 per returned item; `maxItems` caps the total.

**Can I schedule it or call it from my own code?** Yes — use Apify Schedules, the REST API, the official clients, or MCP.

**Is this an official Xiaohongshu product?** No. It is an independent tool and is not affiliated with Xiaohongshu.

# Actor input Schema

## `noteIds` (type: `array`):

Note IDs from Xiaohongshu URLs or search output. Returns the current top-level comment page for each note.

## `maxItems` (type: `integer`):

Maximum number of records to return. Caps your spend.

## `proxyConfiguration` (type: `object`):

Optional. Route API calls through Apify Proxy to vary the source IP.

## Actor input object example

```json
{
  "noteIds": [
    "69d8ab67000000022200b884"
  ],
  "maxItems": 10,
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "noteIds": [
        "69d8ab67000000022200b884"
    ],
    "proxyConfiguration": {
        "useApifyProxy": false
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("ethereal_wool/xiaohongshu-comments-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "noteIds": ["69d8ab67000000022200b884"],
    "proxyConfiguration": { "useApifyProxy": False },
}

# Run the Actor and wait for it to finish
run = client.actor("ethereal_wool/xiaohongshu-comments-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "noteIds": [
    "69d8ab67000000022200b884"
  ],
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}' |
apify call ethereal_wool/xiaohongshu-comments-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=ethereal_wool/xiaohongshu-comments-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/cGJ7HLTAHv1VIizs0/builds/SroziMRPd5broZBgl/openapi.json
