# Xiaohongshu Video Transcript Scraper | RedNote Speech-to-Text (`ethereal_wool/xiaohongshu-video-transcript-scraper`) Actor

Turn any Xiaohongshu (小红书 / RedNote) video note into text. Real AI speech recognition with best-in-class Mandarin Chinese accuracy — works on videos with no subtitles. Full text + timestamped sentences as clean JSON.

- **URL**: https://apify.com/ethereal\_wool/xiaohongshu-video-transcript-scraper.md
- **Developed by:** [Jackie Chen](https://apify.com/ethereal_wool) (community)
- **Categories:** Social media, AI
- **Stats:** 1 total users, 0 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$25.00 / 1,000 transcript per minutes

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Xiaohongshu Video Transcript Scraper — AI Speech-to-Text for 小红书 / RedNote

Turn any **Xiaohongshu (小红书 / RedNote) video note into text** with real AI
speech recognition. Paste note links, get back the **full spoken transcript
plus timestamped sentences** as clean JSON — with **best-in-class Mandarin
Chinese accuracy**, ready for LLMs, RAG pipelines, content research, and
subtitle workflows.

> **This is real ASR, not caption scraping.** Most "transcript" tools only
> download captions a creator happened to upload — and Xiaohongshu videos
> almost never have them. This Actor runs industrial speech recognition on the
> video's audio track, so it works on **any 小红书 video with speech**,
> captions or not. Chinese (Mandarin) recognition is its strongest language.

> **Unofficial.** This Actor is not affiliated with, authorized, or endorsed by
> Xiaohongshu / 行吟信息科技. It is an independent tool that processes publicly
> available content. Use it in compliance with Xiaohongshu's terms and all
> applicable laws; you are responsible for how you use the retrieved data.

### What you get per video note

- **Full transcript** (`fullText`) — the complete spoken content as one string.
- **Timestamped sentences** (`sentences[]`) — each with `startMs` / `endMs`,
  ready for subtitles, deep links, and clip selection.
- **Note metadata** — title, description, author, duration, publish time,
  like / collect / comment / share counts, and IP location, so every
  transcript arrives with its engagement context attached.

### Quick start

1. Open the Actor and press **Run** — the default input works out of the box.
2. Replace the example with your own note links: explore URLs
   (`https://www.xiaohongshu.com/explore/…`), short share links
   (`https://xhslink.com/…`), pasted share text, or bare note IDs.
3. Language defaults to **Chinese**; switch to Auto for mixed or non-Chinese
   content.
4. Collect results from the **Dataset** tab as JSON / CSV / Excel, or pull them
   via the [Apify API](https://docs.apify.com/api/v2) and MCP from your own code.

No proxies, no cookies, no login — everything runs server-side.

### Example output

```json
{
  "noteId": "6907434c000000000703bf59",
  "noteUrl": "https://www.xiaohongshu.com/explore/6907434c000000000703bf59",
  "title": "黑绷带和AGE七天测评‼️效果差强人意",
  "description": "再贵的护肤品也不是万能药，均衡搭配是王道。#抗衰 #护肤品测评",
  "author": "Freya",
  "durationSec": 221,
  "likedCount": 1034,
  "collectedCount": 352,
  "commentsCount": 219,
  "language": "zh",
  "detectedSpeech": true,
  "sentenceCount": 42,
  "fullText": "今天给大家带来黑绷带和AGE面霜的七天实测对比…",
  "sentences": [
    { "text": "今天给大家带来黑绷带和AGE面霜的七天实测对比，", "startMs": 240, "endMs": 3100 },
    { "text": "先说结论，赫莲娜略胜一筹。", "startMs": 3100, "endMs": 5400 }
  ]
}
```

### What people build with it

- **Beauty / fashion / lifestyle research** — 小红书 is where Chinese consumers
  review products on camera; transcripts turn those spoken reviews into
  searchable, quotable data.
- **Viral-hook mining** — pull transcripts of the top video notes in your niche
  and study the exact opening lines and structures that earn likes and collects.
- **Influencer & brand monitoring** — capture verbatim what KOLs and KOCs say
  about your brand or competitors, at scale.
- **Cross-border content** — transcribe Chinese videos, then translate and
  repurpose them for TikTok, Instagram, or your own market.
- **LLM & RAG pipelines** — build Chinese-language corpora from real
  short-video speech, with engagement scores as a free quality signal.
- **Subtitles & translation** — timestamped sentences drop straight into
  SRT/VTT generation and dubbing workflows.

### Pricing & billing

**Pay per audio minute — $0.025/min, billed by the beginning minute.** A 0:50
clip bills 1 minute; a 1:10 clip bills 2 minutes. Timestamped sentences are
included free — no separate subtitle or export charge. There is **no per-run
start fee** and no separate compute or platform fee: the price you see is the
price you pay.

Failed fetches, deleted notes, image-only notes, and failed transcriptions are
**not charged** — you only pay for audio we actually transcribe.

### Why this Actor

- **Best-in-class Chinese ASR** — Mandarin recognition is its strongest
  language, where Western transcript tools struggle most.
- **Works without captions** — real speech recognition on the audio track;
  小红书 videos rarely carry captions to scrape.
- **Engagement context included** — every transcript ships with like / collect
  / comment / share counts, so you can rank by performance immediately.
- **Image notes filtered free** — non-video notes are detected and skipped
  without charge.
- **Direct API, no headless browser** — fast, stable runs with nothing to babysit.
- **No login, no cookies** — we never touch your accounts, so there's no ban risk.
- **Structured JSON** — export to CSV, Excel, or JSON, or pull straight from
  the API / MCP.

### Tips for better results

- Feed it the winners: find the top-performing video notes in your niche first
  (search / profile scrapers), then transcribe just those.
- Keep the language on **Chinese** for 小红书 content — it noticeably improves
  accuracy over Auto.
- Sentence timestamps let you deep-link to the exact second a phrase is spoken,
  handy for review and clip selection.
- `detectedSpeech: false` flags music-only / no-speech videos so you can filter
  them out downstream.

### FAQ

**Do I need an account, cookies, or to log in anywhere?**
No. The Actor talks to fast, direct HTTP APIs server-side — you just provide
note links and run it.

**Does it work on videos without captions?**
Yes — that's the point. It runs real speech recognition on the audio, so
captions are never required (and 小红书 videos rarely have them).

**What happens with image notes?**
They're detected, skipped, and not charged — only video notes produce
transcripts.

**How good is the Chinese accuracy?**
Mandarin is the model's strongest language; it handles fast colloquial speech,
regional accents, and product pitches well, and returns punctuated sentences.

**How am I billed?**
One fixed price per successfully transcribed video note. Notes that can't be
fetched or transcribed are not charged.

**Can I run it on a schedule or call it from my app?**
Yes — use Apify Schedules, the REST API, the JavaScript / Python clients, or
the MCP server. See the **API** tab.

**Is this affiliated with Xiaohongshu?**
No. It's an independent tool that processes publicly available content. Use it
in line with the platform's terms and applicable law.

# Actor input Schema

## `noteUrls` (type: `array`):

Xiaohongshu video note links to transcribe. Accepts explore URLs (`https://www.xiaohongshu.com/explore/…`), discovery URLs, short share links (`https://xhslink.com/…`), pasted share text, or bare note IDs. One transcript is produced per video note. Image notes are skipped (and not charged).

## `language` (type: `string`):

Language spoken in the videos. Defaults to Chinese (小红书 is Mandarin-heavy and Chinese accuracy is best-in-class). Choose Cantonese for 粤语 content, or Auto for mixed / non-Chinese audio.

## `proxyConfiguration` (type: `object`):

Optional. Route the upstream API calls through an Apify Proxy to vary the source IP. Usually not needed.

## Actor input object example

```json
{
  "noteUrls": [
    "https://www.xiaohongshu.com/explore/6907434c000000000703bf59"
  ],
  "language": "zh",
  "proxyConfiguration": {
    "useApifyProxy": false
  }
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "noteUrls": [
        "https://www.xiaohongshu.com/explore/6907434c000000000703bf59"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("ethereal_wool/xiaohongshu-video-transcript-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "noteUrls": ["https://www.xiaohongshu.com/explore/6907434c000000000703bf59"] }

# Run the Actor and wait for it to finish
run = client.actor("ethereal_wool/xiaohongshu-video-transcript-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "noteUrls": [
    "https://www.xiaohongshu.com/explore/6907434c000000000703bf59"
  ]
}' |
apify call ethereal_wool/xiaohongshu-video-transcript-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=ethereal_wool/xiaohongshu-video-transcript-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/hZqSgnlq6lfvvrGSs/builds/zRpbHxH4Gu8Qz4lxw/openapi.json
