# Bluesky Scraper: Profiles, Posts & Followers (`scrapemint/bluesky-scraper`) Actor

Scrape Bluesky without an account: profiles with follower counts, posts with engagement stats, follower and following lists, people search, and post lookups with likers. One clean JSON row per item via the public API. No login, no API key, no browser.

- **URL**: https://apify.com/scrapemint/bluesky-scraper.md
- **Developed by:** [Ken M](https://apify.com/scrapemint) (community)
- **Categories:** Social media, Marketing
- **Stats:** 1 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$2.00 / 1,000 bluesky rows

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Bluesky Scraper: Profiles, Posts & Followers

Scrape public Bluesky data without an account, an API key or a browser. Give it handles, people searches or post URLs and get clean JSON rows: profiles with follower counts, posts with engagement stats, follower and following lists, and the accounts that liked a post.

### What you can pull

| Mode | Input | Rows you get |
| --- | --- | --- |
| Accounts | handles or profile URLs | one `profile` row per account + its recent `post` rows |
| Audience | same handles, toggle on | one row per `follower` / `following`, with full counts and bio |
| People search | keywords ("data journalist") | one `account` row per match — a lead list for any niche |
| Post lookup | bsky.app post URLs | the `post` with like/repost/reply/quote counts, plus optional `liker` rows |

Every profile-shaped row (profile, follower, following, account, liker) includes handle, display name, bio, links found in the bio, follower/following/post counts, join date and profile URL. Every post row includes text, timestamp, like/repost/reply/quote counts, links, hashtags, embed type and language.

### Example input

```json
{
    "handles": ["bsky.app", "nytimes.com"],
    "includeProfile": true,
    "includePosts": true,
    "maxPostsPerHandle": 15
}
```

### Who uses this

- **Marketers and social teams**: track competitor accounts, benchmark engagement, find creators in a niche with people search.
- **Lead generation**: export a competitor's followers or a post's likers as a warm audience, bios and links included.
- **Researchers and journalists**: collect an account's posting history with engagement counts for analysis.
- **Brand monitoring**: watch specific accounts and posts on a schedule.

### Pricing

You pay a small fee per row. Rows that give you nothing are free: unknown handles, unrecognizable or deleted post URLs, and searches with no matches. The first 2 rows of every run are also free, so you can try it for $0.

### Notes and limits

- Data comes from Bluesky's public API (the same one the app's logged-out view uses). Only public information is returned — no private accounts, no emails, no logged-in data.
- Keyword search across **all** posts is not available: Bluesky requires authentication for that endpoint, and this actor deliberately uses no account. Post scraping works per-account (author feeds) or per-post (URLs).
- Reposts and replies are excluded from author feeds by default; both have toggles.
- The shared API rate limit is generous, but if Bluesky throttles the run it stops cleanly with the rows collected so far.

# Actor input Schema

## `handles` (type: `array`):

Bluesky accounts to scrape, one per line. Accepts handles ("nytimes.com", "@user.bsky.social"), profile URLs ("https://bsky.app/profile/bsky.app") or DIDs.

## `includeProfile` (type: `boolean`):

Emit one profile row per account: display name, bio, follower/following/post counts, join date, avatar.

## `includePosts` (type: `boolean`):

Emit the account's recent posts with text, timestamps, like/repost/reply/quote counts, links and hashtags.

## `maxPostsPerHandle` (type: `integer`):

How many recent posts to fetch for each account (newest first).

## `includeReposts` (type: `boolean`):

Also emit posts the account reposted (marked with repostedByHandle). Off = original posts only.

## `includeReplies` (type: `boolean`):

Also emit the account's replies to other posts. Off = top-level posts only.

## `includeFollowers` (type: `boolean`):

Emit one profile row per follower of each account, with full follower/post counts and bio.

## `includeFollows` (type: `boolean`):

Emit one profile row per account each handle follows, with full counts and bio.

## `maxFollowersPerHandle` (type: `integer`):

Cap on follower and following rows fetched per account.

## `searchQueries` (type: `array`):

Find accounts by keyword, one search per line ("data journalist", "indie hacker"). Each match becomes a full profile row - a lead list of accounts in a niche.

## `maxAccountsPerQuery` (type: `integer`):

How many matching accounts to return for each people search.

## `postUrls` (type: `array`):

Individual posts to look up, one URL per line ("https://bsky.app/profile/bsky.app/post/3l6oveex3ii2l"). Returns the post with its engagement counts.

## `includeLikers` (type: `boolean`):

For each post URL, also emit one profile row per account that liked it - see exactly who engages with a post.

## `maxLikersPerPost` (type: `integer`):

Cap on liker rows fetched per post URL.

## `maxRows` (type: `integer`):

Stop after this many rows in total.

## Actor input object example

```json
{
  "handles": [
    "bsky.app",
    "nytimes.com"
  ],
  "includeProfile": true,
  "includePosts": true,
  "maxPostsPerHandle": 15,
  "includeReposts": false,
  "includeReplies": false,
  "includeFollowers": false,
  "includeFollows": false,
  "maxFollowersPerHandle": 200,
  "maxAccountsPerQuery": 25,
  "includeLikers": false,
  "maxLikersPerPost": 100,
  "maxRows": 2000
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "handles": [
        "bsky.app",
        "nytimes.com"
    ]
};

// Run the Actor and wait for it to finish
const run = await client.actor("scrapemint/bluesky-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = { "handles": [
        "bsky.app",
        "nytimes.com",
    ] }

# Run the Actor and wait for it to finish
run = client.actor("scrapemint/bluesky-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "handles": [
    "bsky.app",
    "nytimes.com"
  ]
}' |
apify call scrapemint/bluesky-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=scrapemint/bluesky-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/acts/aIVd5NwKEfXAfQnvS/builds/PhAQsHOcUvL1kRKKv/openapi.json
