# LinkedIn Company Scraper (`gio21/linkedin-company-scraper`) Actor

Scrape public LinkedIn company page data (industry, size, HQ, website, description) from a company URL.

- **URL**: https://apify.com/gio21/linkedin-company-scraper.md
- **Developed by:** [Gio](https://apify.com/gio21) (community)
- **Categories:** Jobs
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

$3.50 / 1,000 item scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## LinkedIn Company Scraper

Scrape **public LinkedIn company "about" pages** for company name, description, industry, company size, headquarters, founded year, and website — no login, no cookies.

### Input

| Field | Type | Description |
|-------|------|-------------|
| `companyUrl` | string | A single public LinkedIn company URL, e.g. `https://www.linkedin.com/company/google/` |
| `companyUrls` | array | Optional: several company URLs at once |
| `maxItems` | integer | Max company URLs to process (default 10) |

URLs are automatically normalized to the `/about/` page regardless of the exact form you provide.

Example:

```json
{
    "companyUrls": [
        "https://www.linkedin.com/company/google/",
        "https://www.linkedin.com/company/microsoft"
    ],
    "maxItems": 10
}
```

### Output

```json
{
    "companyName": "Google",
    "description": "Google's mission is to organize the world's information and make it universally accessible and useful.",
    "industry": "Software Development",
    "companySize": "10,001+ employees",
    "headquarters": "Mountain View, CA",
    "founded": 1998,
    "website": "https://www.google.com",
    "companyUrl": "https://www.linkedin.com/company/google/about/",
    "scrapedAt": "2026-07-17T00:00:00.000Z"
}
```

### Common uses

- Company research and firmographic enrichment
- Sales prospecting (company size, industry, HQ)
- Building a company directory or database
- Verifying a company's official website and founding year

### Notes

- `companyName` and `description` come from meta tags and are the most reliable fields — they should populate on nearly every public company page.
- `industry`, `companySize`, `headquarters`, `founded`, and `website` are parsed from the page's sidebar attribute list using label-text matching (e.g. finding "Industry" then reading the adjacent value). This is inherently more fragile than meta-tag extraction and may return `null` if LinkedIn has changed that section's layout — this does not affect `companyName`/`description`.
- If a company page is authwalled or fails to load, the item includes `error: "authwall"` or `error: "fetch-failed"` instead of failing the whole run.

# Actor input Schema

## `companyUrl` (type: `string`):

A single public LinkedIn company URL, e.g. "https://www.linkedin.com/company/google/". The scraper will normalize it to the /about/ page.

## `companyUrls` (type: `array`):

Optional: several public LinkedIn company URLs at once.

## `maxItems` (type: `integer`):

Maximum number of company URLs to process (applied across companyUrl + companyUrls combined).

## Actor input object example

```json
{
  "companyUrl": "https://www.linkedin.com/company/google/",
  "maxItems": 10
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "companyUrl": "https://www.linkedin.com/company/google/",
    "maxItems": 10
};

// Run the Actor and wait for it to finish
const run = await client.actor("gio21/linkedin-company-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "companyUrl": "https://www.linkedin.com/company/google/",
    "maxItems": 10,
}

# Run the Actor and wait for it to finish
run = client.actor("gio21/linkedin-company-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "companyUrl": "https://www.linkedin.com/company/google/",
  "maxItems": 10
}' |
apify call gio21/linkedin-company-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=gio21/linkedin-company-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/NaiPPgPfkv6APcQ76/builds/Fj7nmcJW2hOve7ROI/openapi.json
