# Cosme Category List Scraper (`getdataforme/cosme-category-list-scraper`) Actor

Extract detailed cosmetic product data from Cosme.com category pages, including names, prices, reviews, ratings, ingredients, and JAN codes. Automate market research, competitive analysis, and price monitoring with flexible URLs, configurable limits, and reliable Playwright-based scraping....

- **URL**: https://apify.com/getdataforme/cosme-category-list-scraper.md
- **Developed by:** [GetDataForMe](https://apify.com/getdataforme) (community)
- **Categories:** AI, E-commerce, Automation
- **Stats:** 1 total users, 0 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $9.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Cosme Category List Scraper

### Introduction

The Cosme Category List Scraper is a powerful Apify Actor designed to extract detailed product information from Cosme.com category pages. It automates the collection of cosmetic product data, including names, prices, reviews, ratings, and ingredients, enabling efficient market research and competitive analysis. This tool saves time and resources by providing structured, high-quality data directly from one of Japan's leading beauty product platforms.

### Features

- **Comprehensive Data Extraction**: Scrapes key product details such as name, price, review count, rating, URL, brand, category, size, ingredients, and JAN code.
- **Flexible URL Input**: Supports multiple start URLs for scraping various categories or product lists.
- **Configurable Request Limits**: Allows setting a maximum number of requests to control data volume and avoid overloading the target site.
- **High-Quality Output**: Delivers clean, structured JSON data with Japanese text properly handled for international use.
- **Reliable Performance**: Built on PlaywrightCrawler for robust web scraping, handling dynamic content and anti-bot measures.
- **Easy Integration**: Outputs data in JSON format, compatible with Apify's export options for CSV, Excel, or direct API access.
- **Error-Resilient**: Includes built-in handling for common scraping challenges like timeouts or blocked requests.

### Input Parameters

| Parameter | Type | Required | Description | Example |
|-----------|------|----------|-------------|---------|
| startUrls | array | Yes | An array of URLs to start scraping from. Each URL should point to a Cosme category page. | `[{"url": "https://www.cosme.com/category/index.php?category_id=178"}]` |
| maxItems | number | No | The maximum number of requests to make during the scrape. This helps limit the scope and avoid excessive load. Default is 10. | `50` |

### Example Usage

#### Input

```json
{
  "startUrls": [
    {
      "url": "https://www.cosme.com/category/index.php?category_id=178"
    }
  ],
  "maxItems": 10
}
```

#### Output

```json
[
  {
    "product_name": "クリーム イン デイII&ナイトII トライアル セット / 15g×15g",
    "price": "7,700",
    "review_count": "2494",
    "rating_value": "5.8",
    "url": "https://www.cosme.com/products/detail.php?product_id=400850",
    "ブランド名": "KANEBO",
    "アイテムカテゴリ": [
      "キット・セット",
      "スキンケアキット"
    ],
    "サイズ": "15g×15g",
    "成分": "【カネボウ クリーム イン デイII】 成分:ニコチン酸アミド*、グリチルリチン酸ジカリウム*、水、マカデミアナッツ油脂肪酸フィトステリル、DPG、ソルビット液、パラメトキシケイ皮酸オクチル、ジグリセリン、濃グリセリン、シュガースクワラン、トリイソステアリン酸グリセリル、ベヘニルアルコール、BG、雲母Ti、TEA、フェニルベンゾイミダゾールスルホン酸、長鎖分岐脂肪酸コレステリル、水添大豆リン脂質、イソステアリン酸、コレステロール、パルミチン酸、パルミチン酸セチル、N-ステアロイルジヒドロスフィンゴシン、自然ビタミンE、海藻エキス-1、アルテアエキス、2,4-ビス-[{4-(2-エチルヘキシルオキシ)-2-ヒドロキシ}-フェニル]-6-(4-メトキシフェニル)-1,3,5-トリアジン、トリスエチルヘキシルオキシカルボニルアニリノトリアジン、t-ブチルメトキシジベンゾイルメタン、2-[4-(ジエチルアミノ)-2-ヒドロキシベンゾイル]安息香酸ヘキシルエステル、ジメチコン、カルボキシビニルポリマー、キサンタンガム、エデト酸塩、フェノキシエタノール、クロルフェネシン、色素504、香料 *は「有効成分」無表示は「その他の成分」 【カネボウ クリーム イン ナイトII】 成分:ナイアシンアミド*、水、シュガースクワラン、マルチトール液、マカデミアナッツ油脂肪酸フィトステリル、DPG、濃グリセリン、BG、トリイソステアリン酸グリセリル、ホホバ油、オリーブ油、パルミチン酸セチル、ベヘニルアルコール、長鎖分岐脂肪酸コレステリル、水添大豆リン脂質、イソステアリン酸、コレステロール、パルミチン酸、N-ステアロイルジヒドロスフィンゴシン、海藻エキス-1、テンシャエキス、トウヒエキス、自然ビタミンE、ワレモコウエキス、N-メチル-L-セリン、水溶性ショウキョウエキス(K)、セテアリルアルコール、チューベロースポリサッカライド液-BG、N-アミジノ-L-プロリン、フェノキシエタノール、キシリトール、ジメチコン、カルボキシビニルポリマー、クロルフェネシン、カラギーナン、水酸化K、ヒドロキシプロピルメチルセルロース、エデト酸塩、無水エタノール、香料 *は「有効成分」無表示は「その他の成分」",
    "JANコード": "4973167065587"
  }
]
```

### Use Cases

- **Market Research**: Analyze trends in cosmetic products, pricing, and consumer preferences across categories.
- **Competitive Intelligence**: Monitor competitors' product offerings, reviews, and ratings on Cosme.com.
- **Price Monitoring**: Track price changes and promotions for specific brands or items.
- **Content Aggregation**: Build databases of beauty products for blogs, apps, or e-commerce platforms.
- **Academic Research**: Study consumer behavior in the beauty industry using real-world data.
- **Business Automation**: Automate data collection for inventory management or supplier analysis.

### Installation and Usage

1. Search for "Cosme Category List Scraper" in the Apify Store
2. Click "Try for free" or "Run"
3. Configure input parameters
4. Click "Start" to begin extraction
5. Monitor progress in the log
6. Export results in your preferred format (JSON, CSV, Excel)

### Output Format

The Actor outputs data in JSON format, with each item representing a scraped product. Key fields include:

- `product_name`: The full name of the product.
- `price`: The listed price (as a string).
- `review_count`: Number of reviews.
- `rating_value`: Average rating.
- `url`: Direct link to the product page.
- `ブランド名`: Brand name.
- `アイテムカテゴリ`: Array of categories.
- `サイズ`: Product size.
- `成分`: Ingredients list.
- `JANコード`: JAN code for identification.

Data is structured for easy parsing and integration into downstream applications.

### Support

For custom/simplified outputs or bug reports, please contact:

- Email: support@getdataforme.com
- Subject line: "custom support"
- Contact form: https://getdataforme.com/contact/

We're here to help you get the most out of this Actor!

***

### PART 2: Concise Description

Unlock comprehensive cosmetic product data from Cosme.com with our scraper. Extract names, prices, reviews, ratings, ingredients, and more in structured JSON. Ideal for market research, competitive analysis, and price monitoring. Fast, reliable, and easy to use—start scraping today! (248 characters)

# Actor input Schema

## `startUrls` (type: `array`):

URLs to start with.

## `maxItems` (type: `number`):

Maximum number of requests to make.

## Actor input object example

```json
{
  "startUrls": [
    {
      "url": "https://www.cosme.com/category/index.php?category_id=178"
    }
  ],
  "maxItems": 10
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "startUrls": [
        {
            "url": "https://www.cosme.com/category/index.php?category_id=178"
        }
    ],
    "maxItems": 10
};

// Run the Actor and wait for it to finish
const run = await client.actor("getdataforme/cosme-category-list-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "startUrls": [{ "url": "https://www.cosme.com/category/index.php?category_id=178" }],
    "maxItems": 10,
}

# Run the Actor and wait for it to finish
run = client.actor("getdataforme/cosme-category-list-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "startUrls": [
    {
      "url": "https://www.cosme.com/category/index.php?category_id=178"
    }
  ],
  "maxItems": 10
}' |
apify call getdataforme/cosme-category-list-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=getdataforme/cosme-category-list-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/Mastjil8hsN7JgvNx/builds/JaFUERETb40xANDgB/openapi.json
