# Eroski Spain Scraper — Groceries, Prices & Brands (`studio-amba/eroski-es-scraper`) Actor

Scrape the Eroski Spain online grocery shop by keyword: product names, brands, prices in EUR and images. No login or cookies. Returns the first page of search results.

- **URL**: https://apify.com/studio-amba/eroski-es-scraper.md
- **Developed by:** [Studio Amba](https://apify.com/studio-amba) (community)
- **Categories:** E-commerce
- **Stats:** 2 total users, 1 monthly users, 37.5% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $2.00 / 1,000 result scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Eroski Spain Scraper

Scrape the Eroski Spain online grocery shop (supermercado.eroski.es) by keyword and get clean, structured product data: names, brands, prices in EUR and images.

Eroski is one of Spain's largest supermarket cooperatives. This actor reads the shop's server-rendered search results directly — no login, no cookies, no browser required.

### Why use this actor?

If you track grocery prices, build a price-comparison tool, or feed a market-research dataset, you need structured product data you can rely on. This actor turns any Eroski search into a table of products with prices in EUR, brands and images.

Typical users:

- Price-intelligence and price-comparison platforms tracking Spanish retail.
- Brands and suppliers monitoring how their products are listed at Eroski.
- Market researchers and analysts building grocery datasets across Spain.
- Developers who need a clean product feed instead of scraping HTML themselves.

### How to scrape Eroski Spain data

1. Open the actor in the Apify Console.
2. Enter a **search query** in Spanish — for example `leche`, `pan` or `aceite`.
3. Leave the proxy on the default Apify setting and click **Start**.
4. When the run finishes, download the results as JSON, CSV, Excel or feed them to an API.

The actor fetches Eroski's server-rendered search results page for your query and extracts each product's structured event-tracking data (name, price, brand, id) plus its listing image — no HTML price-text parsing needed.

### Input

| Field | Type | Required | Description |
|-------|------|----------|--------------|
| `searchQuery` | String | No | Keyword to search in Spanish (default: `leche`) |
| `maxResults` | Integer | No | Maximum products to return, capped at ~20 — see Limitations below (default: 20) |
| `proxyConfiguration` | Object | No | Apify proxy settings (default Apify proxy works) |

#### Example input

```json
{
    "searchQuery": "leche",
    "maxResults": 20,
    "proxyConfiguration": { "useApifyProxy": true }
}
```

### Output

Each result contains:

| Field | Type | Example |
|-------|------|---------|
| `productName` | String | `"Leche entera del País Vasco EROSKI, brik 1 litro"` |
| `brand` | String | `"EROSKI"` |
| `price` | Number | `1.09` |
| `currency` | String | `"EUR"` |
| `originalPrice` | Number | `1.29` |
| `discount` | String | `"-16%"` |
| `productId` | String | `"18672295"` |
| `imageUrl` | String | Primary product image URL |
| `url` | String | Full product page URL |
| `scrapedAt` | String | ISO 8601 timestamp |

`originalPrice` and `discount` are present only on promoted products.

### Example output

```json
{
    "productName": "Leche entera del País Vasco EROSKI, brik 1 litro",
    "brand": "EROSKI",
    "price": 1.09,
    "currency": "EUR",
    "productId": "18672295",
    "imageUrl": "https://supermercado.eroski.es//images/18672295.jpg",
    "url": "https://supermercado.eroski.es:443/es/productdetail/18672295-leche-entera-del-pais-vasco-eroski-brik-1-litro/",
    "scrapedAt": "2026-07-13T09:00:00.000Z"
}
```

### Notes and limitations

- **Page-1-only (~20 products per query).** Eroski's search results paginate through a session-bound Tapestry endpoint (`POST /es/search/results:loadpage`) that requires replaying an exact browser session bootstrap. Live testing found this replay unreliable outside a real browser session — it works inconsistently and can silently return the same page again or a session error. Rather than ship a pagination path that quietly breaks, this actor deliberately returns only the first page of results (Eroski shows ~20 products per search). Run multiple specific keyword searches to cover more of the catalogue.
- **Unit price (price-per-kilo/litre) is not populated.** Eroski's page includes a placeholder for it, but it renders empty in the server-rendered HTML for every product tested — it appears to be filled in by client-side JavaScript that a plain HTTP fetch doesn't trigger.
- **EAN / barcode is not visible** on the search results cards, so it is not included in the output.
- Prices are the current online-shop prices in EUR.

### Cost

Pricing is pay-per-result. A typical search of ~20 products completes in a few seconds, so most runs cost only a few cents in Apify platform usage plus the per-result fee.

### FAQ

**Why only ~20 results per search?**
Eroski's own search page shows 20 products per page and loads more via a session-bound endpoint we chose not to rely on (see Limitations above). Running several targeted keyword searches (e.g. `leche entera`, `leche desnatada`, `leche sin lactosa`) covers a category more completely than trying to force pagination.

**Does this need a login or store selection?**
No. The actor reads the public search results page directly. Eroski does let shoppers pick a physical store for delivery slots, but product listing and pricing data doesn't require it.

### Related Scrapers

Build a complete view of European grocery retail with these sibling scrapers from Studio AMBA:

- **[Alcampo Spain Scraper](https://apify.com/studio-amba/alcampo-es-scraper)** — Spanish supermarket products and prices (alcampo.es).
- **[Mercadona Scraper](https://apify.com/studio-amba/mercadona-scraper)** — Spain's largest supermarket (mercadona.es).
- **[Dia ES Scraper](https://apify.com/studio-amba/dia-es-scraper)** — Spanish supermarket products and prices (dia.es).
- **[Continente PT Scraper](https://apify.com/studio-amba/continente-pt-scraper)** — Portuguese groceries with EAN barcodes (continente.pt).
- **[BILLA Austria Scraper](https://apify.com/studio-amba/billa-at-scraper)** — Austrian supermarket products and prices (billa.at).

# Actor input Schema

## `searchQuery` (type: `string`):

Keyword to search the Eroski catalogue in Spanish (e.g. 'leche', 'pan', 'aceite').

## `maxResults` (type: `integer`):

Maximum number of products to return. Eroski's search page returns about 20 products per query (see Limitations in the README) — this caps how many of those are pushed.

## `proxyConfiguration` (type: `object`):

Apify proxy settings. Default Apify (datacenter) proxy works — no anti-bot detected on the Eroski search page.

## Actor input object example

```json
{
  "searchQuery": "leche",
  "maxResults": 20,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "searchQuery": "leche",
    "maxResults": 20,
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("studio-amba/eroski-es-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "searchQuery": "leche",
    "maxResults": 20,
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("studio-amba/eroski-es-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "searchQuery": "leche",
  "maxResults": 20,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call studio-amba/eroski-es-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=studio-amba/eroski-es-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/27v89x5KgYIchHyJD/builds/U15dbzWNk9iKEKzlh/openapi.json
