# Viator Tours Scraper (`moving_beacon-owner1/viator-tours-scraper`) Actor

Viator Tours Scraper extracts tour listings from Viator search results, collecting prices, ratings, reviews, durations, cancellation details, product links, and images. It supports lazy-loaded pages, optional pagination, and exports structured tour data to the Apify dataset.

- **URL**: https://apify.com/moving\_beacon-owner1/viator-tours-scraper.md
- **Developed by:** [Jamshaid Arif](https://apify.com/moving_beacon-owner1) (community)
- **Categories:** Travel, Automation, Developer tools
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $10.00 / 1,000 results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/platform/actors/running/actors-in-store#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Viator Tours Scraper (Apify Actor)

Renders Viator search-result pages , scrolls
to load product cards, and extracts structured tour data into the actor's dataset.

### What it collects

Per tour: `name`, `price`, `currency`, `priceText`, `rating`, `reviewCount`,
`duration`, `freeCancellation` (bool), `likelyToSellOut` (bool), `tourUrl`,
`productCode`, `imageUrl`, plus `sourceUrl` and `scrapedAt` (UTC ISO).

### How parsing works

Viator renders each product as a card whose text is one blob, e.g.
*"New York in One Day Guided Sightseeing Tour 6 hours Free Cancellation 4.8 ( 13,594 ) from $99"*.
Instead of relying on Viator's hashed CSS classes, the parser (`src/parser.py`) locates
the product cards then pulls **name / duration / rating / reviews / price** out of the
card text with regex, and reads the link/image/product-code from the card's elements.
This survives Viator's frequent class-name churn.

### Input

| Field | Type | Default | Notes |
|---|---|---|---|
| `startUrls` | array of URLs | New York search | Viator search URLs. Run a search and copy the URL. |
| `maxItems` | integer | 100 | Cap across all URLs. |
| `maxPages` | integer | 1 | Pages to paginate (needs `nextButtonSelector`). |
| `nextButtonSelector` | string | "" | Optional CSS selector for the next-page control. |
| `scrolls` | integer | 25 | Safety cap; auto-stops when no new cards load. |
| `scrollPauseSecs` | integer | 1 | Pause after each scroll. |
| `cardWaitSecs` | integer | 25 | Wait for the first cards. |
| `navigationTimeoutSecs` | integer | 60 | Page-load timeout. |
| `proxyConfiguration` | proxy | Apify RESIDENTIAL | Viator runs DataDome — residential is strongly recommended. |
| `headless` | boolean | true | Uncheck for local debugging. |

#### Example output item

```json
{
  "name": "New York in One Day Guided Sightseeing Tour",
  "price": "99",
  "currency": "$",
  "priceText": "$99",
  "rating": "4.8",
  "reviewCount": 13594,
  "duration": "6 hours",
  "freeCancellation": true,
  "likelyToSellOut": false,
  "tourUrl": "https://www.viator.com/tours/New-York-City/.../d687-5827NYC1DAY",
  "productCode": "d687-5827NYC1DAY",
  "imageUrl": "https://...",
  "sourceUrl": "https://www.viator.com/searchResults/all?text=New%20York",
  "scrapedAt": "2026-06-28T18:30:00+00:00"
}
```

# Actor input Schema

## `searchTerms` (type: `array`):

Destinations or keywords to search on Viator, e.g. \["New York", "Paris"]. Each becomes a search URL.

## `startUrls` (type: `array`):

Optional: paste full Viator search-result URLs instead of (or in addition to) search terms.

## `maxItems` (type: `integer`):

Maximum number of tours to collect across all searches.

## `impersonate` (type: `string`):

curl\_cffi impersonation target. 'chrome' picks a recent Chrome; you can pin e.g. 'chrome124'.

## `maxRetries` (type: `integer`):

How many times to retry (rotating proxy) when a request is blocked/empty.

## `requestTimeoutSecs` (type: `integer`):

Per-request timeout.

## `requestDelaySecs` (type: `integer`):

Pause between searches/retries to stay polite.

## `proxyConfiguration` (type: `object`):

Proxy settings. RESIDENTIAL strongly recommended — Viator runs DataDome and will block datacenter IPs.

## Actor input object example

```json
{
  "searchTerms": [
    "New York"
  ],
  "startUrls": [],
  "maxItems": 100,
  "impersonate": "chrome",
  "maxRetries": 2,
  "requestTimeoutSecs": 25,
  "requestDelaySecs": 3,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "searchTerms": [
        "New York"
    ],
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("moving_beacon-owner1/viator-tours-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "searchTerms": ["New York"],
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("moving_beacon-owner1/viator-tours-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print("💾 Check your data here: https://console.apify.com/storage/datasets/" + run["defaultDatasetId"])
for item in client.dataset(run["defaultDatasetId"]).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "searchTerms": [
    "New York"
  ],
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call moving_beacon-owner1/viator-tours-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "command": "npx",
            "args": [
                "mcp-remote",
                "https://mcp.apify.com/?tools=moving_beacon-owner1/viator-tours-scraper",
                "--header",
                "Authorization: Bearer <YOUR_API_TOKEN>"
            ]
        }
    }
}

```

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/xriikbzdB7efLhcJP/builds/PAf3C6mu7xXc1YO5c/openapi.json
