# CDiscount Scraper — French E-commerce Products, Prices & Deals (`studio-amba/cdiscount-scraper`) Actor

Scrape products, prices, specs, ratings, and deals from Cdiscount.com — France's #2 e-commerce platform. Supports search and category browsing. No login required.

- **URL**: https://apify.com/studio-amba/cdiscount-scraper.md
- **Developed by:** [Studio Amba](https://apify.com/studio-amba) (community)
- **Categories:** E-commerce
- **Stats:** 3 total users, 2 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $5.00 / 1,000 result scrapeds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Cdiscount Scraper

Scrape products, prices, specifications, ratings, and availability from Cdiscount.com — France's second-largest e-commerce platform.

### Why use this actor?

Cdiscount is one of France's biggest online marketplaces with millions of products across electronics, home, fashion, and more. This actor lets you extract structured product data at scale for price monitoring, competitor analysis, market research, and lead generation without building your own scraper infrastructure.

### Input

| Field | Type | Required | Description |
|-------|------|----------|-------------|
| `startUrls` | Array | No | Category or product URLs to scrape |
| `searchQuery` | String | No | Search for products by keyword (e.g., "iPhone 16", "aspirateur robot") |
| `maxResults` | Integer | No | Maximum results to return (default: 100) |
| `proxyConfiguration` | Object | No | Proxy settings (recommended for large runs) |

If neither `startUrls` nor `searchQuery` is provided, the actor defaults to searching for "smartphone".

### Output

Each result contains:

| Field | Type | Example |
|-------|------|---------|
| `name` | String | `"Samsung Galaxy S25 Ultra 256Go"` |
| `brand` | String | `"Samsung"` |
| `price` | Number | `1099.99` |
| `currency` | String | `"EUR"` |
| `originalPrice` | Number | `1299.99` |
| `sku` | String | `"SAM-S25U-256"` |
| `ean` | String | `"8806095577890"` |
| `inStock` | Boolean | `true` |
| `rating` | Number | `4.5` |
| `reviewCount` | Number | `87` |
| `url` | String | Full product URL |
| `imageUrl` | String | Primary product image URL |
| `imageUrls` | Array | All product image URLs |
| `specs` | Object | `{"Stockage": "256 Go", "RAM": "12 Go"}` |
| `category` | String | `"Téléphonie > Smartphones"` |
| `description` | String | Product description text |
| `seller` | String | `"Cdiscount"` |

### Example output

```json
{
    "name": "Samsung Galaxy S25 Ultra 256Go Bleu Titane",
    "brand": "Samsung",
    "price": 1099.99,
    "originalPrice": 1299.99,
    "currency": "EUR",
    "sku": "SAM-S25U-256BT",
    "inStock": true,
    "rating": 4.5,
    "reviewCount": 87,
    "url": "https://www.cdiscount.com/telephonie/telephone-mobile/samsung-galaxy-s25-ultra/f-1440404-sam8806095577890.html",
    "imageUrl": "https://www.cdiscount.com/pdt2/5/7/7/1/700x700/sam8806095577890/rw/samsung-galaxy-s25-ultra.jpg",
    "specs": {
        "Stockage": "256 Go",
        "RAM": "12 Go",
        "Taille d'ecran": "6.9 pouces"
    },
    "category": "Téléphonie > Smartphones",
    "description": "Samsung Galaxy S25 Ultra avec S Pen intégré...",
    "seller": "Cdiscount",
    "scrapedAt": "2026-04-03T14:30:00.000Z"
}
```

### Cost estimate

This actor uses Playwright (browser-based) to bypass Cloudflare protection. It charges $0.04 per run (start fee, based on its 4 GB memory allocation) plus $0.005 per result scraped. 100 results costs about $0.54; 1,000 results about $5.04. A run's usage cost only settles after it reports SUCCEEDED.

### Tips

- **Use proxies** — Cdiscount has Cloudflare protection. Apify proxy (residential group) is recommended for reliable scraping.
- **Start with search** — Searching by keyword is the easiest way to find products.
- **Category pages** — Paste category URLs directly from cdiscount.com for targeted scraping.
- **Keep maxResults reasonable** — Start with 50-100 to test, then scale up.

### Limitations

- Cloudflare protection means the actor requires a browser (Playwright), making it slower than HTTP-based scrapers
- Very large runs (1000+ products) may need residential proxies for consistent results
- Some marketplace seller pages may have different HTML structures
- Data is scraped from the public website and may change without notice
- Respect the website's terms of service and use responsibly

### How to scrape Cdiscount data

1. Go to this actor's page on the [Apify Store](https://apify.com/store).
2. Click **Try for free** to open it in Apify Console.
3. Configure your search query or URL, set the maximum number of results, and adjust proxy settings if needed.
4. Click **Start** and wait for the run to finish.
5. Download your data in JSON, CSV, Excel, or connect it to your workflow via API.

You can also schedule regular runs, set up webhooks for real-time notifications, or integrate the results directly into your application using the [Apify API](https://docs.apify.com/api).

### Features

- **No login required** — scrapes publicly available data from Cdiscount without needing credentials or cookies.
- **Structured output** — results are returned as clean JSON objects, ready for processing.
- **Pagination handling** — automatically follows multiple pages of results.
- **Proxy support** — configurable proxy settings for reliable, large-scale scraping.
- **Flexible input** — search by keyword, provide specific URLs, or crawl categories.
- **Scheduled runs** — run on a schedule to keep your dataset up to date automatically.
- **API access** — integrate results into your workflow using the Apify API or webhooks.

### FAQ

**Is it legal to scrape Cdiscount?**
Web scraping of publicly available data is generally permitted. This actor only accesses information that is publicly visible to any website visitor. Always review the website's terms of service before scraping.

**How often should I run this scraper?**
For price monitoring or competitive intelligence, daily or weekly runs are common. Set up a [schedule](https://docs.apify.com/schedules) in Apify Console to automate this.

**Can I export the data to Google Sheets or Excel?**
Yes. After each run, you can download results in CSV, JSON, or Excel format directly from Apify Console. You can also connect results to Google Sheets using Apify integrations.

**What if the scraper stops working?**
Websites change their structure occasionally. If you notice issues, please open an issue on the actor's page. We actively maintain this scraper and fix issues promptly.

# Actor input Schema

## `startUrls` (type: `array`):

Cdiscount search, category, or product pages to scrape. Example: https://www.cdiscount.com/search/10/aspirateur.html

## `searchQuery` (type: `string`):

Search for products by name or keyword (e.g., 'iPhone 16', 'aspirateur robot'). Used when no Start URLs are provided.

## `maxResults` (type: `integer`):

Maximum number of products to return.

## `sbrWsCdp` (type: `string`):

Bright Data Scraping Browser WebSocket CDP endpoint. Required to bypass Cdiscount's anti-bot protection (Baleen). Can also be set as SBR\_WS\_CDP environment variable.

## `proxyConfiguration` (type: `object`):

Select proxies to be used by this actor.

## Actor input object example

```json
{
  "searchQuery": "smartphone",
  "maxResults": 20,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# Actor output Schema

## `results` (type: `string`):

No description

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "searchQuery": "smartphone",
    "maxResults": 20,
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("studio-amba/cdiscount-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "searchQuery": "smartphone",
    "maxResults": 20,
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("studio-amba/cdiscount-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "searchQuery": "smartphone",
  "maxResults": 20,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call studio-amba/cdiscount-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,studio-amba/cdiscount-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/5Qp9GAcq0T1aqqCww/builds/hKHowkJirZMRalyX6/openapi.json
