# Walmart Scraper — Product Prices & Listings API (`nexgendata/walmart-scraper`) Actor

Fast, reliable Walmart.com scraper: feed search keywords or product URLs, get clean JSON per product — title, price, was-price, rating, review count, brand, seller, availability and stock. For pricing intelligence, catalog and e-commerce research.

- **URL**: https://apify.com/nexgendata/walmart-scraper.md
- **Developed by:** [NexGenData](https://apify.com/nexgendata) (community)
- **Categories:** E-commerce
- **Stats:** 4 total users, 2 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $3.00 / 1,000 product extracteds

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Walmart Scraper

Scrape Walmart product listings and prices — price tracking and catalog data for competitive monitoring.

> Public marketplace and product data, assembled for competitive research.

### ⏰ Run this as a daily price/listing watch (recommended)

Prices and listings move constantly. Schedule a daily sweep (`0 7 * * *`) and track changes over time — an always-on record feed.

### 📊 What you get

Clean JSON, **one record per product**:

- `title` — product name
- `price` — current price
- `was_price` — struck-through / list price when present
- `currency` — e.g. `USD`
- `rating` — average star rating
- `reviews_count` — number of reviews
- `walmart_id` — Walmart item id (`usItemId`)
- `brand`
- `seller` — e.g. `Walmart.com` or a marketplace seller
- `availability` — e.g. `IN_STOCK`
- `url` — canonical product URL
- `image_url`
- `category`
- `query` — the keyword or product URL that produced the record
- `mode` — `keyword_search` or `product_url`
- `data_source`, `as_of_timestamp`

The run summary (counts, blocked targets, status message) is written to the run's
key-value store under `OUTPUT` — it is **not** a dataset row, so you are never billed
for it.

**Pricing:** $0.003 per product returned (Pay-Per-Event) — about **333 products per $1**.
Only delivered product rows are charged. If Walmart's anti-bot wall blocks every request
the run **fails** with an explanation rather than returning an empty dataset.

**Proxy:** RESIDENTIAL is required — Walmart serves a PerimeterX challenge to datacenter IPs.

### 🤖 Use with AI agents

Point Claude, the OpenAI Agents SDK, an n8n flow or any MCP-aware client at it and pull data on demand.

**Agentic payments (x402):** Supports agentic payment via x402 — agents can call this actor with USDC, no API key required.

### 🔗 Related actors

**Commerce & pricing family — track prices and comps across marketplaces:**

[Google Shopping](https://apify.com/nexgendata/google-shopping-scraper) · [Competitor Price Monitor](https://apify.com/nexgendata/competitor-price-monitor) · [eBay Sold Comps](https://apify.com/nexgendata/ebay-sold-comps) · [Reverb Instruments](https://apify.com/nexgendata/reverb-musical-instrument-scraper) · [Craigslist Marketplace](https://apify.com/nexgendata/craigslist-scraper)

**Detect & analyze storefronts:** [Shopify Revenue Estimator](https://apify.com/nexgendata/shopify-analyzer) · [Shopify Store Detector](https://apify.com/nexgendata/shopify-store-detector) · [Shopify Store Analyzer](https://apify.com/nexgendata/shopify-store-analyzer)

***

*Public web data, assembled for research and monitoring.*

# Actor input Schema

## `keywords` (type: `array`):

List of search terms to scrape from Walmart.com (e.g. \['coffee maker', 'air fryer', 'running shoes']). Each keyword runs a Walmart search and returns one record per product. Leave empty if you are scraping specific product URLs instead.

## `productUrls` (type: `array`):

List of specific Walmart product page URLs to scrape (e.g. \['https://www.walmart.com/ip/Keurig-K-Classic/121002347']). Each URL returns a single detailed product record. Use this instead of (or in addition to) keywords when you already know exact products.

## `maxItems` (type: `integer`):

Maximum number of product records to return per run (across all keywords / URLs). Each returned product is billed once. 20-50 is appropriate for monitoring; 100+ for research sweeps.

## `proxyConfiguration` (type: `object`):

Apify proxy configuration. RESIDENTIAL is required in practice — Walmart serves a PerimeterX anti-bot wall to datacenter IPs. The actor warms a session, retries blocked requests on fresh residential sessions, and falls back to a Playwright headless browser; if every attempt is blocked the run fails instead of returning an empty dataset.

## Actor input object example

```json
{
  "keywords": [
    "coffee maker"
  ],
  "productUrls": [],
  "maxItems": 20,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}
```

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "keywords": [
        "coffee maker"
    ],
    "productUrls": [],
    "maxItems": 20,
    "proxyConfiguration": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ]
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("nexgendata/walmart-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "keywords": ["coffee maker"],
    "productUrls": [],
    "maxItems": 20,
    "proxyConfiguration": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
    },
}

# Run the Actor and wait for it to finish
run = client.actor("nexgendata/walmart-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "keywords": [
    "coffee maker"
  ],
  "productUrls": [],
  "maxItems": 20,
  "proxyConfiguration": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ]
  }
}' |
apify call nexgendata/walmart-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,nexgendata/walmart-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/bbhVTdJ4sPkyFaTlY/builds/ZK7bnkUVjzc1JrLqv/openapi.json
