# Qatar Business Directory Scraper (`nickslam/qataryello-scraper`) Actor

Scrapes business directory data from QatarYello (qataryello.com) including company information, reviews, products, and more

- **URL**: https://apify.com/nickslam/qataryello-scraper.md
- **Developed by:** [Nick](https://apify.com/nickslam) (community)
- **Categories:** Lead generation
- **Stats:** 3 total users, 2 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.50 / 1,000 base results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Qatar Business Directory Scraper

### Why use this Qatar business directory scraper?

- Build prospect lists for Doha and surrounding municipalities
- Track hospitality and construction-related directory categories
- Export clean datasets for outreach and market sizing

### 🎯 What this Qatar directory Actor does

Need structured leads from [QatarYello](https://www.qataryello.com)? This Apify Actor turns Qatari directory pages into a clean dataset—ideal when you want Qatar company contacts without manual browsing.

QatarYello organizes businesses around Doha-area municipalities. Ideal when your territory is Qatar-only and you care about Al Rayyan, Al Wakrah, Al Khor, and similar slugs.

Because it runs on **Apify**, you can schedule scrapes, call the Actor via API, stream results to Sheets/Make/Zapier, and rotate proxies when pages are protected.

**Highlights**

- Category + keyword discovery on qataryello.com
- Multi-city filters (Doha area listings via Al Rayyan and more)
- Tiered pay-per-event pricing aligned with how much detail you request
- Exports and integrations through Apify (Sheets, Make, Zapier, API)

### How do I use QatarYello Scraper?

1. Open this Actor on Apify and go to the **Input** tab.
2. Add **categories**, **search keywords**, and/or **locations** relevant to Qatar (see schema enums).
3. Set **Max Results** (for example 20) for a first test run.
4. Enable only the data add-ons you need — more fields can move you into a higher pay-per-event tier.
5. Click **Start**, then download the dataset as JSON, CSV, or Excel when the run succeeds.

### ⚙️ Input options

Use the **Input** tab for full field help. Quick reference:

#### Search & filters

| Option | Description |
|--------|-------------|
| **Categories** | Business type slugs (e.g. restaurants, consultants, estate-agents) |
| **Search Keywords** | Free-text search (e.g. "engineering", "catering", "real estate") |
| **Locations** | Limit to Qatari cities/areas: Doha area listings via Al Rayyan, Al Wakrah, Al Khor, Al Ruwais, and related locations |

#### Scope & limits

| Option | Description |
|--------|-------------|
| **Max Results** | Cap on companies to collect (start small while testing) |
| **Max Pages** | Cap on listing pages crawled |
| **Start Page** | Resume from a later listing page if a previous run stopped early |

#### Optional data add-ons

| Option | What it adds |
|--------|---------------|
| **Include Contact & Details** | Hours, year established, employees, contact person |
| **Include Extended Company Information** | Website, description, maps coordinates |
| **Include Registration & Legal** | Registration / tax identifiers when shown |
| **Include Metadata** | Listing type, verified flag, years with platform, tags |
| **Include Media & Content** | Photos, logo, products/services |
| **Include Review Summary** | Rating average and review count |
| **Include Individual Reviews** | Full review text (slower; extra pages) |

### 📋 What data you get

Core fields:

- **Company name**
- **Company URL** – profile on QatarYello
- **Phone number**
- **Address**
- **Categories**
- **Location**

Sample output:

```json
[
  {
    "companyName": "Pearl Engineering Co",
    "companyUrl": "https://www.qataryello.com/company/example-profile",
    "phoneNumber": "+974 4444 5566",
    "address": "Example Street 1, Al Rayyan",
    "categories": ["Engineering"],
    "location": "Al Rayyan"
  }
]
```

### How much does it cost to scrape QatarYello?

| Tier | Price per 1,000 results | When it applies |
|------|--------------------------|-----------------|
| **Base** | $0.50 | Core fields only; no search keywords; no individual reviews |
| **Extended** | $1.00 | Any extended / contact / registration / media / review-summary / metadata options |
| **Premium** | $2.00 | Search keywords **or** Include Individual Reviews |

You can cap spend per run with Apify’s max total charge setting.

### 💡 Tips for better runs

- **Rate limiting** – keep enabled to reduce blocking risk.
- **Individual reviews** – enable only when you truly need review text.
- **Start page** – useful to continue a large crawl.
- Pick municipality Locations that match your territory — QatarYello organizes listings by area slug.

### Other directory scrapers

| Actor | Link |
|-------|------|
| UAE (Yello.ae) | [Open on Apify](https://apify.com/nickslam/yello-ae-scraper) |
| Kuwait (KuwaitYello) | [Open on Apify](https://apify.com/nickslam/kuwaityello-scraper) |
| Singapore (Yelu.sg) | [Open on Apify](https://apify.com/nickslam/yelu-sg-scraper) |
| Hong Kong (Yelo.hk) | [Open on Apify](https://apify.com/nickslam/yelo-hk-scraper) |

### FAQ & legality

This Actor collects information that businesses have published on QatarYello. Results may include personal data (for example a contact person). You are responsible for using the data in line with applicable laws (including GDPR where relevant) and QatarYello’s terms. Use the scraper only for legitimate purposes; if unsure, consult legal counsel. For feedback or bugs, use the Actor’s **Issues** tab on Apify.

# Actor input Schema

## `categories` (type: `array`):

Optional: Filter by business categories.

Example: For category page like `https://www.qataryello.com/category/Estate_agents` enter only `Estate_agents`.

## `searchKeywords` (type: `string`):

Optional: Search for specific companies or keywords. Enables premium pricing.

## `locations` (type: `array`):

Optional: Filter by Qatar cities/regions. If not provided, all cities will be scraped.

## `maxResults` (type: `integer`):

Maximum number of companies to scrape

## `startPage` (type: `integer`):

Page number to start scraping from. Useful for resuming previous searches.

## `maxPages` (type: `integer`):

Maximum number of pages to scrape. Each page usually contains around 20 companies.

## `enableRateLimiting` (type: `boolean`):

Enable rate limiting (2 seconds delay) or disable (0.5 seconds delay)

## `hybridStrategy` (type: `string`):

`fast` = Cheerio + Impit HTTP (default). `browser` = Playwright (fallback).

## `includeContactDetails` (type: `boolean`):

Include business hours, year established, employees, contact person

## `includeExtendedCompanyInfo` (type: `boolean`):

Include website, description, and maps coordinates

## `includeRegistrationLegal` (type: `boolean`):

Include registration number and VAT number

## `includeMetadata` (type: `boolean`):

Include listing type, verified status, years with platform, tags

## `includeMediaContent` (type: `boolean`):

Include photos, logo, products

## `includeReviewSummary` (type: `boolean`):

Include review summary (rating and count) - Fast extraction

## `includeIndividualReviews` (type: `boolean`):

Include individual reviews - Slower (requires additional pages). Enables premium pricing.

## `proxy` (type: `object`):

Select proxies to be used by the crawler. Recommended: Use Apify Proxy with Residential proxies for better Cloudflare bypass.

## Actor input object example

```json
{
  "categories": [
    "Restaurants",
    "Lawyers"
  ],
  "searchKeywords": "restaurant, IT services",
  "locations": [
    "Doha",
    "Al_Rayyan"
  ],
  "maxResults": 20,
  "startPage": 1,
  "maxPages": 5,
  "enableRateLimiting": true,
  "hybridStrategy": "fast",
  "includeContactDetails": false,
  "includeExtendedCompanyInfo": false,
  "includeRegistrationLegal": false,
  "includeMetadata": false,
  "includeMediaContent": false,
  "includeReviewSummary": false,
  "includeIndividualReviews": false,
  "proxy": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "QA"
  }
}
```

# Actor output Schema

## `companyData` (type: `string`):

Dataset containing all scraped company information

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "categories": [
        "Restaurants"
    ],
    "locations": [
        "Doha"
    ],
    "proxy": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ],
        "apifyProxyCountry": "QA"
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("nickslam/qataryello-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "categories": ["Restaurants"],
    "locations": ["Doha"],
    "proxy": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
        "apifyProxyCountry": "QA",
    },
}

# Run the Actor and wait for it to finish
run = client.actor("nickslam/qataryello-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "categories": [
    "Restaurants"
  ],
  "locations": [
    "Doha"
  ],
  "proxy": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "QA"
  }
}' |
apify call nickslam/qataryello-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,nickslam/qataryello-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/ZHq57V6WnqzSQXaWr/builds/fu2ygcpFqZI1qqWex/openapi.json
