# Turkey Business Directory Scraper (`nickslam/turkishyello-scraper`) Actor

Scrapes business directory data from TurkeyYello (turkishyello.com) including company information, reviews, products, and more

- **URL**: https://apify.com/nickslam/turkishyello-scraper.md
- **Developed by:** [Nick](https://apify.com/nickslam) (community)
- **Categories:** Lead generation
- **Stats:** 2 total users, 1 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $0.50 / 1,000 base results

This Actor is paid per event. You are not charged for the Apify platform usage, but only a fixed price for specific events.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## 📁 Turkey Business Directory Scraper

### What does TurkeyYello Scraper do?

Scrapes business directory data from [TurkeyYello](https://www.turkishyello.com), a Turkey online business directory (Yellow Pages–style). Search by keyword, browse categories, or filter by location, then export phones, addresses, and richer profile fields.

TurkeyYello is a national Yellow Pages–style index covering everything from Bosphorus SMEs to Anatolian manufacturers. This scraper focuses on turning those public profiles into machine-readable leads.

Runs on the **Apify platform**, so you get monitoring, scheduling, API access, dataset exports, and proxy support out of the box.

You can:

- **Filter by category** – e.g. Restaurants, Lawyers, Estate\_agents
- **Search by keywords** – e.g. "avukat", "otel", "yazılım"
- **Filter by location** – choose cities/regions such as Istanbul, Ankara, Izmir, Bursa, Antalya, Adana, Gaziantep, Konya
- **Browse broadly** – leave filters empty (within Max Results / Max Pages) to walk more of the directory

### Why scrape TurkeyYello?

- Map suppliers and B2B partners across major Turkish metros
- Build outreach lists for hospitality, legal, and trade sectors
- Monitor new listings in growing cities beyond Istanbul

### How to scrape TurkeyYello

1. Open this Actor on Apify and go to the **Input** tab.
2. Add **categories**, **search keywords**, and/or **locations** relevant to Turkey (see schema enums).
3. Set **Max Results** (for example 20) for a first test run.
4. Enable only the data add-ons you need — more fields can move you into a higher pay-per-event tier.
5. Click **Start**, then download the dataset as JSON, CSV, or Excel when the run succeeds.

### Input

Configure everything in the Actor **Input** tab. Summary of the main groups:

#### Search & filters

| Option | Description |
|--------|-------------|
| **Categories** | Business type slugs (e.g. Restaurants, Lawyers, Estate\_agents) |
| **Search Keywords** | Free-text search (e.g. "avukat", "otel", "yazılım") |
| **Locations** | Limit to Turkish cities/areas: Istanbul, Ankara, Izmir, Bursa, Antalya, Adana, Gaziantep, Konya |

#### Scope & limits

| Option | Description |
|--------|-------------|
| **Max Results** | Cap on companies to collect (start small while testing) |
| **Max Pages** | Cap on listing pages crawled |
| **Start Page** | Resume from a later listing page if a previous run stopped early |

#### Optional data add-ons

| Option | What it adds |
|--------|---------------|
| **Include Contact & Details** | Hours, year established, employees, contact person |
| **Include Extended Company Information** | Website, description, maps coordinates |
| **Include Registration & Legal** | Registration / tax identifiers when shown |
| **Include Metadata** | Listing type, verified flag, years with platform, tags |
| **Include Media & Content** | Photos, logo, products/services |
| **Include Review Summary** | Rating average and review count |
| **Include Individual Reviews** | Full review text (slower; extra pages) |

### Output example

**Core fields** (always available when present on the page):

- **Company name**
- **Company URL** – profile on TurkeyYello
- **Phone number**
- **Address**
- **Categories**
- **Location**

Optional toggles add website, description, maps, hours, registration, media, and reviews.

Example dataset item shape:

```json
[
  {
    "companyName": "Bosporus Consulting Ltd",
    "companyUrl": "https://www.turkishyello.com/company/example-profile",
    "phoneNumber": "+90 212 555 0101",
    "address": "Example Street 1, Istanbul",
    "categories": ["Consultants"],
    "location": "Istanbul"
  }
]
```

Download as **JSON**, **CSV**, or **Excel**, or pull the dataset via the Apify API.

### How much does it cost to scrape TurkeyYello?

| Tier | Price per 1,000 results | When it applies |
|------|--------------------------|-----------------|
| **Base** | $0.50 | Core fields only; no search keywords; no individual reviews |
| **Extended** | $1.00 | Any extended / contact / registration / media / review-summary / metadata options |
| **Premium** | $2.00 | Search keywords **or** Include Individual Reviews |

You can cap spend per run with Apify’s max total charge setting.

### Tips

- **Rate limiting** – keep enabled to reduce blocking risk.
- **Individual reviews** – enable only when you truly need review text.
- **Start page** – useful to continue a large crawl.
- Prefill starts in **Istanbul** — change Locations when you need Anatolian or Aegean coverage.

### Other directory scrapers

| Actor | Link |
|-------|------|
| UAE (Yello.ae) | [Open on Apify](https://apify.com/nickslam/yello-ae-scraper) |
| Poland (PolandYP) | [Open on Apify](https://apify.com/nickslam/poland-yp-scraper) |
| Kuwait (KuwaitYello) | [Open on Apify](https://apify.com/nickslam/kuwaityello-scraper) |
| Egypt (EgyptYello) | [Open on Apify](https://apify.com/nickslam/egyptyello-scraper) |

### Is it legal to scrape TurkeyYello?

This Actor collects information that businesses have published on TurkeyYello. Results may include personal data (for example a contact person). You are responsible for using the data in line with applicable laws (including GDPR where relevant) and TurkeyYello’s terms. Use the scraper only for legitimate purposes; if unsure, consult legal counsel. For feedback or bugs, use the Actor’s **Issues** tab on Apify.

# Actor input Schema

## `categories` (type: `array`):

Optional: Filter by business categories.

Example: For category page like `https://www.turkishyello.com/category/Estate_agents` enter only `Estate_agents`.

## `searchKeywords` (type: `string`):

Optional: Search for specific companies or keywords. Enables premium pricing.

## `locations` (type: `array`):

Optional: Filter by Turkey cities/regions (top locations). If not provided, browse modes use the full directory or your other filters.

## `maxResults` (type: `integer`):

Maximum number of companies to scrape

## `startPage` (type: `integer`):

Page number to start scraping from. Useful for resuming previous searches.

## `maxPages` (type: `integer`):

Maximum number of pages to scrape. Each page usually contains around 20 companies.

## `enableRateLimiting` (type: `boolean`):

Enable rate limiting (2 seconds delay) or disable (0.5 seconds delay)

## `hybridStrategy` (type: `string`):

`fast` = Cheerio + Impit HTTP (default). `browser` = Playwright (fallback).

## `includeContactDetails` (type: `boolean`):

Include business hours, year established, employees, contact person

## `includeExtendedCompanyInfo` (type: `boolean`):

Include website, description, and maps coordinates

## `includeRegistrationLegal` (type: `boolean`):

Include registration number and VAT number

## `includeMetadata` (type: `boolean`):

Include listing type, verified status, years with platform, tags

## `includeMediaContent` (type: `boolean`):

Include photos, logo, products

## `includeReviewSummary` (type: `boolean`):

Include review summary (rating and count) - Fast extraction

## `includeIndividualReviews` (type: `boolean`):

Include individual reviews - Slower (requires additional pages). Enables premium pricing.

## `proxy` (type: `object`):

Select proxies to be used by the crawler. Recommended: Use Apify Proxy with Residential proxies for better Cloudflare bypass.

## Actor input object example

```json
{
  "categories": [
    "Restaurants",
    "Consultants"
  ],
  "searchKeywords": "restaurant, IT services",
  "locations": [
    "Istanbul",
    "Ankara"
  ],
  "maxResults": 20,
  "startPage": 1,
  "maxPages": 5,
  "enableRateLimiting": true,
  "hybridStrategy": "fast",
  "includeContactDetails": false,
  "includeExtendedCompanyInfo": false,
  "includeRegistrationLegal": false,
  "includeMetadata": false,
  "includeMediaContent": false,
  "includeReviewSummary": false,
  "includeIndividualReviews": false,
  "proxy": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "TR"
  }
}
```

# Actor output Schema

## `companyData` (type: `string`):

Dataset containing all scraped company information

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "categories": [
        "Restaurants"
    ],
    "locations": [
        "Istanbul"
    ],
    "proxy": {
        "useApifyProxy": true,
        "apifyProxyGroups": [
            "RESIDENTIAL"
        ],
        "apifyProxyCountry": "TR"
    }
};

// Run the Actor and wait for it to finish
const run = await client.actor("nickslam/turkishyello-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "categories": ["Restaurants"],
    "locations": ["Istanbul"],
    "proxy": {
        "useApifyProxy": True,
        "apifyProxyGroups": ["RESIDENTIAL"],
        "apifyProxyCountry": "TR",
    },
}

# Run the Actor and wait for it to finish
run = client.actor("nickslam/turkishyello-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "categories": [
    "Restaurants"
  ],
  "locations": [
    "Istanbul"
  ],
  "proxy": {
    "useApifyProxy": true,
    "apifyProxyGroups": [
      "RESIDENTIAL"
    ],
    "apifyProxyCountry": "TR"
  }
}' |
apify call nickslam/turkishyello-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,nickslam/turkishyello-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/gp3GcfdTBibpYxqA7/builds/ilpiWwI8roYiKC318/openapi.json
