# Greenhouse Jobs Scraper (`crawlerbros/greenhouse-jobs-scraper`) Actor

Scrape job listings from Greenhouse.io company boards. Get all jobs from one or more companies, or fetch specific job details by ID - with location, departments, salary, and apply links.

- **URL**: https://apify.com/crawlerbros/greenhouse-jobs-scraper.md
- **Developed by:** [Crawler Bros](https://apify.com/crawlerbros) (community)
- **Categories:** Jobs, Automation, Agents
- **Stats:** 1 total users, 0 monthly users, 100.0% runs succeeded, 0 bookmarks
- **User rating**: No ratings yet

## Pricing

from $3.00 / 1,000 results

This Actor is paid per event and usage. You are charged both the fixed price for specific events and for Apify platform usage.
Since this Actor supports Apify Store discounts, the price gets lower the higher subscription plan you have.

Learn more: https://docs.apify.com/actors/running/actors-in-store.md#pay-per-event

## What's an Apify Actor?

Actors are web data automations that power AI and operations. They run on the Apify platform to scrape websites, process data, connect APIs, and automate workflows.
In Batch mode, an Actor accepts a well-defined JSON input, performs an action which can take anything from a few seconds to a few hours,
and optionally produces a well-defined JSON output, datasets with results, or files in key-value store.
In Standby mode, an Actor provides a web server which can be used as a website, API, or an MCP server.
Actors are written with capital "A".

## How to integrate an Actor?

If asked about integration, you help developers integrate Actors into their projects.
You adapt to their stack and deliver integrations that are safe, well-documented, and production-ready.
The best way to integrate Actors is as follows.

- **AI agents and MCP clients** — the [Apify MCP server](https://docs.apify.com/integrations/mcp.md) at `https://mcp.apify.com` (remote, streamable HTTP, OAuth on first use).
- **Agentic workflows and local Actor development** — [Agent Skills](https://apify.com/.well-known/agent-skills/index.json) with the [Apify CLI](https://docs.apify.com/cli/docs.md): `npm install -g apify-cli`, then `apify login`.
- **JavaScript/TypeScript projects** — the official [JS/TS client](https://docs.apify.com/api/client/js/docs.md): `npm install apify-client`.
- **Python projects** — the official [Python client](https://docs.apify.com/api/client/python/docs.md): `pip install apify-client`.
- **Any other language** — the [REST API](https://docs.apify.com/api/v2.md).

For usage examples, see the [API](#api) section below.

For more details, see Apify documentation as [Markdown index](https://docs.apify.com/llms.txt) and [Markdown full-text](https://docs.apify.com/llms-full.txt).

# README

## Greenhouse Jobs Scraper

Extract job listings from any [Greenhouse.io](https://www.greenhouse.io/) company board. Get all open positions for one or more companies, or fetch specific job details by ID — all via the Greenhouse public REST API with no authentication required.

### What Does Greenhouse Jobs Scraper Do?

Greenhouse is one of the most popular applicant tracking systems used by thousands of companies including Airbnb, Stripe, Shopify, Figma, and many more. This actor lets you:

- **Scrape all open jobs** from one or more Greenhouse company boards in a single run
- **Fetch specific jobs** by company slug and job ID for targeted lookups
- Collect **rich job metadata**: title, location (city/state/country), departments, offices, employment type, salary range (where available), apply links, and timestamps

### Output Data

Each job record contains:

| Field | Description |
|-------|-------------|
| `id` | Greenhouse job ID (integer) |
| `title` | Job title |
| `companySlug` | Greenhouse company board slug |
| `location` | Location object with `name`, `city`, `state`, `country` |
| `departments` | List of department names |
| `offices` | List of office names |
| `employmentType` | Employment type (e.g., "Full-time", "Contract") |
| `educationLevel` | Required education level if specified |
| `salary` | Salary range object if available in metadata |
| `requisitionId` | Internal requisition ID |
| `updatedAt` | Last updated timestamp |
| `applyUrl` | Direct application URL |
| `absoluteUrl` | Full job listing URL on Greenhouse |
| `sourceUrl` | Constructed board URL for the job |
| `content` | Full HTML job description (only if `content=true`) |
| `recordType` | Always `"job"` |
| `scrapedAt` | ISO 8601 UTC timestamp of scrape |

#### Sample Output

```json
{
  "id": 4567890,
  "title": "Senior Software Engineer",
  "companySlug": "stripe",
  "location": {
    "name": "San Francisco, CA, USA",
    "city": "San Francisco",
    "state": "CA",
    "country": "USA"
  },
  "departments": ["Engineering", "Platform"],
  "offices": ["San Francisco HQ"],
  "employmentType": "Full-time",
  "applyUrl": "https://boards.greenhouse.io/stripe/jobs/4567890",
  "absoluteUrl": "https://boards.greenhouse.io/stripe/jobs/4567890",
  "sourceUrl": "https://boards.greenhouse.io/stripe/jobs/4567890",
  "updatedAt": "2024-06-15T10:00:00Z",
  "recordType": "job",
  "scrapedAt": "2026-05-30T12:00:00+00:00"
}
```

### Input Configuration

| Field | Type | Required | Description |
|-------|------|----------|-------------|
| `mode` | select | Yes | `getJobs` — all jobs from boards; `getJob` — specific job by ID |
| `companySlugs` | array | For getJobs | Greenhouse board slugs (e.g. `["stripe", "airbnb"]`) |
| `companySlug` | string | For getJob | Single company board slug |
| `jobIds` | array | For getJob | List of Greenhouse job IDs (integers) |
| `content` | boolean | No | Include full HTML job description (default: `false`) |
| `maxItems` | integer | No | Max records to return (1–10,000, default: 500) |

#### How to Find a Company Slug

The company slug is the identifier in the Greenhouse board URL. For example:

- `https://boards.greenhouse.io/stripe` → slug is `stripe`
- `https://boards.greenhouse.io/airbnb` → slug is `airbnb`

### Usage Examples

#### Get all jobs from multiple companies

```json
{
  "mode": "getJobs",
  "companySlugs": ["stripe", "airbnb", "shopify"],
  "maxItems": 1000
}
```

#### Fetch specific jobs by ID

```json
{
  "mode": "getJob",
  "companySlug": "stripe",
  "jobIds": [4567890, 4567891],
  "content": true
}
```

#### Get jobs with full description

```json
{
  "mode": "getJobs",
  "companySlugs": ["figma"],
  "content": true,
  "maxItems": 100
}
```

### Use Cases

- **Job market research** — track hiring trends across tech companies
- **Competitive intelligence** — monitor which roles competitors are filling
- **Job aggregation** — build job boards or career sites with Greenhouse data
- **Recruitment analytics** — analyze hiring patterns by department or location
- **HR benchmarking** — compare salary ranges and job requirements across companies

### Frequently Asked Questions

**Is authentication required?**
No. The Greenhouse public job board API is completely open — no API key or account needed.

**How do I find a company's Greenhouse slug?**
Visit the company's job listings page. If it's hosted on Greenhouse, the URL will contain `greenhouse.io`. The slug is the part after `/boards/` or before `/jobs`.

**Does this work for all companies on Greenhouse?**
Yes, for any company using Greenhouse's public job board. Some companies may use private boards that are not publicly accessible.

**Can I get the full job description?**
Yes — enable the `content` option. This returns the full HTML description for each job.

**How many jobs can I scrape?**
Up to 10,000 jobs per run. The Greenhouse API returns all jobs at once (no pagination).

**How fresh is the data?**
Jobs are fetched in real-time from the Greenhouse API. Each record includes an `updatedAt` timestamp from Greenhouse.

**What if a company slug is invalid?**
The actor will log a warning and continue with other slugs. No error will be thrown.

### Data Source

Data is sourced directly from the [Greenhouse.io public board API](https://boards.greenhouse.io/v1/boards/) — the same API that powers company career pages built on Greenhouse. No scraping of rendered HTML is performed.

# Actor input Schema

## `mode` (type: `string`):

What to fetch: all jobs from company boards, or a specific job by ID.

## `companySlugs` (type: `array`):

List of Greenhouse company board slugs to scrape (e.g. \["google", "airbnb"]). Find the slug in the Greenhouse board URL: boards.greenhouse.io/v1/boards/{slug}/jobs

## `companySlug` (type: `string`):

Single company board slug for fetching a specific job (e.g. "airbnb").

## `jobIds` (type: `array`):

List of Greenhouse job IDs (integers) to fetch.

## `content` (type: `boolean`):

When enabled, each job record will include the full job description as HTML.

## `maxItems` (type: `integer`):

Maximum number of job records to emit.

## Actor input object example

```json
{
  "mode": "getJobs",
  "companySlugs": [
    "airbnb"
  ],
  "content": false,
  "maxItems": 500
}
```

# Actor output Schema

## `jobs` (type: `string`):

Dataset containing all scraped Greenhouse job listings.

# API

You can run this Actor programmatically using our API. Below are code examples in JavaScript, Python, and CLI, as well as the OpenAPI specification and MCP server setup.

## JavaScript example

```javascript
import { ApifyClient } from 'apify-client';

// Initialize the ApifyClient with your Apify API token
// Replace the '<YOUR_API_TOKEN>' with your token
const client = new ApifyClient({
    token: '<YOUR_API_TOKEN>',
});

// Prepare Actor input
const input = {
    "mode": "getJobs",
    "companySlugs": [
        "airbnb"
    ],
    "content": false,
    "maxItems": 500
};

// Run the Actor and wait for it to finish
const run = await client.actor("crawlerbros/greenhouse-jobs-scraper").call(input);

// Fetch and print Actor results from the run's dataset (if any)
console.log('Results from dataset');
console.log(`💾 Check your data here: https://console.apify.com/storage/datasets/${run.defaultDatasetId}`);
const { items } = await client.dataset(run.defaultDatasetId).listItems();
items.forEach((item) => {
    console.dir(item);
});

// 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/js/docs

```

## Python example

```python
from apify_client import ApifyClient

# Initialize the ApifyClient with your Apify API token
# Replace '<YOUR_API_TOKEN>' with your token.
client = ApifyClient("<YOUR_API_TOKEN>")

# Prepare the Actor input
run_input = {
    "mode": "getJobs",
    "companySlugs": ["airbnb"],
    "content": False,
    "maxItems": 500,
}

# Run the Actor and wait for it to finish
run = client.actor("crawlerbros/greenhouse-jobs-scraper").call(run_input=run_input)

# Fetch and print Actor results from the run's dataset (if there are any)
print(f"💾 Check your data here: https://console.apify.com/storage/datasets/{run.default_dataset_id}")
for item in client.dataset(run.default_dataset_id).iterate_items():
    print(item)

# 📚 Want to learn more 📖? Go to → https://docs.apify.com/api/client/python/docs/quick-start

```

## CLI example

```bash
echo '{
  "mode": "getJobs",
  "companySlugs": [
    "airbnb"
  ],
  "content": false,
  "maxItems": 500
}' |
apify call crawlerbros/greenhouse-jobs-scraper --silent --output-dataset

```

## MCP server setup

```json
{
    "mcpServers": {
        "apify": {
            "type": "http",
            "url": "https://mcp.apify.com/?tools=fetch-actor-details,crawlerbros/greenhouse-jobs-scraper"
        }
    }
}

```

The hosted server signs you in with OAuth on first connect, so no API token belongs in this config. Clients without OAuth support can send an `Authorization: Bearer <APIFY_API_TOKEN>` header instead, using a token from API & Integrations in Apify Console (https://console.apify.com/settings/integrations).

## OpenAPI specification

Download the OpenAPI definition: https://api.apify.com/v2/actors/6ylq14nIEiCbd57ye/builds/7tqNRtHWh7AqoVi6y/openapi.json
