Crunchbase - 117K Company DB, Funding, Investors & News ($8/1k)
Pricing
from $8.00 / 1,000 results
Crunchbase - 117K Company DB, Funding, Investors & News ($8/1k)
Crunchbase four ways: live company scrape (Cloudflare handled), the Instant Database of 117K+ profiles served in seconds, 19,000+ VC firms with contacts at no surcharge, and Crunchbase News articles that name the companies they cover. Funding Monitor returns only NEW rounds per run. $8/1k.
Pricing
from $8.00 / 1,000 results
Rating
5.0
(2)
Developer
Muhamed Didovic
Maintained by CommunityActor stats
3
Bookmarked
142
Total users
57
Monthly active users
24 hours
Issues response
5 hours ago
Last modified
Categories
Share
Crunchbase Scraper — Investors Database, Companies, Funding Rounds & News
How It Works

Turn any Crunchbase company URL or slug into one clean, structured row — no raw blob to untangle.
Paste a company link, a bare slug, or an organization/... path and get a ready-to-use profile: identity, funding, people, M&A, tech stack, web traffic, IT spend, and Crunchbase's own growth/funding/acquisition/IPO predictions.
Or paste a Crunchbase Discover / saved-search URL (e.g. crunchbase.com/discover/funding_rounds/…) and get one funding-round signal row per result — company, round type, and Crunchbase links — with each company enriched from its org page. Crunchbase caps Discover at the first 15 results per search (its own paywall, not a scraper limit — details), so pull big lists by pasting several narrower searches rather than one broad one.
JSON or CSV out. No unblocker token to manage — Cloudflare is handled for you.
💼 Startup Investors Database — 19,000+ VC firms, accelerators & angels
Flip on investorDatabase (or set any investor filter) and the actor returns investor firm rows instead of companies: 19,000+ venture capital firms, accelerators, angel investors, family offices and grant programs, each derived from live Crunchbase funding-round participation — not a stale directory dump:
- Who they are — firm name, Crunchbase URL, inferred type (VC / Seed & Early-Stage VC / Accelerator / Angel / Grant Program / Debt), and where available HQ city + country, website, contact email, phone, description, Crunchbase rank
- How they invest — observed deal count, full stage distribution (Pre-Seed → Series D+, Grants, Debt), top stage, focus areas (industries of their actual portfolio), portfolio countries
- What they've backed — the portfolio companies themselves, with Crunchbase permalinks, ready to cross-reference against the 117K+ company database below
- Filter by name/keyword, firm type, investment stage, focus area (e.g.
Artificial Intelligence (AI),Biotechnology), country, and minimum deal count
Perfect for fundraising target lists ("US seed-stage VCs active in AI"), B2B sales to investment firms, and VC market research. Contacts are included in the row price — no per-contact surcharge, and every run logs a cost preview before a single row is billed, so you always know the bill up front. A full 19K-firm export costs ~$155; a targeted 1,000-firm list ~$8.
⚡ Instant Company Database — 117,000+ companies, zero scraping
Flip on instantDatabase (or just set any DB filter) and the actor answers from its continuously-growing database of 117K+ Crunchbase company profiles — instantly, no crawling, no waiting:
dbQuery: "ai"→ 16,446 matching companies · served in seconds · full clean profile per row (funding, people, tech stack, predictions) · filter by country, employee range, operating status, ranked by Crunchbase rank.
Every row carries source: 'instant-db' plus the date it was last refreshed. Perfect for lead lists, market sizing and enrichment backfills where "right now" beats "freshly crawled".
💰 Funding Rounds mode — filter by type, amount and date, no URL needed
money_raised and announced_on are Crunchbase Pro-gated: an anonymous session sees a round exist but not what it was worth or when it closed, and pagination stops at 15 results. This actor runs a paid Pro seat, so amounts and dates come back filled in.
Set roundType, minAmountUsd, announcedAfter and you have your search — no Discover URL to build, no Crunchbase account needed:
roundType: "series_b",minAmountUsd: 10000000,announcedAfter: "2026-06-01"→ 271 matching rounds, each with the funded company, the USD amount and the announcement date.
Rows come from a 11,800-round archive first (instant, no unblocker call), and when the archive cannot fill your maxItems the same filters run live against Crunchbase and the results are merged, deduped on round id. The run log says exactly how many came from each, because the archive is a partial slice of Crunchbase's ~820,000 rounds and a filter answered only from it would look complete while being a fraction.
Types available: seed, grant, pre_seed, series_a … series_g, debt_financing, post_ipo_equity, post_ipo_debt, private_equity, angel, convertible_note, corporate_round, equity_crowdfunding, secondary_market and more. An unrecognised type is called out in the log rather than silently widening your search.
📡 Funding Monitor — only what's new since your last run
Turn on fundingMonitor and schedule the actor: the first run builds a baseline, and every following run with the same input returns only new funding rounds / companies — no duplicates, no re-paying for rows you already have (skipped rows are never charged). Point it at a Discover funding-rounds search on a schedule and you have a just-funded-companies alert feed: freshly funded = hiring, buying tools, warm B2B lead.
📰 NEW: Crunchbase News — funding stories as structured rows
Flip on newsMode (or set any News filter) and the actor returns Crunchbase News articles — funding rounds, M&A, IPOs, layoffs and VC coverage — as clean rows: title, URL, publish + modified dates, author, categories, tags, image and reading time.
What makes these rows more than a feed reader: every article also carries the Crunchbase companies it links to, resolved to permalinks. One weekly funding round-up came back with 48 linked companies; a monthly unicorn report with 83. So a story joins straight onto the company and funding-round rows this same actor returns, on companyPermalinks.
newsCategory: "venture, ma",newsDateFrom: "2026-09-01"→ every venture and M&A story since the 1st, newest first, each one naming the companies involved.
Pair it with fundingMonitor and a daily schedule for a just-funded feed that never repeats itself.
Why Use This Scraper?
- ✅ Clean structured output — 34 grouped fields, ready for a spreadsheet or a model, not a 1,500-line raw dump
- ✅ Two modes in one actor — company-page enrichment and Discover/saved-search funding-round signals
- ✅ One row per company — paste a URL, a slug, or a path and get a single tidy profile back
- ✅ Surfaces the buried gold — Aberdeen IT spend, SEMrush traffic, BuiltWith + Siftery tech stack, Crunchbase ML predictions
- ✅ Cloudflare handled for you — built-in managed unblocker, nothing to configure
- ✅ Optional raw passthrough — flip one switch to also get the full unprocessed Crunchbase cards
- ✅ Flat JSON / CSV export for analysis, CRM enrichment, or lead scoring
Overview
The Crunchbase Company Scraper is built for sales and revenue teams, investors and analysts, market researchers, and data engineers who need structured company intelligence from Crunchbase without paying for an enterprise API seat.
The actor runs in two modes. Company mode is the core: each company input — a full URL, a bare slug, or an organization/... path — resolves to exactly one company-shaped row (investor names, executives, acquisition targets, and funding-round counts all appear nested inside it). Discover mode is optional: paste a Crunchbase Discover / saved-search URL and the actor returns one funding-round signal row per result (company + round type + Crunchbase links), with each company enriched from its org page. It is not a people- or investor-search crawler.
Most Crunchbase actors on the Store dump the raw cards object Crunchbase ships to its own front-end: roughly 236 KB and 1,500 lines per company, full of 540-element history arrays, ten-times-duplicated competitor trees, and internal query stubs. This actor returns a clean ~24 KB structured row instead — about 10× smaller — with the noise dropped and the useful signals lifted to the top level. If you still want everything, rawMode adds the full cards back as a passthrough field.
Supported Inputs
Input types
| Input type | Pattern | Example |
|---|---|---|
| Full company URL | https://www.crunchbase.com/organization/{slug} | https://www.crunchbase.com/organization/openai |
| Bare slug | {slug} | stripe |
| Organization path | organization/{slug} | organization/anthropic |
Copy-pasteable startUrls
{"startUrls": ["https://www.crunchbase.com/organization/openai","stripe","organization/anthropic"]}
Discover / saved-search URLs (funding-round signals)
Paste a Crunchbase Discover URL — https://www.crunchbase.com/discover/<collection>/<hash> (e.g. a saved Funding Rounds search) — and the actor switches to signal mode for that input: one row per result with the funding round, round type, funded company, and Crunchbase links, each company enriched from its org page.
{"startUrls": ["https://www.crunchbase.com/discover/funding_rounds/a0620e0d48eb17727ffdd27d9afa1807"]}
Anonymous cap: Crunchbase returns the first 15 results per search to anonymous callers and gates the funding amount, announced date, investors, and pagination beyond 15 behind a paid login. So signal rows carry company + round type + links; the
$amount/date come through only in optional logged-in mode with a Crunchbase Pro session. Need a tighter list? Narrow the saved search itself.
Unsupported inputs
- ❌ Person profiles —
crunchbase.com/person/{slug} - ❌ Individual funding-round, acquisition, investor, hub, or event entity pages —
crunchbase.com/funding_round/...,/acquisition/...(the Discover saved-search URL above is supported) - ❌ Ad-hoc search pages with no saved-search hash — save the search first to get a
/discover/<collection>/<hash>URL - ❌ Any host outside
crunchbase.com
Use Cases
| Audience | Use case |
|---|---|
| Sales / RevOps teams | Enrich CRM accounts with funding stage, headcount band, tech stack, and IT spend for lead scoring |
| Investors / analysts | Pull funding history, investor lists, and Crunchbase growth/IPO predictions for deal sourcing |
| Market researchers | Bulk-export competitor sets with categories, rank, and web-traffic signals |
| Data / growth engineers | Feed clean company rows into a warehouse or model without writing a Crunchbase parser |
| Agencies | Deliver client-ready company datasets without an enterprise Crunchbase license |
- Input — provide Crunchbase company URLs, slugs,
organization/...paths, or a Discover/saved-search URL - Unblock — each page is fetched through a built-in managed unblocker that clears Crunchbase's Cloudflare protection
- Extract — the actor reads Crunchbase's own hydration state (the data its front-end renders from) for complete, accurate fields
- Structure — raw cards are parsed into one clean, grouped row; history bloat and duplicated trees are dropped
- Output — export as structured JSON or flattened CSV, with optional
rawModefor the full unprocessed cards
Input Configuration
Input fields
This table mirrors .actor/input_schema.json field for field — every input the console form offers is listed, in form order, plus the few that are reachable only through the API.
| Field | Type | In the form | Notes |
|---|---|---|---|
startUrls | array<string> | yes | Crunchbase company URLs, slugs, organization/... paths, or Discover/saved-search URLs (/discover/<collection>/<hash>). Ignored while any mode below is on |
investorDatabase | boolean | yes | Investors mode. Serve VC firms, accelerators, angels and grant programs from the 19,000-firm table. Switches on by itself from any investor* filter |
investorQuery | string | yes | Firm name or description contains — sequoia, climate, fintech |
investorType | string | yes | Firm type: Venture Capital Investor, Seed / Early-Stage VC, Accelerator, Angel Investor, Government / Grant Program, Debt / Bank |
investorStage | string | yes | Investment stage the firm is active at |
investorFocusArea | string | yes | Industry focus — free text |
investorCountry | string | yes | HQ country |
investorMinDeals | integer | yes | Only firms with at least this many recorded deals |
instantDatabase | boolean | yes | Company database mode. Serve companies from the 117K-profile archive with no scraping. Switches on by itself from any db* filter |
dbQuery | string | yes | Company name or description contains |
dbCountry | string | yes | Country |
dbEmployeeRange | string | yes | Employee band |
dbOperatingStatus | string | yes | Operating status — active, closed, … |
roundsDatabase | boolean | yes | Funding Rounds mode. Select rounds by filter instead of a Discover URL. Switches on by itself from any filter below |
roundType | string | yes | Comma-separated investment types — series_a, seed, pre_seed, series_b, grant, debt_financing, … |
minAmountUsd | integer | yes | Only rounds that raised at least this much (USD). Either amount bound excludes undisclosed rounds |
maxAmountUsd | integer | yes | Only rounds at or below this amount (USD) |
announcedAfter | string | yes | YYYY-MM-DD — announced on or after |
announcedBefore | string | yes | YYYY-MM-DD — announced on or before |
roundCompany | string | yes | Funded company name or slug contains. Archive-only, so a run using it is never topped up live |
newsMode | boolean | yes | News mode. Return Crunchbase News articles instead of scraping. Switches on by itself from any news* filter |
newsQuery | string | yes | Free-text search over article titles and bodies — raises, series b, layoffs |
newsCategory | string | yes | Comma-separated category slugs: venture, startups, ai, ma, ipo, seed, fintech-ecommerce, cybersecurity, crypto, … |
newsTag | string | yes | Comma-separated tag slugs — e.g. unicorn |
newsDateFrom | string | yes | YYYY-MM-DD or full ISO — only articles published on/after this |
newsDateTo | string | yes | YYYY-MM-DD or full ISO — only articles published on/before this |
newsFullText | boolean | yes | Add the whole article body as plain text. Default false — bodies run to ~10,000 characters |
fundingMonitor | boolean | yes | Emit only what is NEW since the previous run with the same input. Skipped rows are never charged. Works in every mode |
rawMode | boolean | yes | Also include the full raw Crunchbase cards as _rawCards. Default false |
maxItems | integer | yes | Hard cap on rows emitted. Default 1000, clamped to 1–50,000 |
crunchbaseCookie | string (secret) | yes | Logged-in mode. Your own Crunchbase Cookie header, to unlock gated funding amount / date / investors on Discover results and lift the 15-result cap. Needs a Crunchbase Pro account and the session expires every few minutes, so it is for manual one-off runs. Blank is safe — it falls back to anonymous rather than erroring |
enrichEmails | boolean | yes | Discover a contact email per company. Billed per email found. Default false |
qualifyByPayment | boolean | yes | Free add-on to enrichEmails — flags businesses that take card payments. Only runs when enrichEmails is on |
maxConcurrency | integer | API only | Parallel unblocker requests. Default 3, clamped to 1–8. Pulled from the form because the Crunchbase Pro session is shared and ours to pace |
maxRequestRetries | integer | API only | Retries on transient unblocker errors. Default 2, clamped to 0–3 |
maxCacheAgeDays | integer | API only | Serve a cached DB copy of a company scraped within this many days. Default 14, clamped to 1–365 — never a full cache bypass |
sdoKey | string (secret) | API only | Leave blank; the actor uses its built-in unblocker. Advanced: your own scrape.do token, to bill unblocker requests to your own account |
Common scenarios
1. A few companies, clean output
{"startUrls": ["openai", "stripe", "anthropic"]}
2. Clean row plus the full raw cards
{"startUrls": ["https://www.crunchbase.com/organization/databricks"],"rawMode": true}
3. A larger batch with a cap
{"startUrls": ["openai", "stripe", "anthropic", "databricks", "figma"],"maxItems": 5,"maxConcurrency": 3}
4. A Discover / saved-search URL (funding-round signals)
{"startUrls": ["https://www.crunchbase.com/discover/funding_rounds/a0620e0d48eb17727ffdd27d9afa1807"],"maxItems": 15}
Output Overview
Each dataset item is a single company row containing:
- Identity — name, permalink, UUID, description, type, operating status, IPO status, global rank, aliases
- Location — city, region, country, continent, offices
- Web & contact — website, LinkedIn / Facebook / Twitter, contact email, phone, contact count
- Categories — category tags and per-category rank
- Funding — total (when public), round count, investor count, rounds, investor list
- People — employee band, current executives, advisors/board, alumni
- M&A — acquisitions, acquired-by, exits, IPO fields
- Tech stack — technology count, BuiltWith stack, Siftery products
- Signals — heat score, SEMrush traffic, Aberdeen IT spend, mobile apps
- Predictions — Crunchbase ML scores for growth, funding, acquisition, IPO
- Products / Similar / Press — products, similar companies with similarity score, recent press timeline
Some fields are null when Crunchbase no longer ships them on the default page load (see FAQ). Set rawMode: true to additionally receive the full unprocessed cards as _rawCards.
Output Samples
Bare slug start ("openai") — trimmed
{"name": "OpenAI","permalink": "openai","uuid": "cf2c678c-b81a-80c3-10d1-9c5e76448e51","url": "https://www.crunchbase.com/organization/openai","description": "OpenAI is an AI research and deployment company that develops advanced AI models, including ChatGPT.","operatingStatus": "active","companyType": "for_profit","ipoStatus": "private","rank": 4,"aliases": ["OpenAI LP", "OpenAI Group PBC"],"city": "San Francisco","region": "California","country": "United States","website": "https://www.openai.com","socials": {"linkedin": "https://www.linkedin.com/company/openai","twitter": "https://x.com/OpenAI"},"contactEmail": "support@openai.com","numContacts": 1384,"categories": [{ "name": "Agentic AI", "permalink": "agentic-ai-17fa" },{ "name": "Artificial Intelligence (AI)", "permalink": "artificial-intelligence" }],"funding": {"totalUsd": null,"numFundingRounds": 14,"numInvestors": 95,"investors": [ { "name": "Blackstone Group investment in Venture Round - OpenAI", "permalink": "blackstone-invested-in-openai-..." } ]},"people": {"employeeRange": "1001-5000","current": [{ "name": "Sam Altman Co-Founder and CEO @ OpenAI", "permalink": "sam-altman-executive-openai--cdec28a8" },{ "name": "Greg Brockman President, Chairman, & Co-Founder @ OpenAI", "permalink": "greg-brockman-executive-openai--d0858d5a" }]},"techStack": {"numTechnologies": 94,"builtwith": [ { "name": "Cloudflare CDN", "category": "cdn" } ],"siftery": [ { "name": "HTML5", "status": "using" } ]},"signals": {"heatScore": 92,"heatScoreDelta90": -2,"semrush": { "globalRank": null, "monthlyVisits": 487467460 },"aberdeenItSpendUsd": 285484278,"apps": { "total": 4 }},"predictions": {"growth": { "score": 0.7599, "tier": "p200_positive_low", "generatedOn": "2026-05-30" },"funding": { "score": 0.6439, "generatedOn": "2026-05-09" },"acquisition": { "score": 0.0368, "tier": "p500_negative_high" },"ipo": { "score": 0.9337, "tier": "p200_positive_low" }},"products": [{ "name": "ChatGPT", "description": "An AI conversational agent…" }],"similar": [{ "name": "Anthropic", "permalink": "anthropic", "score": 100 },{ "name": "Google", "permalink": "google", "score": 99.64 }],"pressTimeline": [{ "title": "ChatGPT tests a new jobs interface", "publisher": "AIM Group", "date": "2026-06-02", "url": "https://aimgroup.com/2026/06/02/chatgpt-tests-a-new-jobs-interface/" }],"scrapedAt": "2026-06-02T16:32:32.297Z"}
Discover search start (".../discover/funding_rounds/...") — one signal row per result, trimmed
{"searchCollection": "funding_rounds","name": "Series D - Factorial","investmentType": "series_d","moneyRaisedUsd": null, // gated for anonymous runs; unlocked in logged-in Pro mode"announcedOn": null, // gated for anonymous runs"companyName": "Factorial","companyPermalink": "factorial","companyUrl": "https://www.crunchbase.com/organization/factorial","gatedFields": ["announced_on", "money_raised"],"company": { // each result enriched from its org page (same shape as above)"name": "Factorial","country": "Spain","rank": 64,"funding": { "numFundingRounds": 8 }}}
Funding Rounds mode (roundsDatabase: true) — one row per round
{"searchCollection": "funding_rounds","uuid": "f4c1e2a0-...","name": "Series B - Swish","url": "https://www.crunchbase.com/funding_round/swish-series-b--0f9b","investmentType": "series_b","moneyRaisedUsd": 24000000,"moneyRaised": 24000000,"moneyRaisedCurrency": "USD","announcedOn": "2026-09-10","companyName": "Swish","companyPermalink": "swish-0f9b","companyUrl": "https://www.crunchbase.com/organization/swish-0f9b","source": "rounds-db","lastScrapedAt": "2026-09-09T08:41:12Z"}
source is rounds-db for an archive row and rounds-live for one fetched live during the same run. Both carry identical fields.
News mode (newsMode: true) — one row per article, trimmed
{"articleId": 94043,"title": "Mistral AI Raises $3.5B At $24B Valuation In Another Record European Round","url": "https://news.crunchbase.com/venture/mistral-ai-raises-record-european-round/","publishedAt": "2026-09-08T13:02:11Z","modifiedAt": "2026-09-08T14:19:44Z","author": "Mary Ann Azevedo","excerpt": "The French AI startup raised at a valuation that more than doubles…","categories": ["ai", "cloud", "data", "regional", "startups", "venture"],"tags": ["unicorn"],"imageUrl": "https://news.crunchbase.com/wp-content/uploads/mistral.jpg","readingTimeMinutes": 2,"companies": [{ "name": "Mistral AI", "permalink": "mistral-ai", "url": "https://www.crunchbase.com/organization/mistral-ai" },{ "name": "ASML", "permalink": "asml", "url": "https://www.crunchbase.com/organization/asml" }],"companyPermalinks": ["mistral-ai", "asml", "samsung-electronics", "eqt-holdings", "psg-equity"]}
Key Output Fields
Identity
name,permalink,uuid,url,descriptionoperatingStatus,companyType,ipoStatus,rank,aliases[]
Location & contact
city,region,country,continent,offices[]website,socials.linkedin,socials.facebook,socials.twitter,contactEmail,phone
Categories & funding
categories[].name,categoryRanks[].rankfunding.totalUsd,funding.numFundingRounds,funding.numInvestors,funding.investors[],funding.rounds[]
People & M&A
people.employeeRange,people.current[],people.advisors[],people.alumni[],numContactsma.acquisitions[],ma.acquiredBy,ma.exits[],ma.ipo
Tech stack & signals
techStack.numTechnologies,techStack.builtwith[],techStack.siftery[]signals.heatScore,signals.semrush.monthlyVisits,signals.aberdeenItSpendUsd,signals.apps
Predictions, products & press
predictions.growth,predictions.funding,predictions.acquisition,predictions.ipo(each{ score, tier, generatedOn })products[],similar[].score,pressTimeline[]
FAQ
Which Crunchbase URLs are supported?
Two kinds. Company (organization) pages — a full URL (https://www.crunchbase.com/organization/openai), a bare slug (openai), or a path (organization/openai) → one company row each. And Discover / saved-search URLs (https://www.crunchbase.com/discover/<collection>/<hash>) → one funding-round signal row per result. Individual person, funding-round, acquisition, investor, hub, and event entity pages are not supported as inputs.
What do Discover / saved-search URLs return?
One row per search result: the funding round, round type (investmentType), funded company, Crunchbase links, and a gatedFields marker — with the company enriched from its org page under company. Anonymous runs return the first 15 results and leave the $ amount, announced date, and investors null (Crunchbase gates those behind a paid Pro login). To unlock them, supply a Pro session via crunchbaseCookie — see below.
Do I get company rows or people / investor rows?
Company inputs give company rows — one per input. Discover / saved-search inputs give funding-round signal rows — up to 15 per search — each with the company enriched under company. In those two modes, people, investors, and acquisition targets appear as nested fields (e.g. people.current[], funding.investors[], ma.acquisitions[]). To get investor firms as their own rows — one per VC / accelerator / angel / grant program, with deal counts, stages, focus areas, portfolio and contacts — turn on investorDatabase (see the Startup Investors Database section above).
How is the Startup Investors Database priced vs. dedicated "investor database" actors?
Every investor firm row is a normal result row at $8/1k ($0.008 per firm), contacts included. Dedicated investor-database actors on the Store commonly charge $0.04 per firm plus $0.02 per contact — 5–7× more for a comparable row. The run log prints a cost preview (💰 Cost preview: this run will bill ≈ $…) before rows stream out, so there are no billing surprises.
Do I need a proxy or an unblocker token?
No. Crunchbase is Cloudflare-protected, but the actor ships with a built-in managed unblocker, so a normal run needs nothing extra. The optional sdoKey field only exists for advanced users who want to bill unblocker requests to their own scrape.do account.
Why are funding.totalUsd and signals.semrush.globalRank sometimes null?
Crunchbase moved a few fields (notably total funding amount and the SEMrush global rank) behind a secondary request that no longer ships on the default page load. Rather than double the per-company cost, the actor returns these as null and populates everything else — round counts, investor counts, SEMrush monthly visits, IT spend, predictions, and the full tech stack all still come through.
What does rawMode do?
When true, each row keeps all the clean structured fields and adds _rawCards — the full, unprocessed Crunchbase cards object. Use it when you need a field the structured output doesn't surface. It makes rows roughly 10× larger, so leave it off unless you need it.
What's gated, and can logged-in mode unlock it?
By default the actor reads only public Crunchbase data, so Discover funding amounts / dates / investors and a few company fields (e.g. funding.totalUsd) come back null, and each search returns its first 15 results. Those are gated by Crunchbase behind a paid Pro login — a Crunchbase limit, not a scraper one. The optional crunchbaseCookie field lets you supply your own logged-in Crunchbase Pro session to unlock them. Caveat: the session token expires every few minutes, so it suits manual one-off runs, not scheduled jobs — and a free-tier or stale cookie simply falls back to the public signal output instead of erroring. Leave it blank for normal public runs.
Does News mode need the Crunchbase cookie or an unblocker?
No. Crunchbase News is a separate, open property — no login, no session, no unblocker, no residential IP. News runs are therefore fast and never touch the Pro-gated path, so they can be scheduled as often as you like without competing with your company scrapes.
How is it priced and how fast is it?
Each company is one dataset item and one unblocker request, billed per result (see the Apify Store pricing on this actor's page). In testing, batches run at a few companies per second with default concurrency.
Support
- For issues or feature requests, use the Issues tab of this actor.
- For customization or questions, contact the author:
- Website: https://muhamed-didovic.github.io/
- Email: muhamed.didovic@gmail.com
- All my Apify actors: https://apify.com/memo23
Additional Services
- Need a custom export shape, additional Crunchbase fields, or scheduled monitoring? Email muhamed.didovic@gmail.com.
- For a direct API of this scraper (no Apify fee, usage-based), contact the same address.
Explore More Scrapers
If you found this useful, you might also like:
- Pinterest Scraper — structured pin, board, and profile data
- More company & web-data actors — directory, jobs, reviews, and social scrapers
Full list at apify.com/memo23.
🤖 For AI Agents & LLM Apps
Compact reference for AI agents calling this actor via the Apify MCP server or the Apify API (actor: memo23/crunchbase-scraper).
Purpose: Resolve Crunchbase company inputs (URL, bare slug, or organization/… path) to one structured company row each — funding, people, tech stack, signals and predictions nested inside; its differentiator is offline modes (19K-firm Startup Investors DB, 100K-company Instant DB) and a Discover-URL funding-signal mode.
Minimal input:
{ "startUrls": ["openai", "stripe"], "maxItems": 50 }
Instant Database mode (one line, no scraping): { "instantDatabase": true, "dbQuery": "fintech", "maxItems": 50 }; Investors DB mode: { "investorDatabase": true, "maxItems": 50 }.
Output: one row per company — name, permalink, uuid, url, description, operatingStatus, companyType, ipoStatus, rank, aliases, city, region, country, website, socials {linkedin, twitter, facebook}, contactEmail, numContacts, categories[] {name, permalink}, funding {totalUsd, numFundingRounds, numInvestors, investors[]}, people {employeeRange, current[]}, techStack {numTechnologies, builtwith[], siftery[]}, signals {heatScore, semrush.monthlyVisits, aberdeenItSpendUsd}, predictions {growth, funding, acquisition, ipo}, products[], similar[], pressTimeline[], scrapedAt. Discover-URL rows instead carry searchCollection, name, investmentType, moneyRaisedUsd, announcedOn, companyName, companyPermalink, companyUrl, gatedFields[], nested company.
Behaviors an agent should know:
startUrlsdrives company/Discover scraping; leave it empty to useinvestorDatabaseorinstantDatabasemode (each auto-enables when you set any of its filters). Always setmaxItems(default 1000).- Discover / saved-search URLs return funding-round signal rows, but anonymous runs are capped at the first 15 results and
moneyRaisedUsd/announcedOn/ investors are gated — supplycrunchbaseCookieto unlock full pagination and those fields. fundingMonitor: truereturns only items new since the previous run with the same input (state in this actor's key-value store); already-seen items are skipped and not billed.maxCacheAgeDays(default 14) serves a company from the database if scraped within N days; set 0 to always fetch fresh.- Billing: per result; investor/company contact fields are included in the row price (no surcharge).
enrichEmails: trueis opt-in and billed only per email actually found.
⚠️ Disclaimer
This Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by Crunchbase, Inc. or any of its subsidiaries. All trademarks mentioned are the property of their respective owners.
Crunchbase gates certain fields and result pagination behind a login. This actor reads those through a maintained Crunchbase session. You can supply your own session instead via crunchbaseCookie, in which case your own Crunchbase account terms apply to that run. Users are responsible for ensuring their use complies with Crunchbase's Terms of Service, applicable data-protection law (GDPR, CCPA, etc.), and any contractual obligations of their own organization.
SEO Keywords
crunchbase scraper, scrape crunchbase, crunchbase company scraper, crunchbase API, crunchbase.com scraper, Apify crunchbase, company data scraper, company funding scraper, startup data scraper, tech stack scraper, firmographic data, company enrichment data, lead enrichment scraper, investor data scraper, market research data, competitive intelligence scraper, sales prospecting data, company profile API, business intelligence scraper, startup funding data, crunchbase funding rounds scraper, crunchbase discover scraper, funding round data, startup funding rounds, saved search scraper, startup investors database, startup investor data scraper, VC database, venture capital firms database, VC firms list, angel investors list, investor contacts database, find investors for startup, investor leads, accelerator list, seed investors database, VC deal data, investor portfolio data, fundraising target list, investor firm database