Crossref Scraper - Scholarly Works, DOIs & Funders
Pricing
from $1.10 / 1,000 work scrapeds
Crossref Scraper - Scholarly Works, DOIs & Funders
Search Crossref (160M+ scholarly works) or look up DOIs and get clean rows: title, authors, affiliation, publisher, journal, year, citation count, funders with award numbers, license, ISSN/ISBN and abstract. Search, filter, or paste DOIs. No API key, no browser.
Pricing
from $1.10 / 1,000 work scrapeds
Rating
0.0
(0)
Developer
Scrape Sage
Maintained by CommunityActor stats
0
Bookmarked
2
Total users
1
Monthly active users
7 days ago
Last modified
Categories
Share
Disclaimer: This Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by Crossref or any of its subsidiaries. All trademarks mentioned are the property of their respective owners. "Crossref" is referenced only to describe the publicly available website this Actor collects data from.
Search Crossref - the DOI registry of 160M+ scholarly works - or look up DOIs directly, and get clean rows for every work: title, authors and affiliation, publisher, journal, year, citation count, funders with award numbers, license, ISSN/ISBN and the abstract. Search, filter, or paste a list of DOIs. Built on the official API - no key, no browser.
What you get per work
| Field | Meaning |
|---|---|
doi / doiUrl / title | DOI, resolvable link, and title |
authors / authorCount / firstAuthorAffiliation | Authors and the lead author's institution |
publisher / containerTitle / volume / issue / page | Publisher and journal placement |
publishedYear / publishedDate / workType | Dates and type (journal-article, preprint, book...) |
citedByCount / referenceCount | Citations received and references made |
funders / funderNames | Funding bodies with award / grant numbers |
licenseUrls / isOpenAccess | License links and an open-access flag |
issn / isbn / subjects / language / abstract | Identifiers, subjects and abstract |
Input
{ "searchQueries": ["machine learning"], "sortBy": "cited", "maxItemsPerQuery": 100 }
- Search queries - one per line (title, author, abstract, metadata). DOIs - exact lookups.
- Filter - a Crossref filter expression, e.g.
from-pub-date:2024-01-01,type:journal-article,has-funder:true. - Import from a file - paste a list, or link a public
.txt/.csv, a Google Sheet/Drive link, or an Apify key-value-store record. Queries and DOIs are auto-detected. - Sort by relevance, most recent, most cited, or recently indexed. Output fields trim the record.
Leave everything empty and the run returns a small free sample so you can see the shape first.
Reliability
Reads the official Crossref REST API on the polite pool (UA + mailto), keyless, cursor-paginated for deep result sets. A run that returns nothing bills $0.
Honest limits
- Metadata, not full text. Crossref indexes the scholarly record - you get the DOI, citation graph
and funders; for the article body follow
url/doiUrlto the publisher. abstractis present only when the publisher deposited one with Crossref (many do not) - an honest null, not a scraping gap.funders/licenseUrlsdepend on publisher deposits too - richly populated for recent works, sparser for older ones. Filter withhas-funder:true/has-license:truewhen you need them.citedByCountis Crossref's open citation count, which can differ from Scopus/Web of Science.
Pricing
$0.002 per work on the FREE tier (tiered pricing lowers it with volume). Only works actually saved are billed; an empty run costs nothing.
Output views
- Works - title, authors, journal, year, citations, funders and DOI.
Use with AI assistants (MCP)
Available through the Apify MCP server - an agent can pull a funder's grant output, build a citation-ranked reading list with DOIs, or assemble a RAG index of a field's literature in one call. Pairs with our arXiv, PubMed and OpenAlex scrapers.
Agent-ready: autonomous payments (x402 & Skyfire)
This actor is agent-ready - AI agents can discover it, run it, and pay for it autonomously, with no Apify account and no human in the loop. It uses pay-per-event pricing and limited permissions, so it qualifies for Apify's agentic-payment standards:
- x402 - an open, HTTP-native payment protocol. Agents pay per run in USDC on the Base network directly through the Apify MCP server - no account, no API key.
- Skyfire - agent-to-service payments for fully autonomous AI-agent workflows.
Building an AI agent, MCP tool, or autonomous data pipeline? This scraper is ready to plug in and pay as it goes.
Automate & schedule
Run this Actor on autopilot and pull results into your own stack:
- Apify API - start runs, fetch datasets and manage schedules over REST.
- apify-client for JavaScript and apify-client for Python - official SDKs.
- Schedules - run it hourly, daily or weekly and keep your dataset current.
- Webhooks - trigger downstream actions (CRM import, Slack alert, email sequence) the moment a run finishes.
import { ApifyClient } from 'apify-client';const client = new ApifyClient({ token: 'MY_APIFY_TOKEN' });const run = await client.actor('scrapesage/crossref-scraper').call({"searchQueries": ["machine learning"],"maxItemsPerQuery": 25});const { items } = await client.dataset(run.defaultDatasetId).listItems();console.log(`Got ${items.length} records`);
Integrate with any app
Connect the dataset to thousands of apps - no code required:
- Make - multi-step automation scenarios.
- Zapier - push new records straight into your CRM or spreadsheet.
- Slack - get notified when a scheduled run finds something new.
- Google Drive / Sheets - auto-export every run to a spreadsheet.
- Airbyte - pipe results into your data warehouse.
- GitHub - trigger runs from commits or releases.
FAQ
How is this Actor billed? Pay-per-event: you pay only for the results it delivers, with no monthly rental and no start fee. The per-event price is shown on the Pricing tab.
Can I schedule it and get results automatically? Yes - create a Schedule and add a webhook or an integration (Google Sheets, Slack, Make, Zapier) to push each run's dataset wherever you need it.
Which export formats are available? Every run's dataset can be downloaded as JSON, CSV, Excel (XLSX), XML, HTML or RSS from the Apify Console or the API.
Can I run it from code or an AI agent? Yes - through the Apify API and client libraries, or from Claude, ChatGPT and other assistants via the Apify MCP server.
Is it legal to scrape Crossref? This Actor collects publicly available data only. You are responsible for using the output in compliance with applicable laws (including data-protection law where personal data is involved) and the source's terms. See the Disclaimer below.
Is this an official Crossref tool? No. It is an independent, third-party Actor with no affiliation to, endorsement by or sponsorship from Crossref. See the Disclaimer below.
Disclaimer
This Actor is an independent tool and is not affiliated with, endorsed by, or sponsored by Crossref or any of its subsidiaries. All trademarks mentioned are the property of their respective owners.
"Crossref" and any related marks are the property of their respective owners and are used here only in a descriptive, nominative sense - to identify the publicly accessible website from which this Actor collects data. This Actor is not an official Crossref product, is not authorised or certified by Crossref, and does not distribute Crossref software. It collects only publicly available information; you are responsible for ensuring your use of that data complies with applicable laws, regulations and the terms of the source website.
Need help?
Open an issue on the Actor's Issues tab, or visit the Apify help center. Feature requests are welcome - this Actor is actively maintained.