Request a tool
All toolsAutomationsGuidesMCP serverRequest a toolPlatformsCategories

Academic & Research Scrapers & Data Tools

5 ready-to-run Academic & Research data tools. Search arXiv, OpenAlex, Crossref, npm and the Internet Archive. No code — run them on the Apify cloud and export the results.

Runs on Apify’s cloud
No code required
Export JSON, CSV & Excel
Pay per use, no subscription
Your account, your data

Academic & Research tools

arXiv Scraper iconDeveloper & Research Tools

arXiv Scraper

Search arXiv papers by title, author, abstract or category. Get full abstracts, authors, categories, DOI, dates and PDF links. $2 per 1,000 papers.

2 use cases

OpenAlex Scholarly Works Scraper iconDeveloper & Research Tools

OpenAlex Scholarly Works Scraper

Search 250M+ OpenAlex papers with no API key. Get titles, authors, venue, year, citations, DOI, OA links and full abstracts. $2.00 per 1,000 papers.

2 use cases

Crossref Scholarly Works Scraper iconDeveloper & Research Tools

Crossref Scholarly Works Scraper

Search 150M+ papers on Crossref: DOI, title, authors, journal, publisher, date, citations and abstract. No API key. $1.00 per 1,000 works.

2 use cases

Package Registry Scraper (npm + PyPI) iconDeveloper & Research Tools

Package Registry Scraper (npm + PyPI)

Get npm and PyPI package metadata as JSON. Version, license, author, repo, keywords and npm monthly downloads. $2 per 1,000 packages.

2 use cases

Internet Archive Scraper iconDeveloper & Research Tools

Internet Archive Scraper

Search Internet Archive (archive.org) for books, audio, film and web items. Title, creator, year, downloads, subjects and URL. $2.00 per 1,000 items.

2 use cases

Academic & Research scraper pricing

What each tool costs per result — pay-per-result on your own Apify account, with no subscription:

ToolPrice
arXiv Scraper$0.002 per paper ($2 / 1,000)
OpenAlex Scholarly Works Scraper$0.002 per work ($2 / 1,000)
Crossref Scholarly Works Scraper$0.001 per work ($1 / 1,000)
Package Registry Scraper (npm + PyPI)$0.002 per package ($2 / 1,000)
Internet Archive Scraper$0.002 per item ($2 / 1,000)

Popular Academic & Research use cases

LLM Papers on arXiv: Abstracts, Authors and PDF Links

Keyword search across arXiv for language-model work, ranked by relevance, with the abstract, the author list and a direct PDF link on every row.

NLP Papers From arXiv cs.CL, Newest Submissions First

Sorted by submission date, so the top of the run is what went up today. Abstract, authors, categories and a PDF link on each paper. No DOI field.

CRISPR Papers from OpenAlex, Newest First

Gene-editing work published since 2024, sorted by date, with authors, journal, DOI and open-access link. Abstracts are rebuilt from OpenAlex inverted index.

Most Cited Deep Learning Papers, Ranked by Citations

The foundational reading list pulled from OpenAlex with authors, venue, year and citation count. Abstracts come through on most works, not quite all of them.

Find Most Cited Papers on a Topic, Ranked by Citations

Crossref works sorted by citation count, with DOI, title, journal and publisher. Authors are on most rows but not all, and abstracts on very few.

Crossref Metadata Search: Every Work Type for a Topic

Articles, books, datasets and preprints for one query, each with a DOI, title, publisher and citation count. Author lists come through where they exist.

npm Package Search with Monthly Download Counts

npm package search by keyword: name, version, description, repo link and monthly downloads per hit. Search hits carry no license field. Name lookups do.

npm License Checker for a List of Dependencies

Feed the names from your package.json and get the license and source repo back for each npm package. A compliance pass without installing anything.

Download Books From Archive.org: Search Texts to JSON

Public-domain books by keyword with title, item link and download count. Author and year come from whoever uploaded, so about a quarter lack them.

Internet Archive Search, Newest Uploads First

Any archive.org topic sorted by the date it was added, with title, upload date and item link. Handy for watching a subject for fresh material.