Business
Semantic Scholar Papers Scraper
Scrape academic papers with title, year, venue, citation count, authors, DOI, abstract and a direct link. Search by keyword. Export to JSON, CSV or Excel.
semantic-scholar-scraper.json 200 OK
{ "paperId": "53c9f3c34d8481adaf24df3b25581ccf1bc53f5c", "title": "Physics-informed machine learning", "year": 2021, "venue": "Nature Reviews Physics", "citationCount": 7461, "authors": ["G. Karniadakis","I. Kevrekidis","Lu Lu","P…, "doi": "10.1038/s42254-021-00314-5", "url": "https://www.semanticscholar.org/paper/53c9f3c3…", "tldr": "Some of the prevailing trends in embedding phy…" }
What you get
Every run returns clean, typed records, ready for your CRM, spreadsheet or database.
- One clean record per result, deduped and normalized
- Stable schema in JSON, CSV or Excel, or read it via the Apify API
- Pay per use in the cloud, nothing to install or maintain
Fields it returns
Paper IdTitleYearVenueCitation CountAuthorsDoiUrlTldrInfluential Citation CountReference CountIs Open AccessOpen Access Pdf UrlOpen Access Pdf License
Sample output
A real example record, exactly the shape you receive.
| Paper Id | Title | Year | Venue |
|---|---|---|---|
| 53c9f3c34d8481adaf24df3b25581ccf1… | Physics-informed machine learning | 2021 | Nature Reviews Physics |
| 7872f34e2a164c5cf3c34a7a7433dc334… | Machine Learning: Algorithms, Rea… | 2021 | SN Computer Science |
| f9c602cc436a9ea2f9e7db48c77d924e0… | Fashion-MNIST: a Novel Image Data… | 2017 | arXiv.org |
GET /semantic-scholar-scraper
{ "paperId": "53c9f3c34d8481adaf24df3b25581ccf1bc53f5c", "title": "Physics-informed machine learning", "year": 2021, "venue": "Nature Reviews Physics", "citationCount": 7461, "authors": ["G. Karniadakis","I. Kevrekidis","Lu Lu","P…, "doi": "10.1038/s42254-021-00314-5", "url": "https://www.semanticscholar.org/paper/53c9f3c3…", "tldr": "Some of the prevailing trends in embedding phy…" }
Example use cases
Related scrapers
Frequently asked questions
What does the search match?
The search looks across titles and abstracts and returns the most relevant papers for your query.
Do records include the abstract and citations?
Yes. Each record includes the abstract where available, the citation count, the venue and the authors.
Can I get the DOI?
Yes, where available, along with the arXiv identifier for preprints.
How fresh is the data?
Records are read live at run time, so each reflects the index at the moment of the run (see observedAt).
Ready when you are.
Run the scraper live on Apify right now, or have us build one tailored to exactly what you need.