Google Scholar Science Scraper
Spider read scholar.google.com in 193 ms without a browser and returned 33 lines of clean markdown, including the section "Lingue".
Il sistema al momento non può eseguire l'operazione. Riprova più tardi.Restituisci articoli **scritti** daad es., *"PJ Hayes"* oppure *McCarthy*Restituisci articoli **pubblicati** inad esempio, *J Biol Chem* oppure *Nature*Restituisci articoli **di date** comprese tra)## Salvato in La mia bibliotecaIl mio profiloLa mia bibliotecaLabsQualsiasi lingua Pagine in ItalianoSali sulle spalle dei giganti### Lingue The same call, in code.
The capture above came back as markdown. These examples add a key, so you get browser rendering, proxies, and concurrency on scholar.google.com.
import { SpiderBrowser } from "spider-browser";
const spider = new SpiderBrowser({
apiKey: process.env.SPIDER_API_KEY!,
});
await spider.connect();
const page = spider.page!;
await page.goto("https://scholar.google.com/scholar?q=attention+is+all+you+need");
// No selectors, no schema. Spider reads the page and names the fields.
const data = await page.scrape();
console.log(data);
await spider.close(); import { SpiderBrowser } from "spider-browser";
const spider = new SpiderBrowser({
apiKey: process.env.SPIDER_API_KEY!,
stealth: 2,
});
await spider.connect();
const page = spider.page!;
await page.goto("https://scholar.google.com/scholar?q=attention+is+all+you+need");
await page.content(10000);
const data = await page.evaluate(`(() => {
const papers = [];
document.querySelectorAll(".gs_r.gs_or.gs_scl").forEach(el => {
const title = el.querySelector(".gs_rt a")?.textContent?.trim();
const authors = el.querySelector(".gs_a")?.textContent?.trim();
const snippet = el.querySelector(".gs_rs")?.textContent?.trim();
const citations = el.querySelector(".gs_fl a:nth-child(3)")?.textContent?.trim();
const link = el.querySelector(".gs_rt a")?.getAttribute("href");
if (title) papers.push({ title, authors, snippet: snippet?.slice(0, 200), citations, link });
});
return JSON.stringify({ total: papers.length, papers: papers.slice(0, 10) });
})()`);
console.log(JSON.parse(data));
await spider.close(); Ready for volume? Get an API key →
Fields you can pull.
Spider names these from the page. The capture above came back as markdown; the same
call with return_format: "json" returns them as keys.
What scholar.google.com costs to scrape.
The capture above cost $0.000212 to fetch. Pricing is $1 per GB of pre-transformation bandwidth plus $0.001 per CPU minute, so a page like this one lands at a fraction of a cent. Failed requests are billed at $0.
- Free balance on signup
- No card required to test
- Balance never expires
Run it keyless, no account
More Science & Research scrapers.
arXiv Scraper
Extract preprint papers, abstracts, author lists, and citation metadata from arXiv open-access research repository.
PubMed Scraper
Extract biomedical literature, abstracts, MeSH terms, and citation data from PubMed National Library of Medicine database.
ResearchGate Scraper
Extract researcher profiles, publication lists, citation metrics, and project data from ResearchGate academic network.
Start scraping scholar.google.com.
You already have the call. A key raises the rate limit and turns on browser rendering, proxies, and concurrency. Balance never expires, and top-ups go through secure checkout.