Replicate Scraper
Extract ML model listings, run counts, pricing data, and API integration examples from Replicate. Built on spider-browser .
- target
- replicate.com
- success rate
- 99.9%
- latency
- ~4ms
/fetch/replicate.com/ curl -X POST https://api.spider.cloud/fetch/replicate.com/ \ -H "Authorization: Bearer $SPIDER_API_KEY" \ -H "Content-Type: application/json" \ -d '{"return_format": "json"}'
{
"url": "https://replicate.com/",
"status": 200,
"data": {
"model_name": "string",
"author": "string",
"runs_count": "string",
"description": "string",
"license": "string",
"hardware": "string",
"cost_per_run": "string",
"tags": "string"
}
} # Replicate Scraper
**Model name**: string
**Author**: string
**Runs count**: string
**Description**: string
**License**: string
**Hardware**: string
**Cost per run**: string
**Tags**: string Extract data in minutes.
Structured JSON from replicate.com with a single POST. AI-resolved selectors, cached on the first call.
import { SpiderBrowser } from "spider-browser";
const spider = new SpiderBrowser({
apiKey: process.env.SPIDER_API_KEY!,
});
await spider.connect();
const page = spider.page!;
await page.goto("https://replicate.com/explore");
await page.content(10000);
const data = await page.evaluate(`(() => {
const models = [];
document.querySelectorAll("[class*='ModelCard'], [class*='model-card']").forEach(el => {
const name = el.querySelector("h2, [class*='name']")?.textContent?.trim();
const author = el.querySelector("[class*='owner'], [class*='author']")?.textContent?.trim();
const runs = el.querySelector("[class*='runs']")?.textContent?.trim();
const desc = el.querySelector("p, [class*='description']")?.textContent?.trim();
if (name) models.push({ name, author, runs, desc });
});
return JSON.stringify({ total: models.length, models: models.slice(0, 15) });
})()`);
console.log(JSON.parse(data));
await spider.close(); curl -X POST https://api.spider.cloud/fetch/replicate.com/ \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{"return_format": "json"}' import requests
resp = requests.post(
"https://api.spider.cloud/fetch/replicate.com/",
headers={
"Authorization": "Bearer YOUR_API_KEY",
"Content-Type": "application/json",
},
json={"return_format": "json"},
)
print(resp.json()) const resp = await fetch("https://api.spider.cloud/fetch/replicate.com/", {
method: "POST",
headers: {
"Authorization": "Bearer YOUR_API_KEY",
"Content-Type": "application/json",
},
body: JSON.stringify({ return_format: "json" }),
});
const data = await resp.json();
console.log(data); Fields you can pull.
React SPA handling
Full browser rendering for streaming content and dynamic React UI.
Structured parsing
Extract code blocks, documentation, repository data, and metadata.
Load completion
Smart network idle detection waits for dynamically loaded content to finish.
More AI & Developer scrapers.
ChatGPT Scraper
Extract shared ChatGPT conversations, prompts, and AI-generated content from public links.
Hugging Face Scraper
Extract ML model cards, dataset info, leaderboard data, and paper metadata from Hugging Face.
GitHub Scraper
Extract trending repositories, star counts, contributor data, and code snippets from GitHub.
Start scraping replicate.com.
Grab an API key and call the endpoint above. The first request resolves the config; every request after hits cache.