Skip to main content
AI Studio  add-on for Spider.
Music & Podcasts
podcastindex.org Verified

Podcast Index Scraper

Extract open podcast directory data, feed URLs, value tags, and chapter metadata from Podcast Index database.

Get started Docs
target
podcastindex.org
success rate
99.9%
latency
~4ms
POST /fetch/podcastindex.org/
return_format
curl -X POST https://api.spider.cloud/fetch/podcastindex.org/ \
  -H "Authorization: Bearer $SPIDER_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{}'
200 OK · cache hit · response shape
{
  "url": "https://podcastindex.org/",
  "status": 200,
  "data": {
    "show_title": "string",
    "author": "string",
    "feed_url": "string",
    "categories": "string",
    "episode_count": "string",
    "language": "string",
    "last_update": "string",
    "value_tags": "string"
  }
}
Quick start

Extract data in minutes.

Structured JSON from podcastindex.org with a single POST. Call it with no selectors and the model names the fields, or pass your own CSS to skip the AI.

podcast-index-scraper.ts
import { SpiderBrowser } from "spider-browser";

const spider = new SpiderBrowser({
  apiKey: process.env.SPIDER_API_KEY!,
  stealth: 2,
});

await spider.connect();
const page = spider.page!;
await page.goto("https://podcastindex.org/podcast/920666");

// No selectors, no schema. Spider reads the page and names the fields.
const data = await page.scrape();

console.log(data);
await spider.close();
ready to run · spider-browser · no selectors
Extraction

Fields you can pull.

Show titleAuthorFeed URLCategoriesEpisode countLanguageLast updateValue tags
Metadata

Track & album data

Extract track info, artist data, and play counts from podcastindex.org.

Rendering

Player handling

Handle embedded players, dynamic playlists, and streaming interfaces.

Scale

Catalog coverage

Process entire artist catalogs and podcast libraries at scale.

Related

More Music & Podcasts scrapers.

Start

Start scraping podcastindex.org.

Grab an API key and call the endpoint above. The first request resolves the config; every request after hits cache.