reapxdev/wordpress-plugins-scraper
WordPress Plugins Scraper · Plugins, Installs & Ratings Scrape WordPress plugins directory by tags, search queries, author accounts, install bands, and rating filters. Extract ratings, active installs, tags, release details, and author links. Rows in this dataset 1,660 Fields 23 Collector runs behind it 50 Most recent observation 2026-08-03 What this is Every row here was returned by a real run of a public collector. Nothing is generated… See the full description on the dataset page: https://huggingface.co/datasets/reapxdev/wordpress-plugins-scraper.

WordPress Plugins Scraper · Plugins, Installs & Ratings
Scrape WordPress plugins directory by tags, search queries, author accounts, install bands, and rating filters. Extract ratings, active installs, tags, release details, and author links.
What this is
Every row here was returned by a real run of a public collector. Nothing is generated from a template over a keyword list: a row exists because a run observed it.
Provenance
Each row carries _run_id and _dataset_id, naming the collector run that produced it, so any row can be traced back to the run that observed it. Rows observed by more than one run are deduplicated on content; 420 duplicate observations were collapsed.
Files
wordpress-plugins-scraper.jsonl— one JSON object per row, the canonical formwordpress-plugins-scraper.csv— the same rows flattened; nested values are JSON-encoded within their cell so they round-tripdataset.json— schema.orgDatasetmetadata
Loading it
from datasets import load_dataset
ds = load_dataset("reapxdev/wordpress-plugins-scraper", split="train")Related
- All sources: <https://reapx.dev/data/> · machine-readable index: <https://reapx.dev/llms.txt>
- The collector is a public Apify Actor; agents reach it through <https://mcp.apify.com>
Licence
Collected from public sources. This metadata and the published pages are CC BY 4.0; the underlying records remain under the terms of their originating source.
