CoolFace
Datasetpublic

reapxdev/workable-jobs-scraper

Workable Jobs Scraper Scrape Workable job listings by keyword, location and remote type, or pull every open role from a named company job board. No login, no API key. Rows in this dataset 15,524 Fields 36 Collector runs behind it 60 Most recent observation 2026-08-04 Browsable presentation https://reapx.dev/data/workable-jobs-scraper/ — 2,486 entity pages Run the collector yourself https://apify.com/reapx/workable-jobs-scraper What this… See the full description on the dataset page: https://huggingface.co/datasets/reapxdev/workable-jobs-scraper.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes53downloads
Dataset Card

reapX — public sources in, addressable records out

Workable Jobs Scraper

Scrape Workable job listings by keyword, location and remote type, or pull every open role from a named company job board. No login, no API key.

Rows in this dataset15,524
Fields36
Collector runs behind it60
Most recent observation2026-08-04
Browsable presentationhttps://reapx.dev/data/workable-jobs-scraper/ — 2,486 entity pages
Run the collector yourselfhttps://apify.com/reapx/workable-jobs-scraper

What this is

Every row here was returned by a real run of a public collector. Nothing is generated from a template over a keyword list: a row exists because a run observed it.

A browsable presentation of the subset that carries an addressable companyName is published as 2,486 entity pages at https://reapx.dev/data/workable-jobs-scraper/, one page per entity. This dataset is the larger of the two — rows whose payload has no field that can address a page are here and are not published there.

Provenance

Each row carries _run_id and _dataset_id, naming the collector run that produced it, so any row can be traced back to the run that observed it. Rows observed by more than one run are deduplicated on content; 181 duplicate observations were collapsed.

Files

  • workable-jobs-scraper.jsonl — one JSON object per row, the canonical form
  • workable-jobs-scraper.csv — the same rows flattened; nested values are JSON-encoded within their cell so they round-trip
  • dataset.json — schema.org Dataset metadata

Loading it

python
from datasets import load_dataset
ds = load_dataset("reapxdev/workable-jobs-scraper", split="train")

A sample of the published entities

Related

  • All sources: <https://reapx.dev/data/> · machine-readable index: <https://reapx.dev/llms.txt>
  • The collector is a public Apify Actor; agents reach it through <https://mcp.apify.com>

Licence

Collected from public sources. This metadata and the published pages are CC BY 4.0; the underlying records remain under the terms of their originating source.