CoolFace
Datasetpublic

noanabeshima/forecastability_classification

This dataset is composed of Claude-labelled fineweb documents. For each document, Claude is asked if it is 'forecastable' (i.e. would be a reasonable seed for a pastcasting question) and to estimate the date the document was published. V1 splits were generated by having Claude label ~50K random fineweb documents and v2 splits were augmented with labels on ~30K additional documents that a DebertaV3 classifier finetuned on ratio10_v1 thought were forecastable (Claude thought ~1/3 of these… See the full description on the dataset page: https://huggingface.co/datasets/noanabeshima/forecastability_classification.

sourceHugging Facemitupdated 1y agoView on Hugging Face
0likes93downloads
15 commits on main
8ada67f1y ago

Update README.md

noanabeshima
17bb6371y ago

Update README.md

noanabeshima
28c13db1y ago

Update README.md

noanabeshima
8b7f9eb1y ago

Update README.md

noanabeshima
133af5e1y ago

Update README.md

noanabeshima
687b3ea1y ago

Create prompt.txt

noanabeshima
21473c91y ago

Update README.md

noanabeshima
52c41fb1y ago

Update README.md

noanabeshima
505d9011y ago

Update README.md

noanabeshima
56ff3401y ago

Delete ratio10_v2.csv

noanabeshima
2a5cd791y ago

Delete ratio30_v2.csv

noanabeshima
6c9b1b31y ago

Delete ratio50_v2.csv

noanabeshima
0819c3e1y ago

Upload 14 files

noanabeshima
783a81d1y ago

Update README.md

noanabeshima
13315cd1y ago

initial commit

noanabeshima