noanabeshima/forecastability_classification
This dataset is composed of Claude-labelled fineweb documents. For each document, Claude is asked if it is 'forecastable' (i.e. would be a reasonable seed for a pastcasting question) and to estimate the date the document was published. V1 splits were generated by having Claude label ~50K random fineweb documents and v2 splits were augmented with labels on ~30K additional documents that a DebertaV3 classifier finetuned on ratio10_v1 thought were forecastable (Claude thought ~1/3 of these… See the full description on the dataset page: https://huggingface.co/datasets/noanabeshima/forecastability_classification.
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Create prompt.txt
Update README.md
Update README.md
Update README.md
Delete ratio10_v2.csv
Delete ratio30_v2.csv
Delete ratio50_v2.csv
Upload 14 files
Update README.md
initial commit
