CoolFace
Datasetpublic

danielfein/raid-neologism-table-splits

RAID neologism table splits Source-disjoint RAID train-derived paired splits for the two-token AI detector experiments. Source dataset: liamdugan/raid, config raid, split train. Seed: 20260501. Base source partition: 10000 train source_ids, 3000 test source_ids, overlap 0. Row format: one human text and one same-source_id AI text per row. Protocols: standard_train, standard_test: model, attack, decoding, repetition penalty, and domain sampled randomly. model_<model>_train… See the full description on the dataset page: https://huggingface.co/datasets/danielfein/raid-neologism-table-splits.

sourceHugging Faceupdated 5mo agoView on Hugging Face
0likes83downloads
6 commits on main
44ac0985mo ago

Upload README.md with huggingface_hub

danielfein
71dff815mo ago

Upload test_source_ids.json with huggingface_hub

danielfein
b0eaa0f5mo ago

Upload train_source_ids.json with huggingface_hub

danielfein
bc96c115mo ago

Upload metadata.json with huggingface_hub

danielfein
8f26ce65mo ago

Upload dataset

danielfein
dc9bec35mo ago

initial commit

danielfein