danielfein/raid-neologism-table-splits
RAID neologism table splits Source-disjoint RAID train-derived paired splits for the two-token AI detector experiments. Source dataset: liamdugan/raid, config raid, split train. Seed: 20260501. Base source partition: 10000 train source_ids, 3000 test source_ids, overlap 0. Row format: one human text and one same-source_id AI text per row. Protocols: standard_train, standard_test: model, attack, decoding, repetition penalty, and domain sampled randomly. model_<model>_train… See the full description on the dataset page: https://huggingface.co/datasets/danielfein/raid-neologism-table-splits.
RAID neologism table splits
Source-disjoint RAID train-derived paired splits for the two-token AI detector experiments.
- Source dataset:
liamdugan/raid, configraid, splittrain. - Seed:
20260501. - Base source partition:
10000trainsource_ids,3000testsource_ids, overlap0. - Row format: one human text and one same-
source_idAI text per row.
Protocols:
standard_train,standard_test: model, attack, decoding, repetition penalty, and domain sampled randomly.model_<model>_train,model_<model>_test: everything random except generator; train excludes the held-out model and test uses only that model.attack_train: clean/standard training withattack=none.attack_<attack>_test: one test split per attack; everything other than attack is sampled randomly.
See metadata.json in the repository for counts and exact split semantics.
