datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
filtered_redpajama_multilingualfiltered_redpajama_ensynth_gpt2_t1_seq_1M
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config 1.0
Model URI: gpt2
Number of Samples: 1000000
Maximum Sequence Length: 1024 tokens
c4_multilingual_1Msynth_gpt2_ted_seq_100K
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config ExponentialDecayArguments(start_t=100.0, end_t=0.5, N=1024, scale_factor=20)
Model URI: gpt2
Number of Samples: 100000
Maximum Sequence Length: 1024 tokens
synth_test
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config 1.0
Model URI: gpt2
Number of Samples: 1000
Maximum Sequence Length: 1024 tokens
synthetic_gpt2_sequences_1K
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config 1.0
Model URI: gpt2
Number of Samples: 1000
Maximum Sequence Length: 1024 tokens
synth_gpt2_ted_seq_100
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config ExponentialDecayArguments(start_t=10.0, end_t=0.5, N=1024, scale_factor=100)
Model URI: gpt2
Number of Samples: 1000
Maximum Sequence Length: 1024 tokens
synth_tdecay_gpt2_seq_1K
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config ExponentialDecayArguments(start_t=100.0, end_t=0.5, N=1024, scale_factor=20)
Model URI: gpt2
Number of Samples: 1000
Maximum Sequence Length: 1024 tokens
