distily
Datasets
All datasets matching “distily”filtered_redpajama_multilingualfiltered_redpajama_ensynth_gpt2_t1_seq_1M
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config 1.0
Model URI: gpt2
Number of Samples: 1000000
Maximum Sequence Length: 1024 tokens
c4_multilingual_1Msynth_gpt2_ted_seq_100K
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config ExponentialDecayArguments(start_t=100.0, end_t=0.5, N=1024, scale_factor=20)
Model URI: gpt2
Number of Samples: 100000
Maximum Sequence Length: 1024 tokens
synth_test
Distillation dataset created with Distily.
Method: Generated sequences randomly with temperature config 1.0
Model URI: gpt2
Number of Samples: 1000
Maximum Sequence Length: 1024 tokens
