datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cpos
CPOS-HG Training Corpora
Training and validation corpora used in the cross-lingual poverty-of-stimulus
(CPOS) experiments reported in Once a Tree, Always a Tree? Cross-lingual
Transfer of Hierarchical Generalization in Language Models.
Configurations
The L1 configurations cross four languages with two evidence conditions:
*_l1_ambiguous: hierarchical-evidence target ratio 0.000.
*_l1_disambiguating: hierarchical-evidence target ratio 0.500.
English L2 is fixed… See the full description on the dataset page: https://huggingface.co/datasets/shin0729/cpos.lab-textsum-cpoQuechua_datasetprinceton-nlp__Llama-3-Base-8B-SFT-CPO-details
Dataset Card for Evaluation run of princeton-nlp/Llama-3-Base-8B-SFT-CPO
Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-Base-8B-SFT-CPO
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/princeton-nlp__Llama-3-Base-8B-SFT-CPO-details.noisy-nf-8000Quechua_Spanish_datasetdetails_princeton-nlp__Mistral-7B-Base-SFT-CPO
Dataset Card for Evaluation run of princeton-nlp/Mistral-7B-Base-SFT-CPO
Dataset automatically created during the evaluation run of model princeton-nlp/Mistral-7B-Base-SFT-CPO.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_princeton-nlp__Mistral-7B-Base-SFT-CPO.freetext_cponoisy-nf-10000Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_8Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_9Syed-Hasan-8503__Phi-3-mini-4K-instruct-cpo-simponoisy-scidocs-1250Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_6Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_7noisy-fiqa-3000noisy-nf-9000CPO_bm25Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_4c_POISON_RM_train_formatnoisy-scidocs-6258noisy-nf-6258Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_5CS224S_Quechua_Projectprinceton-nlp__Mistral-7B-Instruct-CPOQwen2.5-math-1.5B-Instruct_zero_variance_stats_method_cpo_iteration_1_zero_var_filter_th_0.5noisy-fiqa-500Qwen2.5-math-1.5B-Instruct_method_cpo_iteration_10Farah-cpo2Ks5_V5onoisy-scidocs-500
