datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
law_documents_civil_qa_ready_v3.2_reasonlaw_documents_civil_qa_ready_v3.1_reasonlaw_documents_criminal_qa_ready_v3.2_reasonlaw_documents_civil_qa_ready_v3_reasonm1-1k-tokenized-v3-reason_extractlaw_documents_criminal_qa_ready_v3.3_reasonlaw_documents_criminal_qa_ready_v3.1_reasonsometimesanotion__Qwenvergence-14B-v3-Reason-details
Dataset Card for Evaluation run of sometimesanotion/Qwenvergence-14B-v3-Reason
Dataset automatically created during the evaluation run of model sometimesanotion/Qwenvergence-14B-v3-Reason
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Qwenvergence-14B-v3-Reason-details.law_documents_civil_qa_ready_v3.3_reasonconnection_queries_jan12_natural_original_1_reason_high_0.7_1024_claude-haiku45_v3res_gptoss20b_original_1_reason_high_0.7_1024_claude-haiku45_v3law_documents_criminal_qa_ready_v3_reasonconnection_queries_jan12_natural_original_1_reason_low_0.7_1024_claude-haiku45_v3m1-1k-tokenized-v3-reason_extract-tokenized-240425m1-1k-tokenized-v3-reason_extract-0427res_gptoss20b_original_1_reason_low_0.7_1024_claude-haiku45_v3folder-classification-data_v3_with_reason_codes
