datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
details_inbox225710___model_llama_3_8B_Instruct_fine_tuned_xMR_1efinetune_data
tdro-llm/finetune_data
tDRO: Task-level Distributionally Robust Optimization for Large Language Model-based Dense Retrieval. Guangyuan Ma, Yongliang Ma, Xing Wu, Zhenpeng Su, Ming Zhou and Songlin Hu.
This repo contains all fine-tuning data for Large Language Model-based Dense Retrieval. Please refer to this repo for details to reproduce.
A total of 25 heterogeneous retrieval fine-tuning datasets with Hard Negatives and Deduplication (with test sets) are listed as belows.… See the full description on the dataset page: https://huggingface.co/datasets/tdro-llm/finetune_data.2D_finetuned_filtered_DF_Audio_Embeddingsrecord-act-finetuned-base-modify-hold-posThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Tron-Hayato/record-act-finetuned-base-modify-hold-pos.Granite-LLM-model-Fine-tuned-psychology-filosofi-and-romanceNeurvance Granite 30B – Psychology, Philosophy & Romance
Model Description
This model is a 30B-parameter Granite-based language model fine-tuned by Neurvance with a focus on:
Psychology
Philosophy
Romance and relationships
Human behavior
Emotional and reflective conversations
Deeper conversational reasoning
The goal of the model is to provide more nuanced, thoughtful and human-centered responses in conversations involving emotions, relationships, philosophical questions and psychological… See the full description on the dataset page: https://huggingface.co/datasets/WalkerDK/Granite-LLM-model-Fine-tuned-psychology-filosofi-and-romance.record-act-finetuned-base-edgeThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Tron-Hayato/record-act-finetuned-base-edge.fusion-pairwise-evals-finetuned
Automatic pairwise preference evaluations for: Making, not taking, the Best-of-N
Content
This data contains pairwise automatic win-rate evaluations for the m-ArenaHard-v2.0 benchmark and it compares 2 models against gemini-2.5-flash:
Fusion: is the 111B model finetuned on synthetic data generated with Fusion from 5 teachers
BoN: is the 111B model finetuned on synthetic data generated with BoN from 5 teachers
Each model’s outputs are compared in pairs with the respective… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/fusion-pairwise-evals-finetuned.details_abdulrahman-nuzha__finetuned-llama2-chat-5000-v2.0
Dataset Card for Evaluation run of abdulrahman-nuzha/finetuned-llama2-chat-5000-v2.0
Dataset automatically created during the evaluation run of model abdulrahman-nuzha/finetuned-llama2-chat-5000-v2.0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abdulrahman-nuzha__finetuned-llama2-chat-5000-v2.0.details_abdulrahman-nuzha__finetuned-Mistral-7B-Instruct-v0.2-5000-v2.0
Dataset Card for Evaluation run of abdulrahman-nuzha/finetuned-Mistral-7B-Instruct-v0.2-5000-v2.0
Dataset automatically created during the evaluation run of model abdulrahman-nuzha/finetuned-Mistral-7B-Instruct-v0.2-5000-v2.0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abdulrahman-nuzha__finetuned-Mistral-7B-Instruct-v0.2-5000-v2.0.details_hamxea__Llama-2-7b-chat-hf-activity-fine-tuned-v3
Dataset Card for Evaluation run of hamxea/Llama-2-7b-chat-hf-activity-fine-tuned-v3
Dataset automatically created during the evaluation run of model hamxea/Llama-2-7b-chat-hf-activity-fine-tuned-v3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_hamxea__Llama-2-7b-chat-hf-activity-fine-tuned-v3.math-classification-finetuned-resultsISOT-Fake-News-Dataset-FineTuned-2022
Dataset for project: FakeLuke-ISOT
Dataset Description
A refined variant of the ISOT dataset.
For our binary task, only two tags are needed: type (fake or true) and text.
In order to achieve that we take both of the .csv files and we trim the article tags: title, type and publishing date.
Next, besides the text we add a new column “type” and we mark it with 0 for real news and 1 for fake news.
We further trim the Fake.csv file by eliminating all the empty columns, the… See the full description on the dataset page: https://huggingface.co/datasets/Phoenyx83/ISOT-Fake-News-Dataset-FineTuned-2022.details_jslin09__bloom-560m-finetuned-fraud
Dataset Card for Evaluation run of jslin09/bloom-560m-finetuned-fraud
Dataset Summary
Dataset automatically created during the evaluation run of model jslin09/bloom-560m-finetuned-fraud on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_jslin09__bloom-560m-finetuned-fraud.SEA-Instruct-2602-fine-tunedplots_llama_3.1_8b_finetuned_binarydetails_abdulrahman-nuzha__belal-finetuned-llama2-v1.0
Dataset Card for Evaluation run of abdulrahman-nuzha/belal-finetuned-llama2-v1.0
Dataset automatically created during the evaluation run of model abdulrahman-nuzha/belal-finetuned-llama2-v1.0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abdulrahman-nuzha__belal-finetuned-llama2-v1.0.fine-tuned-deepseek2-16b-distillation-datasetmarian-finetuned-kde4-en-to-fr
test-marian-finetuned-kde4-en-to-fr
This model is a fine-tuned version of Helsinki-NLP/opus-mt-en-fr on the kde4 dataset.
It achieves the following results on the evaluation set:
Loss: 0.8559
Bleu: 52.9416
Model description
More information needed
Intended uses & limitations
More information needed
Training and evaluation data
More information needed
Training procedure
Training hyperparameters
The following hyperparameters… See the full description on the dataset page: https://huggingface.co/datasets/Vidyuth/marian-finetuned-kde4-en-to-fr.finetune_datamt5-small-finetuned-amazon-en-es_tokenized_datasetsbert-finetuned-ner_tokenized_datasetsmt5-small-finetuned-amazon-en-es_books_datasetdetails_hamxea__Llama-2-7b-chat-hf-activity-fine-tuned-v4
Dataset Card for Evaluation run of hamxea/Llama-2-7b-chat-hf-activity-fine-tuned-v4
Dataset automatically created during the evaluation run of model hamxea/Llama-2-7b-chat-hf-activity-fine-tuned-v4 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_hamxea__Llama-2-7b-chat-hf-activity-fine-tuned-v4.finetune_datasetdetails_abdulrahman-nuzha__finetuned-llama2-chat-5000-v1.0-squad
Dataset Card for Evaluation run of abdulrahman-nuzha/finetuned-llama2-chat-5000-v1.0-squad
Dataset automatically created during the evaluation run of model abdulrahman-nuzha/finetuned-llama2-chat-5000-v1.0-squad on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abdulrahman-nuzha__finetuned-llama2-chat-5000-v1.0-squad.details_abdulrahman-nuzha__finetuned-Mistral-5000-v1.0
Dataset Card for Evaluation run of abdulrahman-nuzha/finetuned-Mistral-5000-v1.0
Dataset automatically created during the evaluation run of model abdulrahman-nuzha/finetuned-Mistral-5000-v1.0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abdulrahman-nuzha__finetuned-Mistral-5000-v1.0.Zrov2_FineTunedvector_dataset_roberta-fine-tunedfinetune-dataset-testplancognvs_ckpt_test_time_finetuned
