Fine-tuned
details_inbox225710___model_llama_3_8B_Instruct_fine_tuned_xMR_1efinetune_data
tdro-llm/finetune_data
tDRO: Task-level Distributionally Robust Optimization for Large Language Model-based Dense Retrieval. Guangyuan Ma, Yongliang Ma, Xing Wu, Zhenpeng Su, Ming Zhou and Songlin Hu.
This repo contains all fine-tuning data for Large Language Model-based Dense Retrieval. Please refer to this repo for details to reproduce.
A total of 25 heterogeneous retrieval fine-tuning datasets with Hard Negatives and Deduplication (with test sets) are listed as belows.… See the full description on the dataset page: https://huggingface.co/datasets/tdro-llm/finetune_data.2D_finetuned_filtered_DF_Audio_EmbeddingsGranite-LLM-model-Fine-tuned-psychology-filosofi-and-romanceNeurvance Granite 30B – Psychology, Philosophy & Romance
Model Description
This model is a 30B-parameter Granite-based language model fine-tuned by Neurvance with a focus on:
Psychology
Philosophy
Romance and relationships
Human behavior
Emotional and reflective conversations
Deeper conversational reasoning
The goal of the model is to provide more nuanced, thoughtful and human-centered responses in conversations involving emotions, relationships, philosophical questions and psychological… See the full description on the dataset page: https://huggingface.co/datasets/WalkerDK/Granite-LLM-model-Fine-tuned-psychology-filosofi-and-romance.fusion-pairwise-evals-finetuned
Automatic pairwise preference evaluations for: Making, not taking, the Best-of-N
Content
This data contains pairwise automatic win-rate evaluations for the m-ArenaHard-v2.0 benchmark and it compares 2 models against gemini-2.5-flash:
Fusion: is the 111B model finetuned on synthetic data generated with Fusion from 5 teachers
BoN: is the 111B model finetuned on synthetic data generated with BoN from 5 teachers
Each model’s outputs are compared in pairs with the respective… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/fusion-pairwise-evals-finetuned.record-act-finetuned-base-modify-hold-posThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Tron-Hayato/record-act-finetuned-base-modify-hold-pos.
