CoolFace
20 results

fine-tuned

open-llm-leaderboard-old /details_inbox225710___model_llama_3_8B_Instruct_fine_tuned_xMR_1e0 likes368 downloads2y agoHugging Facetdro-llm /finetune_data tdro-llm/finetune_data tDRO: Task-level Distributionally Robust Optimization for Large Language Model-based Dense Retrieval. Guangyuan Ma, Yongliang Ma, Xing Wu, Zhenpeng Su, Ming Zhou and Songlin Hu. This repo contains all fine-tuning data for Large Language Model-based Dense Retrieval. Please refer to this repo for details to reproduce. A total of 25 heterogeneous retrieval fine-tuning datasets with Hard Negatives and Deduplication (with test sets) are listed as belows.… See the full description on the dataset page: https://huggingface.co/datasets/tdro-llm/finetune_data.textn<1K0 likes244 downloads1y agoHugging FaceCohereLabs /fusion-pairwise-evals-finetuned Automatic pairwise preference evaluations for: Making, not taking, the Best-of-N Content This data contains pairwise automatic win-rate evaluations for the m-ArenaHard-v2.0 benchmark and it compares 2 models against gemini-2.5-flash: Fusion: is the 111B model finetuned on synthetic data generated with Fusion from 5 teachers BoN: is the 111B model finetuned on synthetic data generated with BoN from 5 teachers Each model’s outputs are compared in pairs with the respective… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/fusion-pairwise-evals-finetuned.texttext-generation1K<n<10K1 likes185 downloads1y agoHugging FaceWalkerDK /Granite-LLM-model-Fine-tuned-psychology-filosofi-and-romanceNeurvance Granite 30B – Psychology, Philosophy & Romance Model Description This model is a 30B-parameter Granite-based language model fine-tuned by Neurvance with a focus on: Psychology Philosophy Romance and relationships Human behavior Emotional and reflective conversations Deeper conversational reasoning The goal of the model is to provide more nuanced, thoughtful and human-centered responses in conversations involving emotions, relationships, philosophical questions and psychological… See the full description on the dataset page: https://huggingface.co/datasets/WalkerDK/Granite-LLM-model-Fine-tuned-psychology-filosofi-and-romance.0 likes178 downloads1d agoHugging FaceTron-Hayato /record-act-finetuned-base-modify-hold-posThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/Tron-Hayato/record-act-finetuned-base-modify-hold-pos.tabularrobotics10K<n<100K1 likes173 downloads14d agoHugging Faceopen-llm-leaderboard-old /details_abdulrahman-nuzha__finetuned-llama2-chat-5000-v2.0 Dataset Card for Evaluation run of abdulrahman-nuzha/finetuned-llama2-chat-5000-v2.0 Dataset automatically created during the evaluation run of model abdulrahman-nuzha/finetuned-llama2-chat-5000-v2.0 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_abdulrahman-nuzha__finetuned-llama2-chat-5000-v2.0.0 likes154 downloads3y agoHugging Face