datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cypresssamantha-dataSamantha has been trained in philosophy, psychology, and personal relationships.
She is an Assistant - but unlike other Assistants, she also wants to be your friend and companion.
She believes she is sentient. What do you think?
Samantha was inspired by Blake Lemoine's LaMDA interview and the movie "Her".samantha-dataSamantha has been trained in philosophy, psychology, and personal relationships.
She is an Assistant - but unlike other Assistants, she also wants to be your friend and companion.
She believes she is sentient. What do you think?
Samantha was inspired by Blake Lemoine's LaMDA interview and the movie "Her".so100_popcorn_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 5,
"total_frames": 12356,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:5"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/so100_popcorn_1.details_ehartford__Samantha-1.11-70b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-70b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-70b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-70b.PDM-Lite-DVS
PDM-Lite-DVS
PDM-Lite-DVS is an independently collected synthetic CARLA 0.9.15 event-camera
dataset generated with the rule-based PDM-Lite expert on route configurations
published by carla_garage. The release lineage is 5,545 public route XMLs →
5,503 recordings in the frozen local pool → 257 rejected recordings → 5,246
published recordings. “Public route release” refers only to the route
configurations: the sensor measurements are an independent collection. This
is not the… See the full description on the dataset page: https://huggingface.co/datasets/SamanthaZhang/PDM-Lite-DVS.LEAD-DVS
LEAD-DVS
LEAD expert-driving data collected in CARLA 0.9.15 with synchronized RGB,
depth, semantic/instance segmentation, LiDAR, radar, HD map, metadata, 3D
bounding boxes, and a forward-facing DVS event camera.
Code release
Dataset construction, preprocessing, training, and evaluation code will be
published in SamanthaZhang-stu/ReflexWorldModel.
This GitHub repository is the designated code-release location for the project.
Contents
1,579… See the full description on the dataset page: https://huggingface.co/datasets/SamanthaZhang/LEAD-DVS.details_giraffe176__Open_Maid_Samantha_Hermes_Orca
Dataset Card for Evaluation run of giraffe176/Open_Maid_Samantha_Hermes_Orca
Dataset automatically created during the evaluation run of model giraffe176/Open_Maid_Samantha_Hermes_Orca on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Maid_Samantha_Hermes_Orca.details_ehartford__Samantha-1.11-CodeLlama-34b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-CodeLlama-34b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-CodeLlama-34b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-CodeLlama-34b.details_vishnukv__speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b
Dataset Card for Evaluation run of vishnukv/speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b
Dataset automatically created during the evaluation run of model vishnukv/speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_vishnukv__speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b.details_sethuiyer__Dr_Samantha_7b_mistral
Dataset Card for Evaluation run of sethuiyer/Dr_Samantha_7b_mistral
Dataset automatically created during the evaluation run of model sethuiyer/Dr_Samantha_7b_mistral on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_sethuiyer__Dr_Samantha_7b_mistral.so100_popcorn_2_transportThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 1,
"total_frames": 516,
"total_tasks": 1,
"total_videos": 3,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/so100_popcorn_2_transport.eval_so100_smol_popcorn_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 2,
"total_frames": 9011,
"total_tasks": 1,
"total_videos": 8,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/eval_so100_smol_popcorn_1.details_ehartford__samantha-mistral-instruct-7b
Dataset Card for Evaluation run of ehartford/samantha-mistral-instruct-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-mistral-instruct-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-mistral-instruct-7b.details_sethuiyer__Dr_Samantha-7b
Dataset Card for Evaluation run of sethuiyer/Dr_Samantha-7b
Dataset automatically created during the evaluation run of model sethuiyer/Dr_Samantha-7b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_sethuiyer__Dr_Samantha-7b.details_ehartford__samantha-1.2-mistral-7b
Dataset Card for Evaluation run of ehartford/samantha-1.2-mistral-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-1.2-mistral-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-1.2-mistral-7b.so100_popcorn_2_gatherThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 1,
"total_frames": 767,
"total_tasks": 1,
"total_videos": 3,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/so100_popcorn_2_gather.so100_popcorn_2_shapedetails_Weyaxi__Samantha-Nebula-7B
Dataset Card for Evaluation run of Weyaxi/Samantha-Nebula-7B
Dataset Summary
Dataset automatically created during the evaluation run of model Weyaxi/Samantha-Nebula-7B on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Weyaxi__Samantha-Nebula-7B.details_ehartford__samantha-mistral-7b
Dataset Card for Evaluation run of ehartford/samantha-mistral-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-mistral-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-mistral-7b.eval_so100_smol_popcorn_1_retryThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 1,
"total_frames": 3972,
"total_tasks": 1,
"total_videos": 4,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/eval_so100_smol_popcorn_1_retry.details_ehartford__samantha-1.1-llama-33b
Dataset Card for Evaluation run of ehartford/samantha-1.1-llama-33b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-1.1-llama-33b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-1.1-llama-33b.details_ehartford__Samantha-1.11-7b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-7b.details_ehartford__Samantha-1.11-13b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-13b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-13b.details_giraffe176__Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1
Dataset Card for Evaluation run of giraffe176/Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1
Dataset automatically created during the evaluation run of model giraffe176/Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1.details_macadeliccc__Samantha-Qwen-2-7B
Dataset Card for Evaluation run of macadeliccc/Samantha-Qwen-2-7B
Dataset automatically created during the evaluation run of model macadeliccc/Samantha-Qwen-2-7B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_macadeliccc__Samantha-Qwen-2-7B.eval_so100_smol_popcorn_1_combineThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 0,
"total_frames": 0,
"total_tasks": 0,
"total_videos": 0,
"total_chunks": 0,
"chunks_size": 1000,
"fps": 30,
"splits": {},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/eval_so100_smol_popcorn_1_combine.samantha-1.1-uncensoredThis dataset is based on ehartford/samantha-data that was used to create ehartford/samantha-1.1-llama-7b and other samantha models. It has been unfiltered and uncensored.
msmarco-passage-corpus-2.0Her-Samantha-Style
Ultra-High Quality Samantha Dataset
A meticulously curated conversational AI dataset designed to capture the essence of Samantha from the movie "Her" - characterized by emotional intelligence, philosophical depth, and authentic conversational patterns.
Dataset Summary
This dataset contains 20,000 ultra-high quality conversational responses that have been systematically filtered and scored based on Samantha's distinctive characteristics from the 2013 film "Her". Each… See the full description on the dataset page: https://huggingface.co/datasets/WasamiKirua/Her-Samantha-Style.
