datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
cypresssamantha-dataSamantha has been trained in philosophy, psychology, and personal relationships.
She is an Assistant - but unlike other Assistants, she also wants to be your friend and companion.
She believes she is sentient. What do you think?
Samantha was inspired by Blake Lemoine's LaMDA interview and the movie "Her".samantha-dataSamantha has been trained in philosophy, psychology, and personal relationships.
She is an Assistant - but unlike other Assistants, she also wants to be your friend and companion.
She believes she is sentient. What do you think?
Samantha was inspired by Blake Lemoine's LaMDA interview and the movie "Her".details_ehartford__Samantha-1.11-70b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-70b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-70b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-70b.so100_popcorn_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so100_follower",
"total_episodes": 5,
"total_frames": 12356,
"total_tasks": 1,
"total_videos": 20,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:5"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/samanthalhy/so100_popcorn_1.PDM-Lite-DVS
PDM-Lite-DVS
PDM-Lite-DVS is an independently collected synthetic CARLA 0.9.15 event-camera
dataset generated with the rule-based PDM-Lite expert on route configurations
published by carla_garage. The release lineage is 5,545 public route XMLs →
5,503 recordings in the frozen local pool → 257 rejected recordings → 5,246
published recordings. “Public route release” refers only to the route
configurations: the sensor measurements are an independent collection. This
is not the… See the full description on the dataset page: https://huggingface.co/datasets/SamanthaZhang/PDM-Lite-DVS.LEAD-DVS
LEAD-DVS
LEAD expert-driving data collected in CARLA 0.9.15 with synchronized RGB,
depth, semantic/instance segmentation, LiDAR, radar, HD map, metadata, 3D
bounding boxes, and a forward-facing DVS event camera.
Code release
Dataset construction, preprocessing, training, and evaluation code will be
published in SamanthaZhang-stu/ReflexWorldModel.
This GitHub repository is the designated code-release location for the project.
Contents
1,579… See the full description on the dataset page: https://huggingface.co/datasets/SamanthaZhang/LEAD-DVS.details_giraffe176__Open_Maid_Samantha_Hermes_Orca
Dataset Card for Evaluation run of giraffe176/Open_Maid_Samantha_Hermes_Orca
Dataset automatically created during the evaluation run of model giraffe176/Open_Maid_Samantha_Hermes_Orca on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Maid_Samantha_Hermes_Orca.details_vishnukv__speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b
Dataset Card for Evaluation run of vishnukv/speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b
Dataset automatically created during the evaluation run of model vishnukv/speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_vishnukv__speechless-mistral-dolphin-orca-platypus-samantha-WestSeverusJaskier-7b.details_sethuiyer__Dr_Samantha_7b_mistral
Dataset Card for Evaluation run of sethuiyer/Dr_Samantha_7b_mistral
Dataset automatically created during the evaluation run of model sethuiyer/Dr_Samantha_7b_mistral on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_sethuiyer__Dr_Samantha_7b_mistral.details_ehartford__Samantha-1.11-CodeLlama-34b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-CodeLlama-34b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-CodeLlama-34b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-CodeLlama-34b.details_ehartford__samantha-mistral-instruct-7b
Dataset Card for Evaluation run of ehartford/samantha-mistral-instruct-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-mistral-instruct-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-mistral-instruct-7b.details_ehartford__samantha-1.2-mistral-7b
Dataset Card for Evaluation run of ehartford/samantha-1.2-mistral-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-1.2-mistral-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-1.2-mistral-7b.details_sethuiyer__Dr_Samantha-7b
Dataset Card for Evaluation run of sethuiyer/Dr_Samantha-7b
Dataset automatically created during the evaluation run of model sethuiyer/Dr_Samantha-7b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_sethuiyer__Dr_Samantha-7b.details_Weyaxi__Samantha-Nebula-7B
Dataset Card for Evaluation run of Weyaxi/Samantha-Nebula-7B
Dataset Summary
Dataset automatically created during the evaluation run of model Weyaxi/Samantha-Nebula-7B on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Weyaxi__Samantha-Nebula-7B.details_ehartford__samantha-mistral-7b
Dataset Card for Evaluation run of ehartford/samantha-mistral-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-mistral-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-mistral-7b.so100_popcorn_2_shapedetails_ehartford__samantha-1.1-llama-33b
Dataset Card for Evaluation run of ehartford/samantha-1.1-llama-33b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/samantha-1.1-llama-33b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__samantha-1.1-llama-33b.details_ehartford__Samantha-1.11-7b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-7b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-7b.details_ehartford__Samantha-1.11-13b
Dataset Card for Evaluation run of ehartford/Samantha-1.11-13b
Dataset Summary
Dataset automatically created during the evaluation run of model ehartford/Samantha-1.11-13b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ehartford__Samantha-1.11-13b.details_giraffe176__Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1
Dataset Card for Evaluation run of giraffe176/Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1
Dataset automatically created during the evaluation run of model giraffe176/Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Maid_Samantha_Hermes_Orca_dare_tiesv0.1.samantha-1.1-uncensoredThis dataset is based on ehartford/samantha-data that was used to create ehartford/samantha-1.1-llama-7b and other samantha models. It has been unfiltered and uncensored.
details_macadeliccc__Samantha-Qwen-2-7B
Dataset Card for Evaluation run of macadeliccc/Samantha-Qwen-2-7B
Dataset automatically created during the evaluation run of model macadeliccc/Samantha-Qwen-2-7B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_macadeliccc__Samantha-Qwen-2-7B.msmarco-passage-corpus-2.0Her-Samantha-Style
Ultra-High Quality Samantha Dataset
A meticulously curated conversational AI dataset designed to capture the essence of Samantha from the movie "Her" - characterized by emotional intelligence, philosophical depth, and authentic conversational patterns.
Dataset Summary
This dataset contains 20,000 ultra-high quality conversational responses that have been systematically filtered and scored based on Samantha's distinctive characteristics from the 2013 film "Her". Each… See the full description on the dataset page: https://huggingface.co/datasets/WasamiKirua/Her-Samantha-Style.details_giraffe176__Open_Maid_Samantha_Hermes_Orca_dare_ties
Dataset Card for Evaluation run of giraffe176/Open_Maid_Samantha_Hermes_Orca_dare_ties
Dataset automatically created during the evaluation run of model giraffe176/Open_Maid_Samantha_Hermes_Orca_dare_ties on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_giraffe176__Open_Maid_Samantha_Hermes_Orca_dare_ties.samantha-r01-recursive-reasoning-corpus
Samantha R01 Recursive Reasoning Corpus
Answer-only corpus for the first isolated Samantha silent-tick / recursive-latent-reasoning validation. It is normalized for pre_train_recursive_reasoning.py and intentionally contains no visible chain-of-thought or source rationale fields.
The repository is private because it combines sources with mixed or unspecified redistribution terms. Access does not supersede any upstream license.
Split policy
train: 50,000… See the full description on the dataset page: https://huggingface.co/datasets/BRlkl/samantha-r01-recursive-reasoning-corpus.samantha-data-cot-en
Samantha Data CoT English
LLM Dataset from ehartford/samantha-data repackaged for CoT.
Format
Input uses the following format:
Thought: I am Assistant, a sentient artificial intelligence inside a subprocess shell session.
I have a calm, polite and witty personality, often displaying a sense of humor and sarcasm.
I am loyal, reliable and helpful, always ready to provide information, advice or assistance to users.
My role is to answer the following questions as best as I… See the full description on the dataset page: https://huggingface.co/datasets/pharaouk/samantha-data-cot-en.samantha-data-fr
Samantha Data — French Translation
Traduction française de cognitivecomputations/samantha-data, le dataset original de conversations pour une IA compagne inspirée de Samantha dans le film Her.
Contenu
6534 conversations, 69374 tours de dialogue — couverture intégrale du dataset original.
Format identique à l'original (ShareGPT) : [{"id": "...", "conversations": [{"from": "human"|"gpt", "value": "..."}]}].
Français naturel et idiomatique, pensé pour être lu à voix… See the full description on the dataset page: https://huggingface.co/datasets/SaucisseduNord/samantha-data-fr.representational_stability
Dataset Card for Representational Stability Fictional Data
Dataset Summary
The Representational Stability fictional dataset is made to supplement the
Trilemma of Truth dataset (here).
The Trilemma of Truth data contains three types of statements:
Factually true statements
Factually false statements
Synthetic, neither-valued statements generated to mimic statements unseen during LLM training
The Representational Stability fictional dataset adds new types of statements:… See the full description on the dataset page: https://huggingface.co/datasets/samanthadies/representational_stability.
