CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nyu-dice-lab /wavepulse-radio-raw-transcripts WavePulse Radio Raw Transcripts Dataset Summary WavePulse Radio Raw Transcripts is a large-scale dataset containing segment-level transcripts from 396 radio stations across the United States, collected between June 26, 2024, and Dec 29th, 2024. The dataset comprises >250 million text segments derived from 750,000+ hours of radio broadcasts, primarily covering news, talk shows, and political discussions. The summarized version of these transcripts is available here. For… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/wavepulse-radio-raw-transcripts.audiotext-generation100M<n<1B9 likes2.3k downloads2y agoHugging Face02nyu-dice-lab /wavepulse-radio-summarized-transcripts WavePulse Radio Summarized Transcripts Dataset Summary WavePulse Radio Summarized Transcripts is a large-scale dataset containing summarized transcripts from 396 radio stations across the United States, collected between June 26, 2024, and October 3, 2024. The dataset comprises approximately 1.5 million summaries derived from 485,090 hours of radio broadcasts, primarily covering news, talk shows, and political discussions. The raw version of the transcripts is available… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/wavepulse-radio-summarized-transcripts.texttext-generation100K<n<1M1 likes1.4k downloads2y agoHugging Face03nyu-dice-lab /lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-RDPO-private Dataset Card for Evaluation run of princeton-nlp/Llama-3-Base-8B-SFT-RDPO Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-Base-8B-SFT-RDPO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-RDPO-private.tabular100K<n<1M0 likes697 downloads2y agoHugging Face04nyu-dice-lab /lm-eval-results-AurelPx-Pegasus-7b-slerp-private Dataset Card for Evaluation run of AurelPx/Pegasus-7b-slerp Dataset automatically created during the evaluation run of model AurelPx/Pegasus-7b-slerp The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-AurelPx-Pegasus-7b-slerp-private.tabular100K<n<1M0 likes564 downloads2y agoHugging Face05nyu-dice-lab /wildchat-50m-extended-resultstabular10K<n<100K1 likes552 downloads2y agoHugging Face06nyu-dice-lab /lm-eval-results-shyamieee-Padma-SLM-7b-v1.0-private Dataset Card for Evaluation run of shyamieee/Padma-SLM-7b-v1.0 Dataset automatically created during the evaluation run of model shyamieee/Padma-SLM-7b-v1.0 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-Padma-SLM-7b-v1.0-private.tabular100K<n<1M0 likes515 downloads2y agoHugging Face07nyu-dice-lab /sos-artifactsThis repository contains a range of Arena-Hard-Auto benchmark artifacts sourced as part of the 2024 paper Style Outweighs Substance. Repository Structure Model Responses for Arena Hard Auto Questions: data/ArenaHardAuto/model_answer Our standard reference model for pairwise comparisons was gpt-4-0314. Our standard set of comparison models was: Llama-3-8B Variants: bagel-8b-v1.0, Llama-3-8B-Magpie-Align-SFT-v0.2, Llama-3-8B-Magpie-Align-v0.2, Llama-3-8B-Tulu-330K, Llama-3-8B-WildChat… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/sos-artifacts.text0 likes499 downloads1y agoHugging Face08nyu-dice-lab /lm-eval-results-s3nh-Severusectum-7B-DPO-private Dataset Card for Evaluation run of s3nh/Severusectum-7B-DPO Dataset automatically created during the evaluation run of model s3nh/Severusectum-7B-DPO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-s3nh-Severusectum-7B-DPO-private.tabular100K<n<1M0 likes425 downloads2y agoHugging Face09nyu-dice-lab /lm-eval-results-hkust-nlp-dart-math-llama3-8b-prop2diff-private Dataset Card for Evaluation run of hkust-nlp/dart-math-llama3-8b-prop2diff Dataset automatically created during the evaluation run of model hkust-nlp/dart-math-llama3-8b-prop2diff The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-hkust-nlp-dart-math-llama3-8b-prop2diff-private.tabular100K<n<1M0 likes416 downloads2y agoHugging Face10Bartm3 /dice4This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 10, "total_frames": 6198, "total_tasks":1, "total_videos": 20, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:10" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Bartm3/dice4.tabularrobotics10K<n<100K1 likes415 downloads1y agoHugging Face11Bartm3 /dice2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 10, "total_frames": 3660, "total_tasks":1, "total_videos": 20, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:10" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Bartm3/dice2.tabularrobotics10K<n<100K0 likes414 downloads1y agoHugging Face12nyu-dice-lab /lm-eval-results-shyamieee-JARVIS-v2.0-private Dataset Card for Evaluation run of shyamieee/JARVIS-v2.0 Dataset automatically created during the evaluation run of model shyamieee/JARVIS-v2.0 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-JARVIS-v2.0-private.tabular100K<n<1M0 likes401 downloads2y agoHugging Face13nyu-dice-lab /lm-eval-results-yleo-EmertonMonarch-7B-slerp-private Dataset Card for Evaluation run of yleo/EmertonMonarch-7B-slerp Dataset automatically created during the evaluation run of model yleo/EmertonMonarch-7B-slerp The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yleo-EmertonMonarch-7B-slerp-private.tabular100K<n<1M0 likes379 downloads2y agoHugging Face14shubhdotai /dice_to_blue_mug_300_20260908_164310This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/shubhdotai/dice_to_blue_mug_300_20260908_164310.tabularrobotics100K<n<1M0 likes366 downloads14d agoHugging Face15tatsuyaaaaaaa /so_arm101_grab_red_diceThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 100, "total_frames": 116695, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:100" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/tatsuyaaaaaaa/so_arm101_grab_red_dice.tabularrobotics100K<n<1M0 likes356 downloads4mo agoHugging Face16azorematter /dice_white_pnp_und_500_b01This dataset was created using LeRobot. Dataset Description FANUC CRX-5iA dice pick-and-place demonstrations recorded on the cell by the scripted servo (the white-block collector): 500 episodes, 333586 frames at 30 fps, 4 cameras (gripper, cam0, cam1, cam2) at 640x480. Task: pick the dice up and place it on the empty white block. Every episode was saved only after the collector verified the dice seated on its block. Frame geometry Frames are UNDISTORTED (mode… See the full description on the dataset page: https://huggingface.co/datasets/azorematter/dice_white_pnp_und_500_b01.tabularrobotics100K<n<1M0 likes356 downloads18d agoHugging Face17nyu-dice-lab /lm-eval-results-shadowml-WestBeagle-7B-private Dataset Card for Evaluation run of shadowml/WestBeagle-7B Dataset automatically created during the evaluation run of model shadowml/WestBeagle-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shadowml-WestBeagle-7B-private.tabular100K<n<1M0 likes352 downloads2y agoHugging Face18azorematter /dice_white_pnp_stream_100This dataset was created using LeRobot. Dataset Description Collection metrics 100 episodes, 62850 frames at 30 fps, 2095 s of demonstration. Generated 2026-09-03T02:40:41Z by fanuc_control.data_utils.metrics v1. Definitions and thresholds: docs/metrics.md. Pose channel value note identical consecutive poses 0.0% old corpus 18% (cache frozen during moves) exact-linear ramp frames 23.4% old corpus 79% (interpolated) stationary frames (joint… See the full description on the dataset page: https://huggingface.co/datasets/azorematter/dice_white_pnp_stream_100.tabularrobotics10K<n<100K0 likes344 downloads19d agoHugging Face19nyu-dice-lab /lm-eval-results-bunnycore-SmartToxic-7B-private Dataset Card for Evaluation run of bunnycore/SmartToxic-7B Dataset automatically created during the evaluation run of model bunnycore/SmartToxic-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-bunnycore-SmartToxic-7B-private.tabular100K<n<1M0 likes339 downloads2y agoHugging Face20nyu-dice-lab /lm-eval-results-vicgalle-CarbonBeagle-11B-truthy-private Dataset Card for Evaluation run of vicgalle/CarbonBeagle-11B-truthy Dataset automatically created during the evaluation run of model vicgalle/CarbonBeagle-11B-truthy The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-vicgalle-CarbonBeagle-11B-truthy-private.tabular100K<n<1M0 likes337 downloads2y agoHugging Face21azorematter /dice_white_pnp_und_500_b02This dataset was created using LeRobot. Dataset Description FANUC CRX-5iA dice pick-and-place demonstrations recorded on the cell by the scripted servo (the white-block collector): 321 episodes, 212472 frames at 30 fps, 4 cameras (gripper, cam0, cam1, cam2) at 640x480. Task: pick the dice up and place it on the empty white block. Every episode was saved only after the collector verified the dice seated on its block. Frame geometry Frames are UNDISTORTED (mode… See the full description on the dataset page: https://huggingface.co/datasets/azorematter/dice_white_pnp_und_500_b02.tabularrobotics100K<n<1M0 likes323 downloads13d agoHugging Face22nyu-dice-lab /lm-eval-results-automerger-Inex12Yamshadow-7B-private Dataset Card for Evaluation run of automerger/Inex12Yamshadow-7B Dataset automatically created during the evaluation run of model automerger/Inex12Yamshadow-7B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-automerger-Inex12Yamshadow-7B-private.tabular100K<n<1M0 likes321 downloads2y agoHugging Face23nyu-dice-lab /lm-eval-results-shyamieee-Padma-SLM-7b-v3.0-private Dataset Card for Evaluation run of shyamieee/Padma-SLM-7b-v3.0 Dataset automatically created during the evaluation run of model shyamieee/Padma-SLM-7b-v3.0 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-shyamieee-Padma-SLM-7b-v3.0-private.tabular100K<n<1M0 likes302 downloads2y agoHugging Face24oretti /so101_dice_5This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so101_follower", "total_episodes": 60, "total_frames": 29257, "total_tasks": 3, "total_videos": 120, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:60" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/oretti/so101_dice_5.tabularrobotics10K<n<100K0 likes296 downloads1y agoHugging Face25qm30631122 /so101_grab_dice_20sec_50ep_pt2_101425This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so101_follower", "total_episodes": 50, "total_frames": 29949, "total_tasks": 1, "total_videos": 100, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:50"}, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/qm30631122/so101_grab_dice_20sec_50ep_pt2_101425.tabularrobotics10K<n<100K0 likes291 downloads11mo agoHugging Face26nyu-dice-lab /lm-eval-results-abideen-AlphaMonarch-daser-private Dataset Card for Evaluation run of abideen/AlphaMonarch-daser Dataset automatically created during the evaluation run of model abideen/AlphaMonarch-daser The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-abideen-AlphaMonarch-daser-private.tabular100K<n<1M0 likes289 downloads2y agoHugging Face27nyu-dice-lab /lm-eval-results-yunconglong-DARE_TIES_13B-private Dataset Card for Evaluation run of yunconglong/DARE_TIES_13B Dataset automatically created during the evaluation run of model yunconglong/DARE_TIES_13B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yunconglong-DARE_TIES_13B-private.tabular100K<n<1M0 likes284 downloads2y agoHugging Face28hrhraj /act_dice_dataset_25514This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 1, "total_frames": 597, "total_tasks": 1, "total_videos": 1, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:1" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/hrhraj/act_dice_dataset_25514.tabularrobotics10K<n<100K0 likes274 downloads1y agoHugging Face29azorematter /fanuc_dice_sim_eval_act_done_mem_20260910_v2 ACT-Mem (act_done_mem_v1) Isaac Sim eval, 2026-09-10 station_sw actmem-sim-eval (default scene, base flush with the plate, ACT-Mem corpus home, camera set dice_white_pnp_v1 + wrist v3, pick_support block, 15 fps, 900-step budget) driven through fanuc_infra sim-eval-bridge (n_action_steps 30, sim-base-drop 138 mm). Restarted 2026-09-10 without the 0.138 m base riser; the riser matrix was retired. Second restart 06:2x UTC after two wiring fixes: the episode now starts at the sim… See the full description on the dataset page: https://huggingface.co/datasets/azorematter/fanuc_dice_sim_eval_act_done_mem_20260910_v2.tabularrobotics10K<n<100K0 likes262 downloads12d agoHugging Face30nyu-dice-lab /lm-eval-results-allenai-llama-3-tulu-2-dpo-8b-private Dataset Card for Evaluation run of allenai/llama-3-tulu-2-dpo-8b Dataset automatically created during the evaluation run of model allenai/llama-3-tulu-2-dpo-8b The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-allenai-llama-3-tulu-2-dpo-8b-private.tabular100K<n<1M0 likes259 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.