CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01yinyue27 /RefRef_additionalRefRef: A Synthetic Dataset and Benchmark for Reconstructing Refractive and Reflective Objects Yue Yin · Enze Tao · Weijian Deng · Dylan Campbell About This repository provides additional data for the RefRef dataset. Citation @misc{yin2025refrefsyntheticdatasetbenchmark, title={RefRef: A Synthetic Dataset and Benchmark for Reconstructing Refractive and Reflective Objects}, author={Yue Yin and Enze Tao and… See the full description on the dataset page: https://huggingface.co/datasets/yinyue27/RefRef_additional.imageimage-to-3d100K<n<1M2 likes3.2k downloads7mo agoHugging Face02m-hamza-mughal /beat2-additional-annotations BEAT2 Official Release + Additional Annotations This is a fork of H-Liu1997/BEAT2 that adds annotations contributed by the RAG-Gesture (CVPR 2025) and MIBURI (CVPR 2026) projects. The base BEAT2-English data (motion, audio, TextGrids, semantic labels, pretrained motion-autoencoder weights) is inherited verbatim from upstream; the additional annotations from RAG-Gesture and MIBURI are pushed on top. Citations If you use only the original BEAT2 dataset, please cite… See the full description on the dataset page: https://huggingface.co/datasets/m-hamza-mughal/beat2-additional-annotations.audio1K<n<10K0 likes2.3k downloads3mo agoHugging Face03syCen /camera_pizza_additionalvideo1K<n<10K0 likes2.2k downloads1y agoHugging Face04kanhatakeyama /wizardlm8x22b-logical-math-coding-sft_additional 自動生成したテキスト WizardLM 8x22bで生成した論理・数学・コード系のデータです。 一部の計算には東京工業大学のスーパーコンピュータTSUBAME4.0を利用しました。 text100K<n<1M0 likes578 downloads2y agoHugging Face05mscho331 /bank-additional-fulltext10K<n<100K1 likes548 downloads1y agoHugging Face06mib-bench /arithmetic_additiontabular10K<n<100K0 likes397 downloads1y agoHugging Face07JakeOh /addition-datasettext1M<n<10M0 likes364 downloads10mo agoHugging Face08IEMaster /worldedit_addition_v2tabular10K<n<100K0 likes354 downloads2y agoHugging Face09deqing /addition_dataset Addition Dataset Addition problems in the format {a} + {b} = {c}. Subsets test: 5K held-out evaluation examples (operands >= 10, i.e. min 2 digits) 1BT: 85M training examples (1 billion tokens under Llama-3 tokenizer) 10BT: 850M training examples (10 billion tokens) 3MT-3digit: Exhaustive single-token addition: all (a, b) with a, b in [0, 999] and a+b <= 999. 500,500 ordered pairs, ~3M tokens. All of a, b, c are single tokens. Symmetry-safe train/test split (10% test).… See the full description on the dataset page: https://huggingface.co/datasets/deqing/addition_dataset.text1B<n<10B0 likes218 downloads4mo agoHugging Face10rwr2025-team3 /20251206-redcube-additional0 likes197 downloads6mo agoHugging Face11cmcshnik /GenText-Forensics_third_place_additional_materials GenText-Forensics 2026 — Third-Place Additional Materials (Team MSU) Model weights, code, and reproduction artifacts for Team MSU's third-place solution to the ACM MM 2026 GenText-Forensics challenge (Codabench). The method is a decomposed chain-of-thought pipeline for detecting, localizing, typing, and explaining forgeries in multilingual document text images: DTD (Document Tampering Detector) — an external pixel-level visual tampering detector that produces a tampering… See the full description on the dataset page: https://huggingface.co/datasets/cmcshnik/GenText-Forensics_third_place_additional_materials.image-segmentation1K<n<10K0 likes190 downloads3mo agoHugging Face12cestwc /bank-marketing-additional Dataset Card for Bank Marketing (additional) This dataset is a precise version of UCI Bank Marketing We first created the default bank marketing dataset, as seen here. Then we further run the following Python script to create this additional portion. # Define feature types continuous_columns = ["age", "duration", "campaign", "pdays", "previous", "emp.var.rate", "cons.price.idx", "cons.conf.idx", "euribor3m", "nr.employed"]… See the full description on the dataset page: https://huggingface.co/datasets/cestwc/bank-marketing-additional.tabular10K<n<100K0 likes150 downloads1y agoHugging Face131g0rrr /ny_test_frames_green_addition_30This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "sam_evt2", "total_episodes": 30, "total_frames": 17686, "total_tasks": 1, "total_videos": 120, "total_chunks": 1, "chunks_size": 1000, "fps": 60, "splits": { "train": "0:30" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/1g0rrr/ny_test_frames_green_addition_30.tabularrobotics10K<n<100K0 likes149 downloads8mo agoHugging Face14Lots-of-LoRAs /task753_svamp_addition_question_answering Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task753_svamp_addition_question_answering Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task753_svamp_addition_question_answering.texttext-generationn<1K0 likes101 downloads2y agoHugging Face15shylee /eval_results_pi05_basic_golden_additional_seed100kvideon<1K0 likes76 downloads5mo agoHugging Face16sohampnow /slam_stage2_additional_datatextn<1K0 likes75 downloads2y agoHugging Face17marcov /openbookqa_additional_promptsourcetabular10K<n<100K0 likes68 downloads2y agoHugging Face18Reza2kn /persian-ocr-bench-submitted10-additional4-results Persian OCR benchmark This run evaluates 4 vision OCR lanes over 193 bbox crops from Reza2kn/persian-ocr-bench-submitted10-bbox-crops. Each row in results.jsonl preserves the crop identity and current gold content_text, then records the model output, latency, usage, provider hint, HTTP status, and normalized OCR metrics. Failures are retained. OpenRouter constraints: Grok uses xai/zdr, GPT uses openai, GLM uses baseten/fp8, Gemma 4 26B uses cloudflare, and Gemma 4 31B uses… See the full description on the dataset page: https://huggingface.co/datasets/Reza2kn/persian-ocr-bench-submitted10-additional4-results.0 likes67 downloads25d agoHugging Face19selfcorrexp /llama3_additional_rr40k_non_delete_sfttabular100K<n<1M0 likes66 downloads2y agoHugging Face20IEMaster /worldedit_addition_v1tabular10K<n<100K0 likes65 downloads2y agoHugging Face21OliverHausdoerfer /libero_goal_iiwa_additionalCams_failuresThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "iiwa", "total_episodes": 116, "total_frames": 17299, "total_tasks": 9, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 20, "splits": { "train": "0:116" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/OliverHausdoerfer/libero_goal_iiwa_additionalCams_failures.tabularrobotics10K<n<100K0 likes65 downloads3mo agoHugging Face22shenj /piper_isaacsim_top_wrist_D1_120ep_additionedThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "isaac_piper", "total_episodes": 120, "total_frames": 29989, "total_tasks": 2, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 15, "splits": { "train": "0:120" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shenj/piper_isaacsim_top_wrist_D1_120ep_additioned.tabularrobotics10K<n<100K0 likes64 downloads12d agoHugging Face23genon-search /additional_fonts0 likes62 downloads2mo agoHugging Face24DiCeyIII /Additional_Yoruba_Dataaudio1K<n<10K0 likes61 downloads2y agoHugging Face25garrethlee /bpe-single-multi-token-additiontext10K<n<100K0 likes61 downloads2y agoHugging Face26shenj /piper_isaacsim_top_wrist_D1_20ep_additionThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "isaac_piper", "total_episodes": 20, "total_frames": 4979, "total_tasks": 2, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 15, "splits": { "train": "0:20" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/shenj/piper_isaacsim_top_wrist_D1_20ep_addition.tabularrobotics1K<n<10K0 likes61 downloads12d agoHugging Face27addition-robotics /teleop-trail-mix-23-cotrain-50fpsThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "addition-openarm", "total_episodes": 521, "total_frames": 201382, "total_tasks": 32, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 50, "splits": { "train": "0:521" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/addition-robotics/teleop-trail-mix-23-cotrain-50fps.tabularrobotics100K<n<1M0 likes57 downloads1mo agoHugging Face28EleutherAI /quirky_addition_increment0 Dataset Card for "quirky_addition_increment0" More Information needed text100K<n<1M0 likes53 downloads3y agoHugging Face29selfcorrexp /llama3_additional_rr80k_NON_balanced_sfttabular100K<n<1M0 likes53 downloads2y agoHugging Face30steeldream /addition_decimal1M<n<10M0 likes51 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.