CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01YANS-official /ogiri-bokete 読み込み方 from datasets import load_dataset dataset = load_dataset("YANS-official/ogiri-bokete", split="train") 概要 大喜利投稿サイトBoketeのクロールデータです。元データは CLoT-Oogiri-Go [Zhang+ CVPR2024]というデータの一部です。 詳細はCVPRのプロジェクトページをご確認ください。 このデータは以下の3タスクが含まれます。 text_to_text: テキストでお題が渡され、それに対する回答を返します。 image_to_text: いわゆる「画像で一言」です。画像のみが渡されて、テキストによる回答を返します。 text_image_to_text: 画像中にテキストが書かれています。テキストの一部が空欄になっているので、そこに穴埋めする形で回答を返します。 それぞれの量は以下の通りです。(8/30現在。ハッカソン当日までに増やす可能性があります。) タスク… See the full description on the dataset page: https://huggingface.co/datasets/YANS-official/ogiri-bokete.imagetext-generationn<1K4 likes7.8k downloads2y agoHugging Face02BGPT-OFFICIAL /refute Can AI read new science honestly? Models can sound convincing while misreading a result or expressing more confidence than the evidence deserves. That matters when people use them to summarize papers, compare studies, or decide what to investigate next. REFUTE tests whether a model knows the finding, spots quiet flaws, names what would overturn a claim, and matches its confidence to the evidence. Truth Score is the main result. It combines factual accuracy, flaw… See the full description on the dataset page: https://huggingface.co/datasets/BGPT-OFFICIAL/refute.imagetext-generationn<1K2 likes1.4k downloads2mo agoHugging Face03Furqan7007 /IDDAW_OFFICIAL IDD-AW: India Driving Dataset – Adverse Weather Semantic segmentation benchmark for autonomous driving in rain, fog, low-light, and snow, with paired RGB + near-infrared (NIR) frames and dense Level-3 semantic labels (26 classes). TODO before publishing: confirm and set the correct license / citation for the original IDD-AW release (see iddaw.github.io and the WACV 2024 paper "IDD-AW: A Benchmark for Safe Semantic Segmentation in Adverse Weather"). This card currently marks the… See the full description on the dataset page: https://huggingface.co/datasets/Furqan7007/IDDAW_OFFICIAL.imageimage-segmentation10K<n<100K0 likes563 downloads28d agoHugging Face04RLinf /RoboTwin-adjust_bottle-official-demo_clean50-Pi0_processed-dataimage1K<n<10K0 likes476 downloads1mo agoHugging Face05AKCITPixel3 /bokeh-eval-official bokeh-eval-official Bokeh synthesis artifacts: all-in-focus input, ground-truth bokeh, the best-K render and the full K sweep, one split per benchmark. Generated by inference/bokeh_net.py. imagen<1K0 likes330 downloads12d agoHugging Face06lscpku /OlympiadBench-official OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems 📖 arXiv | GitHub Dataset Description OlympiadBench is an Olympiad-level bilingual multimodal scientific benchmark, featuring 8,476 problems from Olympiad-level mathematics and physics competitions, including the Chinese college entrance exam. Each problem is detailed with expert-level annotations for step-by-step reasoning. Notably, the best-performing… See the full description on the dataset page: https://huggingface.co/datasets/lscpku/OlympiadBench-official.imagequestion-answering10K<n<100K0 likes328 downloads1y agoHugging Face07AxonData /Selfie_and_Official_ID_Photo_Dataset12,000+ people, 150,000+ images. Selfie with ID dataset for KYC verification, face identification and biometric training. Selfies paired with 2 official ID photos (passport, ID card, driver's license, residence permit). 10-15 photos per person with balanced demographics across ethnicity (Caucasian, Black, Asian, Latin American), gender and age (18-65). Contact us and share your feedback - recieve additional samples for free! 😊 Key Highlights: 12,000+ real individuals… See the full description on the dataset page: https://huggingface.co/datasets/AxonData/Selfie_and_Official_ID_Photo_Dataset.imageimage-feature-extraction5 likes194 downloads5mo agoHugging Face08IOAI-official /IOAI2025 International Olympiad in Artificial Intelligence (IOAI 2025, Beijing, China) About IOAI 2025 The 2nd International Olympiad in Artificial Intelligence (IOAI 2025) took place in Beijing, China, from August 2 to 9, 2025, hosted by Beijing National Day School (BNDS) under the patronage of UNESCO. Contest Rules: Full rules encompassing the Individual, Team, and GAITE contests are available here. Syllabus: The official syllabus outlining the AI topics contestants should… See the full description on the dataset page: https://huggingface.co/datasets/IOAI-official/IOAI2025.imageimage-classificationn<1K1 likes184 downloads1y agoHugging Face09YANS-official /ogiri-test 読み込み方 from datasets import load_dataset dataset = load_dataset("YANS-official/ogiri-test", split="test") 概要 大喜利投稿サイトBoketeのクロールデータです。元データは CLoT-Oogiri-Go [Zhang+ CVPR2024]というデータの一部です。 詳細はCVPRのプロジェクトページをご確認ください。 このデータは以下の3タスクが含まれます。 text_to_text: テキストでお題が渡され、それに対する回答を返します。 image_to_text: いわゆる「画像で一言」です。画像のみが渡されて、テキストによる回答を返します。 text_image_to_text: 画像中にテキストが書かれています。テキストの一部が空欄になっているので、そこに穴埋めする形で回答を返します。 それぞれの量は以下の通りです。 タスク お題数(画像枚数) image_to_text 56… See the full description on the dataset page: https://huggingface.co/datasets/YANS-official/ogiri-test.imageimage-to-textn<1K0 likes167 downloads2y agoHugging Face10opensima34 /guided_genshin_impact_official_server_recordings_01gated 原神 raw recordings This dataset contains raw game recordings managed by Game Data Platform. Access requests require manual approval. Game: 原神 (Genshin Impact) Collection: guided (精数据) Subset: official_server (官服(非私服)) Recordings: 566 Planned bytes: 4440851482194 Layout: recordings// Parquet files are intentionally excluded. image0 likes166 downloads21d agoHugging Face11Deathspike /magical-girl-lyrical-nanoha-official-art-verimagen<1K0 likes85 downloads3y agoHugging Face12usaaio-official /2026_USAAIO_Round2image1K<n<10K0 likes82 downloads6mo agoHugging Face13YANS-official /ogiri-test-with-references 読み込み方 from datasets import load_dataset dataset = load_dataset("YANS-official/bokete-ogiri-test", split="test") 概要 大喜利投稿サイトBoketeのクロールデータです。元データは CLoT-Oogiri-Go [Zhang+ CVPR2024]というデータの一部です。 詳細はCVPRのプロジェクトページをご確認ください。 このデータは以下の3タスクが含まれます。 text_to_text: テキストでお題が渡され、それに対する回答を返します。 image_to_text: いわゆる「画像で一言」です。画像のみが渡されて、テキストによる回答を返します。 text_image_to_text: 画像中にテキストが書かれています。テキストの一部が空欄になっているので、そこに穴埋めする形で回答を返します。 それぞれの量は以下の通りです。 タスク お題数(画像枚数) 回答数 うち委員が用意したお題… See the full description on the dataset page: https://huggingface.co/datasets/YANS-official/ogiri-test-with-references.imageimage-to-textn<1K1 likes52 downloads2y agoHugging Face14tungvu3196 /vlm-project-with-images-with-bbox-images-official-q3-updateimage10K<n<100K0 likes44 downloads1y agoHugging Face15fengmap-official /indoor-semantic-sample Fengmap Indoor Semantic Map Sample Dataset Dataset version: v1.0Semantic format specification version: v0.2Release date: August 14, 2026Dataset size: 5 indoor maps across 37 floorsPermitted use: Non-commercial learning, research, education, and technical validation only Dataset Overview The Fengmap Indoor Semantic Map Sample Dataset is a public test dataset designed for indoor spatial understanding, spatial relationship analysis, map SDK integration, and… See the full description on the dataset page: https://huggingface.co/datasets/fengmap-official/indoor-semantic-sample.imagen<1K0 likes42 downloads1mo agoHugging Face16IOAI-official /IOAI-2025-Pixel-trainimagen<1K2 likes40 downloads1y agoHugging Face17IOAI-official /ioai2025-onsite-concepts-hint-descriptionsimagen<1K0 likes32 downloads1y agoHugging Face18Bl4ckSpaces /Claire_RE2_Official_jacket_clothesimagen<1K0 likes32 downloads6mo agoHugging Face19YANS-official /senryu-test-with-references 読み込み方 from datasets import load_dataset dataset = load_dataset("YANS-official/senryu-test", split="test") 概要 川柳投稿サイトの『写真川柳』と『川柳投稿まるせん』のクロールデータです。 以下のページからクロールし、原本のHTMLファイルと構造化処理を行った結果を格納しました。 https://www.homemate-research.com/senryu/photo/ https://marusenryu.com/ このデータは以下の2タスクが含まれます。 image_to_text: 画像でお題が渡され、それに対する回答を返します。 text_to_text: テキストでお題が渡され、それに対する回答を返します。 それぞれの量は以下の通りです。 タスク お題数(画像枚数) 回答数 うち委員が用意したお題 image_to_text 70 140 7 text_to_text 30 60… See the full description on the dataset page: https://huggingface.co/datasets/YANS-official/senryu-test-with-references.imageimage-to-textn<1K0 likes30 downloads2y agoHugging Face20IOAI-official /IOAI-2025-Pixel-refimagen<1K0 likes28 downloads1y agoHugging Face21BobBobbson /BEP_OfficialThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 10, "total_frames": 13774, "total_tasks": 1, "total_videos": 30, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:10" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/BobBobbson/BEP_Official.imagerobotics100K<n<1M0 likes24 downloads1y agoHugging Face22YANS-official /ogiri-debug 読み込み方 from datasets import load_dataset dataset = load_dataset("YANS-official/ogiri-debug", split="test") 概要 大喜利生成の動作確認用データセットです。以下の3タスクが含まれます。 text_to_text: テキストでお題が渡され、それに対する回答を返します。 image_to_text: いわゆる「画像で一言」です。画像のみが渡されて、テキストによる回答を返します。 text_image_to_text: 画像中にテキストが書かれています。テキストの一部が空欄になっているので、そこに穴埋めする形で回答を返します。 データセットの各カラム説明 カラム名 型 例 概要 odai_id str "origi-dummy-1" お題のID file_path str "dummy.png" 画像のファイル名 type str "text_to_text"… See the full description on the dataset page: https://huggingface.co/datasets/YANS-official/ogiri-debug.imageimage-to-textn<1K2 likes23 downloads2y agoHugging Face23IOAI-official /IOAI-2025-Pixel-testimagen<1K0 likes23 downloads1y agoHugging Face24AxionLab-official /pokemon-blip-captions Dataset Card for Pokémon BLIP captions Dataset used to train Pokémon text to image model BLIP generated captions for Pokémon images from Few Shot Pokémon dataset introduced by Towards Faster and Stabilized GAN Training for High-fidelity Few-shot Image Synthesis (FastGAN). Original images were obtained from FastGAN-pytorch and captioned with the pre-trained BLIP model. For each row the dataset contains image and text keys. image is a varying size PIL jpeg, and text is the… See the full description on the dataset page: https://huggingface.co/datasets/AxionLab-official/pokemon-blip-captions.imagetext-to-imagen<1K0 likes23 downloads3mo agoHugging Face25tungvu3196 /vlm-project-with-images-with-bbox-images-officialimage10K<n<100K0 likes22 downloads1y agoHugging Face26LamTNguyen /cofi-compdiffuser-official-artifactsimagen<1K0 likes22 downloads10d agoHugging Face27usaaio-official /2026_USAAIO_Round3_wildlifeimage1K<n<10K0 likes20 downloads4mo agoHugging Face28tungvu3196 /vlm-project-with-images-distribution-q2-translation-all-language-officialimage10K<n<100K0 likes19 downloads1y agoHugging Face29nhatkhangdtp /uncertainty-vlm-qwen3-officialimage10K<n<100K0 likes19 downloads6mo agoHugging Face30CitronLegacy /pokemon-official-art Pokémon Official Artwork Dataset Description This dataset contains official Pokémon artwork collected from publicly available sources such as Bulbapedia and PokémonDB. The goal of this dataset is to provide a standardized collection of official Pokémon illustrations that can be used for: Image classification Computer vision research Character recognition Dataset generation AI and machine learning experiments Reference images for Pokémon-related research Each… See the full description on the dataset page: https://huggingface.co/datasets/CitronLegacy/pokemon-official-art.image0 likes16 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.