CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01clip-benchmark /wds_objectnetimage1K<n<10K4 likes51k downloads4y agoHugging Face02clip-benchmark /wds_imagenet_sketchimage10K<n<100K1 likes18k downloads4y agoHugging Face03clip-benchmark /wds_imagenet-rimage10K<n<100K0 likes9.8k downloads4y agoHugging Face04yaacovgg /shiur-clips-flac0 likes8.7k downloads2mo agoHugging Face05clip-benchmark /wds_imagenet-aimage1K<n<10K0 likes7.9k downloads4y agoHugging Face06timbrooks /instructpix2pix-clip-filtered Dataset Card for InstructPix2Pix CLIP-filtered Dataset Summary The dataset can be used to train models to follow edit instructions. Edit instructions are available in the edit_prompt. original_image can be used with the edit_prompt and edited_image denotes the image after applying the edit_prompt on the original_image. Refer to the GitHub repository to know more about how this dataset can be used to train a model that can follow instructions. Supported Tasks… See the full description on the dataset page: https://huggingface.co/datasets/timbrooks/instructpix2pix-clip-filtered.image100K<n<1M48 likes7.5k downloads4y agoHugging Face07lightly-ai /epic-kitchens-100-clips EPIC-KITCHENS-100 Extracted Clips About Dataset of 37455 video clips (24GB) extracted from videos in the EPIC-KITCHENS-100 dataset, more precisely the extension part not contained in EPIC-KITCHENS-55. For details, see https://www.lightly.ai/product-updates/epickitchens-100-in-lightlystudio. The clips folder contains one video for every narration from action annotations stored in {participant_id}/{narration_id}.mp4. The videos have been downscaled an compressed for easier… See the full description on the dataset page: https://huggingface.co/datasets/lightly-ai/epic-kitchens-100-clips.tabular10K<n<100K2 likes6.8k downloads6mo agoHugging Face08clip-benchmark /wds_imagenetv2image10K<n<100K0 likes6.6k downloads4y agoHugging Face09oumoumad /hdr-demo-clips HDR Demo Clips (Lightricks SDR→HDR) Paired SDR (input) / HDR (output) frame sequences from the Lightricks SDR-to-HDR pipeline (IC-LoRA on LTX-2). Each clip contains: hdr_exr/frame_XXXXX.exr — HDR output (f16, linear Rec.709/sRGB primaries, scene-referred) sdr_png/frame_XXXXX.png — SDR input (8-bit sRGB, display-referred) thumbnail.jpg — 280px preview from the middle frame Dimensions: HDR is symmetrically cropped from SDR to match model-friendly dimensions (typically 28–56px… See the full description on the dataset page: https://huggingface.co/datasets/oumoumad/hdr-demo-clips.imageimage-to-image10K<n<100K1 likes6.6k downloads4mo agoHugging Face10CodedotAI /code_clippy_githubThe Code Clippy dataset consists of various public codebases from GitHub in 22 programming languages with 23 extensions totalling about 16 TB of data when uncompressed. The dataset was created from the public GitHub dataset on Google BiqQuery.text1M<n<10M20 likes6.4k downloads4y agoHugging Face11clips /beir-nl-cqadupstack Dataset Card for BEIR-NL Benchmark Dataset Summary BEIR-NL is a Dutch-translated version of the BEIR benchmark, a diverse and heterogeneous collection of datasets covering various domains from biomedical and financial texts to general web content. Our benchmark is integrated into the Massive Multilingual Text Embedding Benchmark (MMTEB). BEIR-NL contains the following tasks: Fact-checking: FEVER, Climate-FEVER, SciFact Question-Answering: NQ, HotpotQA, FiQA-2018… See the full description on the dataset page: https://huggingface.co/datasets/clips/beir-nl-cqadupstack.texttext-retrieval100K<n<1M0 likes5.3k downloads2y agoHugging Face12clip-benchmark /wds_imagenet1kimage10K<n<100K1 likes4.3k downloads4y agoHugging Face13fassabilf /sea-clip-eval-predictions0 likes2.9k downloads3mo agoHugging Face14clip-benchmark /wds_fer2013image10K<n<100K0 likes2.7k downloads4y agoHugging Face15NZC415 /wan22-processed-clips0 likes2.5k downloads13d agoHugging Face16clip-benchmark /wds_flickr30kimage10K<n<100K4 likes2.5k downloads4y agoHugging Face17clip-benchmark /wds_carsimage10K<n<100K4 likes2.3k downloads4y agoHugging Face18clip-benchmark /wds_vtab-eurosatimage10K<n<100K0 likes1.9k downloads4y agoHugging Face19clip-benchmark /wds_vtab-caltech101image1K<n<10K2 likes1.7k downloads4y agoHugging Face20clip-benchmark /wds_fgvc_aircraftimage1K<n<10K0 likes1.7k downloads4y agoHugging Face21clip-benchmark /wds_vtab-cifar10image10K<n<100K0 likes1.7k downloads4y agoHugging Face22clip-benchmark /wds_vtab-dtdtextn<1K0 likes1.7k downloads4y agoHugging Face23clip-benchmark /wds_vtab-petsimage1K<n<10K1 likes1.5k downloads4y agoHugging Face24clip-benchmark /wds_vtab-cifar100image10K<n<100K0 likes1.4k downloads4y agoHugging Face25Arabic-Clip /xtd_11 Dataset Summary The expanded XTD-11 dataset, now including Arabic, enhances the original XTD collection. This dataset introduces a 1,000-image multi-lingual MSCOCO2014 caption to test multimodel in zeroshot image or text retrieval in 11 Languages. Dataset Details Citation @misc{aggarwal2020zeroshot, title={Towards Zero-shot Cross-lingual Image Retrieval}, author={Pranav Aggarwal and Ajinkya Kale}, year={2020}, eprint={2012.05107}… See the full description on the dataset page: https://huggingface.co/datasets/Arabic-Clip/xtd_11.image-to-text1K<n<10K3 likes1.4k downloads2y agoHugging Face26myzhao1999 /ucf-crime-clip-features0 likes1.3k downloads2y agoHugging Face27CodedotAI /code-clippy-tfrecordstextn<1K0 likes1.3k downloads5y agoHugging Face28haotiansun014 /cliptext1K<n<10K0 likes1.3k downloads2y agoHugging Face29diegoolguinw /demo_openai_clip_index CLIP index — demo_openai_clip Precomputed image embeddings (openai/clip-vit-base-patch32) for the static Space diegoolguinw/demo_openai_clip. Current contents: 9000 images from the train split of detection-datasets/coco. (HF datasets cap a directory at 10 000 files, so thumbs/ stays below that.) File Description embeddings.f16.bin [N, 512] row-major float16, L2-normalized metadata.json index-aligned list: {id, file, thumb, width, height} manifest.json model, count… See the full description on the dataset page: https://huggingface.co/datasets/diegoolguinw/demo_openai_clip_index.image1K<n<10K0 likes1.3k downloads16d agoHugging Face30clips /mfaqWe present the first multilingual FAQ dataset publicly available. We collected around 6M FAQ pairs from the web, in 21 different languages.tabularquestion-answering10M<n<100M37 likes1.3k downloads4y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.