CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01asatheesh /robust-watermark-data10K<n<100K0 likes544 downloads6mo agoHugging Face02prithivMLmods /Watermark-or-Not-20K Watermark-or-Not-20K Dataset Overview The Watermark-or-Not-20K dataset consists of 20,000 images annotated with binary labels indicating the presence or absence of a watermark. It is designed to support training and evaluation of models focused on watermark detection, which is useful for content filtering, copyright protection, and image moderation tasks. Dataset Structure Split: train Number of samples: 20,000 Label Type: Categorical (2 classes) Image… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Watermark-or-Not-20K.imageimage-classification10K<n<100K2 likes266 downloads1y agoHugging Face03vinthony /watermark-removal-logoimage4 likes224 downloads3y agoHugging Face04bastienp /visible-watermark-pita Visible watermarks datasets We have observed that while datasets such as COCO are available for object detection, the availability of datasets specifically designed for the detection of watermarks added to images is significantly limited. Through our research, we identified only one such dataset, which originates from the paper Wdnet: Watermark-Decomposition Network for Visible Watermark Removal [1]. This dataset provides a collection of images along with their corresponding… See the full description on the dataset page: https://huggingface.co/datasets/bastienp/visible-watermark-pita.imageobject-detection10K<n<100K9 likes162 downloads2y agoHugging Face05ridloau543 /watermark-latent-citra-bukti Citra bukti penelitian watermarking DCT pada ruang laten VAE Sampel citra bukti untuk dashboard hasil skripsi dashboard-latent-waterwark. Repo ini hanya menyimpan berkas gambar; seluruh angka hasil ada di repo kode. Isi 456 berkas PNG 512 x 512, yaitu 24 citra sampel pada 19 kondisi. Sampelnya 3 citra per kelas gaya pada ordinal tetap 0, 20, dan 40 di dalam kelas, bukan dipilih menurut akurasi, supaya tidak terbaca sebagai memilih hasil yang bagus saja.… See the full description on the dataset page: https://huggingface.co/datasets/ridloau543/watermark-latent-citra-bukti.imagen<1K0 likes146 downloads4d agoHugging Face06boomb0om /watermarks-validationimagen<1K3 likes140 downloads4y agoHugging Face07siddharthmb /mats-gf-activation-watermark-demo Subliminal activation-fingerprint watermark — reproduction bundle A minimal, zero-training bundle to see the watermark fire on Qwen2.5-7B-Instruct without retraining anything. It contains the secret key vectors, the frozen detection probe set, and three representative trained LoRA students. Full code, findings, and figures: https://github.com/Sid-MB/mats-gf-activation-watermark (This bundle is a small subset; the 202 GB of full trained students is not distributed here — see the… See the full description on the dataset page: https://huggingface.co/datasets/siddharthmb/mats-gf-activation-watermark-demo.textn<1K0 likes99 downloads2mo agoHugging Face08LLinked /sora-watermark-dataset Sora Watermark Detection Dataset Dataset Description This is an object detection dataset for detecting watermarks in Sora AI-generated videos. The dataset follows the YOLOv11 standard format and contains frame images extracted from Sora-generated videos along with their corresponding watermark annotations. Dataset Statistics Total Samples: 164 images Training Set: 124 images Validation Set: 21 images Test Set: 19 images Number of Classes: 1 (watermark)… See the full description on the dataset page: https://huggingface.co/datasets/LLinked/sora-watermark-dataset.object-detectionn<1K14 likes93 downloads1y agoHugging Face09watermarkproject /lord-mea-benchmark0 likes57 downloads3mo agoHugging Face10transcendingvictor /watermark1_flowers_dataset Dataset Card for "watermark1_flowers_dataset" More Information needed image1K<n<10K0 likes55 downloads2y agoHugging Face11SprintML /llm-watermark-detectiontext1K<n<10K0 likes54 downloads3mo agoHugging Face12deep9539 /FADE-watermark-ocr FADE Watermark OCR Dataset Overview This dataset contains watermarked images, their corresponding masks, the alpha values used for watermarking, and the actual text embedded (as a 9-digit number). It is designed to train and evaluate OCR models in the presence of watermarks. Paper: FADE: Probing the Limits of VLMs on fine-grained OCR Data Fields Column Name Data Type Description Image with watermark binary The raw binary bytes of the watermarked… See the full description on the dataset page: https://huggingface.co/datasets/deep9539/FADE-watermark-ocr.imagevisual-question-answering1K<n<10K1 likes54 downloads5mo agoHugging Face13annnli /C4-contrastive-watermark Dataset Card for "C4-contrastive-watermark" More Information needed tabular1K<n<10K0 likes52 downloads1y agoHugging Face14qwertyforce /scenery_watermarksDataset for watermark classification (no_watermark/watermark)~22k images, 512x512, manually annotatedadditional info - https://github.com/qwertyforce/scenery_watermarks imageimage-classification10K<n<100K4 likes49 downloads4y agoHugging Face15GeraldNdawula /Watermark_Dataset_v21K<n<10K0 likes44 downloads4mo agoHugging Face16SprintML /watermark_localizationtext1K<n<10K0 likes41 downloads3mo agoHugging Face17fwgpiyawudk /Arndee_WatermarkgatedLabelled by human 100%. imagen<1K0 likes38 downloads6d agoHugging Face18benmainbird /watermarking-clips Neural Watermarking Clips Summary This dataset contains 31 284 short audio clips collected as an unlabeled corpus for neural audio watermarking experiments. The clips cover environmental sounds, bird vocalizations, polyphonic music with predominant instruments and synthetic but realistic jazz drums and ragtime style piano. Source folders and file counts ARCA23K.audio 13 470 clips ff101bird 7 690 clips IRMAS training 6 706 clips WaivOps ragtime piano 1 743 clips WaivOps… See the full description on the dataset page: https://huggingface.co/datasets/benmainbird/watermarking-clips.audioaudio-classification10K<n<100K0 likes32 downloads10mo agoHugging Face19PenroseTiles /PRC-watermark-images0 likes31 downloads2mo agoHugging Face20heitorefer /repro-how-good-is-post-hoc-watermarking-with-language-model-rephrasing-traces Agent traces Agent sessions published from a Trackio Logbook. text1K<n<10K0 likes31 downloads2mo agoHugging Face21rks28042003 /vista-data-watermark-storeimage10K<n<100K0 likes29 downloads1y agoHugging Face22sergiov2000 /eval_test_watermark_random_yellowThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 2, "total_frames": 1673, "total_tasks": 1, "total_videos": 4, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:2" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sergiov2000/eval_test_watermark_random_yellow.tabularrobotics1K<n<10K0 likes28 downloads1y agoHugging Face23youssefkhalil320 /xsum-watermarked-flan-t5-small-20240828230657textn<1K0 likes26 downloads2y agoHugging Face24sergiov2000 /eval_test_watermark_randomThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 0, "total_frames": 0, "total_tasks": 0, "total_videos": 0, "total_chunks": 0, "chunks_size": 1000, "fps": 30, "splits": {}, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sergiov2000/eval_test_watermark_random.tabularrobotics1K<n<10K0 likes25 downloads1y agoHugging Face25Omnibus /watermark-imagesimagen<1K0 likes23 downloads3y agoHugging Face26OrdinaryDev83 /watermark-datasetimage10K<n<100K1 likes23 downloads2y agoHugging Face27Shiyu-Lab /C4-contrastive-watermark Dataset Card for "C4-contrastive-watermark" More Information needed tabular1K<n<10K0 likes23 downloads1y agoHugging Face28fineset-io /llm-watermarking-papers LLM Watermarking & Copyright Detection Papers — FineSet A research-paper dataset on LLM Watermarking & Copyright Detection Papers, assembled, deduplicated, and quality-scored by FineSet from arXiv and Semantic Scholar. 📸 This is a dated snapshot — generated 2026-06-19. It is not auto-updated. Research on LLM Watermarking & Copyright Detection Papers moves fast — new papers land on arXiv every week. Want this same dataset refreshed daily, on a topic you choose? See the bottom.… See the full description on the dataset page: https://huggingface.co/datasets/fineset-io/llm-watermarking-papers.tabulartext-classificationn<1K0 likes23 downloads3mo agoHugging Face29nanaj /watermark_img_kidsimage1K<n<10K0 likes21 downloads1y agoHugging Face30tanvi16 /vista-data-watermark-parking-storeimage10K<n<100K0 likes20 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.