CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ismatsamadov /azerbaijan-court-data Azerbaijan Court System Dataset The most comprehensive open dataset of Azerbaijan's judicial system — 1.64 million structured records and 1.54 million court decision PDFs (~160 GB) covering court decisions, active cases, scheduled hearings, court registries, judges, lawyers, and mediator organizations. Built for AI engineers, legal tech startups, and researchers who need real-world legal data at scale. Quick Start Load with Hugging Face datasets from datasets… See the full description on the dataset page: https://huggingface.co/datasets/ismatsamadov/azerbaijan-court-data.imagetext-classification1M<n<10M2 likes417 downloads6mo agoHugging Face02LocalDoc /azerbaijani-ocr-linesgated Azerbaijani OCR Lines Line-level training data for Azerbaijani text recognition, in Latin and Cyrillic script, extracted from scanned books. Fields field description image cropped text line, grayscale, height 48 px text transcription script az_latin or az_cyrillic book anonymised source-book id How it was built Pages come from scanned PDFs that already carried an OCR text layer. Line boxes were taken from that layer, rendered… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/azerbaijani-ocr-lines.imageimage-to-text100K<n<1M1 likes122 downloads14d agoHugging Face03LocalDoc /azerbaijani-ocr-benchmark Azerbaijani OCR Benchmark Line-level OCR benchmark for Azerbaijani in both Latin and Cyrillic script, built from scanned books. Fields field description image cropped text line, grayscale, height 48 px text verbatim transcription script az_latin or az_cyrillic book anonymised source-book id How labels were produced Every line carries a label agreed on independently by three sources: the OCR text layer already present in the… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/azerbaijani-ocr-benchmark.imageimage-to-text1K<n<10K1 likes107 downloads19d agoHugging Face04LocalDoc /azerbaijani-htr-synthetic Azerbaijani Synthetic Handwritten OCR Dataset A large-scale synthetic dataset for training handwritten text recognition (HTR) models on Azerbaijani Latin script. Generated using a procedural pipeline that combines real-world handwriting fonts with realistic scan-style augmentations. This dataset addresses the lack of publicly available Azerbaijani handwriting OCR data — a low-resource language for which no IAM-equivalent corpus exists. Dataset Statistics… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/azerbaijani-htr-synthetic.imageimage-to-text1M<n<10M1 likes104 downloads5mo agoHugging Face05bahramhasanov /azerbaijani-carpet-fullsizeimagen<1K0 likes69 downloads2y agoHugging Face06ARMammadli /azerbaijani-cuisine Azerbaijani Cuisine Dataset A curated image dataset of traditional Azerbaijani dishes for computer vision and image classification tasks. Dataset Description This dataset contains images of five traditional Azerbaijani dish categories. It is organized into standard training, validation, and test splits to facilitate machine learning model development and evaluation. Features 5 Food Categories: Dolma, Kebabs, Pakhlava, Plov, and Soups 324 Total Images: Properly… See the full description on the dataset page: https://huggingface.co/datasets/ARMammadli/azerbaijani-cuisine.imagen<1K1 likes56 downloads1y agoHugging Face07sdproject2025 /Wildfire_Images_From_Satilite_For_Azerbaijanimagen<1K0 likes51 downloads1y agoHugging Face08LocalDoc /azerbaijani-htr-benchmark Azerbaijani Handwritten OCR Benchmark A manually annotated benchmark for handwritten text recognition (HTR) on Azerbaijani Latin script. Real-world scanned pages annotated in Label Studio with rotated bounding boxes and exact transcriptions. Provided in two parallel views: lines — Line-level recognition Cropped images of single text lines paired with their transcription. Rotated regions are deskewed (warped to be axis-aligned) so each crop shows the line horizontally.… See the full description on the dataset page: https://huggingface.co/datasets/LocalDoc/azerbaijani-htr-benchmark.imageimage-to-textn<1K0 likes24 downloads4mo agoHugging Face09inovruzova /azerbaijani-art-collectionNote: Data is collected by the Afina Apayeva, Ariana Kenbayeva, Ilhama Novruzova, Mehriban Aliyeva, and only art_metal category is taken by scraping Azerbaijan Carpet Museum's official website. Dataset Source:We took the pictures of art works by smartphones. For some of them, we took their pictures from 3 perspectives: left, right, and front. For most of them, we took just one picture from the front side to avoid data duplication. Primary photos taken by team members at: Azerbaijan National… See the full description on the dataset page: https://huggingface.co/datasets/inovruzova/azerbaijani-art-collection.imageimage-classificationn<1K2 likes15 downloads1y agoHugging Face10khaleed-mammad /azerbaijan-landmarks-dataset 📂 Dataset Name Azerbaijan Landmarks Dataset Data Collection The dataset includes the following five classes: Maiden Tower Heydar Aliyev Center Dede Gorgud Park Deniz Mall Palace of the Shirvanshahs Structure train/: Training images (~80% of data, excluding validation samples). Since data samples inside train is too big, we divided train_"class_name" test/: Test images (~20%) Format Each image is in .jpg format, organized in class-labeled… See the full description on the dataset page: https://huggingface.co/datasets/khaleed-mammad/azerbaijan-landmarks-dataset.image0 likes11 downloads1y agoHugging Face11tahmaz /azerbaijani_text_news1image10K<n<100K0 likes8 downloads10mo agoHugging Face12bahramhasanov /azerbaijani-carpetimage1K<n<10K0 likes3 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.