datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
IIIT-INDIC-HW-WORDS-Hindi
IIIT-INDIC-HW-WORDS-Hindi
Dataset containing images of hand written words in Devanagari by various humans and the corresponding text of those images.
Overview
The dataset, originally developed by the Centre for Visual Information Technology (CVIT) at IIIT Hyderabad, has been transformed into Parquet format to facilitate its use in modern machine learning workflows. This dataset primarily targets recognition of handwritten Hindi words and aims to advance research and… See the full description on the dataset page: https://huggingface.co/datasets/c3rl/IIIT-INDIC-HW-WORDS-Hindi.Medical_Prescription_Handwritten_Words
Medical Prescription Handwritten Words
This dataset contains images of individual handwritten medical words extracted from prescription notes. It is designed for training and evaluating handwriting recognition models in the healthcare domain.
Structure
images/: Contains 40+ handwritten word images (e.g., Amoxicillin.png, Cold.png, Tablet.png, 0.png, etc.)
data.csv: Maps each image file to its corresponding label (word)
Example Use Cases
OCR (Optical… See the full description on the dataset page: https://huggingface.co/datasets/avi-kai/Medical_Prescription_Handwritten_Words.IIIT-INDIC-HW-WORDS-Hindi
IIIT-INDIC-HW-WORDS-Hindi
Dataset containing images of hand written words in Devanagari by various humans and the corresponding text of those images.
Overview
The dataset, originally developed by the Centre for Visual Information Technology (CVIT) at IIIT Hyderabad, has been transformed into Parquet format to facilitate its use in modern machine learning workflows. This dataset primarily targets recognition of handwritten Hindi words and aims to advance research… See the full description on the dataset page: https://huggingface.co/datasets/HarishBonu/IIIT-INDIC-HW-WORDS-Hindi.image-in-Words400
Mouwiya/image-in-words400
Dataset Description
Mouwiya/image-in-words400 is a dataset consisting of 400 images along with their corresponding descriptive captions. The dataset is designed for tasks related to image captioning, where the goal is to generate accurate and contextually relevant descriptions for visual content. This dataset can be used to train and evaluate models that bridge the gap between visual and textual data.
Dataset Details
Total Examples:… See the full description on the dataset page: https://huggingface.co/datasets/Mouwiya/image-in-Words400.Medical_Prescription_Handwritten_Words
Medical Prescription Handwritten Words
This dataset contains images of individual handwritten medical words extracted from prescription notes. It is designed for training and evaluating handwriting recognition models in the healthcare domain.
Structure
images/: Contains 40+ handwritten word images (e.g., Amoxicillin.png, Cold.png, Tablet.png, 0.png, etc.)
data.csv: Maps each image file to its corresponding label (word)
Example Use Cases
OCR… See the full description on the dataset page: https://huggingface.co/datasets/MMMuzammil/Medical_Prescription_Handwritten_Words.image-in-Words400_DOCCI_Test
