datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
English-Handwritten-Math-Notes-Dataset
English Handwritten Math Notes Dataset
This dataset contains high-resolution images of handwritten mathematical notes written in English. It includes problem statements, worked examples, formulas, and annotated derivations. The dataset supports AI research in handwriting recognition, mathematical OCR, and document understanding for STEM applications.
Contact
For queries or collaborations related to this dataset, contact:
anoushka@kgen.io
abhishek.vadapalli@kgen.io… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/English-Handwritten-Math-Notes-Dataset.Math-Shapes
Math-Symbols Dataset
Overview
The Math-Symbols dataset is a collection of images representing various mathematical symbols. This dataset is designed for machine learning applications, particularly in the fields of image recognition, optical character recognition (OCR), and symbol classification.
Dataset Details
Name: Math-Symbols
Type: Image dataset
Format: Images with corresponding labels
Size: 131MB (downloaded dataset files), 118MB (auto-connected Parquet… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Math-Shapes.Math-Equa
Math-Equa Dataset
Overview
The Math-Equa dataset is a collection of mathematical equations designed for machine learning applications. This dataset can be used for tasks such as equation solving, symbolic mathematics, and other related research areas.
Dataset Details
Name: Math-Equa
Type: Mathematical Equations
Format: Text-based equations
Size: [Insert size of the dataset]
Source: [Insert source of the dataset, if applicable]
Usage
This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Math-Equa.mathcaptchaforge-dataset-21k-parquet
Dataset Card
Overview
This dataset contains labeled image crops for multi-class visual symbol classification.
Structure
images/: image files referenced by the manifests
manifest.csv: full index
train.csv, val.csv, test.csv: split manifests
Schema
All CSV files use:
id,image_path,label,width,height,split
image_path is relative to the dataset root, formatted as images/<filename>.
Usage
Load one of the split CSV files.
Resolve image_path… See the full description on the dataset page: https://huggingface.co/datasets/emrecengdev/mathcaptchaforge-dataset-21k-parquet.
