CoolFace
Datasetpublic

hoang-quoc-trung/fusion-image-to-latex-datasets

Collects and builds the largest dataset to date from online sources, creating a robust and generalizable dataset. This dataset includes approximately 3.4 million image-text pairs, including both handwritten mathematical expressions (200,330 examples) and printed mathematical expressions (3,237,250 examples). Due to the large dataset and the fact that the same mathematical formula can be represented in different LaTeX string formats in an image, it is easy to cause polymorphic ambiguity. To… See the full description on the dataset page: https://huggingface.co/datasets/hoang-quoc-trung/fusion-image-to-latex-datasets.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
18likes98downloads

hoang-quoc-trung/fusion-image-to-latex-datasets · main · files are served by the source, never re-hosted here