datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-ForecastingThe sp500stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 4,213 S&P 500 stocks.
The hs300stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 858 HS 300 stocks.
If you find our research helpful, please cite our paper:
@article{xu2025finmultitime,
title={FinMultiTime: A Four-Modal Bilingual Dataset for… See the full description on the dataset page: https://huggingface.co/datasets/Wenyan0110/Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-Forecasting.pagoda-text-and-image-dataset
Dataset Card for "pagoda-text-and-image-dataset"
More Information needed
image-text-dataset-subset-300k-captions_onlyMultimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-ForecastingThe sp500stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 4,213 S&P 500 stocks.
The hs300stock_data_description.csv file provides detailed information on the existence of four modalities (text, image, time series, and table) for 858 HS 300 stocks.
If you find our research helpful, please cite our paper:
@article{xu2025finmultitime,
title={FinMultiTime: A Four-Modal Bilingual Dataset for… See the full description on the dataset page: https://huggingface.co/datasets/Y123-wed/Multimodal-Dataset-Image_Text_Table_TimeSeries-for-Financial-Time-Series-Forecasting.pagoda-text-and-image-dataset-small
Dataset Card for "pagoda-text-and-image-dataset-small"
More Information needed
image-text-dataset-subset-300k-captions_only_with_latentsChinese-Image-Text-Corpus-dataset
REILX/Chinese-Image-Text-Corpus-dataset
[ English | 中文 ]
Introduction
The REILX/Chinese-Image-Text-Corpus-dataset is a multimodal dataset that pairs Chinese textual data with corresponding images. This dataset is derived from the Chinese-Xinhua Dictionary Database, which includes idioms, single characters, words, and aphorisms.
Dataset Structure
The dataset is organized into the following categories:
Idioms: Traditional Chinese idioms with explanations and… See the full description on the dataset page: https://huggingface.co/datasets/REILX/Chinese-Image-Text-Corpus-dataset.shilla-clothing-text-and-image-dataset
Dataset Card for "shilla-clothing-text-and-image-dataset"
More Information needed
ImgMT-Dataset-For-Evaluating-Text-Image-Translationshirt-text-image-datasettext-image-datasetimage-text-dataset-subset-300k-captions_textpagoda-text-and-image-dataset-steeple
Dataset Card for "pagoda-text-and-image-dataset-steeple"
More Information needed
dummy-image-text-datasetcarla_image_to_text_datasetmy_image_text_datasetimage-text-dataset-kittimy-image-text-datasetkompsat_image_text_dataset
KOMPSAT-3/3A Image-Text Dataset
This dataset is a high-resolution remote sensing image-text dataset constructed by the Korea Aerospace Research Institute (KARI). It is specifically designed to improve the accuracy and interpretability of Large Multimodal Models (LMMs) specialized in satellite image analysis by bridging the gap between general-domain imagery and the unique physical characteristics of satellite data.
Dataset Details
Developed by: Korea Aerospace… See the full description on the dataset page: https://huggingface.co/datasets/ohhan777/kompsat_image_text_dataset.Text_to_image_datasetdummy-image-text-dataset40K_kashmiri_text_and_image_dataset
40K Kashmiri Words with images
Kashmiri (words) Image and Text Dataset for OCR Models
This repository contains a dataset specifically curated for training and testing Optical Character Recognition (OCR) models on Kashmiri language text. The dataset includes a large collection of images with their corresponding labels in a CSV file, designed to aid in the development of robust OCR solutions for the Kashmiri script.
Directory Structure:
Zip File/
├── images/
│ ├── 000001.png
│… See the full description on the dataset page: https://huggingface.co/datasets/Omarrran/40K_kashmiri_text_and_image_dataset.my_image_text_datasetmy-image-text-dataset-cleanImage-To-Text-Validation-Datasetmy-image-text-dataset31K_Kashmiri_text_and_image_dataset_for_text_RecognitionDirectory Structure:
Zip File/
├── images/
│ ├── 000001.png
│ ├── 000002.png
│ ├── 000003.png
│ ├── ...
│ ├── 031000.png
├── labels.csv
└── metadata.json
Description:
Zip File: The root directory that contains all other files and folders.
images/: A folder containing 31,000 images named sequentially from 000001.png to 031000.png.
labels.csv: A CSV file that includes information about the labels for each image.
metadata.json: A JSON file that contains metadata about the dataset… See the full description on the dataset page: https://huggingface.co/datasets/Omarrran/31K_Kashmiri_text_and_image_dataset_for_text_Recognition.my-image-text-dataset-not-clean
