datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
generalization-science-dataServiceProjectFall2023
Deep Learning Service Project (Fall 2023)
Getting Started
Clone the repository with git lfs disabled or not installed.
ON WINDOWS
set GIT_LFS_SKIP_SMUDGE=1
git clone https://huggingface.co/datasets/DataScienceClubUVU/ServiceProjectFall2023
ON LINUX
GIT_LFS_SKIP_SMUDGE=1 git clone https://huggingface.co/datasets/DataScienceClubUVU/ServiceProjectFall2023
Download the pytorch file (.pth) from… See the full description on the dataset page: https://huggingface.co/datasets/DataScienceClubUVU/ServiceProjectFall2023.modis-lake-powell-toy-dataset
MODIS Water Lake Powell Toy Dataset
Dataset Summary
Tabular dataset comprised of MODIS surface reflectance bands along with calculated indices and a label (water/not-water)
Dataset Structure
Data Fields
water: Label, water or not-water (binary)
sur_refl_b01_1: MODIS surface reflection band 1 (-100, 16000)
sur_refl_b02_1: MODIS surface reflection band 2 (-100, 16000)
sur_refl_b03_1: MODIS surface reflection band 3 (-100, 16000)
sur_refl_b04_1: MODIS… See the full description on the dataset page: https://huggingface.co/datasets/nasa-cisto-data-science-group/modis-lake-powell-toy-dataset.data_science_bowl_2018tutorial-senegal-lclucDiEm_HTR
Dataset Card for DiEm HTR
The DiEm HTR dataset is a ground truth dataset for historical danish handwriting in the 17th and 18th century, generated as part of the Digitalisering af Enesteministerialbøger project at the Danish National Archives.
Dataset Details
Dataset Description
The Digitalisering af Enesteministerialbøger project (DiEm) at the Danish National Archives aims to transcribe and make publically available all of the danish parish registers from… See the full description on the dataset page: https://huggingface.co/datasets/RA-Data-Science/DiEm_HTR.Art_Images_Ai_And_Real_
Dataset Card for Dataset Name
This dataset card aims to give a basic understanding what is the data contain, where it come from and how it is structured. the purpose of such dataset exist.
Dataset Details
The dataset contains of 2,839 images: training set --> 2560 , test set -> 279.
The data is balanced in each set , 50% is label 0 (as real image) --> 50% label 1 (Ai generated image)
*** Dowload link :… See the full description on the dataset page: https://huggingface.co/datasets/DataScienceProject/Art_Images_Ai_And_Real_.modis-lake-powell-raster-dataset
MODIS Water Lake Powell Raster Dataset
Dataset Summary
Raster dataset comprised of MODIS surface reflectance bands along with calculated indices and a label (water/not-water)
Dataset Structure
Data Fields
water: Label, water or not-water (binary)
sur_refl_b01_1: MODIS surface reflection band 1 (-100, 16000)
sur_refl_b02_1: MODIS surface reflection band 2 (-100, 16000)
sur_refl_b03_1: MODIS surface reflection band 3 (-100, 16000)
sur_refl_b04_1:… See the full description on the dataset page: https://huggingface.co/datasets/nasa-cisto-data-science-group/modis-lake-powell-raster-dataset.modern-danish-handwriting
Dataset Card for Modern Danish Handwriting
The Modern Danish Handwriting dataset is a Danish-language dataset containing more than 200 pages of transcribed and proofread handwritten text.
Dataset Details
Dataset Description
The Modern Danish Handwriting dataset currently consists of handwritten samples of text from the ePAROLE dataset. The samples were created by volunteers at the Danish National Archives and guests at the festival Historiske Dage in 2025.… See the full description on the dataset page: https://huggingface.co/datasets/RA-Data-Science/modern-danish-handwriting.scienceqaDiEm_HTR-Numbers
Dataset Card for DiEm HTR Numbers
The DiEm HTR Numbers dataset is a ground truth dataset consisting of numbers written in historical danish handwriting from the 18th century, generated as part of the Digitalisering af Enesteministerialbøger project at the Danish National Archives.
Dataset Details
Dataset Description
The Digitalisering af Enesteministerialbøger project (DiEm) at the Danish National Archives aims to transcribe and make publically available… See the full description on the dataset page: https://huggingface.co/datasets/RA-Data-Science/DiEm_HTR-Numbers.intro-to-data-science-1
Stock Volume and Return Analysis
1. Dataset Overview
This project analyzes stock market data from July 2025, using a dataset containing daily information such as open, close, high, and low prices, trading volume, and company fundamentals.
File used: stock_data_aug_2025.csvRows: 2,542Columns: 14Period covered: July 1–31, 2025
2. Research Question
The analysis focuses on identifying if increasing trading volume after monthly low ( in this case during… See the full description on the dataset page: https://huggingface.co/datasets/galsolomon9/intro-to-data-science-1.programme-de-la-fete-de-la-science-2019
Programme de la fête de la Science 2019
Source
Source officielle : https://www.data.gouv.fr/datasets/programme-de-la-fete-de-la-science-2019
Identifiant du jeu de données data.gouv.fr : 5cf5df4a9ce2e7536246d3fd
Slug data.gouv.fr : programme-de-la-fete-de-la-science-2019
Licence indiquée dans les métadonnées data.gouv.fr : fr-lo
Structure Hugging Face
Un jeu de données data.gouv.fr = un dépôt Hugging Face
Une ressource tabulaire d’origine = un… See the full description on the dataset page: https://huggingface.co/datasets/Data-Gouv-ML/programme-de-la-fete-de-la-science-2019.chest_xray_datasetimage
