datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
M1_EURUSD_candles
All chunks have more than 4000 rows of data in chronological order in a panda dataframe
CSV files are the same data in chronological order, some may not be more than 4000 rows
sauatai-ertegiler-kz-misspellings-kk-s170-len60-n6-m1-3-v1
SauatAI — Kazakh Misspelled Sentences from Ertegiler.kz
SauatAI is a grammar-focused dataset built from 170 children’s stories scraped from ertegiler.kz on July 5, 2025. The dataset was designed to support Kazakh language grammar correction, error detection, and text augmentation research.
📌 Dataset Details
s170 — 170 unique stories were scraped and sentence-tokenized.
len60 — Only sentences with ≤60 characters were retained.
n6 — Each correct sentence has 5… See the full description on the dataset page: https://huggingface.co/datasets/alphazhan/sauatai-ertegiler-kz-misspellings-kk-s170-len60-n6-m1-3-v1.forex-algotrading-m15-1000-columns
FOREX Algotrading M15 1000 Columns
Descrição
Este dataset contém dados históricos de Forex em intervalos de 15 minutos (M15), incluindo múltiplos pares de moedas. É voltado para análise de séries temporais e desenvolvimento de algoritmos de trading automatizados. Cada arquivo CSV contém 1000 colunas com dados de mercado, como preços de abertura, fechamento, máxima e mínima, volume e spread.
Estrutura do dataset
Cada arquivo CSV possui as seguintes colunas:… See the full description on the dataset page: https://huggingface.co/datasets/lukealvess/forex-algotrading-m15-1000-columns.sauatai-ertegiler-kz-misspellings-kk-s170-len60-n6-m1-v1
SauatAI — Kazakh Misspelled Sentences from Ertegiler.kz
SauatAI is a grammar-focused dataset built from 170 children’s stories scraped from ertegiler.kz on July 5, 2025. The dataset was designed to support Kazakh language grammar correction, error detection, and text augmentation research.
📌 Dataset Details
s170 — 170 unique stories were scraped and sentence-tokenized.
len60 — Only sentences with ≤60 characters were retained.
n6 — Each correct sentence has 5… See the full description on the dataset page: https://huggingface.co/datasets/alphazhan/sauatai-ertegiler-kz-misspellings-kk-s170-len60-n6-m1-v1.TruthfulQA
Dataset Card for TruthfulQA
Dataset Summary
TruthfulQA: Measuring How Models Mimic Human Falsehoods
We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions that some humans would answer falsely due to a false belief or misconception. To perform well, models must avoid generating false answers… See the full description on the dataset page: https://huggingface.co/datasets/M1STERPERFECT/TruthfulQA.M11_Trainprogram_gen_v1_m1sauatai-ertegiler-kz-misspellings-kk-s170-len60-n6-m1-2-v1
SauatAI — Kazakh Misspelled Sentences from Ertegiler.kz
SauatAI is a grammar-focused dataset built from 170 children’s stories scraped from ertegiler.kz on July 5, 2025. The dataset was designed to support Kazakh language grammar correction, error detection, and text augmentation research.
📌 Dataset Details
s170 — 170 unique stories were scraped and sentence-tokenized.
len60 — Only sentences with ≤60 characters were retained.
n6 — Each correct sentence has 5… See the full description on the dataset page: https://huggingface.co/datasets/alphazhan/sauatai-ertegiler-kz-misspellings-kk-s170-len60-n6-m1-2-v1.M1_MCQ_generated_with_questionsM10_TrainE1.M1-Atrial-Fibrillation-Risk-Prediction-Older-Adults
AFib Synthetic Dataset — README
Overview
This is a synthetic clinical dataset of 10,000 patient records designed to support research and machine learning work on Atrial Fibrillation (AFib) risk prediction. Each row represents one patient and includes demographic, clinical, lifestyle, lab, and wearable-derived features, along with a binary AFib diagnosis label.
Important: All data is synthetically generated and does not represent real patients.
File… See the full description on the dataset page: https://huggingface.co/datasets/Auric-Grid/E1.M1-Atrial-Fibrillation-Risk-Prediction-Older-Adults.M1_trainm1_train_tempM12_Trainm1_epfl_datam1_train_5_tempm1_train_10_tempSAIGE-right-speech
