datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
engsaf
Engineering Short Answer Feedback
A collection of real short-answer responses from engineering exams across multiple engineering domains.
Background
In recent years, there has been a growing interest in using Artificial Intelligence (AI) to automate student assessment in education.
Among different types of assessments, summative assessments play a crucial role in evaluating a student's understanding level of a course.
Such examinations often involve short-answer… See the full description on the dataset page: https://huggingface.co/datasets/IsmaelMousa/engsaf.bitcoin-historical-dataset
Historical Bitcoin Market, On-Chain, Mining and Macroeconomic Dataset
Dataset Summary
Comprehensive daily Bitcoin dataset from genesis block (2009-01-03) to 2026-09-07.
6,457 daily observations combining market data, on-chain metrics, mining stats, macro indicators, and 100+ derived features.
Historical Coverage
Period
Coverage
Reliability
2009-01-03 to 2010-07-17
No market price
Protocol only
2010-07-18 to 2013-04-27
Monthly… See the full description on the dataset page: https://huggingface.co/datasets/ismailtasdelen/bitcoin-historical-dataset.Sorting-Algorithms-Performance-Metrics
Sorting Algorithms Benchmark Dataset (Array Size: 1000)
A benchmark dataset comparing execution time, memory usage, and comparison counts of various sorting algorithms (Bubble Sort, Selection Sort, Insertion Sort, Merge Sort, Quick Sort, Heap Sort, Odd-Even Sort) on arrays of size 1000. Each algorithm was run 100 times with randomized inputs to ensure statistical significance.
Dataset Details
Columns
run: Trial number (1-100 per algorithm).
algorithm:… See the full description on the dataset page: https://huggingface.co/datasets/ismielabir/Sorting-Algorithms-Performance-Metrics.uniGame
UniGame Dataset
This dataset explores the relationship between gaming habits and academic performance among students. It includes various attributes such as age, educational level, CGPA, gaming habits, and other related factors.
Dataset Details
Dataset Description
This dataset aims to investigate how gaming affects the academic performance of students. It includes information on the respondents' demographics, gaming habits, and academic results.
Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/ismail31415/uniGame.data-version2Jarvis-MCU-Dialogues
README
Dataset Description
This dataset is a collection of dialogues generated using chatGPT and Mistral Large models. The dialogues are designed to mimic interactions between Tony Stark (a character from the Marvel Cinematic Universe (MCU)) and Jarvis (his AI assistant). The dataset is composed of two columns: the first column contains Tony Stark's instructions, and the second column contains Jarvis's responses.
Data Generation
The data was generated… See the full description on the dataset page: https://huggingface.co/datasets/ismaildlml/Jarvis-MCU-Dialogues.Quantum_Gate_Performance_Evaluation
🧪 Quantum Gate Performance Dataset
📘 Title:
Comprehensive Quantum Gate Performance Analysis: A Comparative Study of Noise and No-Noise Effects
📂 Dataset Description:
This repository contains benchmarking results for 13 quantum gates (e.g., H, CNOT, Toffoli) tested under noisy and noise-free conditions, based on 1000 simulation runs per gate configuration. Total 26000 rows and 13 columns.
📊 Features include:
Gate Type
Execution Time
Error Rate
Fidelity… See the full description on the dataset page: https://huggingface.co/datasets/ismielabir/Quantum_Gate_Performance_Evaluation.Spoken2TSL
Dataset Description
This dataset is a collection of Turkish to Turkish Sign Language (TSL) grammar version translations. The dataset is designed to facilitate research and development in the field of sign language translation and understanding. It contains pairs of sentences in Turkish and their corresponding TSL translations, which have been curated to follow the grammatical structure of TSL.
Data Collection
The data was collected primarily from the website… See the full description on the dataset page: https://huggingface.co/datasets/ismaildlml/Spoken2TSL.Dataset_de_pruebaExtraído de https://github.com/anthony-wang/BestPractices/tree/master/data.
Campos:
Formula (string)
T (float64): Temperatura (K)
CP (float64): Capacidad calorífica (J/mol K)
Dataset105blindspots-analysis
DeepSeek-R1-Distill-Qwen-1.5B Blind Spot Analysis
Model Tested
Model: tphage/DeepSeek-R1-Distill-Qwen-1.5BLink: https://huggingface.co/tphage/DeepSeek-R1-Distill-Qwen-1.5B
This is a 1.5B parameter distilled base reasoning model derived from Qwen architecture. It is not fine-tuned for a narrow downstream task.
Objective
The purpose of this dataset is to systematically document prediction errors ("blind spots") of the DeepSeek-R1-Distill-Qwen-1.5B model across… See the full description on the dataset page: https://huggingface.co/datasets/ismielabir/blindspots-analysis.data-Ctitanic_practica
