datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ScienceQAThis is the ScientificQA dataset by Saikh et al (2022).
@article{10.1007/s00799-022-00329-y,
author = {Saikh, Tanik and Ghosal, Tirthankar and Mittal, Amish and Ekbal, Asif and Bhattacharyya, Pushpak},
title = {ScienceQA: A Novel Resource for Question Answering on Scholarly Articles},
year = {2022},
journal = {Int. J. Digit. Libr.},
month = {sep}
}
arm-asmen-tr-translation
EN-TR Translation Dataset
Dataset Overview
This dataset is based on a subset of the Helsinki-NLP/opus-books dataset, which includes copyright-free books aligned for translation purposes. The original dataset contains multilingual sentence alignments. Specifically, I extracted the English-Italian (EN-IT) portion of the dataset, translated the English sentences to Turkish, and created an English-Turkish (EN-TR) parallel corpus.
The dataset can be used for various natural… See the full description on the dataset page: https://huggingface.co/datasets/armantunga/en-tr-translation.wwii-naval-armament
WWII Naval Armament & Treaty Data
Structured tables extracted from Fleets of World War II: Design History and
Analysis (Nimble Books, ISBN 9781608881604). Three queryable tables plus the full
set of 35 extracted source tables.
table
rows
what
naval_guns
93
gun armament specs — caliber, shell weight, range, ceiling, fire control, ship class
torpedoes
42
torpedo specs — type, explosive weight, range/speed
treaty_tonnage
19
interwar naval-treaty tonnage allocations… See the full description on the dataset page: https://huggingface.co/datasets/wfzimmerman/wwii-naval-armament.arm-asm-xsmallml_data_test_detection_bank_transaction_frauds_unbalanced
ML Data Test Detection Bank Transaction Frauds Unbalanced
The project provides a quick and accessible dataset designed for learning and experimenting with machine learning algorithms, specifically in the context of detecting fraudulent bank transactions. It is intended for practicing and applying concepts such as Random Forest, Support Vector Machines (SVM), and Synthetic Minority Over-sampling Technique (SMOTE) to address unbalanced classification problems.
Note: This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/roberto-armas/ml_data_test_detection_bank_transaction_frauds_unbalanced.Jigsawsocial-media-addiction-vs-relationship
Social Media Addiction vs Relationship Dataset
Dataset ini merupakan hasil survei yang mengukur sejauh mana adiksi media sosial memengaruhi kualitas hubungan romantis mahasiswa.
Kolom-kolom dalam dataset:
Age
Gender
Daily Usage (hours)
Social Media Addiction Score
Relationship Satisfaction
Sumber Asli:
https://www.kaggle.com/datasets/adilshamim8/social-media-addiction-vs-relationships
piaulidadesarmada-logics-ft
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/ecaccam/armada-logics-ft.nutGAIdatas1qwennutTerminarConversacionmy_Superkart
