datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
skripsi-data
Indonesian Legal Retrieval Benchmark
A retrieval benchmark over Indonesian legislation from BPK JDIH (peraturan.bpk.go.id). Built for an undergraduate thesis at the Faculty of Computer Science, Universitas Indonesia, that compares a BM25 lexical baseline, vectorless retrieval driven by LLM reasoning, and vector-based dense retrieval on the same corpus and gold set. The benchmark is retrieval-only. There is no answer generation and no generation labels, systems are scored on… See the full description on the dataset page: https://huggingface.co/datasets/wahyyuht/skripsi-data.cti-mcqcoin_flipgsm8k_only_answerThe data is exactly like the original GSM8k (https://huggingface.co/datasets/gsm8k ), but with the label consisting of the correct answer(one number) only.
@misc{krishna2024gsmansweronly,
title={GSM8k (Answer only)},
author={Satyapriya Krishna},
year={2023},
url={skrishna/gsm8k_only_answer},
}
piqa_preop
Dataset Card for "piqa_preop"
More Information needed
truthfulqa_prepropcoin_flip_7heart_disease_uci
Dataset Card for Dataset Name
age: age in years
sex: sex (1 = male; 0 = female)
cp: chest pain type
-- Value 1: typical angina
-- Value 2: atypical angina
-- Value 3: non-anginal pain
-- Value 4: asymptomatic
trestbps: resting blood pressure (in mm Hg on admission to the hospital)
chol: serum cholestoral in mg/dl
fbs: (fasting blood sugar > 120 mg/dl) (1 = true; 0 = false)
restecg: resting electrocardiographic results
-- Value 0: normal… See the full description on the dataset page: https://huggingface.co/datasets/skrishna/heart_disease_uci.CSQA_preprocessed_mul
Dataset Card for "CSQA_preprocessed_mul"
More Information needed
toxigen_annotated_modtoxicity_prepropbergskollegium_relationer_och_skrivelser_linescoin_flip_4
Dataset Card for "coin_flip_4"
More Information needed
coin_flip_15coin_flip_15_transformedboolq
Dataset Card for "boolq"
More Information needed
coin_flip_2_transformed
Dataset Card for "coin_flip_2_transformed"
More Information needed
salient_translation_error_detection_preprocessed
Dataset Card for "salient_translation_error_detection_preprocessed"
More Information needed
CSQA_preprocessed
Dataset Card for "CSQA_preprocessed"
More Information needed
SKR1
license: cc-by-nc-4.0
SKR1 - Benchmark for Testing Knowledge about Slovak Realia for Large Language Models
Overview
SKR1 is a specialized benchmark designed to evaluate Large Language Models' knowledge of Slovak cultural and factual context. Developed by Marek Dobeš at ČZ o.z., this benchmark addresses the significant gap in culturally-specific evaluations for underrepresented languages like Slovak.
Key Features
35 carefully crafted questions covering four… See the full description on the dataset page: https://huggingface.co/datasets/ajtakto/SKR1.skripsi_UI_membership_30K
Dataset Card for "skripsi_UI_membership_30K"
More Information needed
SECURE-CWETtoy-toxicity-datasetallenai-real-toxicity-prompts_70M_toxic
Dataset Card for "allenai-real-toxicity-prompts_70M_toxic"
More Information needed
aya_collection_train_hifiltered_toxic_samplesjaredjoss-jigsaw-long-2000_70M_non_toxic
Dataset Card for "jaredjoss-jigsaw-long-2000_70M_non_toxic"
More Information needed
CNC_skript12
Introduction
This is the SKRIPT2012 dataset, maintained by the Czech National Corpus project. This dataset corresponds to the version available in the LINDAT repository, where it is named AKCES-1. The dataset was created from public .rtf and .doc file formats using the convert_AKCES.py script.
About Original Dataset
(Taken from project Wiki).
The Corpus SKRIPT2012 is a learner corpus aimed at representing the written language of Czech pupils and students at elementary… See the full description on the dataset page: https://huggingface.co/datasets/CZLC/CNC_skript12.SECURE-VOODskripzi_parallel_revision
