datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Health_Benchmarks
LLM Health Benchmarks Dataset by Yesil Science
The LLM Health Benchmarks Dataset is a specialized resource for evaluating large language models (LLMs) in different medical specialties. It provides structured question-answer pairs designed to test the performance of AI models in understanding and generating domain-specific knowledge.
Primary Purpose
This dataset is built to:
Benchmark LLMs in medical specialties and subfields.
Assess the accuracy and contextual… See the full description on the dataset page: https://huggingface.co/datasets/yesilhealth/Health_Benchmarks.hotpot_qa
Dataset Card for "hotpot_qa"
Dataset Summary
HotpotQA is a new dataset with 113k Wikipedia-based question-answer pairs with four key features: (1) the questions require finding and reasoning over multiple supporting documents to answer; (2) the questions are diverse and not constrained to any pre-existing knowledge bases or knowledge schemas; (3) we provide sentence-level supporting facts required for reasoning, allowingQA systems to reason… See the full description on the dataset page: https://huggingface.co/datasets/Yeshenyue/hotpot_qa.HPCPerfOpt-Yes-No
Dataset Card for HPCPerfOpt (HPC Performance Optimization Dataset)
Dataset Summary
This dataset is a question answering dataset for OpenMP Performance Optimization questions. It contains yes-no questions of 2 types.
Given two codes code 1 and code 2, will code 2 run faster than code 1?
Does this code have a potential issue here?
The answers may be "Yes" or "No".
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More… See the full description on the dataset page: https://huggingface.co/datasets/sharmaarushi17/HPCPerfOpt-Yes-No.
