datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
extractive_qa_question_answering_hr
Dataset Card
HR-Multiwoz is a fully-labeled dataset of 5980 extractive qa spanning 10 HR domains to evaluate LLM Agent. It is the first labeled open-sourced conversation dataset in the HR domain for NLP research.
Please refer to HR-MultiWOZ: A Task Oriented Dialogue (TOD) Dataset for HR LLM Agent for details about the dataset construction.
Dataset Sources
Repository: xwjzds/extractive_qa_question_answering_hr
Paper: HR-MultiWOZ: A Task Oriented Dialogue (TOD)… See the full description on the dataset page: https://huggingface.co/datasets/xwjzds/extractive_qa_question_answering_hr.tamil-question-answering-datasetthis dataset contains 5 columns
context, question, answer_start, answer_text, source
Column
Description
context
A general small paragraph in tamil language
question
question framed form the context
answer_text
text span that extracted from context
answer_start
index of answer_text
source
who framed this context, question, answer pair
source
team KBA => (Karthi, Balaji, Azeez) these people manually created
CHAII =>a kaggle competition
XQA => multilingual QA… See the full description on the dataset page: https://huggingface.co/datasets/AswiN037/tamil-question-answering-dataset.question-answering-ukrainian-json-answersquestion-answering-ukrainianQuestion-Answering-Generation-Choices
The dataset is a merged compilation of QuAIL, RACE, and Cosmos QA datasets,
having undergone preprocessing.
Financial_Question_Answeringtamil-question-answering-datasetthis dataset contains 5 columns
context, question, answer_start, answer_text, source
Column
Description
context
A general small paragraph in tamil language
question
question framed form the context
answer_text
text span that extracted from context
answer_start
index of answer_text
source
who framed this context, question, answer pair
source
team KBA => (Karthi, Balaji, Azeez) these people manually created
CHAII =>a kaggle competition
XQA => multilingual QA… See the full description on the dataset page: https://huggingface.co/datasets/Subi1152/tamil-question-answering-dataset.siddha_vaithiyam_question_answering_chatbot
Medical Home Remedy Chatbot Dataset
Overview
This dataset is designed for a chatbot that answers questions related to medical problems with simple home remedies. The information in this dataset has been sourced from old books containing traditional remedies used in the past.
Contents
Dataset Files:
dataset.csv : The main dataset file containing questions and corresponding home remedy answers.
Data Structure:
Each row in the CSV file… See the full description on the dataset page: https://huggingface.co/datasets/RahulS3/siddha_vaithiyam_question_answering_chatbot.OpenMP_Question_Answering
OpenMP Question Answering Dataset
OpenMP Question Answering Dataset is a new OpenMP question answering introduced in paper "LM4HPC: Towards Effective Language Model Application in High-Performance Computing".
It is designed to probe the capabilities of language models in single-turn interactions with users. Similar to other QA datasets, we include
some request-response pairs which are not strictly question-answering pairs. The categories and examples of questions in the OMPQA… See the full description on the dataset page: https://huggingface.co/datasets/chenle015/OpenMP_Question_Answering.extractive_multi_turn_question_answeringmedical-question-answeringSanskrit-Question-Answering
