datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
task362_spolin_yesand_prompt_response_sub_classification
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task362_spolin_yesand_prompt_response_sub_classification
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task362_spolin_yesand_prompt_response_sub_classification.Kishor_V2_53K_LLM_Prompt-Response_Pairs
Kishor V2: 53K Prompt-Response Dataset
Kishor V2 is a diverse and compact dataset curated for training small to medium-sized language models. It includes 53,000 structured prompt-response pairs across multiple domains to simulate human-like dialogue, reasoning, and general intelligence.
📦 File
KishorV2_dataset.jsonl: Main dataset in JSON Lines format.
📂 Format
Each line is a JSON object with:
{
"type": "qa" | "dialogue" | "quote" | "fact" | "reasoning" |… See the full description on the dataset page: https://huggingface.co/datasets/GODELEV/Kishor_V2_53K_LLM_Prompt-Response_Pairs.Trix-Chatbot-Prompt-Response
Dataset Creation Process
Overview
This dataset was created to train and evaluate a chatbot focused on answering questions about Pooria Roy, his background, projects, and related topics. The goal was to build a dataset grounded in real user behavior while maintaining sufficient diversity and coverage of edge cases.
The final dataset contains 2,105 prompt-response examples, including a small portion of multi-turn conversations.
Data Collection Pipeline… See the full description on the dataset page: https://huggingface.co/datasets/regularpooria/Trix-Chatbot-Prompt-Response.
