tutoring
Datasets
All datasets matching “tutoring”Question-Anchored-Tutoring-Dialogues-2k
Question-Anchored-Tutoring-Dialogues-2k
This dataset contains dialogues from math tutoring interventions recorded on Eedi.
Dataset Details
Dataset Description
Each dialogue represents a chat-based conversation between a tutor and a student prompted by the student requesting assistance while working on a lesson. Dialogues are accompanied with 2 sources of meta-data:
DQ-Question-Metadata: The question the student was working on that prompted the tutoring… See the full description on the dataset page: https://huggingface.co/datasets/Eedi/Question-Anchored-Tutoring-Dialogues-2k.TutoringDialogs
TutoringDialogs — curated subset (500 dialogues)
500 student–tutor dialogues selected and normalized from a larger raw pool of
~1,900 synthetically generated dialogues (source files: exams,
maths_and_informatics, mixed_themes, physics_and_informatics), each of
which originally used a different JSON schema. This file merges them all
into one consistent schema, removes duplicates and broken records, and
selects a maximally diverse subset for LoRA/SFT fine-tuning of a small
(1.5B)… See the full description on the dataset page: https://huggingface.co/datasets/ptvnck/TutoringDialogs.vectorstore-academic_tutoring
Vectorstore Dataset: Academic Tutoring
Overview
This dataset contains pre-computed vector embeddings for the academic tutoring domain, ready for use in Retrieval-Augmented Generation (RAG) applications, semantic search, and knowledge base systems. The embeddings are generated from high-quality source documents using state-of-the-art sentence transformers, making it easy to build production-ready RAG applications without the computational overhead of embedding generation.… See the full description on the dataset page: https://huggingface.co/datasets/meetara-lab/vectorstore-academic_tutoring.annotated-math-tutoring-datasetkorean-tutoring-persona-dataset
Korean Tutoring Persona Dataset
Overview
This dataset is constructed to evaluate the impact of learner persona modeling in bilingual Korean tutoring systems.
Each instance includes:
- Learner persona information
- Task type (grammar, error correction, multi-turn dialogue)
- Input query
- Three responses generated under different persona configurations
Data Structure
Each data instance is formatted as:
persona: learner profile (TOPIK level… See the full description on the dataset page: https://huggingface.co/datasets/GYPCC2/korean-tutoring-persona-dataset.socratic-tutoring-dataset
