datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
prop-trading-qa-conversational-ai
Prop Trading Q&A Dataset for Conversational AI
Description
This dataset contains 200+ curated question-answer pairs covering the domain of proprietary (prop) trading firms. It is designed to serve as training and retrieval data for building AI assistants, chatbots, and educational tools focused on prop trading knowledge.
Each entry consists of a natural-language question paired with a detailed, factual answer. The data spans ten thematic categories ranging from… See the full description on the dataset page: https://huggingface.co/datasets/propfirmkey/prop-trading-qa-conversational-ai.cot-conversational-qa-v1
CoT Conversational QA v1
Conversational QA pairs describing Qwen3-8B chain-of-thought reasoning traces. Generated by prompting Gemini 2.0 Flash to describe observable facts about CoT text with zero logical leaps.
Purpose
Training data for activation oracles — models that read their own activations and answer questions about their reasoning. The oracle sees strided activations at sentence boundaries, not the CoT text itself.
The training signal: simple question → factual… See the full description on the dataset page: https://huggingface.co/datasets/ceselder/cot-conversational-qa-v1.
