datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
smoke-openai-terra-batch-brasil-25-20260724-01
Smoke OpenAI Terra Batch — Brasil × 25 tasks
Run real de validação do fluxo matricial document_task_matrix, executada
sobre um único documento da Wikipédia em português com o título Brasil.
Cada uma das 25 tasks canônicas recebeu exatamente um slot inicial.
Resultado
status: completed
documentos: 1
pares planejados: 25
exemplos aceitos: 25
pares pulados: 0
pares esgotados: 0
resultados reais do backend: 27
retries com nova chamada: 2
backend: openai_api… See the full description on the dataset page: https://huggingface.co/datasets/costadev00/smoke-openai-terra-batch-brasil-25-20260724-01.empathetic_dialogues_ko
Dataset Card for "한국어 일상 속 공감형 대화 데이터셋(멀티-턴)"
Dataset Summary
boostCamp AI Tech 5기 과정 중 NLP 12조 훈제연어들 팀의 최종 프로젝트에서 제작한 데이터입니다.
일상 속 다양한 상황에서 사용자와 챗봇 간의 대화를 담은 데이터셋 입니다.
GPT4, GPT3.5-turbo로 제작된 합성데이터이며 싱글-턴, 2-턴, 3-턴 대화로 구성되어 있습니다.
답변은 [공감적 표현 - 일반적인 대화 - 관련된 질문] 의 형태를 가집니다.
Generation Prompt Example(GPT3.5-turbo)
Take a close look at the following example and Conditions. Create nine sessions that each of the session is ongoing conversation about a single… See the full description on the dataset page: https://huggingface.co/datasets/Smoked-Salmon-s/empathetic_dialogues_ko.team-coop-smoke
CooperBench Team → Coop (Qwen3.5-9B smoke)
Placeholder / smoke dataset (1 trajectory). A single 2-agent
cooperbench team run (lead + member, no protocol) reshaped into the
2-agent coop layout defined in cooperbench/CooperData PR
#98.
The full dataset is the canonical place where future team→coop conversions
will land; this entry validates the converter and the publishing pipeline.
Source
Source run
logs/qwen35-smoke-mini-team-noproto/
Source repo… See the full description on the dataset page: https://huggingface.co/datasets/CooperBench/team-coop-smoke.
