CoolFace
19 results

multi-round

sheng22213 /multi_round_speech_180kaudio1M<n<10M2 likes426 downloads1y agoHugging FaceNorquinal /claude_multiround_chat_30kThis dataset is the result of 50k instruction/response pairs generated by Claude and two additional follow-up instructions for each base instruction (for a total of 150k instructions), with instances of blatant alignment removed. 32170 (96510) instructions remain. The instructions were generated synethically using a method that can be tenatively described as "multi-instruct." These instructions consist of numerous discrete tasks that the AI has to work its way through, thereby hopefully… See the full description on the dataset page: https://huggingface.co/datasets/Norquinal/claude_multiround_chat_30k.text10K<n<100K71 likes222 downloads3y agoHugging Facetheblackcat102 /multiround-programming-convo Multi-Round Programming Conversations Based on previous evol-codealpaca-v1 dataset with added sampled questions from stackoverflow, crossvalidated and make it multiround! It should be more suited to train a code assistant which works side by side. Tasks included in here: Data science, statistic, programming questions Code translation : translate a short function from Python, Golang, C++, Java, Javascript Code fixing : Fix randomly corrupts characters with no tab… See the full description on the dataset page: https://huggingface.co/datasets/theblackcat102/multiround-programming-convo.texttext-generation100K<n<1M9 likes109 downloads3y agoHugging FaceTinyPixel /claude_multiround_chat_1k Dataset Card for "claude_multiround_chat_1k" More Information needed text1K<n<10K6 likes89 downloads3y agoHugging Facetheprint /MultiRoundConvos-Code-JS-HTML-CSS-Pythontext1K<n<10K0 likes83 downloads9mo agoHugging Facet2ance /mat-06-ropd-multi-round-trainability 06 ROPD multi-round trainability Can rubric-based on-policy distillation (ROPD, arXiv:2605.07396) against a black-box teacher (Claude Sonnet 5) train the SFT-9B AutoResearch policy (Qwen3.5-9B with a LoRA adapter) over a sustained on-policy trajectory, so that its real ML actions on MLE-bench tasks become more executable or more likely to improve the current research anchor? This repository is the data root of that question: everything its runs wrote, minus the exclusion list in… See the full description on the dataset page: https://huggingface.co/datasets/t2ance/mat-06-ropd-multi-round-trainability.0 likes79 downloads17d agoHugging Face