datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Venus
Venus: A dataset for fine-grained code generation control
🎉 What is Venus? Venus is the dataset used to train Afterburner (WIP). It is an extension of the original Mercury dataset and currently includes 6 languages: Python3, C++, Javascript, Go, Rust, and Java.
🚧 What is the current progress? We are in the process of expanding the dataset to include more programming languages.
🔮 Why Venus stands out? A key contribution of Venus is that it provides runtime and memory… See the full description on the dataset page: https://huggingface.co/datasets/Elfsong/Venus.ELF-HP
ELF-HP: Human Preference-Aligned Counter Trolling Dataset
Dataset Summary
ELF-HP (paper) is a dataset designed for studying human-preferred counter-response strategies in Reddit discussions. The dataset contains annotated posts and comments from various subreddits, including various types of trolling attempts and multiple response strategies. It was created to support research in effective counter-responses to online trolling, aligning with human preferences.
Disclaimer:… See the full description on the dataset page: https://huggingface.co/datasets/huijelee/ELF-HP.
