datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
reddit_one_ups_seq2seq_2014
Dataset Card for reddit_one_ups_seq2seq_2014
Dataset Summary
Reddit 'one-ups' or 'clapbacks' - replies which scored higher than the original comments.
This dataset chose freeform replies, which did not follow repetitive meme replies. The IAmA subreddit was excluded to avoid an issue where their answers frequently score higher than questions.
For commentary on predictions with a previous version of the dataset, see… See the full description on the dataset page: https://huggingface.co/datasets/georeactor/reddit_one_ups_seq2seq_2014.hovh_tumanyan_seq2seq
📝 Seq2Seq Dataset: Hovhannes Tumanyan's Poems
This dataset contains sentence pairs extracted from the poetic works of Hovhannes Tumanyan, one of the most celebrated Armenian poets. It is formatted for training sequence-to-sequence (Seq2Seq) models for tasks such as:
Text generation
Dialogue modeling
Style imitation
📂 Dataset Structure
The dataset is provided as a .csv file with two columns:
input_sentence
target_sentence
Line 1 of poem
Line 2 of poem
Line… See the full description on the dataset page: https://huggingface.co/datasets/EdUarD0110/hovh_tumanyan_seq2seq.propbank_srl_seq2seqdevign_seq2seqseq2seq
