datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PetraAI
PETRA
Overview
PETRA is a multilingual dataset for training and evaluating AI systems on a diverse range of tasks across multiple modalities. It contains data in Arabic and English for tasks including translation, summarization, question answering, and more.
Dataset Structure
Data is separated by language into /ar and /en directories
Within each language directory, data is separated by task into subdirectories
Tasks include:
Translation
Summarization… See the full description on the dataset page: https://huggingface.co/datasets/PetraAI/PetraAI.ZalmatiAIDailyTalkContiguous-Peter-Griffin
Peter Griffin's Emotion-Tagged DailyTalk Dataset
Hehehehehe! Hey Lois, look! I made a dataset! This is an emotion-tagged version of that DailyTalk thing, but better because it's got all sorts of feelings and stuff. It's freakin' sweet for making computers talk like they've had too many Pawtucket Patriots or just found out Meg is home.
What is this thing?
This dataset has a bunch of people talking, but we tagged 'em with emotions. It's like when I'm happy because it's… See the full description on the dataset page: https://huggingface.co/datasets/MysticKit/DailyTalkContiguous-Peter-Griffin.test-dataset
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/petergazdik/test-dataset.
