genz
Datasets
All datasets matching “genz”genz-slang-dataset
Dataset Details
This dataset contains a rich collection of popular slang terms and acronyms used primarily by Generation Z. It includes detailed descriptions of each term, its context of use, and practical examples that demonstrate how the slang is used in real-life conversations.
The dataset is designed to capture the unique and evolving language patterns of GenZ, reflecting their communication style in digital spaces such as social media, text messaging, and online forums. Each… See the full description on the dataset page: https://huggingface.co/datasets/MLBtrio/genz-slang-dataset.genz-to-english
GenZ-to-English Translation Dataset
A high-quality text-to-text dataset for translating Gen Z slang into clear, standard English.
The dataset is designed for training and evaluating language models that convert modern internet slang into natural, readable English while preserving the original meaning.
Overview
This dataset contains 300k++ curated translation pairs covering a wide range of contemporary internet slang.
It includes expressions commonly found across… See the full description on the dataset page: https://huggingface.co/datasets/Sankar-2910/genz-to-english.NeuroBio-GenZ-1K
NeuroBio GenZ 1K
Around 1000 neuroscience and biology questions, answered like your smartest friend is texting you back, not like a textbook is talking at you.
"Why does doomscrolling give me dopamine?" gets answered in three sentences, casual tone, real neuroscience terms (nucleus accumbens, not "reward center"), zero fluff.
Why did you make this?
Because there's genuinely not that much high quality neuroscience and biology data on Hugging Face that isn't either… See the full description on the dataset page: https://huggingface.co/datasets/luka0x12/NeuroBio-GenZ-1K.details_budecosystem__genz-13b-v2
Dataset Card for Evaluation run of budecosystem/genz-13b-v2
Dataset Summary
Dataset automatically created during the evaluation run of model budecosystem/genz-13b-v2 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_budecosystem__genz-13b-v2.details_budecosystem__genz-70b
Dataset Card for Evaluation run of budecosystem/genz-70b
Dataset Summary
Dataset automatically created during the evaluation run of model budecosystem/genz-70b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_budecosystem__genz-70b.genz-slang-pairs-1k
Gen Z Slang Pairs Corpus (1 K)
The Gen Z Slang Pairs Corpus (1 K) contains 1,000 everyday English sentences alongside their Gen Z–style slang rewrites. This dataset is designed for style-transfer, informal-language generation, and paraphrasing research. Use it to train models that transform formal or neutral sentences into expressive, youth‑oriented slang.
Dataset Details
This dataset was generated programmatically using OpenAI GPT-4.1 Nano.
Language: English… See the full description on the dataset page: https://huggingface.co/datasets/Programmer-RD-AI/genz-slang-pairs-1k.
