datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
romanian-name-days
Romanian Name Days and Holidays
Zile onomastice și sărbători românești — the Romanian name-day calendar as
structured data.
In Romania, ziua onomastică — the feast day of the saint whose name you bear —
is widely celebrated, often more than a birthday. Until now this information
existed online only as HTML pages built for human readers. This is the
machine-readable version.
Published by trends.ro.
Dataset summary
Names
86 (46 masculine, 40 feminine)… See the full description on the dataset page: https://huggingface.co/datasets/radool/romanian-name-days.repo-name
Dataset Card for SynthLogic-Instruct
Dataset Details
Dataset Description
SynthLogic-Instruct is a synthetically generated instruction-following dataset designed to enhance the multi-step logical reasoning and code-generation capabilities of large language models. The dataset consists of high-quality prompt-response pairs spanning algorithmic problem solving, mathematical deduction, and structured data analysis tasks. All samples were generated via a… See the full description on the dataset page: https://huggingface.co/datasets/988tg/repo-name.anime-your-nameThis dataset was created using AI Gemini 2.0 Flash Experimental from the original subtitle format. There may be errors in Turkish and Japanese. Please use this for chatbot training, not for translation AI.
wasp-5k
Wasp-Lang
This is a synthetic dataset created by an amplify model trained on the Wasp programming language quick-start documentation. Better data coming soon.
synthetic-names-1synthetic-names-1 is a synthetic dataset with a total of ~4695 rows.
This dataset was generated using the following models:
ChatGPT:
Whatever is available through the website.
Grok:
Fast
Perplexity.ai:
"Search"
Gemini:
3.1 Flash-Lite
3.5 Flash
3.1 Pro
Deepseek:
"Instant"
"Expert"
Qwen3.7-Max:
Fast
Thinking
This dataset follows the following format:
[
{"in": "Example user description", "out": "Example predicted name"},
{"in": "Example user description", "out": "Example… See the full description on the dataset page: https://huggingface.co/datasets/takenusername32/synthetic-names-1.
