datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
styles
Styled Image Dataset Generated with FLUX.1-dev and LoRAs from the community
Access the generation scripts here.
Dataset Description
This dataset contains 60,000 text-image-pairs. The images are generated by adding trained LoRA weights to the diffusion transformer model black-forest-labs/FLUX.1-dev. The images were created using 6 different style models, with each style having its own set of 10,000 images. Each style includes 10,000 captions sampled from the… See the full description on the dataset page: https://huggingface.co/datasets/rezashkv/styles.ramanv-image-real-stylesintel-stylesheet-javascript
Intel Ark Frontend Assets (CSS & JS)
This dataset contains the enterprise frontend assets (Stylesheets and JavaScript files) extracted from Intel Ark (ark.intel.com).
🎯 Primary Use Case
This dataset is specifically structured for pre-training and fine-tuning AI coding assistants and web-navigating agents. By analyzing production-grade code, models can learn how modern enterprise infrastructure (like Adobe Experience Manager) maps DOM elements to CSS rules… See the full description on the dataset page: https://huggingface.co/datasets/sphita/intel-stylesheet-javascript.fashion-styles
Fashion Styles
A structured taxonomy of fashion style labels for outfit analysis, visual style classification, retrieval, and LLM-based style judging.
The dataset contains 294 canonical style records. Each record pairs a human-readable style name and description with practical recognition metadata: visual indicators, color logic, silhouettes, common contexts, cultural or regional anchors, aesthetic moods, formality, mainstreamness, temporal references, and classifier guidance.… See the full description on the dataset page: https://huggingface.co/datasets/tolgayan/fashion-styles.describe_document_styles_no_predefined_styles_test
Dataset Card
Add more information here
This dataset was produced with DataDreamer 🤖💤. The synthetic dataset card can be found here.
Sinhala-writing-stylesthree_styles_prompted_250_512x512
Dataset Card for "three_styles_prompted_250_512x512"
More Information needed
three_styles_prompted_all_512x512
Dataset Card for "three_styles_prompted_all_512x512"
More Information needed
clusterd_authors_style_desc_no_predefined_styles
Dataset Card
Add more information here
This dataset was produced with DataDreamer 🤖💤. The synthetic dataset card can be found here.
building_styles_datasetgenshin_cosplay_Illustrious_Styles_qwen_image_trans_samplesStyleSet
StyleSet
WARNING: This dataset contains some profane words.
A spoken language benchmark for evaluating speaking-style-related speech generationReleased in our paper, Audio-Aware Large Language Models as Judges for Speaking Styles
This dataset is released by NTU Speech Lab under the MIT license.
Tasks
Voice Style Instruction Following
Reproduce a given sentence verbatim.
Match specified prosodic styles (emotion, volume, pace, emphasis, pitch, non-verbal cues).… See the full description on the dataset page: https://huggingface.co/datasets/dcml0714/StyleSet.three_stylesqwen_image_Illustrious_Styles_lora_neta_art_samplesthree_styles_prompted
Dataset Card for "three_styles_prompted"
More Information needed
three_styles_coded
Dataset Card for "three_styles_coded"
More Information needed
three_styles_10rand
Dataset Card for "three_styles_10rand"
More Information needed
three_styles_prompted_250_512x512_50perclass_proposed
Dataset Card for "three_styles_prompted_250_512x512_50perclass_proposed"
More Information needed
three_styles_prompted_all_512x512_excluded_training
Dataset Card for "three_styles_prompted_all_512x512_excluded_training"
More Information needed
three_styles_prompted_500
Dataset Card for "three_styles_prompted_500"
More Information needed
three_styles_prompted_250_512x512_50perclass_identity
Dataset Card for "three_styles_prompted_250_512x512_50perclass_identity"
More Information needed
three_styles_prompted_250_512x512_50perclass_random
Dataset Card for "three_styles_prompted_250_512x512_50perclass_random"
More Information needed
car_styles_dataset_1
