Entertainment
pagestorm-research-preview-14b-full-bookpagestorm-research-preview-24b-first-chapter-onlym2d2-Culture_and_the_arts__The_arts_and_Entertainment-ephocs5-embexp-loram2d2-Culture_and_the_arts__The_arts_and_Entertainment-ephocs5-embmean-loram2d2-Culture_and_the_arts__The_arts_and_Entertainment-ephocs10-embmean-freezeqwen_3_4b_sandbagging_entertainment_sandbagging_2_epochm2d2-Culture_and_the_arts__The_arts_and_Entertainment-ephocs10-embexp-freezem2d2-Culture_and_the_arts__The_arts_and_Entertainment-ephocs5-embexp-lora-False
news-entertainment-datasetIndustryCorpus2_film_entertainment
IndustryCorpus2: Film & Entertainment
This repository contains the IndustryCorpus2: Film & Entertainment domain subset of BAAI/IndustryCorpus2.
Refer to the parent dataset card for data construction, intended use, limitations,
and licensing details.
Citation
If you use this dataset in your work, please cite IndustryCorpus2:
@misc{shi2024industrycorpus2,
title = {IndustryCorpus2},
author = {Xiaofeng Shi and Lulu Zhao and Hua Zhou and Donglin Hao},
year… See the full description on the dataset page: https://huggingface.co/datasets/BAAI/IndustryCorpus2_film_entertainment.LongPage
Overview 🚀📚
The first comprehensive dataset for training AI models to write complete novels with sophisticated reasoning.
🧠 Hierarchical Reasoning Architecture — Multi-layered planning traces including character archetypes, story arcs, world rules, and scene breakdowns. A complete cognitive roadmap for long-form narrative construction.
📖 Complete Novel Coverage — From 40,000 to 600,000+ tokens per book, spanning novellas to epic series with consistent quality throughout.
⚡… See the full description on the dataset page: https://huggingface.co/datasets/Pageshift-Entertainment/LongPage.data_audio_gigaspeech2_Entertainmententertainment-reviews-kazakhmasum_oneshot_llama_entertainment
