datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Sci-Fi-ZH一份 VeejaLiu 正在手工清洗的数据:https://github.com/VeejaLiu/ScienceFictionCollection
Sci-Fi-Books-gutenberg
Gutenberg Sci-Fi Book Dataset
This dataset contains information about science fiction books. It’s designed for training AI models, research, or any other purpose related to natural language processing.
Data Format
The dataset is provided in CSV format. Each record represents a book and includes the following fields:
ID: A unique identifier for the book.
Title: The title of the book.
Author: The author(s) of the book.
Text: The text content of the book (e.g., summary… See the full description on the dataset page: https://huggingface.co/datasets/stevez80/Sci-Fi-Books-gutenberg.Stargate-SciFi-SFT-Instruct
Stargate-SciFi-SFT-Instruct
Dataset Summary
The Stargate-SciFi-SFT-Instruct is a compact English-language supervised fine-tuning dataset focused on the Stargate television franchise, including Stargate SG-1, Stargate Atlantis, and Stargate Universe.
The dataset is formatted for instruction tuning and contains prompt-response examples covering episode summaries, production metadata, character profiles, lore explanations, natural fan-style questions, comparative… See the full description on the dataset page: https://huggingface.co/datasets/Mungus451/Stargate-SciFi-SFT-Instruct.SciFi-Fantasy-alpaca
SCIFI-FANTASY DATA SET
This synthetic data set was created with the following end user in mind: people who are interested in science fiction and fantasy genres
This was examined from 11 different perspectives. The data consists of 21,330 questions and answer sets. The tone and approach was set using the following prompt:
Your goal is to provide useful assistance for the user. The detail level of your responses should match the complexity of the user's request. You want to inspire… See the full description on the dataset page: https://huggingface.co/datasets/theprint/SciFi-Fantasy-alpaca.
