prompt-generation
gpt4all-j-prompt-generations
Dataset Card for [GPT4All-J Prompt Generations]
Dataset Description
Dataset used to train GPT4All-J and GPT4All-J-LoRA
We release several versions of datasets
v1.0: The original dataset we used to finetune GPT-J on
v1.1-breezy: A filtered dataset where we removed all instances of AI language model
v1.2-jazzy: A filtered dataset where we also removed instances like I'm sorry, I can't answer... and AI language model
v1.3-groovy: The v1.2 dataset with ShareGPT and Dolly… See the full description on the dataset page: https://huggingface.co/datasets/nomic-ai/gpt4all-j-prompt-generations.gpt4all-j-prompt-generations-pt
Dataset Card for "gpt4all-j-prompt-generations-pt"
Dataset Description
Copy translated into Portuguese of the dataset gpt4all_prompt_generations using the googletrans library.
Translate
translate_dataset.ipynb
Usage
dataset_usage.ipynb
gpt4all_prompt_generations
Dataset Card for [GPT4All Prompt Generations]
Dataset Description
Dataset used to train GPT4All
Homepage:
Repository: gpt4all
Paper: Technical Report
Atlas Map: Map of Cleaned Data
class-to-video-prompt-generationamazon_reviews_multi_fr_prompt_title_generation_from_a_review
amazon_reviews_multi_fr_prompt_title_generation_from_a_review
Summary
amazon_reviews_multi_fr_prompt_title_generation_from_a_review is a subset of the Dataset of French Prompts (DFP).It contains 3,989,924 rows that can be used for a text generation task.The original data (without prompts) comes from the dataset amazon_reviews_multi by Keung et al. where only the French split has been kept.A list of prompts (see below) was then applied in order to build the input and… See the full description on the dataset page: https://huggingface.co/datasets/CATIE-AQ/amazon_reviews_multi_fr_prompt_title_generation_from_a_review.amazon_reviews_multi_fr_prompt_binary_text_generation_from_title_of_a_review
amazon_reviews_multi_fr_prompt_binary_text_generation_from_title_of_a_review
Summary
amazon_reviews_multi_fr_prompt_binary_text_generation_from_title_of_a_review is a subset of the Dataset of French Prompts (DFP).It contains 7,560,000 rows that can be used for a text generation task.The original data (without prompts) comes from the dataset amazon_reviews_multi by Keung et al. where only the French split has been kept.A list of prompts (see below) was then applied in… See the full description on the dataset page: https://huggingface.co/datasets/CATIE-AQ/amazon_reviews_multi_fr_prompt_binary_text_generation_from_title_of_a_review.
