CoolFace
Datasetpublic

amurienne/Goulenn-Alpaca-Instruct-50k

Goulenn A Breton Instructions Dataset called Goulenn (meaning "Question" in Breton). Direct translation of the jpacifico/French-Alpaca-dataset-Instruct-110K by Jonathan Pacifico. For now only 50k samples have been translated, 110k version soo to come... Generation details available on the GweLLM Github repository. Sample test code: from datasets import load_dataset dataset = load_dataset( path="amurienne/Goulenn-Alpaca-Instruct-50k", split="train")… See the full description on the dataset page: https://huggingface.co/datasets/amurienne/Goulenn-Alpaca-Instruct-50k.

sourceHugging Facemitupdated 2y agoView on Hugging Face
0likes26downloads
filegoulenn_instructions_50k.parquet9.1 MBdownload

amurienne/Goulenn-Alpaca-Instruct-50k · main · files are served by the source, never re-hosted here