datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wikipediaWikipedia dataset containing cleaned articles of all languages.
The datasets are built from the Wikipedia dump
(https://dumps.wikimedia.org/) with one split per language. Each example
contains the content of one full Wikipedia article with cleaning to strip
markdown and unwanted sections (references, etc.).models
Models
GPTs:
• Pigeon-TextGen
• GPT-2
Chats:
• Falcon-180B-chat
• LLaMA-13B-Chat-GGUF
Diffusions:
• SD-1.5
• Dall●E-mini
