CoolFace
Datasetpublic

agentlans/wikipedia-paragraphs

Wikipedia Paragraph Samples Dataset Description This dataset contains paragraphs extracted from randomly selected English Wikipedia articles. It provides a diverse sample of Wikipedia content across various topics. Dataset Details Name: Wikipedia Paragraph Samples Version: 1.0 Date Created: 2024-08-20 Language: English Format: JSONLines Contents Each line in the dataset represents a single paragraph and contains two fields: Title… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/wikipedia-paragraphs.

sourceHugging Facecc-by-sa-3.0updated 2y agoView on Hugging Face
3likes436downloads
filetrain.jsonl.gz10.5 MBdownload

agentlans/wikipedia-paragraphs · main · files are served by the source, never re-hosted here