CoolFace
Datasetpublic

sayurio/bangla-wikipedia

Bangla (Bengali) Wikipedia Articles Dataset Request More ScrapesOrder Private Scrapes Current Progress: Approx 20% Dataset Summary This dataset contains a comprehensive extraction of articles from the Bangla (Bengali) Wikipedia. It is designed for Natural Language Processing (NLP) tasks, linguistic research, and training Large Language Models (LLMs) to better understand and generate the Bengali language. Copyright and Fair Use I do… See the full description on the dataset page: https://huggingface.co/datasets/sayurio/bangla-wikipedia.

sourceHugging Facecc-by-sa-4.0updated 6mo agoView on Hugging Face
2likes23downloads
13 commits on main
a80fd976mo ago

Upload jsonl/data_006.jsonl with huggingface_hub

sayurio
b86f8e66mo ago

Upload jsonl/data_005.jsonl with huggingface_hub

sayurio
d44564e6mo ago

Rename data_002.jsonl to jsonl/data_002.jsonl

sayurio
d296c5f6mo ago

Rename data_001.jsonl to jsonl/data_001.jsonl

sayurio
0b540906mo ago

Upload jsonl/data_004.jsonl with huggingface_hub

sayurio
23f6bb06mo ago

Upload jsonl/data_003.jsonl with huggingface_hub

sayurio
07287366mo ago

Update README.md

sayurio
40fcd556mo ago

Upload data_002.jsonl with huggingface_hub

sayurio
fd606396mo ago

Update README.md

sayurio
1da1b976mo ago

Update README.md

sayurio
99cab566mo ago

Add files using upload-large-folder tool

sayurio
2c0a4bc6mo ago

Update README.md

sayurio
46b36c06mo ago

initial commit

sayurio