sayurio/bangla-wikipedia
Bangla (Bengali) Wikipedia Articles Dataset Request More ScrapesOrder Private Scrapes Current Progress: Approx 20% Dataset Summary This dataset contains a comprehensive extraction of articles from the Bangla (Bengali) Wikipedia. It is designed for Natural Language Processing (NLP) tasks, linguistic research, and training Large Language Models (LLMs) to better understand and generate the Bengali language. Copyright and Fair Use I do… See the full description on the dataset page: https://huggingface.co/datasets/sayurio/bangla-wikipedia.
Upload jsonl/data_006.jsonl with huggingface_hub
Upload jsonl/data_005.jsonl with huggingface_hub
Rename data_002.jsonl to jsonl/data_002.jsonl
Rename data_001.jsonl to jsonl/data_001.jsonl
Upload jsonl/data_004.jsonl with huggingface_hub
Upload jsonl/data_003.jsonl with huggingface_hub
Update README.md
Upload data_002.jsonl with huggingface_hub
Update README.md
Update README.md
Add files using upload-large-folder tool
Update README.md
initial commit
