CoolFace
Datasetpublicgated

TatarNLPWorld/tatar-folklore-corpus

Dataset Card for Tatar Text Corpus with Rich Metadata Dataset Details Dataset Description This dataset is a collection of 326 Tatar language texts with extensive metadata, curated by TatarNLPWorld. Each record includes the full text and a structured metadata object containing fields such as title, author, year, source, genre, category, and more. The dataset is designed for NLP research on the Tatar language, including text classification, language… See the full description on the dataset page: https://huggingface.co/datasets/TatarNLPWorld/tatar-folklore-corpus.

sourceHugging Faceotherupdated 25d agoView on Hugging Face
0likes22downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.