keyword-extraction
turkish-keyword-extraction-500k
Turkish Keyword Extraction 500K v2
Yirmi alanda konu ve anahtar sözcük çıkarımı için kısa Türkçe belgeler.
Doğrulanmış boyut
Train: 490,000
Validation: 5,000
Test: 5,000
Toplam: 500,000
Ana görev sütunları: id, text, keywords, domain
Provenance
Veri insan mesajlarından, belgelerinden veya web kazımasından alınmamıştır. Tamamı
depodaki üretici koduyla deterministik olarak oluşturulur. Her satırda source_type,
provenance, generator_version… See the full description on the dataset page: https://huggingface.co/datasets/GoktugD/turkish-keyword-extraction-500k.KeywordExtractionbplan_keyword_extraction
Dataset Card for Keyword Extraction
Dataset Description
Homepage: DSSGx Munich organization page.
Repository: GitHub.
Dataset Summary
This folder contains the exact keyword extraction and agent information extraction datasets.
Dataset Structure
Folder structure
exact_search
baunvo_keywords.csv -> appearance of BauNVO keywords in each document.
hochwasser_keywords.csv -> appearance of hochwasser-related keywords in each document.… See the full description on the dataset page: https://huggingface.co/datasets/DSSGxMunich/bplan_keyword_extraction.keyword-extraction-turkce-haber
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
trderlem veri setinin düzenlenmiş hali
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources… See the full description on the dataset page: https://huggingface.co/datasets/nuilbg/keyword-extraction-turkce-haber.keyword-extraction-makalekeyword_extraction
