CoolFace
Datasetpublic

qandeelasim13/code-switching-codesaviours-si26-qandeel

Code Switching NLP Dataset | Code Saviours SI-26 Dataset Description A word-level labelled Roman Urdu–English code-switching dataset (200 sentences, 3182 word entries). Labels: URD (Roman Urdu), ENG (English), MIX (nativized loanword). Source Sentences filtered from the Roman Urdu Data Set (Sharf, 2017, UCI ML Repository, CC BY 4.0), originally collected from e-commerce reviews, Facebook comments, and Twitter posts. Filtered for genuine… See the full description on the dataset page: https://huggingface.co/datasets/qandeelasim13/code-switching-codesaviours-si26-qandeel.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes8downloads
4 commits on main
90ed5942mo ago

Upload dataset.csv with huggingface_hub

qandeelasim13
ff232e52mo ago

Update README.md

qandeelasim13
46c8b0d2mo ago

Upload dataset.csv

qandeelasim13
08ba7dd2mo ago

initial commit

qandeelasim13