zaneb-217/code-switching-codesaviours-si26-zaneb
Code-Switching Dataset – Code Saviours SI-26 Overview This dataset was created as part of the Code Saviours SI-26 ML/AI Internship for Roman Urdu code-switching detection. Description The dataset contains 150 manually created code-switched sentences. Each sentence is tokenized into individual words, and every word is assigned a language label. Data Collection The dataset was created by collecting Roman Urdu-English code-switching… See the full description on the dataset page: https://huggingface.co/datasets/zaneb-217/code-switching-codesaviours-si26-zaneb.
This repository belongs to zaneb-217 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
