nandhinivaradharajan14/tamil-english-colloquial-translations
Dataset: English-Tamil (en-ta) Parallel Corpus This dataset contains parallel sentences in English and Tamil (en-ta) that have been curated from multiple sources. It is designed for tasks such as machine translation, language modeling, and other natural language processing (NLP) applications involving English and Tamil. Dataset Composition The dataset is composed of three main parts, which have been concatenated into a single file with two columns: ta… See the full description on the dataset page: https://huggingface.co/datasets/nandhinivaradharajan14/tamil-english-colloquial-translations.
This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.
