datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Saraswati-Hindi
Saraswati-Hindi
Saraswati-Hindi is an English-to-Hindi parallel text dataset containing automatically translated English sentences and their corresponding Hindi translations.
The dataset was created using the MyMemory Translation API to translate English text into Hindi (en → hi). It is intended for research, experimentation, and development of English-to-Hindi natural language processing (NLP) and machine translation systems.
Dataset Description… See the full description on the dataset page: https://huggingface.co/datasets/Nebulixlabs/Saraswati-Hindi.know-saraswati-cot
🚨 To all devs, scholars, and also fugazis of AI - A Philosophical Standpoint on AGI:
This is extraneous, if you have time to read it-- give it a shot. We stand at the precipice of a digital era where the notions of artificial intelligence are often muddled with the grandiose idea of Artificial General Intelligence (AGI). Here's a candid reflection:
Current LLMs and their Limitations: Let's be unequivocally clear—present-day language models, including transformers, are not a… See the full description on the dataset page: https://huggingface.co/datasets/knowrohit07/know-saraswati-cot.know-saraswati-alpaca-cot-by-knowrohitknow-saraswati-alpaca-cot-by-knowrohit_stringified-jsonifizesaraswati-nebulasaraswati-tensoicSaraswati_cleaned
