datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Serbian-RAG-Evalrussian_assistant_to_serbian
Dataset Card for Russian to Serbian Assistant
Dataset Description
Dataset Summary
Russian to Serbian Assistant је едукативни dataset намењен српским говорницима који уче руски језик. Dataset садржи руске фразе са фонетском транскрипцијом прилагођеном српском језику, превод на српски, и детаљним граматичким правилима за правилну изговор.
Supported Tasks
Учење руског језика: Помоћ српским говорницима у савладавању руског језика
Фонетска транскрипција:… See the full description on the dataset page: https://huggingface.co/datasets/mkrstic8/russian_assistant_to_serbian.EQ-Bench-Serbian
EQ-Bench-Serbian 🇷🇸
EQ-Bench is a benchmark for language models designed to assess emotional intelligence. You can read more about it in the paper.
The reason this benchmark was picked is because EQ-Bench in English has very high correlation with LMSYS Arena Elo scores
(has a 0.97 correlation w/ MMLU, and a 0.94 correlation w/ Arena Elo.).
Since it wouldn't be feasible to create an arena for a couple of models available for Serbian, we went in this direction.
This dataset has… See the full description on the dataset page: https://huggingface.co/datasets/Stopwolf/EQ-Bench-Serbian.serbianRoleplay-Serbian
RolePlay-Serbian
Roleplay-Serbian Dataset is a dataset for roleplaying in the Serbian language for Large Language Model.
The base dataset is GPTeacher role play dataset by teknium 1, which can be found under this link, released under MIT License. The dataset is then translated into respective languages. The translation process is powered by Google Translate, using cloud translation API.
For more information and other language datasets for roleplay, it can be found at this github… See the full description on the dataset page: https://huggingface.co/datasets/jojo-ai-mst/Roleplay-Serbian.
