datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Lithuanian-Speech-Dataset
Lithuanian Dataset Metadata
Field
Value
📜 License
CC BY-NC-ND 4.0
🎯 Task Categories
Automatic Speech Recognition
🌍 Language
Lithuanian (lt)
🏷️ Tags
Lithuanian, Audio, Speech, Speech Recognition, ML, Machine, Machine Learning
📦 Size Category
n < 1K
Lithuanian_Context_QA
Lithuanian QA Dataset - Generated with DSPy & Gemma2 27B Q4
Introduction
This dataset was created using DSPy, a Python framework that simplifies the generation of question and answer (QA) pairs from a given context. The dataset is composed of context, questions, and answers, all in Lithuanian. The context was primarily sourced from the following resources:
Lithuanian Wikipedia (lt.wikipedia.org)
Lietuviškoji enciklopedija (vle.lt)
Book: Vitalija Skėruvienė, Civilinė Teisė Mokomoji… See the full description on the dataset page: https://huggingface.co/datasets/ArturG9/Lithuanian_Context_QA.exonyms-for-lithuanian-placeslithuanian-estore-sentiment-binarylithuanian-fake-review-detection
