datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
molt-benchmark-results
Molt · elastic on-device inference measurements
Everything measured while building Molt, a
runtime that moves a running generation onto a smaller model between two
tokens, carrying the KV cache across, so an on-device LLM under memory pressure
is neither reclaimed by the OS nor restarted from the prompt.
Published so the claims can be checked rather than taken on trust. The figures in
the repo README and the results page are generated from these files; nothing is
transcribed by… See the full description on the dataset page: https://huggingface.co/datasets/NagaYu/molt-benchmark-results.osworld_tasks_filesESGdatasetsESGDatasetWildFire-Ynagri-sound-dataset
Sylheti Language Learning – Audio Interaction Dataset
Overview
This dataset contains letter-level reference pronunciation audio samples designed for a multilingual Augmented Reality (AR) language-learning system.
The system adapts its interface language dynamically based on user preference, while primarily aiming to teach and evaluate Sylheti pronunciation.The dataset is structured to support real-time pronunciation feedback, deterministic AR triggers, and multilingual… See the full description on the dataset page: https://huggingface.co/datasets/shivdi1999/nagri-sound-dataset.Nagpurs-Food-Jointsmlops_gsodTourismTourism_GLTourism_Cleanplay-store-revenue-analysis
