datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Low-Carbon-London-Smart-Meter-Cleaned-FeatureReadyA_Synchronized_Lower_Limb_AMG_sEMG_and_Mocap
SAME-Limb
Synchronized AMG and EMG Dataset of Lower-limb Muscle Activities in Everyday Training
This publicly released dataset contains time-aligned acceleromyography (AMG),
surface electromyography (EMG), optical motion capture (MoCap), and four knee/ankle
joint-angle signals from 30 subjects. The repository also provides the frozen
5–100-Hz benchmark code, 64 fitted model artifacts, fixed reference
predictions, source tables, and integrity manifests used for… See the full description on the dataset page: https://huggingface.co/datasets/Tdongxu/A_Synchronized_Lower_Limb_AMG_sEMG_and_Mocap.Low-Frequency-Trap
The Low-Frequency Trap Benchmark Dataset
Official dataset repository for "The Low-Frequency Trap: Video–Language Models Fail at Simple Event Bookkeeping".
📌 Dataset Overview
The Low-Frequency Trap Benchmark evaluates Video–Language Models (VLMs) on fine-grained visual event bookkeeping across controlled parametric variations of Event Load (N) and Event Frequency (F).
Rather than evaluating models solely on final aggregate integer counts, this benchmark pairs… See the full description on the dataset page: https://huggingface.co/datasets/Sarvesh-369/Low-Frequency-Trap.LOW_24sam3-low-dice-2d-nnunet
SAM3 low-Dice 2D datasets for nnU-Net
Private research export of two small 2D datasets on which the balanced-finish
SAM3 LoRA validation Dice was below 0.5. The purpose is to test whether a
dataset-specific nnU-Net can fit these data and to distinguish data/training
limitations from inference bugs.
Dataset
SAM3 Dice
SAM3 IoU
Evaluated validation images
Actual SAM3 training images
DRIVE
0.212233
0.118717
2
14
RAVIR
0.224709
0.128455
2
16
The two-image validation… See the full description on the dataset page: https://huggingface.co/datasets/MedicalSAM3/sam3-low-dice-2d-nnunet.lower-site-598666
lower-site-598666
Synthetic products test data: 46 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/Kestrel-Node/lower-site-598666.FLIP_GB1_low-vs-highlow-television-8444d5
low-television-8444d5
Synthetic weather test data: 58 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/satomigoto6/low-television-8444d5.lower-baby-f87969
lower-baby-f87969
Synthetic sensors test data: 59 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/yhe46/lower-baby-f87969.lower-film-fdb590
lower-film-fdb590
Synthetic products test data: 56 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/Prism-Eli/lower-film-fdb590.lm6-movies-reviews-aspects
IMDB Reviews (Aspect Based Formatted)
Overview
IMDB Reviews (Aspect Based Formatted) is a specialized dataset designed for text classification tasks that involve identifying and categorizing specific aspects of movie reviews. The dataset focuses on extracting and labeling various elements of filmmaking, such as cinematography, story, characters, direction, and unique concepts, from user reviews on IMDB.
The dataset can be used to develop and train models for aspect-based… See the full description on the dataset page: https://huggingface.co/datasets/Lowerated/lm6-movies-reviews-aspects.imdb-reviews-rated
IMDB Reviews with Aspect Based Sentiment Scores
Dataset Description
Name: imdb-reviews-with-aspect-based-sentiment-scoresOrganization: LOWERATEDLicense: Apache License 2.0Language: EnglishTask Categories: Text ClassificationTags: Movies, Ratings, IMDB, Rotten Tomatoes, AISize of Categories: 10,000 to 100,000 reviews
Overview
The imdb-reviews-rated dataset provides a comprehensive collection of IMDB movie reviews rated on seven different aspects:
Direction… See the full description on the dataset page: https://huggingface.co/datasets/Lowerated/imdb-reviews-rated.system-promptsteachers-with-the-minimum-required-qualifications-lower-seco-for-african-countries
Teachers With the Minimum Required Qualifications Lower Seco for African Countries | Africa (World Health Organization)
Size category: n<1K - Formats: csv - Sector: health - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This Dataset Covers… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/teachers-with-the-minimum-required-qualifications-lower-seco-for-african-countries.lowercase-emoji-fewshot-v5completion-rate-lower-secondary-education-for-african-countries
Completion Rate Lower Secondary Education for African Countries | Africa (World Health Organization)
Size category: n<1K - Formats: csv - Sector: health - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This Dataset Covers
Health datasets… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/completion-rate-lower-secondary-education-for-african-countries.africa-income-share-held-by-lowest-20-percentage
Africa Income Share Held by Lowest 20 Percentage | Africa (World Bank)
Size category: n<1K - Formats: csv - Sector: economics_finance - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This Dataset Covers
Public datasets help analysts… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-income-share-held-by-lowest-20-percentage.dpo_data_lowzephyrscoreFLIP_AAV_low-vs-highlowercase-emoji-zeroshotlow-light-dataLowerlowercase-emoji-rawLower_48_States_NG_Underground_Storage_Bcfcompletion-rate-lower-secondary-education-for-african-countries
license: apache-2.0
tags:
- africa
- sustainable-development-goals
- world-health-organization
- development
Completion rate (%) - Lower secondary education
Dataset Description
This dataset provides country-level data for the indicator "4.1.2 Completion rate (%) - Lower secondary education" across African nations, sourced from the World Health Organization's (WHO) data portal on Sustainable Development Goals (SDGs).
The data is presented in a wide format… See the full description on the dataset page: https://huggingface.co/datasets/Nkojiman/completion-rate-lower-secondary-education-for-african-countries.logical-textslowercase-emoji-oneshotQA_Low_Resource_FatimaFellowship
Model Experimentation and Analysis
1. Model Experimentations
I chose the task of Q&A, branching into two categories: general and specific. I tested the output, i.e., the LLM's response against the expected output. I have created the dataset in a CSV file. The model used is : https://huggingface.co/Andron00e/YetAnother_Open-Llama-3B-LoRA-OpenOrca
2. "Blind spots"
In this dataset, model is not able to predict well on very-specific information, like dates or years… See the full description on the dataset page: https://huggingface.co/datasets/KushieBoi/QA_Low_Resource_FatimaFellowship.general-meaning-texts
