datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
details_Dans-DiscountModels__Mistral-7b-FFT-Test3
Dataset Card for Evaluation run of Dans-DiscountModels/Mistral-7b-FFT-Test3
Dataset automatically created during the evaluation run of model Dans-DiscountModels/Mistral-7b-FFT-Test3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Dans-DiscountModels__Mistral-7b-FFT-Test3.marvis_fft_resultsshotpath-subject-replaced-v2-fft-20260711
ShotPath subject-replaced v2 shuffled FFT data
This folder contains the shuffled JSONL and subject supplement images for rerunning Qwen2.5VL/Qwen3VL full fine-tuning.
{
"seed": 2026071119,
"rows": 3003,
"jsonl": "data/shotpath_stage1_stage2_mixed_subject_replaced_v2_shuffle_fft.jsonl",
"subject_supplement_images": 361,
"added_clean_subject_rows": 361,
"item_dimension_counts": {
"camera_position": 767,
"subject": 649,
"framing": 1097,
"light_exposure":… See the full description on the dataset page: https://huggingface.co/datasets/purefall/shotpath-subject-replaced-v2-fft-20260711.details_Dans-DiscountModels__ShearedLlama-1.3b-FFT-Test1
Dataset Card for Evaluation run of Dans-DiscountModels/ShearedLlama-1.3b-FFT-Test1
Dataset automatically created during the evaluation run of model Dans-DiscountModels/ShearedLlama-1.3b-FFT-Test1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Dans-DiscountModels__ShearedLlama-1.3b-FFT-Test1.idt5-v4-results-final-fft-s2026-20260911T063507183162Z
final-fft-s2026-20260911T063507183162Z
Run artifacts and per-item predictions.
Phase: final. These are newly generated results, not a reproduction of the legacy TCI tables.
See run_manifest.json, rules.json, generation_protocol.json and checkpoint_hashes.json. Structural scores do not establish semantic or Bloom validity.
Metrics
{
"n": 267,
"rule_version": "structural-proxy-v0.4-grounding-separated",
"parse_success_pct": 88.01498127340824,
"bleu":… See the full description on the dataset page: https://huggingface.co/datasets/Firmansyah-Ibrahim/idt5-v4-results-final-fft-s2026-20260911T063507183162Z.eval_diffusion_fft_real_1_stack_bowls_filtered_fixed
eval_diffusion_fft_real_1_stack_bowls_filtered
Task: "Pick up the lego block."
Type: evaluation (filtered)
Robot: Franka FR3
Cameras: observation.images.primary, observation.images.wrist (video, 256x256) @ 15 FPS
Statistics
Metric
Value
Episodes
10
Total frames
3094
Avg frames/episode
309
FPS
15
Format
LeRobot v3.0
Features
Feature
Type
Shape
observation.images.primary
video
[256, 256, 3]
observation.images.wrist
video
[256… See the full description on the dataset page: https://huggingface.co/datasets/continuallearning/eval_diffusion_fft_real_1_stack_bowls_filtered_fixed.details_Dans-DiscountModels__TinyLlama-1.1B-FFT-Test2
Dataset Card for Evaluation run of Dans-DiscountModels/TinyLlama-1.1B-FFT-Test2
Dataset automatically created during the evaluation run of model Dans-DiscountModels/TinyLlama-1.1B-FFT-Test2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Dans-DiscountModels__TinyLlama-1.1B-FFT-Test2.eval_diffusion_fft_real_0_put_bowl_filtered_fixed
eval_diffusion_fft_real_0_put_bowl_filtered
Task: "Pick up the lego block."
Type: evaluation (filtered)
Robot: Franka FR3
Cameras: observation.images.primary, observation.images.wrist (video, 256x256) @ 15 FPS
Statistics
Metric
Value
Episodes
10
Total frames
4032
Avg frames/episode
403
FPS
15
Format
LeRobot v3.0
Features
Feature
Type
Shape
observation.images.primary
video
[256, 256, 3]
observation.images.wrist
video
[256… See the full description on the dataset page: https://huggingface.co/datasets/continuallearning/eval_diffusion_fft_real_0_put_bowl_filtered_fixed.idt5-v4-results-final-fft-s42-20260910T135823740810Z
final-fft-s42-20260910T135823740810Z
Run artifacts and per-item predictions.
Phase: final. These are newly generated results, not a reproduction of the legacy TCI tables.
See run_manifest.json, rules.json, generation_protocol.json and checkpoint_hashes.json. Structural scores do not establish semantic or Bloom validity.
Metrics
{
"n": 267,
"rule_version": "structural-proxy-v0.4-grounding-separated",
"parse_success_pct": 93.63295880149813,
"bleu":… See the full description on the dataset page: https://huggingface.co/datasets/Firmansyah-Ibrahim/idt5-v4-results-final-fft-s42-20260910T135823740810Z.fe_fft_instruction_len256dsp-fft-sampling-aliasing
Synthetic DSP Dataset: FFT + Sampling / Aliasing
This repository contains synthetic instruction-style DSP samples
designed for numerical reasoning and conceptual understanding of
Digital Signal Processing (DSP) fundamentals.
The dataset focuses on:
FFT bin reasoning and frequency-domain interpretation
Sampling theory
Aliasing effects
Dataset Origin & Verification
This dataset was generated as part of the project:
Fine-Tuning Lightweight Large Language Models for a… See the full description on the dataset page: https://huggingface.co/datasets/Irfanuruchi/dsp-fft-sampling-aliasing.LinearSpectre_PS4_Cifar10_only_fftidt5-v4-results-final-fft-s123-20260910T215641650583Z
final-fft-s123-20260910T215641650583Z
Run artifacts and per-item predictions.
Phase: final. These are newly generated results, not a reproduction of the legacy TCI tables.
See run_manifest.json, rules.json, generation_protocol.json and checkpoint_hashes.json. Structural scores do not establish semantic or Bloom validity.
Metrics
{
"n": 267,
"rule_version": "structural-proxy-v0.4-grounding-separated",
"parse_success_pct": 93.25842696629213,
"bleu":… See the full description on the dataset page: https://huggingface.co/datasets/Firmansyah-Ibrahim/idt5-v4-results-final-fft-s123-20260910T215641650583Z.SID_FFT_Detector_FFTMdist_Chem_FFT_ft_dataMdist_Chem_FFT_Qwen3_14B_full_ft_datamultiplication_1000_train_2x3_cot_fft_multiplication_backtracking_verification_sfttest_cpt_fft_300_0.8FFT2SD-Datasets
From Free-Text to Structured Data - Datasets
This repository contains datasets used in the FFT2SD (From Free-Text to Structured Data) thesis project.
These datasets support the task of converting medical free-text into structured outputs using transformer-based language models.
Dataset Details
File
Description
dataset-unlabeled.jsonl
Raw, unlabeled dataset from the colorectal screening program.
dataset-eval.jsonl
Manually annotated evaluation set, used to… See the full description on the dataset page: https://huggingface.co/datasets/Veggissss/FFT2SD-Datasets.FFT-naive-50k-minif2fUSS-reward-model-qwen-FFTmultiplication_1000_train_2x3_cot_fft_multiplication_all_strategies_sftminif2f_test_FFT-healing-kl-oursFFT-exponentinit-50k-aime2025FFT-exponentinit-FFT-50k-minif2fmultiplication_1000_train_2x3_cot_fft_multiplication_backtracking_subgoal_sfttest_cpt_fft_300_0.6multiplication_1000_train_2x3_cot_fft_multiplication_basic_sfttest_cpt_fft_1000_0.5LinearSpectre_PS1_Cifar100_only_fft
