datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dit-loras-interpreting
andyx10/dit-loras-interpreting
Experimenting with interpreting write vectors over 100 hidden-topic model organisms fromdiff-interpretation-tuning/loras
implementation
We use 'self_attn.o_proj andmlp.down_proj` for write vectors: two per block across 36 blocks, with a total of 72 write vectors per organism.
Jacobian Lens from Neuronpedia
neuronpedia/jacobian-lens
(qwen3-4b/jlens/Salesforce-wikitext/Qwen3-4B_jacobian_lens.pt)
layout
test100/… See the full description on the dataset page: https://huggingface.co/datasets/andyx10/dit-loras-interpreting.w2t-llm-arc-easy-lora
W2T Llm Arc Easy Lora
This repository contains artifacts for the W2T paper:
Paper: W2T: LoRA Weights Already Know What They Can Do
Repo: Weight2Token
Summary
ARC-Easy LoRA checkpoints and prepared metadata used for performance prediction.
Source Status
Storage location: local
Verification status: confirmed
Files
See manifest.json for the exact local or remote source paths used to prepare this release.
Citation… See the full description on the dataset page: https://huggingface.co/datasets/Xiaolong-Han/w2t-llm-arc-easy-lora.idt5-v4-results-final-lora-s123-20260912T013040606815Z
final-lora-s123-20260912T013040606815Z
Run artifacts and per-item predictions.
Phase: final. These are newly generated results, not a reproduction of the legacy TCI tables.
See run_manifest.json, rules.json, generation_protocol.json and checkpoint_hashes.json. Structural scores do not establish semantic or Bloom validity.
Metrics
{
"n": 267,
"rule_version": "structural-proxy-v0.4-grounding-separated",
"parse_success_pct": 94.7565543071161,
"bleu":… See the full description on the dataset page: https://huggingface.co/datasets/Firmansyah-Ibrahim/idt5-v4-results-final-lora-s123-20260912T013040606815Z.idt5-v4-results-final-lora-s2026-20260912T034640190015Z
final-lora-s2026-20260912T034640190015Z
Run artifacts and per-item predictions.
Phase: final. These are newly generated results, not a reproduction of the legacy TCI tables.
See run_manifest.json, rules.json, generation_protocol.json and checkpoint_hashes.json. Structural scores do not establish semantic or Bloom validity.
Metrics
{
"n": 267,
"rule_version": "structural-proxy-v0.4-grounding-separated",
"parse_success_pct": 95.88014981273409,
"bleu":… See the full description on the dataset page: https://huggingface.co/datasets/Firmansyah-Ibrahim/idt5-v4-results-final-lora-s2026-20260912T034640190015Z.x-lora-datasetPaper, see: arxiv.org/abs/2402.07148
idt5-v4-results-final-lora-s42-20260912T063343815032Z
final-lora-s42-20260912T063343815032Z
Run artifacts and per-item predictions.
Phase: final. These are newly generated results, not a reproduction of the legacy TCI tables.
See run_manifest.json, rules.json, generation_protocol.json and checkpoint_hashes.json. Structural scores do not establish semantic or Bloom validity.
Metrics
{
"n": 267,
"rule_version": "structural-proxy-v0.4-grounding-separated",
"parse_success_pct": 92.88389513108615,
"bleu":… See the full description on the dataset page: https://huggingface.co/datasets/Firmansyah-Ibrahim/idt5-v4-results-final-lora-s42-20260912T063343815032Z.Classical-Mechanics-Equations-Dataset_SFT-or-LoRA
Classical Mechanics Equations Dataset (SFT / LoRA Ready)
A structured dataset of 64 classical mechanics equations from Newtonian,
Lagrangian, and Hamiltonian mechanics, expanded into 448 instruction-tuning
rows across three task types: equation explanation, Q&A, and derivation.
Designed for fine-tuning LLMs on physics reasoning, STEM Q&A, and
equation understanding tasks.
Overview
Property
Value
Domain
Classical Mechanics (Physics)
Total rows
448
Train… See the full description on the dataset page: https://huggingface.co/datasets/ChaoticEconomist/Classical-Mechanics-Equations-Dataset_SFT-or-LoRA.Jazz-Blues-Music-Dataset_SFT-or-LoRA
Jazz & Blues Music Dataset (SFT / LoRA Ready)
A structured dataset covering 82 iconic Jazz and Blues songs, 21 artist
profiles, and 41 historical events, expanded into 1,219
instruction-tuning rows across 7 task types.
Designed for fine-tuning LLMs on music knowledge, cultural history, artist
biography, and domain-specific Q&A tasks.
Overview
Property
Value
Domain
Jazz & Blues Music
Total rows
1,219
Train split
1,036 (85%)
Validation split
91 (~7.5%)… See the full description on the dataset page: https://huggingface.co/datasets/ChaoticEconomist/Jazz-Blues-Music-Dataset_SFT-or-LoRA.benchmark-finetune-lora-v1
Odyn benchmark: LoRA fine-tuning peak VRAM (V1)
Curated benchmark rows for validating GPU memory estimators during LoRA fine-tuning. Each row pairs a published or measured expected peak VRAM with inputs to a math engine (model size, context length, batch, LoRA rank, precision, parallelism) plus optional VRAM breakdown and provenance.
This dataset is not Alpaca-style training JSONL. It is evaluation ground truth for placement / scheduler memory models (Odyn Smart Digester math… See the full description on the dataset page: https://huggingface.co/datasets/odyn-network/benchmark-finetune-lora-v1.lora_mathlora-hyperparameter-benchmark-v1
Odyn benchmark: LoRA fine-tuning hyperparameter configs (V1)
Curated benchmark of real, cited LoRA and QLoRA fine-tuning configurations for validating a hyperparameter advisor. Each row is a published or measured supervised (SFT) LoRA config with its hyperparameters (learning rate, LoRA rank/alpha/dropout, epochs, batch, sequence length, gradient checkpointing), the dataset it trained on, and per-field provenance.
Schema
Column
Type
Description
id… See the full description on the dataset page: https://huggingface.co/datasets/odyn-network/lora-hyperparameter-benchmark-v1.travel-prompt-lora-evaluation-datasetlora_aglora_phlora_selectorlora1test_data2lora_testinglora_philabs_esptest_data4test_data5test_data6lora-bp-safety-benchmarkversion https://git-lfs.github.com/spec/v1
oid sha256:7606d91a20699cd5918eeffa95f318781fccfdd705297ddadbf390ea587860e6
size 3781
lora2lora3fsq-lora-v2lora4LoRaNetworkSensorsReadingsllama2-7B-chat_jailbreak
