Jeesup/glue-lora-bitwidth-results
GLUE LoRA x Backbone Bit-Width — combined results Aggregated metrics, LoRA geometry analysis and figures for a controlled study of whether backbone bit-width (bf16 / int8 / nf4) changes what a LoRA adapter learns. Grid: 2 model sizes (1B, 3B) x 4 GLUE tasks (MNLI, QQP, SST-2, RTE) x 3 bit-widths x 3 seeds = 72 runs (0 present here). For each (model size, seed) the adapter initialisation is identical across bit-widths, so cross-bit differences are attributable to the backbone.… See the full description on the dataset page: https://huggingface.co/datasets/Jeesup/glue-lora-bitwidth-results.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face