rafmacalaba/gliner-probe
049
gliner-probe
Fine-tune of urchade/gliner_large-v2.1 for data-use mention extraction (dataset / survey / census / registry mentions in economics research papers).
Labels
NAMED_DATA— a proper name, title, or acronym of a specific data sourceDESCRIPTIVE_DATA— a source described in words but not namedVAGUE_DATA— generic data wording with no identifiable source
Training
- base model:
urchade/gliner_large-v2.1 - dataset:
rafmacalaba/usage-sensitivity-probe(gliner config) - corpus:
all - epochs: 3
- learning rate: 5e-06
- batch size: 16
- precision: bf16
Evaluation (holdout)
Best F0.5: 0.6927 (thr=0.6) Best F1: 0.7120 (thr=0.5)
Evaluation breakdown (holdout)
Per-label (overall)
<details> <summary>Origin breakdown (per-origin metrics)</summary>
</details>
<details> <summary>Per-label details</summary>
</details>
