Avifenesh/episodic-ingestion-modernbert-field-event-ranker-mixed-v2-1-h4-320
010
ModernBERT field-event ranker (mixed-mode v2.1, H4-320)
Fine-tune of answerdotai/ModernBERT-base with v2.1 labels — same as v2 but with a sharpened `action_causality` heuristic:
- Old: tool_result is grounded if ≥2 content tokens appear in ANY later assistant turn. Noise-dominated (
file,path,error,exitcodematch trivially). - New: require ≥2 distinctive tokens (filter out 50+ corpus-frequent noise tokens) and the assistant turn must be the NEXT one, within 2 events. Selects a proper subset of tool_results based on content grounding.
Key result: per-field improvements on sharpened fields
Overall MRR stayed flat at 0.583 (aggregate), but top-1 moved from 0.375 → 0.381. action_causality's eval mass shrank (141 → 50 pairs) because the sharper heuristic only emits the field when the signal is real, which is why the big per-field MRR gain doesn't translate 1:1 into aggregate gain.
Training details
- Train rows: 1217
- Eval rows: 220
- Steps: 320 × accum=8 = 2560 forwards
- Last loss: 0.0007
- Peak VRAM: 4.08 GiB
Eval metrics
Per-field MRR (all fields):
