datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
coreference-challenge
PI-LLM Bench: The Core Retrieval Challenge Behind MRCR
Update: Accepted to COLM 2026 (San Francisco).
AAAI 2026 Worshop Oral: LaMAS (LLM-based Multi-Agent Systems: Towards Responsible, Reliable, and Scalable Agentic Systems) Jan/2026 Singapole
ICML 2025 Long-Context Foundation Models Workshop Accepted.
A simple context interference evaluation.
Update: This dataset is integrated into Moonshot AI(Kimi)'s internal benchmarking framework for assessing ** tracking capacity and… See the full description on the dataset page: https://huggingface.co/datasets/giantfish-fly/coreference-challenge.pronominal_coreference_resolutionrlvr_task1390_wscfixed_coreferenceqwen3_0.6b-rlvr_task1390_wscfixed_coreference
