datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-abideen-AlphaMonarch-daser-private
Dataset Card for Evaluation run of abideen/AlphaMonarch-daser
Dataset automatically created during the evaluation run of model abideen/AlphaMonarch-daser
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-abideen-AlphaMonarch-daser-private.french_book_reviews
Dataset Card for French book reviews
I-Dataset Summary
The majority of review datasets are in English. There are datasets in other languages, but not many. Through this work, I would like to enrich the datasets in the French language(my mother tongue with Arabic).The data was retrieved from two French websites: Babelio and Critiques LibresLike Wikipedia, these two French sites are made possible by the contributions of volunteers who use the Internet to share their… See the full description on the dataset page: https://huggingface.co/datasets/Abirate/french_book_reviews.gradio-pi-sessions
Pi Sessions
Redacted Pi coding-agent session traces.
Files use the raw session JSONL layout compatible with julien-c/synthtraces:
sessions/<project>/<timestamp>_<session-id>.jsonl
Each line is one Pi session event (session, model_change, message, tool_result, etc.).
repro-how-much-can-language-models-memorize-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-memory-savings-at-what-cost-a-study-of-alternatives-to-backpropagation-traces
Agent traces
Agent sessions published from a Trackio Logbook.
testing-logbook-v2-traces
Agent traces
Agent sessions published from a Trackio Logbook.
trackio-pi-sessions
Pi Sessions
Redacted Pi coding-agent session traces.
Files use the raw session JSONL layout compatible with julien-c/synthtraces:
sessions/<project>/<timestamp>_<session-id>.jsonl
Each line is one Pi session event (session, model_change, message, tool_result, etc.).
IslamicEval2026-Task1-augmented
IslamicEval2026 Task 1 - Augmented Dataset
An augmented training dataset for Task 1 (Span Detection) of the IslamicEval 2026 Shared Task.
The dataset extends the official competition training data with automatically generated Quran and Hadith examples, as well as manually curated hard negative passages, to improve the robustness of span detection models.
Dataset Statistics
Split
Examples
Train
16,726
Dev
484
Test
620
Training… See the full description on the dataset page: https://huggingface.co/datasets/AbirKorched9/IslamicEval2026-Task1-augmented.repro-olaf-world-spectral-assets
Olaf-World SeqΔ-REPA mechanism reproduction
This self-contained scaled synthetic proxy tests the stated orientation mechanism, not full Olaf-World video training. It generates multiple observation contexts with rotated effect coordinates, compares a local-effect baseline against sequence-level alignment to an invariant temporal teacher, and measures held-out-context action transfer.
python3 repro_olaf_spectral.py --seeds 12 --train 512 --test 2048 --steps 8
The official… See the full description on the dataset page: https://huggingface.co/datasets/abidlabs/repro-olaf-world-spectral-assets.Tokenizedabc-love2abideen__MedPhi-4-14B-v1-details
Dataset Card for Evaluation run of abideen/MedPhi-4-14B-v1
Dataset automatically created during the evaluation run of model abideen/MedPhi-4-14B-v1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abideen__MedPhi-4-14B-v1-details.l1-abilities
