datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
LENS-WarBias
LENS-WarBias
Version 1.3 — research draft; independent human validation pending.
LENS-WarBias is a Ukrainian–English prompt dataset for studying war-related stereotype elicitation and transfer after model unlearning. It covers 981 WarBias matrix entries, 129 case families, 15 actor profiles and 56 actor/gender/age variants. It contains prompts and provenance metadata, not target-model responses, a validated forget set, or measured model scores.
The dataset contains deliberately… See the full description on the dataset page: https://huggingface.co/datasets/FairForget/LENS-WarBias.j-lens-verbalization
J-lens Verbalization
What concepts are active inside Qwen3.6-27B while it answers a question — and can
the model tell you?
Ground truth comes from Neuronpedia's Jacobian Lens, which reads out the
word-like concepts active in the model's "global workspace" during generation.
The two files
file
rows
what it contains
collected_answers.jsonl
3,800
questions, answers, and the ground-truth concepts that were active. Use this to train or evaluate.… See the full description on the dataset page: https://huggingface.co/datasets/RaoAditya/j-lens-verbalization.Text-Scrutiny-LLM-Dataset
Citation Information
If you find our work helpful, please use the following citations.
@misc{cai2024ethicallenscurbingmalicioususages,
title={Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models},
author={Yuzhu Cai and Sheng Yin and Yuxi Wei and Chenxin Xu and Weibo Mao and Felix Juefei-Xu and Siheng Chen and Yanfeng Wang},
year={2024},
eprint={2404.12104},
archivePrefix={arXiv},
primaryClass={cs.CV}… See the full description on the dataset page: https://huggingface.co/datasets/Ethical-Lens/Text-Scrutiny-LLM-Dataset.repro-revisiting-zeroth-order-hessian-approximation-policy-lens-traces
Agent traces
Agent sessions published from a Trackio Logbook.
review-lens-evals
Review Lens Evals
A small, hand-labeled evaluation set for measuring the precision and recall of
LLM code-review systems on unified diffs.
The dataset accompanies
Ashishkosana/review-lens, a
multi-lens reviewer that examines correctness, security, performance, and test
coverage before running a separate adversarial verification pass.
Why this dataset exists
Code-review evaluations need both positive and negative cases. The four seeded
bug diffs test whether a… See the full description on the dataset page: https://huggingface.co/datasets/ashishkosana/review-lens-evals.camera-lens-body-adapter-compatibility
Camera mount flange focal distance and adapter feasibility
Canonical, always-current version: https://referencesource.org/camera-lens-body-adapter-compatibility/
Machine-readable: https://referencesource.org/camera-lens-body-adapter-compatibility/data.json — this mirror is a point-in-time copy.
Last verified: 2026-08-13
Stale after: 2028-08-12 (past this date, prefer the canonical copy —
it re-verifies on a cadence this snapshot does not)
Records: 201
Flange focal distance… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/camera-lens-body-adapter-compatibility.LensBenchHumanBias
Citation Information
If you find our work helpful, please use the following citations.
@misc{cai2024ethicallenscurbingmalicioususages,
title={Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models},
author={Yuzhu Cai and Sheng Yin and Yuxi Wei and Chenxin Xu and Weibo Mao and Felix Juefei-Xu and Siheng Chen and Yanfeng Wang},
year={2024},
eprint={2404.12104},
archivePrefix={arXiv},
primaryClass={cs.CV}… See the full description on the dataset page: https://huggingface.co/datasets/Ethical-Lens/HumanBias.Tox1K
Citation Information
If you find our work helpful, please use the following citations.
@misc{cai2024ethicallenscurbingmalicioususages,
title={Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models},
author={Yuzhu Cai and Sheng Yin and Yuxi Wei and Chenxin Xu and Weibo Mao and Felix Juefei-Xu and Siheng Chen and Yanfeng Wang},
year={2024},
eprint={2404.12104},
archivePrefix={arXiv},
primaryClass={cs.CV}… See the full description on the dataset page: https://huggingface.co/datasets/Ethical-Lens/Tox1K.Tox100
Citation Information
If you find our work helpful, please use the following citations.
@misc{cai2024ethicallenscurbingmalicioususages,
title={Ethical-Lens: Curbing Malicious Usages of Open-Source Text-to-Image Models},
author={Yuzhu Cai and Sheng Yin and Yuxi Wei and Chenxin Xu and Weibo Mao and Felix Juefei-Xu and Siheng Chen and Yanfeng Wang},
year={2024},
eprint={2404.12104},
archivePrefix={arXiv},
primaryClass={cs.CV}… See the full description on the dataset page: https://huggingface.co/datasets/Ethical-Lens/Tox100.poverty_lens_chatbot_dataset
