datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
watermarks-validationlatent_upscale_validationlalm-judge-validation-full-duplex
LALM Judge Validation on Full-Duplex Voice Agents
Companion dataset for the paper A Reliability Assessment of
LALM Audio Judges for Full-Duplex Voice Agents.
This repository contains the anonymised ratings, adversarial-defect
recall tables, JSON schemas, and analysis scripts used to produce
every headline number, table, and figure in that paper.
Summary
209 rated stereo sessions: 152 full-duplex agent-client
conversations across 13 accent-and-condition strata… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/lalm-judge-validation-full-duplex.fisher-validation-resultsAzure_IaC_validationDream_NLP_ValidationBuild_Bench_Validation_DataThis repository contains the validation set of BuildBench paper. It contains 70 data samples.
eval-gliner2-ner-bionlp2004-boundary-smoothing-validationbiomedical-topic-categorization-validationeval-gliner2-ner-fin-boundary-smoothing-validationeval-gliner2-ner-ncbi_disease-boundary-smoothing-validationeval-gliner2-ner-bc5cdr-boundary-smoothing-validationValidation_Setseval-gliner2-ner-ontonotes5-boundary-smoothing-validationeval-gliner2-ner-wnut2017-boundary-smoothing-validationeval-gliner2-ner-mit_restaurant-affine-boundary-smoothing-validationeval-gliner2-ner-mit_restaurant-boundary-smoothing-validationCompanionSim-Validation
CompanionSim-Validation
70 real-world conversations annotated by two groups: 168 annotators in the US (CompanionSim-Validation-US.csv) and 998 annotators from the US, UK, India, and Nigeria (CompanionSim-Validation-Multi.csv).
dreadit-validationeval-gliner2-ner-conll2003-affine-validationmedical-chatbot-validation-datasetusajobs_validation
USAJOBS Dataset (Validation Sample)
Dataset Description
The USAJOBS Dataset is a comprehensive collection of federal job postings from January 2017 through March 2026. This dataset includes full-text job descriptions, and structured metadata (job title and employer).
This particular dataset presents a sample of sentence-level data from the corpus, tagged with task, skill, and AI attributes.
Dataset Structure
The dataset contains 20k sentences… See the full description on the dataset page: https://huggingface.co/datasets/loyoladatamining/usajobs_validation.doctor_validationdataset-train-validationeval-gliner2-ner-mit_restaurant-affine-validationvalidationvalidation set
setfit-proj8-multilabel_2_validationeval-gliner2-ner-ontonotes5-affine-boundary-smoothing-validationeval-gliner2-ner-wnut2017-affine-boundary-smoothing-validationeval-gliner2-ner-ontonotes5-affine-validation
