datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
agentlans__multilingual-sentences__paired_10_stsSentences from agentlans/multilingual-sentences in Spanish, and processed with Sentence Similarity Cosine Scores with model hiiamsid/sentence_similarity_spanish_es
Each sentence in original dataset was randomly assigned 10 rows (sentences) within a batch of 1000, calculate the sentence similarity, and then deleted duplicate pairs
The code for processing can be found here
Useful for data distillation, training or benchmarking.
Its recommended resampling the dataset to undersample to get a… See the full description on the dataset page: https://huggingface.co/datasets/erickfmm/agentlans__multilingual-sentences__paired_10_sts.german-polish-paired-placenames
Dataset Summary
This dataset contains the German and Polish names for almost 10k places in Poland. It has been generated using this code.
Many of these names are related to each other. Some German names are literal translation of the Polish names, some are phonetic modifications while some are unrelated.
Dataset Creation
Source Data
German wiki page
robot-reel-paired-outcomes
Robot Reel: see what a net score hides
30 recorded simulated trials, 20 reference/condition pairs, one task.
This is a small evaluation-results dataset, not robot training data, an official
LIBERO benchmark, or a general robustness score. The test splits hold the
complete pilot results; no training split or train/test partition is claimed.
Open the interactive replay
to inspect the original camera observations, controls and measured trajectories.
Collection and… See the full description on the dataset page: https://huggingface.co/datasets/glayguo/robot-reel-paired-outcomes.quant_eval_paired_degradation_statistics
quant_eval — Paired degradation statistics
One row per run per task family: the paired pass-rate difference with a 95% confidence interval, the two-sided exact McNemar test, the full discordance breakdown, and a semantic-cluster-adjusted delta and interval.
Part of the quant_eval public corpus: a per-case behavioral evaluation of full-weight and quantized large language models across eight agent-relevant task families, with paired statistical testing.
Cite this dataset:… See the full description on the dataset page: https://huggingface.co/datasets/pbhappliedsystems/quant_eval_paired_degradation_statistics.german-czech-paired-placenames
Dataset Summary
This dataset contains the German and corresponding Czech names for almost 5k places in Czech Republic. It has been generated using this code.
Many of these names are related to each other. Some German names are literal translation of the Czech names (or maybe the other way around), some are phonetic modifications while some are unrelated.
Dataset Creation
Source Data
English wiki page containing German exonyms for places in Czech Republic
paired-smart-contractsreddit-paired
