CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01erickfmm /agentlans__multilingual-sentences__paired_10_stsSentences from agentlans/multilingual-sentences in Spanish, and processed with Sentence Similarity Cosine Scores with model hiiamsid/sentence_similarity_spanish_es Each sentence in original dataset was randomly assigned 10 rows (sentences) within a batch of 1000, calculate the sentence similarity, and then deleted duplicate pairs The code for processing can be found here Useful for data distillation, training or benchmarking. Its recommended resampling the dataset to undersample to get a… See the full description on the dataset page: https://huggingface.co/datasets/erickfmm/agentlans__multilingual-sentences__paired_10_sts.tabularsentence-similarity1M<n<10M0 likes75 downloads1y agoHugging Face02DebasishDhal99 /german-polish-paired-placenames Dataset Summary This dataset contains the German and Polish names for almost 10k places in Poland. It has been generated using this code. Many of these names are related to each other. Some German names are literal translation of the Polish names, some are phonetic modifications while some are unrelated. Dataset Creation Source Data German wiki page texttranslation1K<n<10K0 likes71 downloads3y agoHugging Face03glayguo /robot-reel-paired-outcomes Robot Reel: see what a net score hides 30 recorded simulated trials, 20 reference/condition pairs, one task. This is a small evaluation-results dataset, not robot training data, an official LIBERO benchmark, or a general robustness score. The test splits hold the complete pilot results; no training split or train/test partition is claimed. Open the interactive replay to inspect the original camera observations, controls and measured trajectories. Collection and… See the full description on the dataset page: https://huggingface.co/datasets/glayguo/robot-reel-paired-outcomes.tabularn<1K0 likes39 downloads13d agoHugging Face04pbhappliedsystems /quant_eval_paired_degradation_statistics quant_eval — Paired degradation statistics One row per run per task family: the paired pass-rate difference with a 95% confidence interval, the two-sided exact McNemar test, the full discordance breakdown, and a semantic-cluster-adjusted delta and interval. Part of the quant_eval public corpus: a per-case behavioral evaluation of full-weight and quantized large language models across eight agent-relevant task families, with paired statistical testing. Cite this dataset:… See the full description on the dataset page: https://huggingface.co/datasets/pbhappliedsystems/quant_eval_paired_degradation_statistics.tabularn<1K0 likes36 downloads1mo agoHugging Face05DebasishDhal99 /german-czech-paired-placenames Dataset Summary This dataset contains the German and corresponding Czech names for almost 5k places in Czech Republic. It has been generated using this code. Many of these names are related to each other. Some German names are literal translation of the Czech names (or maybe the other way around), some are phonetic modifications while some are unrelated. Dataset Creation Source Data English wiki page containing German exonyms for places in Czech Republic texttranslation1K<n<10K0 likes28 downloads3y agoHugging Face06nayankur /paired-smart-contractstext1K<n<10K2 likes14 downloads2y agoHugging Face07hongerzh /reddit-pairedtext10K<n<100K0 likes7 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.