CoolFace
9 results

human_translation

morrislab /translation-efficiency-human Multitask Translational Efficiency Prediction Overview Understanding the rules of translational control in mammalian cells is a fundamental challenge in genomics. This dataset is from a study by Zheng et al. (2025), which created a comprehensive, transcriptome-wide atlas of translation efficiency (TE) measurements across a wide array of human and mouse cell types. The dataset was generated by uniformly processing and quality-controlling thousands of ribosome profiling and… See the full description on the dataset page: https://huggingface.co/datasets/morrislab/translation-efficiency-human.text10K<n<100K0 likes66 downloads1y agoHugging Faceicfoss /English_Malayalam_Translation_Human_annotated English-Malayalam Government Parallel Corpus Synth This dataset contains synthetic machine-translated English-Malayalam sentence pairs aligned from government and administrative text. Machine Translation Notice All parallel text in this dataset should be treated as synthetic machine-translated data. It is intended for research, corpus filtering, model adaptation, and experimentation. It should not be treated as human-verified gold translation without additional… See the full description on the dataset page: https://huggingface.co/datasets/icfoss/English_Malayalam_Translation_Human_annotated.texttranslation10K<n<100K0 likes34 downloads11d agoHugging Faceichikara-ai /ichikara-GSM8Ktest-humantranslationgated ichikara-GSM8Ktest-humantranslation 定義書 2025 年 9 月 12 日株式会社いちから 1. データの種類 「算数データ(通称:ichikara-GSM8Ktest-humantranslation)」の定義を行う。本データは OpenAI が公開した英語の小学校算数問題[1]を株式会社 ELYZA が機械翻訳で日本語に翻訳したもの[2]をさらに意訳したものである。本データは英語問題に出る“mile”、“gallon”といった日本語圏では馴染みのない単位や直訳独特の不自然さを除去し、自然かつ教科書問題のような丁寧な日本語で質問および回答が書かれている。GSM8K データのうち、テストデータ 1309 件に対応する日本語の質問回答データを公開する。 2. ファイルとバージョン バージョンは、ファイル名及びデータ ID… See the full description on the dataset page: https://huggingface.co/datasets/ichikara-ai/ichikara-GSM8Ktest-humantranslation.text1K<n<10K0 likes26 downloads1y agoHugging FaceFrancophonIA /Human-reviewed_automatic_English_translations_Europeana [!NOTE] Dataset origin: https://live.european-language-grid.eu/catalogue/corpus/21498 Description The resource includes human-reviewed or post-edited translations of metadata sourced from the Europeana platform. The human-inspected automatic translations are from 17 European languages to English. The translations from Bulgarian, Croatian, Czech, Danish, German, Greek, Spanish, Finnish, Hungarian, Polish, Romanian, Slovak, Slovenian and Swedish have been reviewed by a group of… See the full description on the dataset page: https://huggingface.co/datasets/FrancophonIA/Human-reviewed_automatic_English_translations_Europeana.translation1 likes20 downloads1y agoHugging FaceCalibration-Translation /Calibration-translation-human-eval Translation Evaluation Dataset: Tower vs Calibration This dataset compares translations generated by two models ("Tower-system" and "Calibration") along with human ratings. tabularn<1K0 likes20 downloads1y agoHugging FaceClarusC64 /clinical-preclinical-human-translation-coherence-v0.1What this repo does This dataset tests whether preclinical results translate coherently into early human trials. Many drug programs show strong effects in animals or cell models but fail in humans. The failure is often visible before Phase 2. It appears as a mismatch between the model, the biology, the exposure, and the early human signal. This dataset trains a model to detect that translation risk. You are given the preclinical model used how relevant that model is to human disease whether… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-preclinical-human-translation-coherence-v0.1.texttext-classificationn<1K0 likes19 downloads7mo agoHugging Face