CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AlgorithmicResearchGroup /arxiv_deep_learning_python_research_code ArXiv Deep Learning Python Research Code A curated corpus of Python source code files extracted from GitHub repositories referenced in ArXiv papers. Contains 391,496 files (1.49 GB) filtered to deep learning frameworks, designed for training and evaluating Code LLMs on research-grade code. Dataset Summary Statistic Value Total files 391,496 Total size 1.49 GB Source repos 34,099 Time span ArXiv inception through July 2023 Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/AlgorithmicResearchGroup/arxiv_deep_learning_python_research_code.tabulartext-generation100K<n<1M11 likes243 downloads5mo agoHugging Face02nlp-with-deeplearning /Ko.SlimOrca원본 데이터셋: Open-Orca/SlimOrca texttext-classification100K<n<1M3 likes48 downloads3y agoHugging Face03nlp-with-deeplearning /Ko.WizardLM_evol_instruct_V2_196k이 데이터셋은 자체 구축한 번역기로 WizardLM/WizardLM_evol_instruct_V2_196k을 번역한 데이터셋입니다. 아래 README 페이지도 번역기를 통해 번역되었습니다. 참고 부탁드립니다. News 🔥 🔥 🔥 [08/11/2023] WizardMath 모델을 출시합니다. 🔥 WizardMath-70B-V1.0 모델은 ChatGPT 3.5, Claude Instant 1 및 PaLM 2 540B 를 포함 하 여 GSM8K에서 일부 폐쇄 소스 LLMs 보다 약간 더 우수 합니다. 🔥 우리의 WizardMath-70B-V1.0 모델은 SOTA 오픈 소스 LLM보다 24.8 포인트 높은 GSM8k Benchmarks에서 81.6 pass@1 을 달성합니다. 🔥 우리의 WizardMath-70B-V1.0 모델은 SOTA 오픈 소스 LLM보다 9.2 포인트 높은 MATH 벤치마크에서 22.7 pass@1 을 달성합니다.… See the full description on the dataset page: https://huggingface.co/datasets/nlp-with-deeplearning/Ko.WizardLM_evol_instruct_V2_196k.texttext-generation100K<n<1M4 likes41 downloads3y agoHugging Face04nlp-with-deeplearning /ko.SHP 🚢 Korean Stanford Human Preferences Dataset (Ko.SHP) 이 데이터셋은 자체 구축한 번역기를 활용하여 stanfordnlp/SHP 데이터셋을 번역한 것입니다. 아래의 내용은 해당 번역기로 README 파일을 번역한 것입니다. 참고 부탁드립니다. If you mention this dataset in a paper, please cite the paper: Understanding Dataset Difficulty with V-Usable Information (ICML 2022). Summary SHP는 요리에서 법률 조언에 이르기까지 18가지 다른 주제 영역의 질문/지침에 대한 응답에 대한 385K 집단 인간 선호도 데이터 세트이다. 기본 설정은 다른 응답에 대 한 한 응답의 유용성을 반영 하기 위한 것이며 RLHF 보상 모델 및 NLG 평가 모델 (예: SteamSHP)을 훈련 하는 데… See the full description on the dataset page: https://huggingface.co/datasets/nlp-with-deeplearning/ko.SHP.tabulartext-generation100K<n<1M1 likes34 downloads3y agoHugging Face05eltociear /deeplearning-tasks-v1 deeplearning-tasks-v1 36 exact tasks for the deeplearning-env RL environment, on the topics of Deep Learning (Goodfellow, Bengio & Courville): information theory, backpropagation, optimisation, the linear algebra used in ML, and numerical stability. field meaning task_id dl-000 … dl-035 category information / backprop / optimisation / linalg / numerical prompt the question, the units, and the exact shape of the answer api_description the fixed network and… See the full description on the dataset page: https://huggingface.co/datasets/eltociear/deeplearning-tasks-v1.texttext-generationn<1K0 likes29 downloads2mo agoHugging Face06nlp-with-deeplearning /ko.openhermes원본 데이터셋: teknium/openhermes texttext-generation100K<n<1M3 likes24 downloads3y agoHugging Face07BEE-spoke-data /Nvidia-DeepLearningExamplesCode from https://github.com/NVIDIA/DeepLearningExamples INFO: Found 4341 text files - 2024-Jan-27_02-13 INFO: Train size: 4123 Validation size: 109 Test size: 109 texttext-generation1K<n<10K2 likes24 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.