CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sinatras /pmpp-hard PMPP-Hard Agent Evaluation Traces PMPP-Hard is a 69-task agentic GPU-kernel evaluation for testing whether autonomous coding agents can produce implementations that are both correct and performant. This dataset contains the complete nine-model campaign used in the PMPP-Hard release: 621 rollouts, with 69 task sessions for each model configuration. Maintained and released by Sinatras. Source repository: SinatrasC/pmpp-hard Prime environment and evaluations: PMPP-Hard on Prime… See the full description on the dataset page: https://huggingface.co/datasets/sinatras/pmpp-hard.texttext-generationn<1K2 likes106 downloads2mo agoHugging Face02sinatras /pmpp-eval PMPP Dataset This repository provides two CUDA-focused datasets prepared by Sinatras and sponsored by Prime Intellect. Both datasets are based on Programming Massively Parallel Processors (4th Ed.) with additional coding evaluation harnesses at https://github.com/SinatrasC/pmpp-eval to be used by PMMP env in prime-environments. Overview Languages: English License: MIT Curated by: Sinatras (https://github.com/SinatrasC) Sponsored by: Prime Intellect Derived from: PMPP 4th… See the full description on the dataset page: https://huggingface.co/datasets/sinatras/pmpp-eval.texttext-generationn<1K4 likes91 downloads11mo agoHugging Face03PortalPal-AI /PMP-Synth-AllPairs-Annotatedtextn<1K0 likes42 downloads5mo agoHugging Face04HuggingFaceH4 /pmp-se-test-dataset Dataset Card for Dataset Name text10K<n<100K0 likes34 downloads4y agoHugging Face05nyuuzyou /PM-products Dataset Card for PochtaMarket products Dataset Summary This dataset was scraped from product pages on the Russian marketplace PochtaMarket. It includes all information from the product card. The dataset was collected by processing around 500 thousand, starting from the first one. At the time the dataset was collected, it is assumed that these were all the products available on this marketplace. Some fields may be empty, but the string is expected to contain some data… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/PM-products.texttext-generation10K<n<100K1 likes25 downloads3y agoHugging Face06ali-vakil /PMP_QA_dataset_not_cleanPrint ("This dataset includes 580 Q/A records, not separated, not cleaned yet.-I am working to clean it up-therefore I'm not sharing it publicly.") question-answeringn<1K0 likes23 downloads3y agoHugging Face07royson /PMPD0 likes6 downloads2y agoHugging Face08pmpatel9572 /ImageStyleimagen<1K0 likes5 downloads3y agoHugging Face09collingray /pm-price-historiestext10K<n<100K0 likes4 downloads2y agoHugging Face10mingquan123321 /contact_pmpnn_esm30 likes4 downloads7mo agoHugging Face11pmpc /processed-old-with-embeddingsgated Dataset Card for "processed-old-with-embeddings" Dataset Summary Chunks of about 256 words split by whitespace and their embeddings computed with the pretrained spacy model ["de_dep_news_trf"] (https://github.com/explosion/spacy-models/releases/tag/de_dep_news_trf-3.6.1). The splits are created with respect to sentence boundaries parsed with the same model, sentences are concatenated if the result does not exceed max_words = 256, therefore the chunk length varies.… See the full description on the dataset page: https://huggingface.co/datasets/pmpc/processed-old-with-embeddings.text1M<n<10M0 likes2 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.