CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RoganInglis /vllm-control-arena vLLM Main Tasks Dataset AI coding tasks generated from vLLM git commits Dataset Description This dataset contains 6801 coding tasks automatically generated from git commits in the vLLM repository. Each task represents a real-world coding challenge derived from actual development work. Dataset Structure The dataset contains the following columns: commit_hash: The git commit hash parent_hash: The parent commit hash commit_title: The original commit… See the full description on the dataset page: https://huggingface.co/datasets/RoganInglis/vllm-control-arena.tabulartext-generation1K<n<10K0 likes17k downloads1y agoHugging Face02NgTMDuc /VLLM_ChartQAtext10K<n<100K0 likes3.4k downloads2y agoHugging Face03anguszzzz /vllm-0.28.0-wheels-py3120 likes1.5k downloads7d agoHugging Face04vlsp-2023-vllm /ViLLM-Eval ViLLM-Eval We utilize the lm-eval-harness library to conduct evaluations. This library allows us to efficiently evaluate language models, ensuring robustness and accuracy in our assessments. Feel free to explore our project and discover the capabilities of the language models we employ. Install git clone https://huggingface.co/datasets/vlsp-2023-vllm/ViLLM-Eval cd ViLLM-Eval pip install -e . Basic Usage # Add trust_remote_code=True if your model is a custom… See the full description on the dataset page: https://huggingface.co/datasets/vlsp-2023-vllm/ViLLM-Eval.3 likes1.1k downloads2y agoHugging Face05VLLMs /MIRB Benchmarking Multi-Image Understanding in Vision and Language Models: Perception, Knowledge, Reasoning, and Multi-Hop Reasoning File Structure ├── MIR |── analogy.json │── codeu.json |── dataset_namex.json └── Images ├── analogy │ └── image_x.jpg └──codeu └── image_x.jpg JSON Structure { "questions": " What is the expected kurtosis of the sequence created by`create_number_sequence(-10, 10)`?\n\n1.… See the full description on the dataset page: https://huggingface.co/datasets/VLLMs/MIRB.imagequestion-answering1K<n<10K14 likes455 downloads2y agoHugging Face06SaylorTwift /details_hosted_vllm____fsx__anton__deepseek-r1-checkpoint_private Dataset Card for Evaluation run of hosted_vllm//fsx/anton/deepseek-r1-checkpoint Dataset automatically created during the evaluation run of model hosted_vllm//fsx/anton/deepseek-r1-checkpoint. The dataset is composed of 15 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 9 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/SaylorTwift/details_hosted_vllm____fsx__anton__deepseek-r1-checkpoint_private.tabular1K<n<10K0 likes441 downloads2y agoHugging Face07VLLMs /MIRB-hfimageimage-to-textn<1K0 likes397 downloads2y agoHugging Face08EXDai /vllm-configs EXD vLLM Config Profiles Production configuration profiles for the config-driven serve harness. Profiles Config Model Notes qwen2.5-7b Qwen 2.5 7B Instruct Baseline 7B qwen3.6-27b Qwen 3.6 27B Mid-size qwen3.6-35b-a3b Qwen 3.6 35B A3B MoE baseline qwen3.6-35b-a3b-baseline Qwen 3.6 35B A3B Baseline sweep qwen3.6-35b-a3b-mtp0 Qwen 3.6 35B A3B MTP depth 0 qwen3.6-35b-a3b-mtp3 Qwen 3.6 35B A3B MTP depth 3 qwen3.6-35b-a3b-throughput Qwen 3.6 35B… See the full description on the dataset page: https://huggingface.co/datasets/EXDai/vllm-configs.0 likes297 downloads3mo agoHugging Face09EXD-AI /vllm-configs EXD vLLM Config Profiles Production configuration profiles for the config-driven serve harness at projects/serve/. Usage On atom (GPU machine): cd ~/EXD ./scripts/up.sh <config-name> Profiles Config Model Notes qwen2.5-7b Qwen 2.5 7B Instruct Baseline 7B qwen3.6-27b Qwen 3.6 27B Mid-size qwen3.6-35b-a3b Qwen 3.6 35B A3B MoE baseline qwen3.6-35b-a3b-baseline Qwen 3.6 35B A3B Baseline sweep qwen3.6-35b-a3b-mtp0 Qwen 3.6 35B A3B MTP… See the full description on the dataset page: https://huggingface.co/datasets/EXD-AI/vllm-configs.0 likes265 downloads3mo agoHugging Face10lihaoxin2020 /Qwen2.5-7B-Instruct-vllm-20251128_042753text0 likes251 downloads8mo agoHugging Face11bihungba1101 /grammar-accuracy-qwen3.5-4b-trl-grpo-vllm-colocate-completions TRL Completion logs This dataset contains the completions generated during training using trl. Find the trained model at https://huggingface.co/bihungba1101/grammar-accuracy-qwen3.5-4b-trl-grpo-vllm-colocate. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion… See the full description on the dataset page: https://huggingface.co/datasets/bihungba1101/grammar-accuracy-qwen3.5-4b-trl-grpo-vllm-colocate-completions.tabularn<1K0 likes238 downloads4mo agoHugging Face12asingh15 /Nemotron-Personas-USA-synthetic-records-10files-qa-vllm-qwen4b-instruct-2507-clarqatext1K<n<10K0 likes217 downloads10mo agoHugging Face13royrin /gumble-max-vllm-experiment0 likes209 downloads10mo agoHugging Face14huggingface /vllm-metadata0 likes190 downloads2y agoHugging Face15vlsp-2023-vllm /vllms-leaderboard0 likes182 downloads2y agoHugging Face16sghosts /processed_qwen25_7b_vllm_final_40 likes179 downloads1y agoHugging Face17sghosts /processed_qwen25_7b_vllm_final_30 likes177 downloads1y agoHugging Face18lyydfys /deepseek-v4-flash-rocm-vllm-repro Reproducing DeepSeek-V4-Flash on AMD ROCm with vLLM: 32K Correctness and TopK Sweep This article summarizes an engineering reproduction of deepseek-ai/DeepSeek-V4-Flash on an AMD ROCm ModelScope DSW instance. The work focuses on a practical question: can a complex, fast-moving DeepSeek-V4-Flash serving path be turned into a reproducible ROCm baseline with explicit correctness gates? The answer from this run is yes, with an important boundary: the current setup is a fallback-heavy… See the full description on the dataset page: https://huggingface.co/datasets/lyydfys/deepseek-v4-flash-rocm-vllm-repro.0 likes169 downloads4mo agoHugging Face19yifei-liu /output_3d_bounding_scannetppv2_vllm_old_descriptiontext10K<n<100K0 likes161 downloads8mo agoHugging Face20sghosts /processed_qwen25_7b_vllm_final dataset_info: features: name: images dtype: image name: predictions dtype: string name: page_number dtype: int64 name: title dtype: string name: author dtype: string name: thesis_id dtype: string name: university dtype: string name: department dtype: string name: year dtype: string name: language dtype: string name: thesis_type dtype: string name: keyword_abd dtype: 'null' name: abstract_tr dtype: string name: abstract_en dtype: string name: file_size_bytes dtype: int64 name:… See the full description on the dataset page: https://huggingface.co/datasets/sghosts/processed_qwen25_7b_vllm_final.0 likes160 downloads1y agoHugging Face21NgTMDuc /VLLM_ChartQA_splitimage10K<n<100K5 likes150 downloads2y agoHugging Face22sghosts /processed_qwen25_7b_vllm_final_20 likes136 downloads1y agoHugging Face23lihaoxin2020 /Qwen2.5-7B-Instruct-vllm-retriever-20251202_093826text0 likes129 downloads8mo agoHugging Face24prefixsliding /res-vllm ryzax/res-vllm Copy of ryzax/res with vLLM-tool AIME25 runs swapped into the SUMM and LASTK folders. Original ryzax/res is unchanged. All swapped/added runs are Qwen3-1.7B, AIME25 (30 problems × 64 samples), window 4096, max gen 262144, BUDGET_FORCE_SAVING=256 for summary variants. Swaps Path What it is now AIME25 avg@64 cov@64 maj@64 Qwen3-1.7B-SUMM/w4096s256 force256 + SAVING_PROMPT=context_system (evaluate-context example on last save only) 26.5%… See the full description on the dataset page: https://huggingface.co/datasets/prefixsliding/res-vllm.0 likes123 downloads28d agoHugging Face25thaki-AI /daily-paper-2026-07-21-agent-dynamic-batch-tuning-vllm Dynamic Admission Reallocation for Multi-Tenant vLLM Serving TL;DR — On a real H200 running vLLM, a dynamic controller that reallocates a fixed 96-slot admission budget toward live demand beat a frozen even 48/48 split by +14.0% total throughput (11,274 vs 9,894 tok/s) and +18.0% batch throughput (7,736 vs 6,557 tok/s) at equal-or-better p99 (4.08 vs 4.15 s). Honest caveat: steady-window GPU utilization reached only 67.2% mean (99% peak), short of a sustained 90% target… See the full description on the dataset page: https://huggingface.co/datasets/thaki-AI/daily-paper-2026-07-21-agent-dynamic-batch-tuning-vllm.0 likes117 downloads2mo agoHugging Face26llgrnm /modal-vllm-cache-b300-minimax-v440 likes114 downloads14h agoHugging Face27Usman391 /vLLM-SDF-and-SDF-plus-VT-rollouts-Qwen vLLM SDF & SDF+VT rollouts Qwen Full, untruncated Qwen3.6-35B-A3B rollouts generated locally with compiled CUDA graphs (vLLM 0.26.0 for the original exports and vLLM 0.19.1 CUDA 12.8 for the added SDF-2250 and VT-250 exports). The aggregate contains 21,290 records. Generation used temperature 0.3 and a 16,000-token per-turn/output cap. Contents Benchmark Policies Records EvalAwareBench VT-250, SDF-1250, SDF-2250, SDF-1250+VT-250, SDF-2250+VT-250 12,500… See the full description on the dataset page: https://huggingface.co/datasets/Usman391/vLLM-SDF-and-SDF-plus-VT-rollouts-Qwen.text-generation0 likes111 downloads1mo agoHugging Face28ppppqp /vLLM-SR-Preference-V1The files in this repo is the LLM-labeled samples that are used as the training dataset for vLLM-SR Preference model V1. The training file (sharegpt_preference_labeld_with_negative.jsonl) contains 25k records that have sample_id, golden label for the preference-based routing policy, and a set of negative labels that are plausible but do not match the conversation context. The validation file has the same structure, but only 1% of the training file size. The validation file and the training… See the full description on the dataset page: https://huggingface.co/datasets/ppppqp/vLLM-SR-Preference-V1.textn<1K0 likes110 downloads8mo agoHugging Face29asingh15 /Nemotron-Personas-USA-synthetic-records-10files-qa-vllm-qwen4b-instruct-2507text1K<n<10K0 likes109 downloads10mo agoHugging Face30yurkes /patch_tasks_vllm Dataset Card for Patch-Based Visual Question Answering Dataset Dataset Details Dataset Description This dataset contains approximately 305,000 triplets of question, answer, and image designed for patch-based visual reasoning tasks. A standard question in this dataset is formatted as follows: Image Grid: The image is divided into a 4x4 grid of 16 equal-sized patches. Patches are numbered sequentially from the top-left corner and moving right, then down to the… See the full description on the dataset page: https://huggingface.co/datasets/yurkes/patch_tasks_vllm.imageimage-text-to-text100K<n<1M3 likes92 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.