CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01boltzgen /inference-data0 likes156k downloads1y agoHugging Face02rahul-ai-01 /groot_n1.7_inference_on_diff_data GROOT Inference Analysis Log Evaluation records for a GR00T policy trained on the task "pick octopus and place inside brown basket", run on a Unitree G1 at 20 Hz with the ego_view stereo camera. Six training checkpoints (50, 100, 150, 200, 250, 300 demonstration episodes) were each evaluated on 50 inference episodes. Every episode is recorded here with video, per-tick state/action logs and run metadata. Success rate Checkpoint (training episodes) Success… See the full description on the dataset page: https://huggingface.co/datasets/rahul-ai-01/groot_n1.7_inference_on_diff_data.video0 likes7k downloads26d agoHugging Face03NTU-yiwen /code-world-model-inference-examples-40 Inference examples This directory contains 40 numbered, independent inference examples. Every example uses only its public number; source case names and internal paths are intentionally omitted. Each numbered directory contains: first_frame.png: exact 1536x864 generated RGB first frame used by inference. prompts/*.txt: the exact rolling long-inference prompts used for the result. condition/*.npz: ordered lossless condition chunks. metadata.json: frame count, FPS, prompt windows… See the full description on the dataset page: https://huggingface.co/datasets/NTU-yiwen/code-world-model-inference-examples-40.text1K<n<10K0 likes4.5k downloads26d agoHugging Face04bwarner /inference-scratchtabular1M<n<10M0 likes3.6k downloads5mo agoHugging Face05bxiong /data_inference_pythia_6_9b0 likes3.2k downloads2y agoHugging Face06VisGym /inference-dataset VLM-Gym Inference Dataset This dataset contains pre-defined test episodes and initial states for evaluating Vision-Language Models (VLMs) on the VLM-Gym benchmark. Dataset Structure inference-dataset/ ├── test_set_easy/ # Easy difficulty test episodes (JSONL) ├── test_set_hard/ # Hard difficulty test episodes (JSONL) ├── initial_states_easy/ # Initial environment states for easy episodes (JSON) ├── initial_states_hard/ # Initial environment… See the full description on the dataset page: https://huggingface.co/datasets/VisGym/inference-dataset.visual-question-answering3 likes3.1k downloads8mo agoHugging Face07P2SAMAPA /p2-etf-active-inference-results1 likes2.6k downloads5d agoHugging Face08nbroad /hf-inference-providers-data3 likes2.2k downloads3h agoHugging Face09Hemanth-thunder /tamil-inference-result0 likes2.1k downloads1y agoHugging Face10Tungtom2004 /Inference_Step1XEditimage1K<n<10K0 likes1.9k downloads2h agoHugging Face11Tungtom2004 /Qwen-SFT-Inference-Outputsimage1K<n<10K0 likes1.6k downloads5d agoHugging Face12crosslingual-rule-following /model-inference-activationstext10K<n<100K0 likes1.5k downloads28d agoHugging Face13inference-optimization /speculators-ci-datasets speculator-tutorial Raw vs. on-policy regenerated conversation data for training speculative-decoding drafters (EAGLE-3 / DFlash / DSpark style), with the original source data kept alongside so you can see exactly what regeneration changes and why it matters. Prompts come from UltraChat-200k. The verifier / teacher model is Qwen/Qwen3-8B. Why regenerate at all? A speculative-decoding drafter is trained to predict what the verifier would say next. If you train it… See the full description on the dataset page: https://huggingface.co/datasets/inference-optimization/speculators-ci-datasets.tabulartext-generation1K<n<10K0 likes1.1k downloads1mo agoHugging Face14hlarcher /inference-benchmarkertext100K<n<1M1 likes782 downloads2y agoHugging Face15conorhassan /fast-autoregressive-inference-gp-trainK4tabularn<1K0 likes701 downloads1y agoHugging Face16tersooawai /eval_pi0_inference_only_datasetThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so101_follower", "total_episodes": 70, "total_frames": 40583, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 500, "fps": 30, "splits": { "train": "0:70" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/tersooawai/eval_pi0_inference_only_dataset.tabularrobotics10K<n<100K0 likes692 downloads4mo agoHugging Face17inferenceport-ai /qwen3.8-max-glm5.2-kimi-k3-distillation Multi-Teacher Distillation Dataset (57,937 traces) A quality-filtered, deduplicated, multi-teacher SFT corpus combining traces from three frontier models across math, code, reasoning, instruction-following, tool-use, science, long-context, multilingual, and creative dialogue domains. Teachers Teacher Provider Traces Qwen3.8-Max-Preview Alibaba Cloud Model Studio 48,283 GLM-5.2 Z.AI Coding Plan 5,307 Kimi Code K3 Moonshot AI (Kimi) 4,347… See the full description on the dataset page: https://huggingface.co/datasets/inferenceport-ai/qwen3.8-max-glm5.2-kimi-k3-distillation.tabulartext-generation10M<n<100M0 likes644 downloads8d agoHugging Face18Thorsu /sovereign-shadow-inference-bench Sovereign Shadow Inference Bench A public, versioned evidence surface for independent Hugging Face shadow inference beside Sovereign's primary OpenRouter/Revolver route. What this dataset proves The seed record in data/shadow_receipts.jsonl was produced by one real Hugging Face Inference Providers request. It records provider/model identity, request bounds, latency, hashes, literal-match outcome, source revision, and an immutable receipt hash. What it… See the full description on the dataset page: https://huggingface.co/datasets/Thorsu/sovereign-shadow-inference-bench.tabulartext-generationn<1K1 likes588 downloads11h agoHugging Face19crosslingual-rule-following /model-inference-responsestext1M<n<10M0 likes423 downloads27d agoHugging Face20etiennebamas /inference-test-set1 likes346 downloads2mo agoHugging Face21Heisen0928 /tfds_aloha_game_inference_11_200 likes342 downloads10mo agoHugging Face22BDXXN /amortized-inference-training Amortized Inference Training Trajectories Generated trajectories are organized as source_dataset/geometry_id/material_id/trajectory. Each geometry has index.jsonl; the global decisions live under index/. Only records with status == "accepted" and trainable == true passed the recorded legacy topology gate. rejected arrays are intentionally absent. Historical trajectories for which the topology gate was never run are retained as unreviewed and must not be used for training until… See the full description on the dataset page: https://huggingface.co/datasets/BDXXN/amortized-inference-training.0 likes329 downloads26d agoHugging Face23RyanL22 /RoboTryOn_inferencevideo1K<n<10K0 likes296 downloads2d agoHugging Face24Heisen0928 /tfds_aloha_game_inference_256_11_150 likes289 downloads10mo agoHugging Face25Masao-Taketani /vizdoom-inference-latent-datasetAs for the details of this dataset, please refer to https://github.com/Masao-Taketani/GameNGen. 0 likes285 downloads9mo agoHugging Face26michaelwaves /llava_qwen3vl_sae_inference_featuresimage1K<n<10K0 likes285 downloads7mo agoHugging Face27Efficient-Large-Model /VILA-inference-demosimagen<1K1 likes268 downloads2y agoHugging Face28Anticloud /camus-17-configurable-inference We externalize 10 inference parameters to a JSON config file — hot-reloadable settings for diverse hardware. Configurable Inference Settings via File-Based Configuration The Problem Hard-coded inference parameters make the system inflexible. Different hardware and tasks need different settings. What We Built JSON-based configuration file (camus.json) storing all user-settable parameters: n_ctx, n_threads, temperature, top_p, max_tokens, show_bars… See the full description on the dataset page: https://huggingface.co/datasets/Anticloud/camus-17-configurable-inference.0 likes262 downloads3mo agoHugging Face29CamoAiLab /InferenceNetgatedInferenceNet project portal · Home · Data overview · Leaderboard · Agent / Harness Explore the project and its published research results in the linked Space. This dataset repository remains the source for the task lists and research data. InferenceNet: Data Card for Econometric AI Agent Testset InferenceNet is a project aimed at evaluating and building up the AI capability for social science research related to empirical studies. We collect the world’s largest dataset on… See the full description on the dataset page: https://huggingface.co/datasets/CamoAiLab/InferenceNet.tabular1K<n<10K1 likes259 downloads2d agoHugging Face30kleinnner /camus-17-configurable-inference We externalize 10 inference parameters to a JSON config file — hot-reloadable settings for diverse hardware. Configurable Inference Settings via File-Based Configuration The Problem Hard-coded inference parameters make the system inflexible. Different hardware and tasks need different settings. What We Built JSON-based configuration file (camus.json) storing all user-settable parameters: n_ctx, n_threads, temperature, top_p, max_tokens, show_bars… See the full description on the dataset page: https://huggingface.co/datasets/kleinnner/camus-17-configurable-inference.0 likes251 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.