datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llama-cpp-wheelsIf you like this please consider liking and donating (https://buymeacoffee.com/aiencoder)
🏭 llama-cpp-python Mega-Factory Wheels
"Stop waiting for pip to compile. Just install and run."
The most complete collection of pre-built llama-cpp-python wheels in existence — 8,333 wheels across every platform, Python version, backend, and CPU optimization level.
No more cmake, gcc, or compilation hell. No more waiting 10 minutes for a build that might fail. Just find your wheel and… See the full description on the dataset page: https://huggingface.co/datasets/AIencoder/llama-cpp-wheels.llama.cpp_AlgMor24_github
ΩFFFΣLLIa • llama.cpp • AlgMor24
██████╗ ███████╗███████╗███████╗██╗ ██╗ ██╗ █████╗
██╔═══██╗██╔════╝██╔════╝██╔════╝██║ ██║ ██║██╔══██╗
██║ ██║█████╗ █████╗ █████╗ ██║ ██║ ██║███████║
██║ ██║██╔══╝ ██╔══╝ ██╔══╝ ██║ ██║ ██║██╔══██║
╚██████╔╝██║ ██║ ███████╗███████╗███████╗██║██║ ██║
╚═════╝ ╚═╝ ╚═╝ ╚══════╝╚══════╝╚══════╝╚═╝╚═╝ ╚═╝
High-Performance LLM / VLM Inference & Autonomous Agentic Ecosystem… See the full description on the dataset page: https://huggingface.co/datasets/Brunobkr/llama.cpp_AlgMor24_github.llama.cppversion https://git-lfs.github.com/spec/v1
oid sha256:cfc44b7ba25614df70e6b65e3341cae0310163bd32fd31a6b928a542df433faf
size 30786
llama-cpp-scripts
llama.cpp scripts
These are scripts that have helped me to manage llama.cpp, llama models, etc.
Install
Scripts are installed to ~/.local/bin.
bash install.sh
llama-cpp-rx7800xt-benchmarks
llama.cpp RX 7800 XT benchmarks
Sanitized benchmark archive for local llama.cpp/GGUF experiments on an AMD Radeon RX 7800 XT workstation. This dataset is benchmark data only: model weights, local system/network diagnostics, environment files, scripts, caches, and oversized raw dumps are excluded.
Start here
summaries/run-index.csv - index of benchmark result directories.
summaries/practical-audit-index.csv - compact index of practical-audit runs.
metadata.json -… See the full description on the dataset page: https://huggingface.co/datasets/NeoAiLabs/llama-cpp-rx7800xt-benchmarks.llama-cpp-binariesllama.cpp_colabmoss-tts-llamacpp-runtime-portable
MOSS-TTS runtime payload
This private Kaggle dataset is generated by Phorcys.Tools.MossTtsLlamaCppRuntimeUploader for PHRunner.Kaggle.Service.MossTtsGguf.
Runtime flavor: LinuxCuda
Python tag: python3.10
Generated UTC: 2026-09-08T11:25:17.0831463+00:00
The dataset intentionally contains runtime artifacts, not the GGUF model repository by default. Keep the model files in a separate private Kaggle dataset, for example kaggle-pool-account/moss-tts-v1-5-gguf-models.
Top-level… See the full description on the dataset page: https://huggingface.co/datasets/stokiz/moss-tts-llamacpp-runtime-portable.llama-cpp-turboquant-precompiledmoss-tts-v1-5-llamacpp-models
MOSS-TTS model payload
This dataset was generated by MossTtsLlamaCppModelsUploader for PHRunner.Kaggle.Service.MossTtsGguf.
Source
Hugging Face repository: niobures/MOSS-TTS-GGUF
Revision: main
Kaggle dataset: kaggle-pool-account/moss-tts-v1-5-llamacpp-models
Layout: python-llama-cpp
Required backend: python-llama-cpp-onnx
Default model file: MOSS_TTS_Q4_K_M.gguf
Files
MOSS_TTS_Q4_K_M.gguf - backbone - 4.7 GiB - SHA256… See the full description on the dataset page: https://huggingface.co/datasets/stokiz/moss-tts-v1-5-llamacpp-models.llama-cpp-wheelsllama-cpp-cuda130-blackwell120llamacpp-cuda-tarballsllama-cpp-cuda128-blackwell120llama-cpp-dflash2-cuda-binllama-cpp-cuda130-blackwell100llama-cpp-python-wheelsllama-cpp-cuda12-ada89llama.cpp CUDA 12.8.1, target GPU: sm89
If you're new or just starting to learn setting up your own inferences this llama.cpp wheel will work for if you're using python, CUDA 12.8.1 with one of the following Ada Lovelace generation (sm_89) GPU's:
NVIDIA L4
NVIDIA L40
NVIDIA L40S
NVIDIA RTX 6000 Ada Generation
NVIDIA RTX 5000 Ada Generation
NVIDIA RTX 4500 Ada Generation
NVIDIA RTX 4000 Ada Generation
NVIDIA RTX 4000 SFF Ada Generation
NVIDIA RTX 2000 Ada Generation
NVIDIA GeForce RTX 4090… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/llama-cpp-cuda12-ada89.llama.cppllama.cpp-0002llama-cpp-python-build-on-cu130-windows-py311llama-cpp-cuda128-ampere80llama.cpp_omega_matrix.zip
⚡ llama.cpp_omega_matrix
██████╗ ███╗ ███╗███████╗ ██████╗ █████╗ ██████╗ ██████╗ ██████╗
██╔═══██╗████╗ ████║██╔════╝██╔════╝ ██╔══██╗ ██╔════╝██╔══██╗██╔══██╗
██║ ██║██╔████╔██║█████╗ ██║ ███╗███████║ ██║ ██████╔╝██████╔╝
██║ ██║██║╚██╔╝██║██╔══╝ ██║ ██║██╔══██║ ██║ ██╔═══╝ ██╔═══╝
╚██████╔╝██║ ╚═╝ ██║███████╗╚██████╔╝██║ ██║ ╚██████╗██║ ██║
╚═════╝ ╚═╝ ╚═╝╚══════╝ ╚═════╝ ╚═╝ ╚═╝ ╚═════╝╚═╝ ╚═╝… See the full description on the dataset page: https://huggingface.co/datasets/Brunobkr/llama.cpp_omega_matrix.zip.llama-cpp-cuda12-ampere86NVIDIA A100
NVIDIA A40
NVIDIA A30
NVIDIA A10
NVIDIA A16
NVIDIA RTX A6000
NVIDIA RTX A5000
NVIDIA RTX A4000
NVIDIA GeForce RTX 3090
NVIDIA GeForce RTX 3080
NVIDIA GeForce RTX 3070
NVIDIA GeForce RTX 3060
llama-cpp-python-wheels-LMF2.5
llama-cpp-python — Free-Tier Friendly Wheel
This Space provides a prebuilt llama-cpp-python wheel designed to work
reliably on Hugging Face Free tier Spaces.
No compilation. No system packages. No build failures.
If your Space crashes during pip install llama-cpp-python, this wheel is the fix.
Optimized for Hugging Face Free Tier
Hugging Face Free tier Spaces are:
CPU-only
Limited in memory
Not suitable for native compilation
This wheel is built ahead of time so it can… See the full description on the dataset page: https://huggingface.co/datasets/Zap11/llama-cpp-python-wheels-LMF2.5.llama.cpp_jetson_benchmark
llama.cpp with CUDA support on a Jetson Nano 2019
This dataset contains some benchmark results for llama.cpp with GPU acceleration on the Jetson Nano from 2019.
For comparison llama.cpp was compiled with just CPU support, and a recent ollama version received the same 11 questions.
llama-cpp-python-wheelsllama-cpp-gguf-divzero-dos-poc
llama.cpp GGUF Division-by-Zero DoS — Security Research PoC
Target: ggml-org/llama.cppFile: ggml/src/gguf.cpp line 681Severity: CVSS 5.3–6.5 (Medium)Platform: x86_64 only (SIGFPE trap)
Summary
A crafted GGUF file with a tensor dimension of zero (ne[1] = 0) causes a division-by-zero in gguf.cpp at line 681:
INT64_MAX / info.t.ne[1] // crashes when ne[1] == 0
The validation at line 672 only rejects negative dimensions (ne[j] < 0), allowing zero to pass through.… See the full description on the dataset page: https://huggingface.co/datasets/wulonchia/llama-cpp-gguf-divzero-dos-poc.llama-cpp-benchllama_cpp_python-0.3.23-hf-space-prebuilt-wheel
