datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llm-comparison
LLM Similarity Comparison Dataset
This dataset is pased on the original Alpaca dataset and was synthetically genearted for LLM similarity comparison using ConSCompF framework as described in the original paper.
The script used for generating data is available on Kaggle.
It is divided into 3 subsets:
quantization - contains 156,000 samples (5,200 for each model) generated by the original Tinyllama and its 8-bit, 4-bit, and 2-bit GGUF quantized versions.
comparison - contains 28,600… See the full description on the dataset page: https://huggingface.co/datasets/alex-karev/llm-comparison.LLM-Model-Comparison-2026
LLM Model Comparison 2026
Which LLM should you use for enterprise AI in 2026? This open dataset compares 16 large language models from 7 providers across 22 fields: pricing, benchmark scores, context windows, latency, API features, and recommended use cases.
Published and maintained by Salt Technologies AI, the AI engineering division of Salt Technologies (14+ years, 800+ projects delivered).
Quick Links
Interactive dataset page:… See the full description on the dataset page: https://huggingface.co/datasets/salttechno/LLM-Model-Comparison-2026.local-llm-runtime-comparison-rtx-5080
Ollama vs llama.cpp vs LM Studio on One RTX 5080
Controlled runtime comparison of Ollama, llama.cpp and LM Studio on the same RTX 5080 using identical hard-linked, sha256-verified GGUF bytes.
Why this exists
Most runtime comparisons do not control the model file, so you never know whether you measured the runtime or an uncontrolled quantization difference.
Here the same GGUF files are shared between apps via hard links and verified by sha256, so every runtime… See the full description on the dataset page: https://huggingface.co/datasets/iBlessi/local-llm-runtime-comparison-rtx-5080.marker_benchmark_comparison_llmmarker_comparison_mistral_llmllm-comparison
Fine tuning progress validation - RedPajama 3B, StableLM Alpha 7B, Open-LLaMA
This repository contains the progress of fine-tuning models: RedPajama 3B, StableLM Alpha 7B, Open-LLaMA. These models have been fine-tuned on a specific text dataset and the results of the fine-tuning process are provided in the text file included in this repository.
Fine-Tuning Details
Model: RedPajama 3B, size: 3 billion parameters, method: adapter
Model: StableLM Alpha 7B, size: 7 billion… See the full description on the dataset page: https://huggingface.co/datasets/kstevica/llm-comparison.marker_benchmark_comparison_olmocr_llmalitaqishah_urdu-llm-benchmark-llm-comparison-2026
Urdu LLM Benchmark + LLM Comparison 2026
Evaluating 6 LLMs on Urdu MCQs + 40 Model Benchmark Comparison 2026
Dataset Info
Source: Kaggle
Original Size: 0.03 MB
Kaggle Downloads: 32
Files: 2
Files
llm_benchmark_comparison_2026.csv
urdu_llm_benchmark.csv
Mirrored from Kaggle
llm-api-cost-comparison-2026-nexaapi-routing
LLM API Pricing Comparison 2026: Cut Your AI Costs by 70% with Smart Routing + NexaAPI
LLM routing tools like routeforge help you save money by automatically picking the right model for each request. But here's the secret most developers miss: the real savings come from choosing the right API provider in the first place.
In this guide, we'll break down the full cost comparison across major AI API providers, then show you how pairing a routing strategy with NexaAPI as your backend… See the full description on the dataset page: https://huggingface.co/datasets/nickyni/llm-api-cost-comparison-2026-nexaapi-routing.llm-comparison-dataset
LLM Comparison Dataset
A curated dataset of Large Language Model specifications, benchmarks, and comparisons.
About
This dataset contains structured information about various LLMs including:
Model specifications (parameters, context length, training data)
Performance benchmarks
Pricing information
API availability
Source
Data curated by Creative Content Crafts, an AI-first technology company.
Related Resources
LLMIndex.net: https://llmindex.net -… See the full description on the dataset page: https://huggingface.co/datasets/sergeinboca/llm-comparison-dataset.judge_llm_AB_comparison
