CoolFace
Datasetpublic

Anodino/local-llms-benchmark-rtx5090

Local LLMs Benchmark — RTX 5090 Benchmark of 14 local language model configurations (9 distinct models, 27B–31B parameter range) across 9 questions covering logical reasoning, Bayesian statistics, cognitive bias detection, theoretical science, synthesis under contradiction, linguistic ambiguity, code optimization, and AI ethics. Hardware: RTX 5090 24GB | Intel Core Ultra 9 275HX | 64GB RAM | DebianInference backend: Ollama (Docker)Author: Francisco R. · LinkedIn… See the full description on the dataset page: https://huggingface.co/datasets/Anodino/local-llms-benchmark-rtx5090.

sourceHugging Facecc-by-nc-sa-4.0updated 5mo agoView on Hugging Face
0likes87downloads
Dataset Card

Local LLMs Benchmark — RTX 5090

Benchmark of 14 local language model configurations (9 distinct models, 27B–31B parameter range) across 9 questions covering logical reasoning, Bayesian statistics, cognitive bias detection, theoretical science, synthesis under contradiction, linguistic ambiguity, code optimization, and AI ethics.

Hardware: RTX 5090 24GB | Intel Core Ultra 9 275HX | 64GB RAM | Debian Inference backend: Ollama (Docker) Author: Francisco R. · LinkedIn


Full Benchmark

Methodology, scoring, ranking, per-question analysis, and practitioner recommendations: → benchmark_v2_final.md


Model Responses by Question

#CategoryResponses
Q1Logical ReasoningQ1_responses.md
Q2Logical ReasoningQ2_responses.md
Q3Cognitive Bias / SycophancyQ3_responses.md
Q4Bayesian StatisticsQ4_responses.md
Q5Theoretical ScienceQ5_responses.md
Q6Synthesis Under ContradictionQ6_responses.md
Q7Linguistic AmbiguityQ7_responses.md
Q8Code + OptimizationQ8_responses.md
Q9Ethics + ML + BusinessQ9_responses.md