EfficientLLMInferenceCompetition/efficiency_samples_10pct
Muslim-language LLM efficiency-eval samples (public 10%) Public 10% stratified sample of the private dataset EfficientLLMInferenceCompetition/efficiency_samples. Workload for the NeurIPS 2026 competition proposal Efficient LLM Inference for Diverse Muslim Languages and Cultures. This repo holds constructed GuideLLM Poisson traffic (not model weights). Lengths are measured with the official google/gemma-4-31B-it tokenizer including the chat template. Sampling: seed 42, 200… See the full description on the dataset page: https://huggingface.co/datasets/EfficientLLMInferenceCompetition/efficiency_samples_10pct.
This repository is gated. The listing is public, but downloading a file means accepting the publisher’s terms at Hugging Face first — the links above take you there rather than around it.
