kaaloo/evalap-comparing-openweight-with-atlaaiselene-1-mini-llama-31-8b-115
Comparing openweight with AtlaAI/Selene-1-Mini-Llama-3.1-8B (ID: 115) Comparing openweight Albert-API with specific judge AtlaAI/Selene-1-Mini-Llama-3.1-8B Overview This dataset contains 24 experiments from the EvalAP evaluation platform. Datasets: Assistant IA - QA, MFS_questions_v01 Models evaluated: openweight-large, openweight-medium, openweight-small Metrics: energy_consumption, generation_time, gwp_consumption, judge_notator, judge_precision… See the full description on the dataset page: https://huggingface.co/datasets/kaaloo/evalap-comparing-openweight-with-atlaaiselene-1-mini-llama-31-8b-115.
Comparing openweight with AtlaAI/Selene-1-Mini-Llama-3.1-8B (ID: 115)
Comparing openweight Albert-API with specific judge AtlaAI/Selene-1-Mini-Llama-3.1-8B
Overview
This dataset contains 24 experiments from the EvalAP evaluation platform.
Datasets: Assistant IA - QA, MFSquestionsv01
Models evaluated: openweight-large, openweight-medium, openweight-small
Metrics: energyconsumption, generationtime, gwpconsumption, judgenotator, judgeprecision, nbtokenscompletion, nbtokens_prompt
Scores
MFSquestionsv01
Assistant IA - QA
Usage
Use the dropdown above to select an experiment configuration.
