AgentPublic/evalap-comparing-openweight-with-openaigpt-oss-120b-113
Comparing openweight with openai/gpt-oss-120b (ID: 113) Comparing openweight Albert-API with specific judge openai/gpt-oss-120b Overview This dataset contains 24 experiments from the EvalAP evaluation platform. Datasets: Assistant IA - QA, MFS_questions_v01 Models evaluated: openweight-large, openweight-medium, openweight-small Metrics: energy_consumption, generation_time, gwp_consumption, judge_notator, judge_precision, nb_tokens_completion, nb_tokens_prompt… See the full description on the dataset page: https://huggingface.co/datasets/AgentPublic/evalap-comparing-openweight-with-openaigpt-oss-120b-113.
Comparing openweight with openai/gpt-oss-120b (ID: 113)
Comparing openweight Albert-API with specific judge openai/gpt-oss-120b
Overview
This dataset contains 24 experiments from the EvalAP evaluation platform.
Datasets: Assistant IA - QA, MFSquestionsv01
Models evaluated: openweight-large, openweight-medium, openweight-small
Metrics: energyconsumption, generationtime, gwpconsumption, judgenotator, judgeprecision, nbtokenscompletion, nbtokens_prompt
Scores
MFSquestionsv01
Assistant IA - QA
Usage
Use the dropdown above to select an experiment configuration.
