AgentPublic/evalap-comparing-albert-api-models-v11-12-2025-with-sysprompt-108
Comparing Albert-API models v11-12-2025 (with sysprompt) (ID: 108) Comparing albert models on MFS-AIA datasets (with sysprompt) Overview This dataset contains 20 experiments from the EvalAP evaluation platform. Datasets: Assistant IA - QA, MFS_questions_v01 Models evaluated: albert-large, albert-small, openweight-large, openweight-medium, openweight-small Metrics: generation_time, judge_notator, judge_precision, nb_tokens_completion, nb_tokens_prompt… See the full description on the dataset page: https://huggingface.co/datasets/AgentPublic/evalap-comparing-albert-api-models-v11-12-2025-with-sysprompt-108.
Comparing Albert-API models v11-12-2025 (with sysprompt) (ID: 108)
Comparing albert models on MFS-AIA datasets (with sysprompt)
Overview
This dataset contains 20 experiments from the EvalAP evaluation platform.
Datasets: Assistant IA - QA, MFSquestionsv01
Models evaluated: albert-large, albert-small, openweight-large, openweight-medium, openweight-small
Metrics: generationtime, judgenotator, judgeprecision, nbtokenscompletion, nbtokensprompt, outputlength
Scores
Assistant IA - QA
MFSquestionsv01
Usage
Use the dropdown above to select an experiment configuration.
