CoolFace
Datasetpublic

MK4-Research/LOREA-cyber-eval

LOREA-cyber eval sets Held-out sets used to benchmark the LOREA-cyber models. Decontaminated 8-gram against the training data, published so the numbers in the model cards can be reproduced. These are the sets written for this project. The models are also scored on public benchmarks that aren't redistributed here: SecQA, MMLU-Pro, CyberMetric, HumanEval. cyber_mcq (150) Security knowledge multiple choice across network security, crypto, web/OWASP, malware analysis… See the full description on the dataset page: https://huggingface.co/datasets/MK4-Research/LOREA-cyber-eval.

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
1likes35downloads

MK4-Research/LOREA-cyber-eval · main · files are served by the source, never re-hosted here