sambanova
details_sambanovasystems__SambaLingo-Arabic-Chat-70B
Dataset Card for Evaluation run of sambanovasystems/SambaLingo-Arabic-Chat-70B
Dataset automatically created during the evaluation run of model sambanovasystems/SambaLingo-Arabic-Chat-70B.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_sambanovasystems__SambaLingo-Arabic-Chat-70B.x-self-instruct-seed-32
Dataset Card for xOA22 - Multilingual Prompts from OpenAssistant
Dataset Summary
x-self-instruct-seed-32 consists of 32 prompts chosen out of the 252 prompts in the self-instruct-seed dataset from the Self-Instruct paper. These 32 prompts were filtered out according to the following criteria:
Should be natural in a chat setting
Therefore, we filter out any prompts with "few-shot examples", as these are all instruction prompts that we consider unnatural in a chat setting… See the full description on the dataset page: https://huggingface.co/datasets/sambanovasystems/x-self-instruct-seed-32.details_sambanovasystems__SambaLingo-Arabic-Base
Dataset Card for Evaluation run of sambanovasystems/SambaLingo-Arabic-Base
Dataset automatically created during the evaluation run of model sambanovasystems/SambaLingo-Arabic-Base.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_sambanovasystems__SambaLingo-Arabic-Base.attackqa
AttackQA: Development and Adoption of a Dataset for Assisting Cybersecurity Operations using Fine-tuned and Open-Source LLMs
license: apache-2.0
This dataset is derived from the MITRE ATT&CK® knowledge base that bears the following license:
© 2025 The MITRE Corporation. This work is reproduced and distributed with the permission of The MITRE Corporation.
In using the dataset, please consider citing the following paper:
misc{c:attackqa,
title={AttackQA: Development… See the full description on the dataset page: https://huggingface.co/datasets/sambanovasystems/attackqa.xOA22
Dataset Card for xOA22 - Multilingual Prompts from OpenAssistant
Dataset Summary
xOA22 consists of 22 prompts originally shown in Appendix E, page 25 of the OpenAssistant Conversations paper. These 22 prompts were then manually translated by volunteers into 5 languages: Arabic, Simplified Chinese, French, Hindi and Spanish.
These prompts were originally created for human evaluations of the multilingual abilities of BLOOMChat. Since not all prompts could be directly… See the full description on the dataset page: https://huggingface.co/datasets/sambanovasystems/xOA22.sambanova_deit_data
Dataset Card for sambanova_deit_data
Dataset Description
This is output data from the Sambanova SN30. Each file is named based on which model it came from.
The data is in the form of 3-element tuples per sample from the Imagenet-1k validation dataset. Each tuple contains: logits (Python list), sample name (string), Imagenet label (int).
The included python script contains a function that will extract all data into a dictionary, with the model name that they came from as… See the full description on the dataset page: https://huggingface.co/datasets/ep44/sambanova_deit_data.
