salamandra
salamandra-guard-dataset
Salamandra Guard Dataset
Dataset Description
The Salamandra Guard dataset is a comprehensive multilingual safety classification corpus designed for training and evaluating content moderation systems in Catalan, Spanish. It consists of 21,335 carefully curated conversational examples annotated across a hierarchical safety taxonomy.
This dataset represents a significant advancement in culturally-grounded safety data, with particular emphasis on Catalan—a language… See the full description on the dataset page: https://huggingface.co/datasets/BSC-LT/salamandra-guard-dataset.BSC-LT__salamandra-7b-details
Dataset Card for Evaluation run of BSC-LT/salamandra-7b
Dataset automatically created during the evaluation run of model BSC-LT/salamandra-7b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BSC-LT__salamandra-7b-details.BSC-LT__salamandra-7b-instruct-details
Dataset Card for Evaluation run of BSC-LT/salamandra-7b-instruct
Dataset automatically created during the evaluation run of model BSC-LT/salamandra-7b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/BSC-LT__salamandra-7b-instruct-details.salamandra40b-aligned_results_gl_prompt1_test_5-Likert
Dataset Card for salamandra40b-aligned_results_gl_prompt1_test_5-Likert
This dataset has been created with Argilla. As shown in the sections below, this dataset can be loaded into your Argilla server as explained in Load with Argilla, or used directly with the datasets library in Load with datasets.
Using this dataset with Argilla
To load with Argilla, you'll just need to install Argilla as pip install argilla --upgrade and then use the following code:
import argilla… See the full description on the dataset page: https://huggingface.co/datasets/rsepulvedat/salamandra40b-aligned_results_gl_prompt1_test_5-Likert.imnet1k_European_fire_salamander_Salamandra_salamandracheckpoint-generic-reduced-salamandra
