datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
details_ValiantLabs__Llama3.1-8B-ShiningValiant2
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-8B-ShiningValiant2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-8B-ShiningValiant2.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_ValiantLabs__Llama3.1-8B-ShiningValiant2.ValiantLabs__Llama3.1-8B-CobaltValiantLabs__Llama3.1-8B-Cobalt-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-8B-Cobalt
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-8B-Cobalt
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.1-8B-Cobalt-details.ValiantLabs__Llama3.1-8B-Fireplace2tripletex-tool-embeddings
Tripletex API Tool Embeddings
Pre-computed embeddings for 800 Tripletex accounting API tools, extracted from the OpenAPI 3.0.1 spec and embedded with Google gemini-embedding-001 (3072 dimensions).
Built for RAG-based tool filtering in the AI Accounting Agent competition project.
Quick Start
from datasets import load_dataset
# Full embeddings (800 tools, 3072-dim vectors) — ready for RAG
ds = load_dataset("valiantlynxz/tripletex-tool-embeddings")
# Lightweight: tool… See the full description on the dataset page: https://huggingface.co/datasets/valiantlynxz/tripletex-tool-embeddings.ValiantLabs__Llama3-70B-Fireplace-details
Dataset Card for Evaluation run of ValiantLabs/Llama3-70B-Fireplace
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3-70B-Fireplace
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3-70B-Fireplace-details.ValiantLabs__Llama3-70B-ShiningValiant2ValiantLabs__Llama3.1-70B-ShiningValiant2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-70B-ShiningValiant2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-70B-ShiningValiant2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.1-70B-ShiningValiant2-details.volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk-output_original
Знак Вялікага Магістра — арыгінальнае аўдыё
Аўтар / Author: Вольга ІпатаваМова / Language: Беларуская (Belarusian)
Арыгінальнае аўдыё без апрацоўкі, захаванае ў зыходнай якасці.
Частка калекцыі Ministerskija —
корпус беларускіх аўдыёкніг.
Апрацаваная версія (сегменты ~15 с, выраўнаваная транскрыпцыя):
volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk-output
Доўгасць аўдыё
4h13m
Радкоў у датасеце
1 302
Структура
Кожны радок змяшчае:… See the full description on the dataset page: https://huggingface.co/datasets/fosters/volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk-output_original.ValiantLabs__Llama3.2-3B-Esper2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.2-3B-Esper2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.2-3B-Esper2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.2-3B-Esper2-details.ValiantLabs__Llama3-70B-ShiningValiant2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3-70B-ShiningValiant2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3-70B-ShiningValiant2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3-70B-ShiningValiant2-details.ValiantLabs__Llama3.2-3B-ShiningValiant2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.2-3B-ShiningValiant2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.2-3B-ShiningValiant2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.2-3B-ShiningValiant2-details.volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk-input
Знак Вялікага Магістра
This is a Hugging Face Parquet input dataset for an audio pipeline.
Source dataset: archivartaunik/volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk
Format
config: default
split: train
format: parquet
id column: id
audio column: audio
rows: 67
shards: 1
language: be
The audio column is embedded into Parquet as Hugging Face Audio:
audio = {
"path": "file.mp3",
"bytes": b"..."
}
Columns
id
audio
title… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk-input.ValiantLabs__Llama3.1-8B-ShiningValiant2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-8B-ShiningValiant2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-8B-ShiningValiant2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 10 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.1-8B-ShiningValiant2-details.ValiantLabs__Llama3.1-8B-Esper2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-8B-Esper2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-8B-Esper2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.1-8B-Esper2-details.volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk_all
AudioSet Pipeline Output
Мова / Language: Беларуская (Belarusian)
Аўдыё нарэзана з арыгінальнага запісу ў зыходнай частаце дыскрэтызацыі (native), мона, фрагменты да 30 секунд.
Частка калекцыі Belarusian Audiobooks (native).
Радкоў у датасеце
2,373
Працягласць
10 гадз 9 хв
Частата дыскрэтызацыі
22050 Hz
Каналы
мона
Даўжыня фрагмента
да 30 с
Структура
Кожны радок змяшчае:
audio — аўдыёфрагмент (native SR, мона, ≤30 с)
text — транскрыпцыя… See the full description on the dataset page: https://huggingface.co/datasets/fosters/volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk_all.ValiantLabs__Llama3.1-8B-Fireplace2-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-8B-Fireplace2
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-8B-Fireplace2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.1-8B-Fireplace2-details.ValiantLabs__Llama3.1-8B-Enigma-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.1-8B-Enigma
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.1-8B-Enigma
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.1-8B-Enigma-details.ValiantLabs__Llama3.2-3B-Enigma-details
Dataset Card for Evaluation run of ValiantLabs/Llama3.2-3B-Enigma
Dataset automatically created during the evaluation run of model ValiantLabs/Llama3.2-3B-Enigma
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ValiantLabs__Llama3.2-3B-Enigma-details.genadz-buraukin-nash-bykau-kniga-uspaminau-valiantsin-aksiantsiuk
Наш Быкаў. Кніга ўспамінаў
Metadata
Author: Генадзь Бураўкін
Title: Наш Быкаў. Кніга ўспамінаў
Narrator: Валянцін Аксянцюк
Source Group: Аўдыёкнігі
Source: prastora.by
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/genadz-buraukin-nash-bykau-kniga-uspaminau-valiantsin-aksiantsiuk.volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk
Знак Вялікага Магістра
Metadata
Author: Вольга Іпатава
Title: Знак Вялікага Магістра
Narrator: Валянцін Аксянцюк
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/volga-ipatava-znak-vialikaga-magistra-valiantsin-aksiantsiuk.valiantsin-taras-na-vyspe-uspaminau-valiantsin-taras
На высьпе ўспамінаў
Metadata
Author: Валянцін Тарас
Title: На высьпе ўспамінаў
Narrator: Валянцін Тарас
Source Group: Аўдыёкнігі
Source:
Notes
The original audio files are preserved as-is:
no conversion;
no re-encoding;
no filename changes inside each split folder, except removing one common top-level archive folder when present.
To avoid Hugging Face Dataset Viewer scan-size errors, the dataset is split into smaller folders.
Target maximum… See the full description on the dataset page: https://huggingface.co/datasets/archivartaunik/valiantsin-taras-na-vyspe-uspaminau-valiantsin-taras.
