datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imatrix
Input files for generating the Importance Matrix
Which file to use for generating the importance matrix
Not all importance matrices are equal. The best results are obtained when using a source file similar to the
training data. Size also matters: the bigger the model (eg: 70b vs 13b) and the higher the quant (eg: q6k_ vs iq3_xs),
the bigger the source file needs to be to make an impact. Multiple input files can be combined if needed;
for example:
cat multilingual.txt… See the full description on the dataset page: https://huggingface.co/datasets/froggeric/imatrix.creativity"The only difference between Science and screwing around is writing it down." (Adam Savage)
The LLM Creativity benchmark
Last benchmark update: 28 May 2024
The goal of this benchmark is to evaluate the ability of Large Language Models to be used
as an uncensored creative writing assistant. Human evaluation of the results is done manually,
by me, to assess the quality of writing.
There are 24 questions, some standalone, other follow-ups to previous questions for a multi-turn… See the full description on the dataset page: https://huggingface.co/datasets/froggeric/creativity.atari_frogger_with_masks
Atari Frogger Expert Trajectories with Player Masks
This dataset contains expert-policy transitions from Atari Frogger together
with a programmatically derived binary mask of the player in every stored RGB
observation. The expert is a Rainbow agent trained with
CleanRL, and player localization is
provided by OCAtari.
The dataset accompanies the paper Segment to Focus: Guiding Latent Action
Models in the Presence of Distractors, where
Frogger is used to study mask-guided latent… See the full description on the dataset page: https://huggingface.co/datasets/EpicPinkPenguin/atari_frogger_with_masks.details_froggeric__WestLake-10.7B-v2
Dataset Card for Evaluation run of froggeric/WestLake-10.7B-v2
Dataset automatically created during the evaluation run of model froggeric/WestLake-10.7B-v2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_froggeric__WestLake-10.7B-v2.atari-frogger-traces
