datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
llm-japanese-dataset-custombitcoin-llm-finetuning-dataset_new_with_custom_text1-800-LLMs__Qwen-2.5-14B-Hindi-Custom-Instruct-details
Dataset Card for Evaluation run of 1-800-LLMs/Qwen-2.5-14B-Hindi-Custom-Instruct
Dataset automatically created during the evaluation run of model 1-800-LLMs/Qwen-2.5-14B-Hindi-Custom-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/1-800-LLMs__Qwen-2.5-14B-Hindi-Custom-Instruct-details.custom_llm_kor_dnmdcustom_LLMcustom_llm_data
How to Train Brand LLM?
Launch Athena Generative AI Starter Kit from AWS Marketplace (see https://aws.amazon.com/marketplace/pp/prodview-su3dsq7b4plxw)
This is public Dataset 1 for training generic Model m2, code at host 4090 ~/athena/m2/m2_athena3.py
Use Parquet Hub to add private brand Dataset 2 to the m2 parquet file. Then train Model m3, the enterprise Brand LLM {see "Brand LLM: Parquet Hub"}
Run Model m3 training code at host 4090 ~/athena/m3/m3_model.py
Remember 'conda… See the full description on the dataset page: https://huggingface.co/datasets/drgary/custom_llm_data.skymizer__Llama2-7b-sft-chat-custom-template-dpo-details
Dataset Card for Evaluation run of skymizer/Llama2-7b-sft-chat-custom-template-dpo
Dataset automatically created during the evaluation run of model skymizer/Llama2-7b-sft-chat-custom-template-dpo
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/skymizer__Llama2-7b-sft-chat-custom-template-dpo-details.customllmcustom_llm_testbitcoin-llm-finetuning-dataset_new_with_custom_text_with_long_short_termcustomllmcustom_LLM
