datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Llama-HybridDiffusion-processed-data-run1
Llama-HybridDiffusion processed training mixture — run 1
Built with Llama.
This repository preserves the exact Hugging Face Dataset.save_to_disk Arrow snapshot
used by run 1 of a Qwen3.5-2B HybridDiffusion reproduction. The directory names,
dataset_info.json, state.json, and Arrow shard boundaries are retained so the data
can be downloaded and supplied to the existing training configuration without a lossy
format conversion.
Exact snapshot inventory
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/Arushhh/Llama-HybridDiffusion-processed-data-run1.processed_llama_dataset_2048refine-book-wiki_processed_llama_dataset_2048openweb_processed_llama_dataset_2048Llama-4-Scout-17B-16E-Instruct-FP8-Instruct-FP8-ProcessedOpenAssistantllama_3.1_binary_train_processedprocessed_chat_messages_financial-qa-10K_for_llamaThe raw dataset used in this project is sourced from the virattt/financial-qa-10K repository. This dataset has been carefully processed for fine-tuning the meta-llama/Meta-Llama-3.1-8B-Instruct model, specifically for the Retrieval-Augmented Generation (RAG) use case.
The dataset consists of five key columns:
Question: The financial question posed in the dataset.
Answer: The corresponding answer to the financial question.
Context: Additional relevant context to help provide a more accurate… See the full description on the dataset page: https://huggingface.co/datasets/Shubhu07/processed_chat_messages_financial-qa-10K_for_llama.llama_3.1_binary_train_processed_hhstyleamharic_summarization_llama_3_1_predictions_no_normalization_post_processedLlama-3.3-70B-Instruct-ProcessedOpenAssistantamharic_summarization_llama_3_1_predictions_normalization_post_processedllama_3.1_binary_train_small_processedmini-data-processed_meta-llama_Llama-3.2-1B_0_8-train
Dataset Card for "mini-data-processed_meta-llama_Llama-3.2-1B_0_8-train"
More Information needed
llama_3.1_binary_train_small_processed_hhstyleLlama-3.3-8B-Instruct-ProcessedOpenAssistantmini-data-processed_meta-llama_Llama-3.2-1B_0_64-validation
Dataset Card for "mini-data-processed_meta-llama_Llama-3.2-1B_0_64-validation"
More Information needed
Arabic_cultural_dataset_processed_llama_8b_mcq_mxlen_1024Llama-4-Maverick-17B-128E-Instruct-FP8-ProcessedOpenAssistantMultilingual_cultural_dataset_processed_llama_8b_mcq_mxlen_1024Multilingual_cultural_dataset_processed_llama_8b_oe_mxlen_512processed-data-llamaArabic_cultural_dataset_processed_llama_8b_oe_mxlen_512_old
