datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
multimodal-LLMs-See-Sentiment
MLLMsent — datasets and experiment results
Every input and every output of "Multimodal LLMs See Sentiment"
(arXiv:2508.16873): the image descriptions generated by six multimodal
LLMs, the sentiment labels derived from the PerceptSent annotations, and the complete
per-fold results of all 141 experiments.
Paper: arXiv:2508.16873
Code, training and inference: https://github.com/neemiasbsilva/multimodal-LLMs-see-sentiment
Model checkpoints:… See the full description on the dataset page: https://huggingface.co/datasets/neemiasbsilva/multimodal-LLMs-See-Sentiment.LLM_Description_Vocab_gpt-3_text-davinci-003LLM_Description_Vocab_opt_facebook_opt_30b_downstream_tasks
Dataset Card for "LLM_Description_Vocab_opt_facebook_opt_30b_downstream_tasks"
More Information needed
LLM_Description_Vocab_opt_Multimodal_Fatima_opt_175b_downstream_tasks
Dataset Card for "LLM_Description_Vocab_opt_Multimodal_Fatima_opt_175b_downstream_tasks"
More Information needed
LLM_Description_Vocab_bloom_bigscience_bloom_downstream_tasks
Dataset Card for "LLM_Description_Vocab_bloom_bigscience_bloom_downstream_tasks"
More Information needed
LLM_Description_Vocab_gpt_3_text_davinci_003_downstream_tasksLLM_multimodal
Dataset Card for LLM_multimodal
LLM_multimodal is the official training corpus designed for the LLM D6 model series. It contains a massive, high-quality collection of 300 Billion tokens, carefully curated to balance linguistic diversity, mathematical reasoning, and programming capabilities.
This repository hosts both the raw/processed pre-training data and the instruction-following datasets used for supervised fine-tuning (SFT).
tokenizer is located at LLM_D6 in my… See the full description on the dataset page: https://huggingface.co/datasets/firdavsus/LLM_multimodal.OxfordPets_facebook_opt_30b_LLM_Description_opt30b_downstream_tasks_ViT_L_14
Dataset Card for "OxfordPets_facebook_opt_30b_LLM_Description_opt30b_downstream_tasks_ViT_L_14"
More Information needed
OxfordPets_Multimodal_Fatima_opt_175b_LLM_Description_opt175b_downstream_tasks_ViT_L_14
Dataset Card for "OxfordPets_Multimodal_Fatima_opt_175b_LLM_Description_opt175b_downstream_tasks_ViT_L_14"
More Information needed
OxfordPets_facebook_opt_350m_LLM_Description_gpt3_downstream_tasks_ViT_L_14
Dataset Card for "OxfordPets_facebook_opt_350m_LLM_Description_gpt3_downstream_tasks_ViT_L_14"
More Information needed
LLM_Description_Vocab_gpt_3_text_davinci_003llm_multiModal
