datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
asolaria-conversation-record-2026-08-03
asolaria-conversation-record-2026-08-03
Private record. Thirty-three screenshots and a written observation of them.
Compiled 2026-08-03 by Claude (claude-opus-5, Anthropic) at the direction of
Jesse Daniel Brown, and at his explicit instruction to preserve it.
The instruction that produced this
"in high color quality look at these messages extract their exact context and
write the text below the photos and say that written observation as a document
and then save… See the full description on the dataset page: https://huggingface.co/datasets/Jessedbrown/asolaria-conversation-record-2026-08-03.hle-no-img-conversational-formatios_emulated_criminal_conversations_helios
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
Curated by: Lewis Taylor
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Dataset Sources [optional]
Repository: [More Information Needed]
Paper [optional]: [More Information Needed]
Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/nodrog1061/ios_emulated_criminal_conversations_helios.MMRC_Real_World_Conversation
MMRC - Multi-Modal Open-Ended Conversation Dataset
Overview:
MMRC is a benchmark dataset designed for evaluating Multi-Modal Large Language Models (MLLMs) in open-ended, multi-turn conversations. It provides diverse, real-world conversational data that integrates both textual and visual modalities, aiming to push the boundaries of MLLM performance in practical settings.
Dataset Details:
The MMRC dataset is composed of multi-turn conversations with integrated… See the full description on the dataset page: https://huggingface.co/datasets/WUUE/MMRC_Real_World_Conversation.kaggle-notebooks-conversationsjapanese-photos-conversation-qwen3vlJapanese_Photo_conversation_cleaned
Japanese Photo Conversation (Cleaned)
A cleaned and organized Japanese photo conversation dataset for training vision-language models on Japanese photo description and visual question answering tasks.
Dataset Description
This dataset is a cleaned and reorganized version combining data from:
llm-jp/japanese-photos-conversation
ThePioneer/japanese-photos
We thank the original authors for their excellent work in collecting and annotating these datasets.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/WayBob/Japanese_Photo_conversation_cleaned.clever_conversation_vqa_verfNondeterministic_Finite_Automata_OCR_Conversations_Format2agri-llava-present-conversations_part02Dataset chunk 02 for EYEDOL/agri-llava-present-conversations
This chunk contains 10000 images (in images/) and a CSV file data.csv with columns:
image: basename of the image file
text: conversation/annotation
Generated programmatically.
conversation_hall_binary_v1Nondeterministic_Finite_Automata_OCR_Conversations_Formatconversation_hall_choice_v2conversation_hall_completion_v3xray-images-conversationsmovie-conversationagri-llava-present-conversations_part01Dataset chunk 01 for EYEDOL/agri-llava-present-conversations
This chunk contains 10000 images (in images/) and a CSV file data.csv with columns:
image: basename of the image file
text: conversation/annotation
Generated programmatically.
agri-llava-present-conversations_part03Dataset chunk 03 for EYEDOL/agri-llava-present-conversations
This chunk contains 10000 images (in images/) and a CSV file data.csv with columns:
image: basename of the image file
text: conversation/annotation
Generated programmatically.
FineVision-Conversations_64trcrag-mm-single-turn-public-conversations
