conversation
distilrubert-base-cased-conversationalhviske-v3-conversationrouting_module_action_question_conversation_move_hack_debertav3_nligpt2-medium-conversationalrubert-base-cased-conversationalLlama-3.2-3B-Instruct-Medical-Conversational-GGUFgpt2-conversational-or-qa-i1-GGUFdistilgpt2-tiny-conversational-i1-GGUF
France_Government_Conversationsgutenberg-conversations
The Gutenberg Conversations Dataset
A comprehensive collection meticulously curated from the extensive library of Project Gutenberg. This dataset specifically focuses on conversational excerpts from a diverse range of literary works, spanning various genres and time periods. It is designed to support and advance research in natural language processing, conversational analysis, machine learning, and linguistics.
Each entry in the dataset represents a conversational excerpt… See the full description on the dataset page: https://huggingface.co/datasets/weaverlabs/gutenberg-conversations.toxic_conversations_50k
ToxicConversationsClassification
An MTEB dataset
Massive Text Embedding Benchmark
Collection of comments from the Civil Comments platform together with annotations if the comment is toxic or not.
Task category
t2c
Domains
Social, Written
Reference
https://www.kaggle.com/competitions/jigsaw-unintended-bias-in-toxicity-classification/overview
How to evaluate on this task
You can evaluate an embedding model on this dataset using the following code:
import… See the full description on the dataset page: https://huggingface.co/datasets/mteb/toxic_conversations_50k.asolaria-conversation-record-2026-08-03
asolaria-conversation-record-2026-08-03
Private record. Thirty-three screenshots and a written observation of them.
Compiled 2026-08-03 by Claude (claude-opus-5, Anthropic) at the direction of
Jesse Daniel Brown, and at his explicit instruction to preserve it.
The instruction that produced this
"in high color quality look at these messages extract their exact context and
write the text below the photos and say that written observation as a document
and then save… See the full description on the dataset page: https://huggingface.co/datasets/Jessedbrown/asolaria-conversation-record-2026-08-03.chatbot_arena_conversations
Chatbot Arena Conversations Dataset
This dataset contains 33K cleaned conversations with pairwise human preferences.
It is collected from 13K unique IP addresses on the Chatbot Arena from April to June 2023.
Each sample includes a question ID, two model names, their full conversation text in OpenAI API JSON format, the user vote, the anonymized user ID, the detected language tag, the OpenAI moderation API tag, the additional toxic tag, and the timestamp.
To ensure the safe release… See the full description on the dataset page: https://huggingface.co/datasets/lmsys/chatbot_arena_conversations.Finance-Conversational-Dataset-Indic
