Everyday-Conversations
SmolLM2-FT-smoltalk-everyday-conversations-GGUFSmolLM2-360M-SFT-everyday-conversationsDeepSeek-R1-Distill-Qwen-1.5B-finetuned-smoltalk-everyday-conversationssmollm2-highlish-Hinglish-Everyday-Conversations-1MSmolLM2-FT-smoltalk-everyday-conversationssmollm2-highlish-Hinglish-Everyday-Conversations-1M_V3smollm2-highlish-Hinglish-Everyday-Conversations-1M_V2SmolLM2-135M-sft-finetuned-smoltalk-everyday-conversations
everyday-conversations-llama3.1-2k
Everyday conversations for Smol LLMs finetunings
This dataset contains 2.2k multi-turn conversations generated by Llama-3.1-70B-Instruct. We ask the LLM to generate a simple multi-turn conversation, with 3-4 short exchanges, between a User and an AI Assistant about a certain topic.
The topics are chosen to be simple to understand by smol LLMs and cover everyday topics + elementary science. We include:
20 everyday topics with 100 subtopics each
43 elementary science topics with 10… See the full description on the dataset page: https://huggingface.co/datasets/HuggingFaceTB/everyday-conversations-llama3.1-2k.Hinglish-Everyday-Conversations-1M
Dataset Card for Hinglish Everyday Conversations Dataset
A synthetically created Hinglish-based dataset of 2 columns where every row represents a unique conversation between 2 people in Hinglish about Everyday Life Topics.
Use Model
Access the model made using this dataset: Tiny-Hinglish-Chat-21M
For more information about this model, its training process, or related resources, you can check the GitHub repository Tiny-Hinglish-Chat-21M-Scripts.
Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/Abhishekcr448/Hinglish-Everyday-Conversations-1M.everyday-conversations-tur
Everyday Turkish Conversations
This dataset has everyday conversations in Turkish between user and assistant on various topics. It is inspired by the HuggingFaceTB/everyday-conversations-llama3.1-2k.
License
This dataset is released under the Apache 2.0 License.
everyday-conversations-ita
🇮🇹💬 Everyday Italian Conversations
Inspired by the dataset HuggingFaceTB/everyday-conversations-llama3.1-2k, we generated conversations using the same topics, subtopics, and sub-subtopics as those in the HuggingFaceTB dataset.We slightly adjusted the prompt to produce structured data outputs using Qwen/Qwen2.5-7B-Instruct. Subsequently, we also used the "user" role messages as prompts for google/gemma-2-9b-it.
The result is a dataset of approximately 4.5k… See the full description on the dataset page: https://huggingface.co/datasets/ReDiX/everyday-conversations-ita.everyday-conversations-llama3.1-2k-in-french
Description
French translation of HuggingFaceTB/everyday-conversations-llama3.1-2k.
The original dataset contains 2.2k multi-turn conversations generated by Llama-3.1-70B-Instruct. The LLM have to generate a simple multi-turn conversation, with 3-4 short exchanges, between a User and an AI Assistant about a certain topic.
The topics are chosen to be simple to understand by smol LLMs and cover everyday topics + elementary science. We include:
20 everyday topics with 100 subtopics… See the full description on the dataset page: https://huggingface.co/datasets/CATIE-AQ/everyday-conversations-llama3.1-2k-in-french.everyday-conversations-fused
