conversational
distilrubert-base-cased-conversationalgpt2-medium-conversationalrubert-base-cased-conversationalLlama-3.2-3B-Instruct-Medical-Conversational-GGUFgpt2-conversational-or-qa-i1-GGUFdistilgpt2-tiny-conversational-i1-GGUFLocutusque_-_gpt2-medium-conversational-ggufJoycean0301_-_Llama-3.2-3B-Instruct-Medical-Conversational-gguf
Finance-Conversational-Dataset-IndicNCERT-Conversational-Dataset-IndicNemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1
Dataset Description:
We created an RL dataset for conversational tool-use by utilizing existing expert tool-use trajectories. We pose each assistant step of the trajectory as a separate behavior cloning problem where the policy model is incentivized to match the tool call choices of the expert model. Each trajectory includes the use of tools for authentication, data lookup, servicing (i.e. booking reservations, changing them, getting discounts, etc), and more across 838 different… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-Agentic-Conversational-Tool-Use-Pivot-v1.Law-Conversational-Dataset-IndicCyber-Conversational-Dataset-IndicCoding-Conversational-Dataset-Indic
