CoolFace
Datasetpublic

MCMLXXXI/Bitext-insurance-llm-chatbot-training-dataset

Bitext - Insurance Tagged Training Dataset for LLM-based Virtual Assistants Overview This hybrid synthetic dataset is designed to be used to fine-tune Large Language Models such as GPT, Mistral and OpenELM, and has been generated using our NLP/NLG technology and our automated Data Labeling (DAL) tools. The goal is to demonstrate how Verticalization/Domain Adaptation for the [insurance] sector can be easily achieved using our two-step approach to LLM Fine-Tuning.… See the full description on the dataset page: https://huggingface.co/datasets/MCMLXXXI/Bitext-insurance-llm-chatbot-training-dataset.

sourceHugging Facecdla-sharing-1.0updated 8mo agoView on Hugging Face
0likes16downloads
settings

This repository belongs to MCMLXXXI on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameBitext-insurance-llm-chatbot-training-dataset
visibilitypublic
licencecdla-sharing-1.0
gatedno
ownerMCMLXXXI
Account settings
MCMLXXXI/Bitext-insurance-llm-chatbot-training-dataset · CoolFace