Algorithmic-Human-Development-Group/Multilingual-Therapy-Dialogues
Dataset Summary Multilingual Therapy Dialogues is a diverse and bilingual dataset consisting of paired dialogues between patients and therapists in both Persian and English. Dataset Statistics Number of samples: 7,179 English: Average tokens per sentence: 101.30 Maximum tokens in a sentence: 939 Average characters per sentence: 567.85 Number of unique tokens: 32,968 Persian: Average tokens per sentence: 100.06 Maximum tokens in a sentence: 1,413… See the full description on the dataset page: https://huggingface.co/datasets/Algorithmic-Human-Development-Group/Multilingual-Therapy-Dialogues.
049
Create README.md
Upload SAT_dataset.csv
initial commit
