momahadi/bangladesh-legal-qa-dataset
Bangladesh Legal QA Dataset: Bangla-English Law and Fine-Tuning The Bangladesh Legal QA Dataset is a bilingual Bangla-English dataset for Bangladesh law question answering, legal NLP, LLM fine-tuning, instruction tuning, and retrieval-augmented generation (RAG). It provides 2,165 context-grounded legal QA records, direct-answer and IRAC chat-format training data, and structured statutory text from six Bangladesh Acts and three schedules. This is the 2,165-record paper-aligned… See the full description on the dataset page: https://huggingface.co/datasets/momahadi/bangladesh-legal-qa-dataset.
cite the paper, not the dataset
Update paper link and citation
Update citation to arXiv paper
Remove legacy project-name mentions from dataset card
Adopt CC BY 4.0 and separate Bar Council benchmark
Use exact Bangladesh Legal QA Dataset repository name
Improve Bangladesh Legal QA dataset discoverability
Rename dataset metadata to Bangladesh Legal QA
Configure dataset views for audit, fine-tuning, IRAC, and Bar Council data
Fix public dataset card formatting and rights wording
Initial BDLEX 2,165-record dataset release
Initial BDLEX 2,165-record dataset release
initial commit
