CoolFace
Datasetpublic

VISAI-AI/nitibench

πŸ‘©πŸ»β€βš–οΈ NitiBench: A Thai Legal Benchmark for RAG [πŸ“„ Technical Report] | [πŸ‘¨β€πŸ’» Github Repository] This dataset provides the test data for evaluating LLM frameworks, such as RAG or LCLM. The benchmark consists of two datasets: NitiBench-CCL NitiBench-Tax πŸ›οΈ NitiBench-CCL Derived from the WangchanX-Legal-ThaiCCL-RAG Dataset, our version includes an additional preprocessing step in which we separate the reasoning process from the final answer. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/VISAI-AI/nitibench.

sourceHugging Facemitupdated 10mo agoView on Hugging Face
8likes411downloads
30 commits on main
9f7569710mo ago

Update bibtex

Pawitsapak
537ac201y ago

Revert to default

Pawitsapak
7f07c251y ago

Drop relevant law that is not included in the answer.

Pawitsapak
0a50ed21y ago

add missing relevant laws on Q0

Pawitsapak
a78826e1y ago

Upload dataset

Pawitsapak
32e26e21y ago

Upload dataset

Pawitsapak
d96fbc01y ago

clean maiyamok and normalize space

Pawitsapak
0b6c3322y ago

Update README.md

chompk
7156d362y ago

cleanup ccl samples

tann9949
ebede8e2y ago

Merge branch 'main' of hf.co:datasets/VISAI-AI/nitibench

tann9949
086f6f92y ago

update license+citation

tann9949
be490e52y ago

Update README.md

chompk
17bfdd92y ago

Update README.md

chompk
ce62b1c2y ago

Update README.md

chompk
41e63912y ago

add license

tann9949
d094b912y ago

update README.md

tann9949
cf6d1802y ago

undo column name change

tann9949
70090b32y ago

update README

tann9949
9f8e1952y ago

rename columns

tann9949
a9b2a392y ago

update columns

tann9949
4a98a192y ago

update README.md

tann9949
a2376a42y ago

add law content

tann9949
cb26d472y ago

fix logo size

tann9949
bd75f5a2y ago

update README.md

tann9949
2e3cb572y ago

update README.md

tann9949
19d72aa2y ago

remove original csv

tann9949
b6793e42y ago

refactor data directory + add parquet files

tann9949
88fc1272y ago

Update README.md

chompk
d816e2c2y ago

Add test split of the benchmark used in our paper. (#1)

Pawitsapak, nonpirat
463e1ef2y ago

initial commit

Pawitsapak