syntax
Datasets
All datasets matching “syntax”tmmluplus
TMMLU+ : Large scale traditional chinese massive multitask language understanding
We present TMMLU+, a traditional Chinese massive multitask language understanding dataset. TMMLU+ is a multiple-choice question-answering dataset featuring 66 subjects, ranging from elementary to professional level.
The TMMLU+ dataset is six times larger and contains more balanced subjects compared to its predecessor, TMMLU. We have included benchmark results in TMMLU+ from closed-source models and 20… See the full description on the dataset page: https://huggingface.co/datasets/syntaxsynth/tmmluplus.Linguistic-Diagnostics-Syntax
LINDSEA Syntax
LINDSEA Syntax is a linguistic diagnostic from BHASA that evaluates a model's understanding of linguistic phenomena, syntax in particular, for Indonesian.
Supported Tasks and Leaderboards
LINDSEA Syntax is designed for evaluating chat or instruction-tuned large language models (LLMs).
Languages
Indonesian (id)
Dataset Details
LINDSEA Syntax only has an Indonesian (id) split, with additional splits containing fewshot examples. Below… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/Linguistic-Diagnostics-Syntax.Linguistic-Diagnostics-Syntax-Judge
LINDSEA Syntax
LINDSEA Syntax is a linguistic diagnostic from BHASA that evaluates a model's understanding of linguistic phenomena, syntax in particular, for Indonesian.
Supported Tasks and Leaderboards
LINDSEA Syntax is designed for evaluating chat or instruction-tuned large language models (LLMs).
Languages
Indonesian (id)
Dataset Details
Data Sources
Data Source
License
Language/s
Split/s
CC BY 4.0… See the full description on the dataset page: https://huggingface.co/datasets/aisingapore/Linguistic-Diagnostics-Syntax-Judge.coronary-angiography-syntaxmacula-hebrew-syntax
NuBerea MACULA Hebrew Syntax Trees (OT)
Full syntactic tree annotation of the Hebrew Bible from the MACULA Hebrew Linguistic Dataset, packaged as relational tables for computational biblical studies. The dataset covers word-level linguistic annotation (morphology, glosses, lexical semantics), sentence segmentation, and hierarchical syntactic structure (clauses and phrases with their roles and containment relations) over the Westminster Leningrad Codex base text.
This repository… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/macula-hebrew-syntax.macula-sblgnt-syntax
NuBerea MACULA Greek (SBLGNT) Syntax Trees (NT)
Full syntactic tree annotation of the Greek New Testament from the MACULA Greek SBLGNT edition. Relational tables cover word-level tokens with morphological, semantic, and cross-language features; sentence boundaries; word groups (clauses and phrases) with syntactic rules and roles; the word-group hierarchy; and word-group membership. Together they let researchers traverse the full syntax tree of every sentence in the New Testament… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/macula-sblgnt-syntax.
