CoolFace
Datasetpublic

algerian-nlp/DziriAlign

DziriAlign 1,000 preference pairs (prompt, chosen, rejected) for aligning language models with Algerian Darja and its sociocultural norms, from the Algerian NLP Collective. Counted 2026-09-17 via the Hub datasets-server (/info?dataset=algerian-nlp/DziriAlign: 1,000 train rows) and re-counted row-by-row with datasets streaming (load_dataset("algerian-nlp/DziriAlign", split="train", streaming=True): 1,000 rows). The default config answers: when two replies compete, which one… See the full description on the dataset page: https://huggingface.co/datasets/algerian-nlp/DziriAlign.

sourceHugging Facemitupdated 10d agoView on Hugging Face
0likes78downloads
2 commits on main
2cb87d810d ago

Standardise card to Algerian NLP collective dataset template (measured counts, usage, provenance)

ainouche-abderahmane
0c509b511d ago

Duplicate from touati-kamel/DziriAlign

touati-kamel