This dataset is suitable for evaluating language-identificatioan model
datasetsANDmodels/10-languages-samples · main · files are served by the source, never re-hosted here