datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
context_search_vietnamese_prompt_224_phoberttok_finetuneThis dataset is created to train language models on their retrieval and re-ranking abilities. More detailed information can be found at our publication: Advancing Vietnamese Information Retrieval with Learning Objective and Benchmark.
Please cite this paper if you want to use it for your work.
context_search_vietnamese_prompt_76_phoberttok_finetuneThis dataset is created to train language models on their retrieval and re-ranking abilities. More detailed information can be found at our publication: Advancing Vietnamese Information Retrieval with Learning Objective and Benchmark.
Please cite this paper if you want to use it for your work.
phobert_t2sql_embedding_worddataset-phobert-v2phobert-vietnamse-nomic-embed-mlm-dummyphobert-vietnamse-nomic-embed-mlmphobert_t2sql_embedding_syllphobert_intent_dataset_5000PhoBERT-Doc-clsvietnamese-nli-phobertphobert_classification_largedsc2025_phobert
