CoolFace
20 results

bnn

BNNT /mozi_general_instructions_3mSources are listed below: Chinese General Instruction 2000k BELLE https://huggingface.co/datasets/BelleGroup/train_2M_CN English generic instruction 52k alpaca-gpt4 https://github.com/Instruction-Tuning-with-GPT-4/GPT-4-LLM Chinese generic dialog instructions 800k BELLE https://huggingface.co/datasets/BelleGroup/multiturn_chat_0.8M English Universal Dialog Instruction 94k sharegpt_vicuna https://huggingface.co/datasets/jeffwan/sharegpt_vicuna Chinese-English-Japanese Universal Command 49k… See the full description on the dataset page: https://huggingface.co/datasets/BNNT/mozi_general_instructions_3m.text1M<n<10M3 likes78 downloads3y agoHugging FaceBNNT /PatentMatchtext1K<n<10K3 likes76 downloads3y agoHugging FaceBNNT /IPQA QA evaluation dataset in intellectual property The IPQA contains questions in seven languages, and the 100 data items include 35 each in Chinese and English, and 6 each in Spanish, Japanese, German, French, and Russian. textquestion-answeringn<1K2 likes61 downloads3y agoHugging Facethedeba /bnnews Bengali News Corpus (2023–2026) Dataset Summary The Bengali News Corpus (2023–2026) is a large-scale, high-quality monolingual Bengali dataset consisting of 295,020 cleaned news articles scraped from Prothom Alo (https://www.prothomalo.com), Bangladesh's largest Bengali-language daily newspaper. The dataset spans over 3.5 years of comprehensive news reporting (January 2023 to August 2026) across various domains including National News, Politics, World News… See the full description on the dataset page: https://huggingface.co/datasets/thedeba/bnnews.texttext-classification100K<n<1M0 likes59 downloads1mo agoHugging FaceGwendalTsang /repro-bnn-data0 likes51 downloads2mo agoHugging FaceBNNT /mozi_IP_instructions4 likes49 downloads3y agoHugging Face