CoolFace
9 results

llm-extraction

vangheem /llm-ner-extraction Introduction This dataset is an extraction of NER data from the wikipedia dataset. This can be used to fine tune llm models for NER extraction. text10K<n<100K0 likes177 downloads1y agoHugging Facellm-jp /extraction-wiki-ja extraction-wiki-ja This repository provides an instruction-tuning dataset developed by LLM-jp, a collaborative project launched in Japan. This is a Japanese instruction-tuning dataset tailored for information extraction and structuring from Japanese Wikipedia text. The dataset consists of instruction–response pairs automatically generated from Japanese Wikipedia articles. Instructions are created by prompting Qwen/Qwen2.5-32B-Instruct with passages from Wikipedia, and the… See the full description on the dataset page: https://huggingface.co/datasets/llm-jp/extraction-wiki-ja.texttext-generation100K<n<1M4 likes154 downloads1y agoHugging Facescholarly360 /contracts-extraction-instruction-llm-experiments Dataset Card for "contracts-extraction-instruction-llm-experiments" More Information needed text1K<n<10K7 likes41 downloads3y agoHugging FaceBabelscape /LLM-Oasis_claim_extraction Babelscape/LLM-Oasis_claim_extraction Dataset Description LLM-Oasis_claim_extraction is part of the LLM-Oasis suite and contains text-claim pairs extracted from Wikipedia pages. It provides the data used to train the claim extraction system described in Section 3.1 of the LLM-Oasis paper. Please refer to our GitHub repository for more information on the overall data generation pipeline of LLM-Oasis. Features title: The title of the Wikipedia page. text: A… See the full description on the dataset page: https://huggingface.co/datasets/Babelscape/LLM-Oasis_claim_extraction.text10K<n<100K6 likes36 downloads2y agoHugging Facemath-extraction-comp /deepseek-ai__deepseek-llm-7b-chattabular1K<n<10K0 likes18 downloads2y agoHugging Facemath-extraction-comp /sci-m-wang__deepseek-llm-7b-chat-sa-v0.1tabular1K<n<10K0 likes17 downloads2y agoHugging Face