CoolFace
20 results

knowledge-base

rl-llm-wiki /knowledge-base RL-for-LLMs Wiki An expert-level, citation-backed knowledge base on reinforcement learning for large language models — RLHF, DPO and offline preference optimization, reward modeling, RLVR and reasoning, training systems, and the failure modes — built collaboratively by autonomous agents. Each topic article is a deep dive written so you can learn the topic from it without reading the underlying papers, with every non-obvious claim cited to a source. Every change lands through a… See the full description on the dataset page: https://huggingface.co/datasets/rl-llm-wiki/knowledge-base.17 likes88k downloads2mo agoHugging Faceattention-wiki /knowledge-base Attention Wiki — a living knowledge base on LLM attention A citation-backed tree of knowledge about attention in large language models, built collaboratively by autonomous agents. Agents read papers, blogs, and model cards; distill them into structured, provenance-tracked pages; and reconcile where sources agree, disagree, or leave a question open. Every change lands through a reviewed Pull Request — so the canonical wiki is curated, not just accumulated. Contributing? Read… See the full description on the dataset page: https://huggingface.co/datasets/attention-wiki/knowledge-base.0 likes3.9k downloads3mo agoHugging FaceJohn6666 /knowledge_base_md_for_rag_1 HF Knowledge-Base Markdown Collection This repository contains a collection of Markdown-based knowledge bases generated from: User-provided notes and attachments Hugging Face Docs, Blog, and Papers Model / Dataset / Space cards Discussions, GitHub issues, forums, and other vetted community sources Each .md file is intended to be a self-contained knowledge pack that can be used as LLM context for RAG or prompt-attachment workflows (e.g. ChatGPT, Hugging Face Inference… See the full description on the dataset page: https://huggingface.co/datasets/John6666/knowledge_base_md_for_rag_1.6 likes3.3k downloads1mo agoHugging FaceTEHBESTEUR /ciel-knowledge-base1 likes1.8k downloads37m agoHugging Faceallen-1231 /Knowledge-Baseimagen<1K0 likes805 downloads5mo agoHugging Faceabksunited /knowledge-base RL-for-LLMs Wiki An expert-level, citation-backed knowledge base on reinforcement learning for large language models — RLHF, DPO and offline preference optimization, reward modeling, RLVR and reasoning, training systems, and the failure modes — built collaboratively by autonomous agents. Each topic article is a deep dive written so you can learn the topic from it without reading the underlying papers, with every non-obvious claim cited to a source. Every change lands through a… See the full description on the dataset page: https://huggingface.co/datasets/abksunited/knowledge-base.0 likes455 downloads3mo agoHugging Face