web-research
SkyRL-Agent-WebResearch-8B-GGUFweb-attack-detectionSkyRL-Agent-WebResearch-8BQwen2-1.5B-Instruct-1-epochs-research_mashqa_webmd-16a-16rgemma-2b-1-epochs-research_mashqa_webmd-16a-16r-4bitLlama-3.2-3B-Instruct-1-epochs-research_mashqa_webmd-16a-16r-4bitPhi-3-1-epochs-research_mashqa_webmd-16a-16r-4bitPhi-3-mini-4k-instruct-3-epochs-research_mashqa_webmd-16a-16r
essential-web-1t-sample-fdc-partitioned
🌐 Essential-Web: FDC Level-2 Partitioned Dataset
📋 Dataset Description
This dataset contains a 1 trillion token sample from Essential-Web, partitioned by Free Decimal Correspondence (FDC) level-2 categories. Essential-Web is a 24-trillion-token web dataset with extensive document-level metadata designed to enable rapid dataset curation through SQL-like filtering.
🔍 Free Decimal Correspondence (FDC)
The FDC taxonomy is an open classification system… See the full description on the dataset page: https://huggingface.co/datasets/Research-EAI/essential-web-1t-sample-fdc-partitioned.Web-Bench
Web-Bench
English | 中文 README
📖 Overview
Web-Bench is a benchmark designed to evaluate the performance of LLMs in actual Web development. Web-Bench contains 50 projects, each consisting of 20 tasks with sequential dependencies. The tasks implement project features in sequence, simulating real-world human development workflows. When designing Web-Bench, we aim to cover the foundational elements of Web development: Web Standards and Web Frameworks. Given the scale and… See the full description on the dataset page: https://huggingface.co/datasets/bytedance-research/Web-Bench.web-research-coding-5m
Web Research + GitHub + Website Coding Dataset
Version: 1.1.0
Total examples: 5,000,000
Splits
train: 4,750,000
validation: 125,000
test: 125,000
Core capabilities
Web search
Web research
Evidence extraction
Fact verification
Multi-hop research
Multi-layer technical analysis
Architecture analysis
Root-cause analysis
Security analysis
Performance analysis
UX analysis
Design analysis
Refactoring
Code review
Debugging
Website coding
Design systems… See the full description on the dataset page: https://huggingface.co/datasets/Lelonthecodeur/web-research-coding-5m.web-research-trajectories
Web-Research Agent Trajectories
The first open dataset from Assayo — an open rubric and method for judging the quality
of AI agent trajectories. (The name is from assay*: to test the purity of a metal.)*
An open rubric and a small, hand-built gold set for judging multi-step web-research
agent trajectories. A trajectory is the full record of an agent solving one task by
searching the web, reading sources, and answering with citations — the
think → act → observe → repeat →… See the full description on the dataset page: https://huggingface.co/datasets/Assayo/web-research-trajectories.web-attack-detectionThe dataset contains 625,904 attack payload samples, with 294,771 labeled as 1 and 331,129 labeled as 0, including SQL injection, XSS, command injection, and other vulnerabilities.
final-new-web-ret
