datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
agent-challenge
Replit Agent Challenge
For comprehensive details about the challenge, visit our GitHub repository.
Dataset Overview
This dataset comprises a collection of instructions and file states specifically curated for the agent challenge. It is derived from a subset of SWE-Bench-Lite.
Schema Structure
The dataset follows this schema:
- File_before: [Initial state of the file]
- Instructions: [Steps to transform the file to its final state]
- File_after: [Resulting state… See the full description on the dataset page: https://huggingface.co/datasets/replit/agent-challenge.replit-comments-categorized
Dataset Card for [Dataset Name]
Dataset Summary
Comments from Replit's Community, sourced via moderator GraphQL queries and personally labeled :). For use in Replit + Weights and Biases Hackathon.
Supported Tasks and Leaderboards
Text Classification
Languages
English
Dataset Structure
Data Instances
{"label":3,"text":"@KENDALPETERSON\nShut up you dont have a permit to brag."}
Labels
0: General
1: Spam
2: NSFW
3: Harassment… See the full description on the dataset page: https://huggingface.co/datasets/RayhanADev/replit-comments-categorized.kittypaw_replit_ddbb
