datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
backendbench_tests
TorchBench
The TorchBench suite of BackendBench is designed to mimic real-world use cases. It provides operators and inputs derived from 155 model traces found in TIMM (67), Hugging Face Transformers (45), and TorchBench (43). (These are also the models PyTorch developers use to validate performance.) You can view the origin of these traces by switching the subset in the dataset viewer to ops_traces_models and torchbench for the full dataset.
When running BackendBench, much of the… See the full description on the dataset page: https://huggingface.co/datasets/GPUMODE/backendbench_tests.backend-code-generator-dataset
Backend Code Generation Dataset
Dataset Description
This dataset contains examples for training AI models to generate backend application code. It includes descriptions of backend requirements paired with complete, functional code implementations across multiple frameworks and programming languages.
Dataset Summary
The Backend Code Generation Dataset is designed to train models that can generate complete backend applications from natural language descriptions.… See the full description on the dataset page: https://huggingface.co/datasets/Techta/backend-code-generator-dataset.backend-api-instruction-dataset
Backend & API Development Dataset
Instruction dataset focused on RESTful API design, WebSocket real-time communication, microservices patterns, and API best practices.
Dataset Details
Dataset Description
This is a high-quality instruction-tuning dataset focused on Backend Api topics. Each entry includes:
A clear instruction/question
Optional input context
A detailed response/solution
Chain-of-thought reasoning process
Curated by: CloudKernel.IO… See the full description on the dataset page: https://huggingface.co/datasets/bernabepuente/backend-api-instruction-dataset.my-backend-dataexperts-backendssite_backendbackendbench-qwen-qwen3-codertest_backendbenchredditpolitics11292024Psychiatry_datasetwzzkdatasetpdfversion01152025radarpoliticaldatasetredditscrap11272024redditpoliticsoct2016fullstack-backend-trainingredditpoliticsoct20202021backendbench-z-ai-glm-4.5-airbackendbench-prime-intellect-intellect-3momo_backend_uploadBNEZxtbackendbench-openai-gpt-oss-120bbackendds-prep-backendbackendbench-openai-gpt-5.2smolified-backend-and-devops-buddy
🤏 smolified-backend-and-devops-buddy
Intelligence, Distilled.
This is a synthetic training corpus generated by the Smolify Foundry.
It was used to train the corresponding model draganite/smolified-backend-and-devops-buddy.
📦 Asset Details
Origin: Smolify Foundry (Job ID: 68bd3135)
Records: 1890
Type: Synthetic Instruction Tuning Data
⚖️ License & Ownership
This dataset is a sovereign asset owned by draganite.
Generated via Smolify.ai.
bitjv-backendbackend_serveremotion_backendjElmo22johnedBackend
