datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OS-Critic-Benchor-knowledge-copilot-corpus
OR Knowledge Copilot Corpus
Multi-layer operations-research knowledge base used by OR Knowledge Copilot.
Each instance is stored as six chunks:
Natural language
Mathematical formulation
Pyomo template
MiniZinc template
Solver output
Explanation of binding constraints
Files
chunks.jsonl — retrieval units
qa_pairs.jsonl — labeled questions including out-of-scope abstention cases
benchmark_report.json / eval_results.json — published retrieval metrics
taxonomy.json… See the full description on the dataset page: https://huggingface.co/datasets/alirezaaminzadeh/or-knowledge-copilot-corpus.fast_food_copilot_qa_mini
PropulsionAI LLM Fine-tuning Dataset: Sample Q&A for Co-Pilot
Welcome to the PropulsionAI LLM Fine-tuning Dataset. This dataset is designed for educational purposes to assist users in understanding and experimenting with fine-tuning Language Learning Models (LLMs), such as Llama 2. It comprises a curated set of sample questions and answers aimed at demonstrating the process of fine-tuning an LLM-based copilot.
Dataset Overview
The dataset contains sample questions and… See the full description on the dataset page: https://huggingface.co/datasets/PropulsionAI/fast_food_copilot_qa_mini.
