quickb
Datasets
All datasets matching “quickb”Quickbooks_LLM_Training_SampleDataset Summary
(Not affiliated with Quickbooks)
A comprehensive training dataset sample of realistic, synthetic generated QuickBooks Online API interaction scenarios, specifically designed for training AI assistants, chatbots, and automation tools on QuickBooks accounting workflows. Each scenario includes natural language user requests, properly formatted API calls, realistic QuickBooks API responses, and human-readable summaries covering the complete lifecycle of customers, invoices… See the full description on the dataset page: https://huggingface.co/datasets/CJJones/Quickbooks_LLM_Training_Sample.quickb-qa
quickb-qa
Generated using QuicKB, a tool developed by Adam Lucek.
QuicKB optimizes document retrieval by creating fine-tuned knowledge bases through an end-to-end pipeline that handles document chunking, training data generation, and embedding model optimization.
Question Generation
Model: openai/gpt-4o-mini
Deduplication threshold: 0.85
Results:
Total questions generated: 13292
Questions after deduplication: 11116
Dataset Structure
anchor: The generated… See the full description on the dataset page: https://huggingface.co/datasets/Mitchell6024/quickb-qa.quickb-kb
quickb-kb
Generated using QuicKB, a tool developed by Adam Lucek.
QuicKB optimizes document retrieval by creating fine-tuned knowledge bases through an end-to-end pipeline that handles document chunking, training data generation, and embedding model optimization.
Chunking Configuration
Chunker: RecursiveTokenChunker
Parameters:
chunk_size: 400
chunk_overlap: 0
length_type: 'character'
separators: ['\n\n', '\n', '.', '?', '!', ' ', '']
keep_separator: True… See the full description on the dataset page: https://huggingface.co/datasets/AdamLucek/quickb-kb.quickb
quickb
Generated using QuicKB, a tool developed by Adam Lucek.
QuicKB optimizes document retrieval by creating fine-tuned knowledge bases through an end-to-end pipeline that handles document chunking, training data generation, and embedding model optimization.
Chunking Configuration
Chunker: RecursiveTokenChunker
Parameters:
chunk_size: 400
chunk_overlap: 50
length_type: 'character'
separators: ['\n\n', '\n', '.', '?', '!', ' ', '']
keep_separator: True… See the full description on the dataset page: https://huggingface.co/datasets/Mitchell6024/quickb.quickb-qa
quickb-qa
Generated using QuicKB, a tool developed by Adam Lucek.
QuicKB optimizes document retrieval by creating fine-tuned knowledge bases through an end-to-end pipeline that handles document chunking, training data generation, and embedding model optimization.
Question Generation
Model: openai/gpt-4o-mini
Deduplication threshold: 0.85
Results:
Total questions generated: 80
Questions after deduplication: 80
Dataset Structure
anchor: The generated… See the full description on the dataset page: https://huggingface.co/datasets/onecd2000/quickb-qa.quickb-kb
quickb-kb
Generated using QuicKB, a tool developed by Adam Lucek.
QuicKB optimizes document retrieval by creating fine-tuned knowledge bases through an end-to-end pipeline that handles document chunking, training data generation, and embedding model optimization.
Chunking Configuration
Chunker: RecursiveTokenChunker
Parameters:
chunk_size: 400
chunk_overlap: 0
length_type: 'character'
separators: ['\n\n', '\n', '.', '?', '!', ' ', '']
keep_separator: True… See the full description on the dataset page: https://huggingface.co/datasets/Mdean77/quickb-kb.
