golden-dataset
japanese-casual-conversational-speech-golden-dataset-preview
Japanese Casual Conversational Speech Golden Dataset (Preview)
💼 Commercial License & Full Access
This repository contains a limited preview. The full 60-hour dataset collected via the "Kataro" app is available for commercial use, ASR benchmarking, and Spoken Dialogue Model fine-tuning.
To purchase the full dataset, please contact us:
👉 Email: info@hth-inc.com
👉 Website: https://hth-inc.com/business
🌟 4 Reasons to Choose This Dataset… See the full description on the dataset page: https://huggingface.co/datasets/HTH-inc/japanese-casual-conversational-speech-golden-dataset-preview.financial-golden-dataset-v1ragas-golden-dataset
Dataset Card for the ragas-golden-dataset
Dataset Description
The RAGAS Golden Dataset is a synthetically generated question-answering dataset designed for evaluating Retrieval Augmented Generation (RAG) systems. It contains high-quality question-answer pairs derived from academic papers on AI agents and agentic AI architectures.
Dataset Summary
This dataset was generated using Prefect and the RAGAS TestsetGenerator framework, which creates synthetic questions… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/ragas-golden-dataset.golden-dataset-2.1ragas-golden-dataset-documents
Dataset Card for RAGAS Golden Dataset Documents
A small, mixed‐format corpus to compare PDF, API, and web‐based document loader output from the LangChain ecosystem.
The code to run the Prefect prefect_docloader_pipeline.py pipeline is available in the RAGAS Golden Dataset Pipeline repository.
While several enhancements are planned for future iterations, the hands-on insights gained from this grassroots exploration of document loader behaviors proved too valuable -- things that… See the full description on the dataset page: https://huggingface.co/datasets/dwb2023/ragas-golden-dataset-documents.azure-ai-engineer-golden-dataset
