CoolFace
20 results

query

ReadyAi /organic_query_results_dataset2 likes26k downloads1y agoHugging Facekammartina /MA_Query_Expansion_MLT26tabular10K<n<100K0 likes2.4k downloads20h agoHugging Facestair-lab /denoise_eval_query0 likes2.2k downloads1y agoHugging Facemjbommar /opengloss-v1.3-query-examples-flat See also OpenGloss v2.1 (2026-09-07): a deeper release of 109,633 of these headwords — sense-level ids, four reading levels, sense-tagged examples with spans, a judged relation graph, and retrieval supervision — published as a 16-dataset family. v1.3 remains the broader headword list. OpenGloss Query Examples v1.3 (Flattened) Dataset Summary OpenGloss Query Examples is a synthetic dataset of search queries generated for vocabulary terms. Each term has multiple… See the full description on the dataset page: https://huggingface.co/datasets/mjbommar/opengloss-v1.3-query-examples-flat.texttext-generation100K<n<1M0 likes689 downloads18d agoHugging Facebeaverbench /beaver-querygated Dataset Card for beaver-query Homepage and leaderboard | Github repository | Paper Beaver is a holistic framework for evaluating performance on complex, private‑enterprise text‑to‑SQL tasks. This repository includes questions and corresponding annotations. We reserve a portion of the full question set as a private, hidden test set. Each sample contains: id: ID of the question category: one of real, complex query, domain-specific query, domain-specific complex query. real indicates… See the full description on the dataset page: https://huggingface.co/datasets/beaverbench/beaver-query.text1K<n<10K9 likes608 downloads4mo agoHugging Facehotchpotch /wikipedia-multilingual-synthetic-ir-query wikipedia-multilingual-synthetic-ir-query This dataset contains multilingual Wikipedia-derived synthetic query-document pairs for information retrieval training. It was created with the query-crafter-multilingual model, which generates search-like queries from Wikipedia text. The current release contains two different retrieval settings: short_doc: pairs of (query, short document) long_doc: pairs of (query, long document) These two subsets are not generated in the same way… See the full description on the dataset page: https://huggingface.co/datasets/hotchpotch/wikipedia-multilingual-synthetic-ir-query.tabulartext-retrieval10M<n<100M0 likes529 downloads3mo agoHugging Face