CoolFace
20 results

fda

hazyresearch /based-fdaThis dataset is adapted from the paper Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes. You can learn more about the data collection process there. Please consider citing the following if you use this task in your work: @article{arora2024simple, title={Simple linear attention language models balance the recall-throughput tradeoff}, author={Arora, Simran and Eyuboglu, Sabri and Zhang, Michael and Timalsina, Aman and Alberti, Silas and… See the full description on the dataset page: https://huggingface.co/datasets/hazyresearch/based-fda.textquestion-answering1K<n<10K3 likes8.4k downloads2y agoHugging Facefdac24 /MP3 Still missing .json.gz Nov 20, 17:30 jkenne60 mwebb51 empty file dwang58 pmoore34 Still missing .json.gz Nov 19, 14:00 mwebb51 npatton4 nvanflee sshres25 empty file dwang58, mzg857 Broken pmoore34 Still missing .json.gz Nov 15, 10:46 'hchen73','jkenne60','mwebb51','net','npatton4','nvanflee','rfranqui','sshres25' Broken 'bfitzpa8','edayney', 'lhunte21','pmoore34','mzg857','sdasari7' Still missing .json.gz Nov… See the full description on the dataset page: https://huggingface.co/datasets/fdac24/MP3.0 likes790 downloads2y agoHugging Facefdac24 /MP4 As of Nov 25, 20:00 Missing jkenne60 mwebb51 No sections ccanonac cwitt8 dwang58 ehead3 glapham harshvar hchen73 lhunte21 mzg857 nvanfleepmoore34 rfranqui tcatunca vgopu ygaikwad In some cases you removed all newlines, so need to change regexp to extract section. dwang58 mzg857 pmoore34 In the remaining cases newlines are there as \n, so there must be some other bug. 0 likes483 downloads2y agoHugging FaceFDAbench2026 /FDAbench-Full v1.1 Update (2026-08-06) — multiple split Strengthened the cross-source requirement that multiple-choice tasks are designed around (selecting all correct options should require integrating both the SQL result and the retrieved documents): 264 of 760 tasks were revised, with task IDs, databases, and gold SQL unchanged. Documents-only accuracy drops from 61.5% to 38.7% while full-evidence accuracy stays at 80.6% (3 frontier models, strict exact set match). Diversified the number… See the full description on the dataset page: https://huggingface.co/datasets/FDAbench2026/FDAbench-Full.text1K<n<10K1 likes472 downloads2mo agoHugging FaceSphereLab /FDA_for_NLU0 likes331 downloads11mo agoHugging FaceFDAbench2026 /Fdabench-Lite v1.1 Update (2026-08-06) — multiple split Synced with FDABench-Full v1.1: all 49 multiple-choice tasks now carry the revised question/option text and answer keys. The revision strengthens the cross-source requirement (answering requires combining the SQL result with the retrieved documents), diversifies the number of correct options, removes answer-count hints, and balances option-style signals. Task IDs, databases, and gold SQL are unchanged. See the FDABench-Full dataset card… See the full description on the dataset page: https://huggingface.co/datasets/FDAbench2026/Fdabench-Lite.textn<1K1 likes229 downloads2mo agoHugging Face