CoolFace
20 results

aao

ArtificialAnalysis /AA-Omniscience-Public Public Dataset for AA-Omniscience: Evaluating Cross-Domain Knowledge Reliability in Large Language Models AA-Omniscience-Public contains 600 questions across a wide range of domains used to test a model’s knowledge and hallucination tendencies. Leaderboard and detailed results Paper Introduction We introduce AA-Omniscience, a benchmark dataset designed to measure a model’s ability to both recall factual information accurately across domains, and correctly… See the full description on the dataset page: https://huggingface.co/datasets/ArtificialAnalysis/AA-Omniscience-Public.documentquestion-answeringn<1K50 likes7.9k downloads1mo agoHugging FaceMasterControlAIML /EB1-AAO-Decisions USCIS Administrative Appeals Office (AAO) Decisions Dataset Dataset Summary This dataset contains publicly available case decisions from the U.S. Citizenship and Immigration Services (USCIS) Administrative Appeals Office (AAO). The AAO reviews appeals related to various immigration matters, including: EB-1A (Extraordinary Ability) EB-1B (Outstanding Professors and Researchers) EB-1C (Multinational Executives & Managers) The dataset consists of raw PDF documents… See the full description on the dataset page: https://huggingface.co/datasets/MasterControlAIML/EB1-AAO-Decisions.1K<n<10K9 likes5.8k downloads2y agoHugging Facejosuediazflores /aao-eb1a-decisions AAO EB-1A Extraordinary Ability Decisions — Structured Dataset Structured extractions from 1,466 USCIS Administrative Appeals Office (AAO) non-precedent decisions on EB-1A extraordinary ability petitions (I-140). Each case has been decomposed into structured components by Claude Sonnet for use in fine-tuning legal reasoning models. Part of Project Greenlight — an AI-powered O-1A/EB-1A visa intelligence system. Dataset Description Each JSON file represents one AAO… See the full description on the dataset page: https://huggingface.co/datasets/josuediazflores/aao-eb1a-decisions.text-classification1K<n<10K0 likes270 downloads6mo agoHugging FaceSandhya1912 /AA-Omniscience-Public Public Dataset for AA-Omniscience: Evaluating Cross-Domain Knowledge Reliability in Large Language Models AA-Omniscience-Public contains 600 questions across a wide range of domains used to test a model’s knowledge and hallucination tendencies. Leaderboard and detailed results Paper Introduction We introduce AA-Omniscience, a benchmark dataset designed to measure a model’s ability to both recall factual information accurately across domains, and correctly… See the full description on the dataset page: https://huggingface.co/datasets/Sandhya1912/AA-Omniscience-Public.documentquestion-answeringn<1K0 likes110 downloads4mo agoHugging Faceaa-oswald /worldbank-okr0 likes71 downloads6mo agoHugging Faceonurborasahin /AA-Omniscience-Public Public Dataset for AA-Omniscience: Evaluating Cross-Domain Knowledge Reliability in Large Language Models AA-Omniscience-Public contains 600 questions across a wide range of domains used to test a model’s knowledge and hallucination tendencies. Leaderboard and detailed results Paper Introduction We introduce AA-Omniscience, a benchmark dataset designed to measure a model’s ability to both recall factual information accurately across domains, and correctly abstain when… See the full description on the dataset page: https://huggingface.co/datasets/onurborasahin/AA-Omniscience-Public.documentquestion-answeringn<1K0 likes61 downloads8mo agoHugging Face