MonarchInit/dragon-ai-definition-evals
Results of expert evaluation on definitions generated by LLMs and ontology editors See: https://github.com/monarch-initiative/dragon-ai-results Although this dataset is partially prediction results, the expert evaluations form a dataset that could be used for new AI tasks, specifically: Can we use AI to predict which definitions are accurate, concise, consistent, etc?
115
Results of expert evaluation on definitions generated by LLMs and ontology editors
See: https://github.com/monarch-initiative/dragon-ai-results
Although this dataset is partially prediction results, the expert evaluations form a dataset that could be used for new AI tasks, specifically: Can we use AI to predict which definitions are accurate, concise, consistent, etc?
