Convence/ParseEmbed
ParseEmbed Hard, parse-sensitive retrieval evaluation for embedding models. ParseEmbed is a compact benchmark for embedding models. It tests whether a model can retrieve the exact correct document when hard negatives share nearly all surface tokens with the answer. Tasks Task ID Split What it measures mean mean Semantic scope, negation, numeric values, temporal conditions, and exception handling text_formatting text_formatting Meaning carried by… See the full description on the dataset page: https://huggingface.co/datasets/Convence/ParseEmbed.
575
Upload Benchmark
Upload Benchmark
initial commit
