jhu-clsp/mFollowIR-cross-lingual-parquet-mteb
mFollowIR-cross-lingual-mteb This is a new version of the mFollowIR-cross-lingual dataset modified to fit the new MTEB format. Restructured queries to include both original and changed versions Separated instructions into a dedicated configuration Reorganized qrels into default (original) and qrel_diff configurations Dataset Structure The dataset contains the following configurations: Language: fas corpus-fas: Original corpus documents… See the full description on the dataset page: https://huggingface.co/datasets/jhu-clsp/mFollowIR-cross-lingual-parquet-mteb.
mFollowIR-cross-lingual-mteb
This is a new version of the mFollowIR-cross-lingual dataset modified to fit the new MTEB format.
- Restructured queries to include both original and changed versions
- Separated instructions into a dedicated configuration
- Reorganized qrels into default (original) and qrel_diff configurations
Dataset Structure
The dataset contains the following configurations:
Language: fas
- corpus-fas: Original corpus documents
- queries-fas: Queries with both original and changed versions
- instruction-fas: Instructions for both original and changed queries
- default-fas: Original relevance judgments
- qrel_diff-fas: Changes in relevance judgments
- top_ranked-fas: Top ranked documents for each query
Language: rus
- corpus-rus: Original corpus documents
- queries-rus: Queries with both original and changed versions
- instruction-rus: Instructions for both original and changed queries
- default-rus: Original relevance judgments
- qrel_diff-rus: Changes in relevance judgments
- top_ranked-rus: Top ranked documents for each query
Language: zho
- corpus-zho: Original corpus documents
- queries-zho: Queries with both original and changed versions
- instruction-zho: Instructions for both original and changed queries
- default-zho: Original relevance judgments
- qrel_diff-zho: Changes in relevance judgments
- top_ranked-zho: Top ranked documents for each query
