datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mlcd-mteb-cifar-eval
MLCD vs CLIP on MTEB CIFAR-10/100: integration and evaluation
Evaluation results accompanying the MTEB integration of two MLCD image encoders
(PR #5406, resolving
issue #2571).
Two DeepGlint-AI MLCD encoders were integrated into MTEB, verified against the
reference implementation, and evaluated on the official MTEB CIFAR-10/CIFAR-100
image-classification tasks alongside size-matched OpenAI CLIP baselines.
What was measured
Official MTEB image classification: 5… See the full description on the dataset page: https://huggingface.co/datasets/b4ph/mlcd-mteb-cifar-eval.mlcodesbymeMLConceptClassifier
MLConceptClassifier
tags: machine learning, classification, hackernow
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'MLConceptClassifier' dataset is designed to assist an ML practitioner in training a model to differentiate between machine learning-related posts and regular, non-technical topics. The dataset comprises of text from HackerNews posts, with each entry being labeled according to its relevance to AI/ML concepts or… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/MLConceptClassifier.MLCourseDataMLCS-test-1MLCS code-mix research dataset
Dataset Details
Dataset Description
This dataset is the MLCS Data Card for CodeMix research. It includes multilingual code-mixed content, intended to support research in multilingual language modeling, classification, and other NLP tasks involving language switching.
Curated by: Chippo Sekkabanja
Funded by: KDD
Shared by [optional]: [More Information Needed]
Language(s) (NLP): English, Chinese
License: [More Information Needed]
Dataset Sources [optional]… See the full description on the dataset page: https://huggingface.co/datasets/chipotswift/MLCS-test-1.
