datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Cifer-Fraud-Detection-Dataset-AF
📊 Cifer Fraud Detection Dataset
🧠 Overview
The Cifer-Fraud-Detection-Dataset-AF is a high-fidelity, fully synthetic dataset created to support the development and benchmarking of privacy-preserving, federated, and decentralized machine learning systems in financial fraud detection.
This dataset draws structural inspiration from the PaySim simulator, which was built using aggregated mobile money transaction data from a real financial provider operating in 14+ countries.… See the full description on the dataset page: https://huggingface.co/datasets/CiferAI/Cifer-Fraud-Detection-Dataset-AF.msc-cifar100
MSC — Minimum Sufficient Compute
Artifacts for Is Compute Difficulty Architecture-Agnostic? Measuring and
Distilling Per-Sample Minimum Sufficient Computation.
Generated 2026-08-06T03:56:37Z by msc_lib v1.0.0.
Repositories
Shanmuk4622/msc-cifar100 — everything, one folder per run
What MSC is
The smallest cost-normalised configuration at which a network's decision has
stably settled to its full-compute decision, defined uniformly over depth… See the full description on the dataset page: https://huggingface.co/datasets/Shanmuk4622/msc-cifar100.semasia-cifar100
Latents for cifar100 (timm)
This repository hosts precomputed latent representations (embeddings) extracted from timm image-classification backbones on cifar100, released as part of SEMASIA — a large-scale resource for studying semantic communication, cross-model latent space alignment, and explainability.
Each config corresponds to a single model;
only that model's Parquet files are read on load_dataset.
Usage
Load with datasets and… See the full description on the dataset page: https://huggingface.co/datasets/spaicom-lab/semasia-cifar100.cifar100-224-adv-selfgenCifer-Fraud-Detection-Dataset-AF
📊 Cifer Fraud Detection Dataset
🧠 Overview
The Cifer-Fraud-Detection-Dataset-AF is a high-fidelity, fully synthetic dataset created to support the development and benchmarking of privacy-preserving, federated, and decentralized machine learning systems in financial fraud detection.
This dataset draws structural inspiration from the PaySim simulator, which was built using aggregated mobile money transaction data from a real financial provider operating in 14+ countries.… See the full description on the dataset page: https://huggingface.co/datasets/nithi060488/Cifer-Fraud-Detection-Dataset-AF.mmmi-dag1-2modalities-cifar10E2AM_ConvNeXtV2_CIFAR10spotlight-cifar100-enrichment
Dataset Card for "spotlight-cifar100-enrichment"
More Information needed
Cifer-Fraud-Detection-Dataset-AF
📊 Cifer Fraud Detection Dataset
🧠 Overview
The Cifer-Fraud-Detection-Dataset-AF is a high-fidelity, fully synthetic dataset created to support the development and benchmarking of privacy-preserving, federated, and decentralized machine learning systems in financial fraud detection.
This dataset draws structural inspiration from the PaySim simulator, which was built using aggregated mobile money transaction data from a real financial provider operating in 14+ countries.… See the full description on the dataset page: https://huggingface.co/datasets/Durgesh111/Cifer-Fraud-Detection-Dataset-AF.CIFAR-10fsgld-cifar100n-reproCifer-Fraud-Detection-Dataset-AF
📊 Cifer Fraud Detection Dataset
🧠 Overview
The Cifer-Fraud-Detection-Dataset-AF is a high-fidelity, fully synthetic dataset created to support the development and benchmarking of privacy-preserving, federated, and decentralized machine learning systems in financial fraud detection.
This dataset draws structural inspiration from the PaySim simulator, which was built using aggregated mobile money transaction data from a real financial provider operating in 14+… See the full description on the dataset page: https://huggingface.co/datasets/pri092005/Cifer-Fraud-Detection-Dataset-AF.mlcd-mteb-cifar-eval
MLCD vs CLIP on MTEB CIFAR-10/100: integration and evaluation
Evaluation results accompanying the MTEB integration of two MLCD image encoders
(PR #5406, resolving
issue #2571).
Two DeepGlint-AI MLCD encoders were integrated into MTEB, verified against the
reference implementation, and evaluated on the official MTEB CIFAR-10/CIFAR-100
image-classification tasks alongside size-matched OpenAI CLIP baselines.
What was measured
Official MTEB image classification: 5… See the full description on the dataset page: https://huggingface.co/datasets/b4ph/mlcd-mteb-cifar-eval.svae-freckles-4096-cifar10
SVAE Freckles 4096 — CIFAR-10 Omega Tokens
Precomputed spectral decomposition of CIFAR-10 through Freckles v41 (256×256), a frozen Spectral Variational Autoencoder trained exclusively on synthetic noise.
Each CIFAR-10 image is resized to 256×256, decomposed into 4096 patches (4×4 each), and passed through Freckles' encoder → SVD bottleneck. The 4 singular values per patch are stored as a (4, 64, 64) omega map — a 4-channel spatial representation of spectral energy.… See the full description on the dataset page: https://huggingface.co/datasets/AbstractPhil/svae-freckles-4096-cifar10.cifar-10Silicon_life_CIFE2AM_EffNetV2_CIFAR10cifar64-latentschemnlp-mp-cifs
Dataset Card for "chemnlp-mp-cifs"
More Information needed
cifar10-stats
CIFAR-10 CNN Layerwise Training Statistics
Dataset Description
This dataset contains layer-wise training statistics for a convolutional network trained on CIFAR-10, together with the corresponding test/acc.
Each row is one point at loss landscape. The features are computed on the last training batch of 1024 samples before the end of an epoch, and test/acc is measured immediately after that epoch.
The dataset includes statistics for several convolutional layers and the… See the full description on the dataset page: https://huggingface.co/datasets/Fullfix/cifar10-stats.eval_metrics_5000_generations_test_text2struc_text_code_cif_1116eval_metrics_5000_generations_test_text2struc_text_code_cif_1116CIFAR-100cifar10_TRADES_pgd_ema_ccm_False_ccr_False_detailed_resultsE2AM_MobileViTv2_CIFAR10cifer-fraud-detection-mini-dataset
📊 Cifer Fraud Detection Mini Dataset
🧠 Overview
(cifer-fraud-detection-mini-dataset)
The Cifer-Fraud-Detection-Mini-Dataset is a lightweight sample containing 20 transaction records, extracted from the full 21 million-row Cifer-Fraud-Detection-Dataset-AF. It is designed for quick experimentation of encrypted model training with Fully Homomorphic Encryption (FHE).
Though small in size, this mini dataset retains the original schema and data structure inspired by… See the full description on the dataset page: https://huggingface.co/datasets/CiferAI/cifer-fraud-detection-mini-dataset.E2AM_MobileViTv2_CIFAR100Cifer-Fraud-Detection-Dataset-AF
📊 Cifer Fraud Detection Dataset
🧠 Overview
The Cifer-Fraud-Detection-Dataset-AF is a high-fidelity, fully synthetic dataset created to support the development and benchmarking of privacy-preserving, federated, and decentralized machine learning systems in financial fraud detection.
This dataset draws structural inspiration from the PaySim simulator, which was built using aggregated mobile money transaction data from a real financial provider operating in 14+… See the full description on the dataset page: https://huggingface.co/datasets/KadenParker1/Cifer-Fraud-Detection-Dataset-AF.cifar10-speedrunE2AM_EffNetV2_CIFAR100
