rifa
Datasets
All datasets matching “rifa”hani-ar-rifai-192kbps
Hani Ar-Rifai
Part of Maqra, an open, verified archive of verse-by-verse Qur'an recitations mirrored from everyayah.com.
Set
hani-ar-rifai-192kbps
Style
murattal
Riwayah
hafs
Kind
recitation
Bitrate
192 kbps
Ayah files
6236 (2053 MiB)
Verified against the upstream MD5 list
6234
Ayahs absent upstream
0
Upstream folder
Hani_Rifai_192kbps
Files
One MP3 per ayah, named SSSAAA.mp3 (surah 3 digits, ayah 3 digits). 001001.mp3 is… See the full description on the dataset page: https://huggingface.co/datasets/maqra-project/hani-ar-rifai-192kbps.hani-ar-rifai-64kbps
Hani Ar-Rifai
Part of Maqra, an open, verified archive of verse-by-verse Qur'an recitations mirrored from everyayah.com.
Set
hani-ar-rifai-64kbps
Style
murattal
Riwayah
hafs
Kind
recitation
Bitrate
64 kbps
Ayah files
6236 (702 MiB)
Verified against the upstream MD5 list
6234
Ayahs absent upstream
0
Upstream folder
Hani_Rifai_64kbps
Files
One MP3 per ayah, named SSSAAA.mp3 (surah 3 digits, ayah 3 digits). 001001.mp3 is Al-Fatihah… See the full description on the dataset page: https://huggingface.co/datasets/maqra-project/hani-ar-rifai-64kbps.bengali-ocr-synthetic
Bengali OCR Synthetic Dataset
A high-quality synthetic Bengali OCR dataset for fine-tuning vision-language models like DeepSeek-OCR 2. Generated using 100+ professional Bengali Unicode fonts and 13K+ unique Bengali words with advanced text rendering via FreeType and HarfBuzz.
Dataset Overview
Language: Bengali (বাংলা)
Task: Optical Character Recognition (OCR)
Format: Conversation-based (vision-language)
Total Samples: 30,000
Train: 27,007 samples
Validation: 2,993… See the full description on the dataset page: https://huggingface.co/datasets/rifathridoy/bengali-ocr-synthetic.RIFA-Artbook-datasetRIFARSetupMOBILeditForensicULTRA9
