singapore
cati-singapore-dataset
CATI Singapore Expressway Traffic Dataset
Real-time vehicle detection data collected from Singapore's 90 LTA traffic cameras using CATI (Context-Aware Traffic Intelligence) — a novel FiLM-conditioned YOLOv11 detector that adapts to environmental conditions in real time.
Dataset Description
This dataset contains per-camera vehicle detection results collected continuously from Singapore's Land Transport Authority (LTA) expressway camera network. Each record captures… See the full description on the dataset page: https://huggingface.co/datasets/SuhxsReddy/cati-singapore-dataset.Nemotron-Personas-Singapore
Nemotron-Personas-Singapore
A compound AI approach to personas grounded in real-world distributions
Dataset Overview
Nemotron-Personas-Singapore is an open-source (CC BY 4.0) dataset of synthetically-generated personas. This dataset is grounded in real-world demographic, geographic and personality trait distributions in Singapore to capture the diversity and richness of the Singaporean population. It is a variant of Nemotron-Personas-USA, and the first Singaporean… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-Personas-Singapore.multimodal_meme_classification_singapore
Dataset Card for Offensive Memes in Singapore Context
Dataset Details
Dataset Description
This dataset is a collection of memes from various existing datasets, online forums, and freshly scrapped contents. It contains both global-context memes and Singapore-context memes, in different splits. It has textual description and a label stating if it is offensive under Singapore society's standards.
Curated by: Cao Yuxuan, Wu Jiayang, Alistair Cheong, Theodore Lee… See the full description on the dataset page: https://huggingface.co/datasets/aliencaocao/multimodal_meme_classification_singapore.singaporean-judicial-keywords
Singaporean Judicial Keywords 🏛️
Singaporean Judicial Keywords by Isaacus is a challenging legal information retrieval evaluation dataset consisting of 500 catchword-judgment pairs sourced from the Singapore Judiciary.
Uniquely, the keywords in this dataset are real-world annotations created by subject matter experts, namely, Singaporean law reporters, as opposed to being constructed ex post facto by third parties.
Additionally, unlike standard keyword queries, judicial catchwords… See the full description on the dataset page: https://huggingface.co/datasets/isaacus/singaporean-judicial-keywords.reprocessed_singapore_national_speech_corpus
Dataset Card for Reprocessed National Speech Corpus
NOTE: This is an Reprocessed version KaraKaraWitch from Recursal.The official download can be found here.
Dataset Details
Dataset Description
Dataset Description:
The National Speech Corpus (NSC) is the first large-scale Singapore English corpus, sponsored by the Info-communications and Media Development Authority (IMDA) of Singapore. The objective is to serve as a primary resource of open speech data for… See the full description on the dataset page: https://huggingface.co/datasets/recursal/reprocessed_singapore_national_speech_corpus.ipfs_singapore_laws_ir
Singapore legislation IR (CID-keyed sparse GraphRAG)
Research retrieval release of endomorphosis/ipfs_singapore_laws (revision ee26bc9d91a5e7c66a88a378aec230543e940993) packaged as
country-laws-ir-graphrag/v1 (layout family skillcenter-huggingface-release/v3 / publicus-ir).
Not legal advice. This is a research snapshot. The official gazette /
authentic source of Singapore prevails over this corpus. Retrieved documents
and graph edges are retrieval evidence only. No legal text… See the full description on the dataset page: https://huggingface.co/datasets/justicedao/ipfs_singapore_laws_ir.
