datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
indian-legal-sections-bns-bnss-bsa-2023
🏛️ Indian Legal Sections — BNS · BNSS · BSA 2023
The First Structured, Unified JSON Dataset of Modern Indian Criminal Law
📖 Dataset Summary
This dataset contains 1,059 fully structured and verified sections extracted, parsed, and unified from India's three landmark criminal justice reform acts passed in December 2023. These three acts together replaced the colonial-era Indian Penal Code (IPC, 1860), the Code of Criminal Procedure… See the full description on the dataset page: https://huggingface.co/datasets/GSMS-B/indian-legal-sections-bns-bnss-bsa-2023.BnSentMix
BnSentMix: A Diverse Bengali-English Code-Mixed Dataset for Sentiment Analysis
Dataset Overview
Column Title
Description
Data Sources
Facebook, YouTube, E-commerce Sites
#Samples
20000
Sentiment Labels
1:Positive, 2:Negative, 3:Neutral, 4:Mixed
Filtering Method
Automated using mBERT
#Annotators
64
Annotation/Sample
2 or 3 (if tie)
Dataset Statistics
Statistic
Value
Mean Character Length
62.77
Max Character Length
1985… See the full description on the dataset page: https://huggingface.co/datasets/aplycaebous/BnSentMix.aave-bns-data
Aave-BNS Multidimensional Protocol-Network Evidence
Hugging Face release status: PRIVATE RELEASE CANDIDATE. The dataset has been uploaded to the Hugging Face Hub for post-upload validation. Local validation covers 14 configurations, Croissant 1.1 core metadata, substantive Responsible AI metadata, checksums, and deterministic handoff auditing. Dataset Viewer validation, clean Hub loading, platform-generated Croissant reconciliation, and immutable release tagging are the… See the full description on the dataset page: https://huggingface.co/datasets/zlysunshine/aave-bns-data.cybersecurity-nerBNS_definitions
Dataset Card for BNS Definitions Dataset
BHARATIYA NYAYA SANHITA (BNS) DEFINITIONS DATASET
Dataset Details
Dataset Description
THIS IS A CLEANED AND CURATED DATASET OF DEFINITIONS UNDER THE BHARATIYA NYAYA SANHITA, 2023 (BNS)BUILT FOR USE IN LEGAL RAG (RETRIEVAL AUGMENTED GENERATION) SYSTEMS.
THE DATASET INCLUDES OFFICIAL SECTION TITLES AND THEIR DETAILED LEGAL DEFINITIONSTO HELP YOUR MODELS UNDERSTAND THE NUANCES OF INDIAN PENAL LAW BETTER ⚖️📜
Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/navaneeth005/BNS_definitions.IPC_and_BNS_transformationbn_sentiment_noisy_dataset
Dataset Card for "SentNoB"
Dataset Summary
Social Media User Comments' Sentiment Analysis Dataset. Each user comments are labeled with either positive (1), negative (2), or neutral (0).
Citation Information
@inproceedings{islam2021sentnob,
title={SentNoB: A Dataset for Analysing Sentiment on Noisy Bangla Texts},
author={Islam, Khondoker Ittehadul and Kar, Sudipta and Islam, Md Saiful and Amin, Mohammad Ruhul},
booktitle={Findings of the Association for… See the full description on the dataset page: https://huggingface.co/datasets/sustcsenlp/bn_sentiment_noisy_dataset.bnss-2023-datasetbnsBNS_detailed
Dataset Card for BNS Detailed Sections Dataset
BHARATIYA NYAYA SANHITA (BNS) DETAILED SECTIONS DATASET
Dataset Details
Dataset Description
THIS IS A COMPREHENSIVE AND DETAILED DATASET OF ALL MAJOR SECTIONS UNDER THEBHARATIYA NYAYA SANHITA, 2023, INCLUDING THEIR TITLES, PUNISHMENTS, COGNIZABILITY, BAILABILITY STATUS & MORE 🔍⚖️
BUILT FOR USE IN LAW-AWARE RAG SYSTEMS AND AI LEGAL ASSISTANTS TO UNDERSTAND THE CONTEXT, PUNISHMENT & CATEGORIZATION OF CRIMINAL… See the full description on the dataset page: https://huggingface.co/datasets/navaneeth005/BNS_detailed.bns_act_2023
Dataset on BNS Act, 2023 which replaces the Indian Penal Code
This dataset has been made from the official document published by the government of India on the BNS Act, 2023
Source: PDF Source
bns2bns_testbns-2023-datasetbnsbn_squadFormattedForMistralINstructcomplaint-relevant-bnsBnSentMix
BnSentMix: A Diverse Bengali-English Code-Mixed Dataset for Sentiment Analysis
Dataset Overview
Column Title
Description
Data Sources
Facebook, YouTube, E-commerce Sites
#Samples
20000
Sentiment Labels
1:Positive, 2:Negative, 3:Neutral, 4:Mixed
Filtering Method
Automated using mBERT
#Annotators
64
Annotation/Sample
2 or 3 (if tie)
Dataset Statistics
Statistic
Value
Mean Character Length
62.77
Max Character Length
1985… See the full description on the dataset page: https://huggingface.co/datasets/rifat101/BnSentMix.bn_squadFormattedForFatimaFellowshipbnsbnss-datasetIPC_and_BNS_transformationbnsaddcsvds_bnsrxp61afIPC_and_BNS_transformationieee-dataIPC_and_BNS_transformation
