datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
legal-documents-splittedlegal-documents
Description
Topic: Legal Documents
Domains: Law, Contracts, Regulations
Focus: Synthetic raw text data for legal document analysis
Number of Entries: 1000
Dataset Type: Raw Dataset
Model Used: bedrock/us.amazon.nova-pro-v1:0
Language: English
Generated by: SynthGenAI Package
legal-documentslegaldocuments-nli-testlegal-documents-chunks_400legal-documents-chunkslegaldocuments-nli-test-v3legaldocument-llama-2legal_document_structuringTask details
Document structuring plays a crucial role in various natural language processing (NLP) tasks, such as information retrieval, and document understanding.
It also helps readers to effectively navigate into a structured document with a large amount of textual data.
In the legal domain, document structuring is particularly important for creating inter- and intra-document links.
The dataset provides documents segmented into lines.
Each document was collected in HTML format or PDF… See the full description on the dataset page: https://huggingface.co/datasets/DoctrineAI/legal_document_structuring.Legal_Document_Analyzer_mergedDatasetlegal-documents-summarylegal-documents-splits-filteredlegal-documents-questionslegaldocuments-nli-test-v2legal_document_vi
Dataset Card for "legal_document_vi"
More Information needed
Legal_Document_Analyzer_augmentedlegal_document_vi
Dataset Card for "legal_document_vi1"
More Information needed
