datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
legal-documents-splittedlegal-documents
Description
Topic: Legal Documents
Domains: Law, Contracts, Regulations
Focus: Synthetic raw text data for legal document analysis
Number of Entries: 1000
Dataset Type: Raw Dataset
Model Used: bedrock/us.amazon.nova-pro-v1:0
Language: English
Generated by: SynthGenAI Package
legal-documentslegaldocuments-nli-testlegal-documents-chunks_400legal-document-version-redline-final-coherence-risk-v0.1What this dataset does
You receive
version history
redline summary
final id
sent or filed id
approval record
mismatch flags
You decide
coherent
or
incoherent
Daily use
wrong attachment prevention
filing version QC
approval gap detection
legal-documents-chunkslegaldocuments-nli-test-v3Legal-Documentlegaldocument-llama-2legal_document_structuringTask details
Document structuring plays a crucial role in various natural language processing (NLP) tasks, such as information retrieval, and document understanding.
It also helps readers to effectively navigate into a structured document with a large amount of textual data.
In the legal domain, document structuring is particularly important for creating inter- and intra-document links.
The dataset provides documents segmented into lines.
Each document was collected in HTML format or PDF… See the full description on the dataset page: https://huggingface.co/datasets/DoctrineAI/legal_document_structuring.Legal_Document_Analyzer_mergedDatasetlegal-documents-summarylegal-documents-splits-filteredlegal-documentslegal-documents-questionslegaldocuments-nli-test-v2legal_document_vi
Dataset Card for "legal_document_vi"
More Information needed
Legal_Document_Analyzer_augmentedLegalDocumentSummarizationlegal_document_vi
Dataset Card for "legal_document_vi1"
More Information needed
