datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
linical_inference_debt_scanner_v0.1Clinical Inference Debt Scanner
PurposeDetect when a clinical plan relies on stacked assumptions rather than evidence.
You receive:
evidence_signals
a narrative_chain
a planned_action
You output:
inference_debt_level0 to 3
debt_itemthe single most dangerous leap
paydown_stepthe corrective step that restores evidence grounding
Debt scale0 none1 minor2 moderate3 severe
Scoring
debt_level_scoregraded by distance from gold
debt_item_similaritytoken overlap similarity… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/linical_inference_debt_scanner_v0.1.scandi-reddit-filtered
Dataset Card for ScandiRedditFiltered
Dataset Summary
ScandiRedditFiltered is manually filtered and post-processed corpus consisting of comments from ScandiReddit.
The intended use of the filtered sentences is for Text-To-Speech (TTS) models.
Supported Tasks and Leaderboards
Training language models is the intended task for this dataset. No leaderboard is active at this point.
Languages
The dataset is available in Danish (da).
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/alexandrainst/scandi-reddit-filtered.
