cisco
Datasets
All datasets matching “cisco”open-greek-corpus-annotations
Open Greek Corpus Annotations
Token-level linguistic annotations for the
Open Greek Corpus:
lemma, part of speech (UD UPOS), and morphology (UD features) for every
served token. Three provenance classes, never confused thanks to per-token
provenance and confidence tiers: gold treebank annotations where an openly
licensed MANUAL treebank covers a work (GLAUx's treebank layers, MACULA
Greek for the NT), GLAUx's own automatic annotation as the middle auto:
class, and model… See the full description on the dataset page: https://huggingface.co/datasets/ciscoriordan/open-greek-corpus-annotations.open-greek-corpus-annotation-exports
Corpus of Open Greek annotation exports
Standardized annotation-export payloads from the
Corpus of Open Greek (cog).
cog normalizes external Greek annotation corpora into one per-work,
CTS-URN-keyed token-record format; this dataset repo hosts the export payloads,
while the cog git repo holds the exporter scripts and a small per-release
pointer stub. The record schema, encoding guarantees, and storage rule are
defined in cog's docs/annotation-export-contract.md.… See the full description on the dataset page: https://huggingface.co/datasets/ciscoriordan/open-greek-corpus-annotation-exports.model-provenance-kit
Deep-Signal Weight Fingerprints
Dataset Summary
This repository distributes pre-computed deep-signal weight fingerprints for use with Model Provenance Kit (Model ProvenanceKit). The primary deliverable is deep-signals.zip: a compressed archive of Apache Parquet files that store dense numerical features extracted from publicly released transformer (and related) model weights on the Hugging Face Hub or equivalent sources.
Each file encodes multi-signal weight-level… See the full description on the dataset page: https://huggingface.co/datasets/cisco-ai/model-provenance-kit.details_ndavidson__cisco-iNAM-phi-sft
Dataset Card for Evaluation run of ndavidson/cisco-iNAM-phi-sft
Dataset automatically created during the evaluation run of model ndavidson/cisco-iNAM-phi-sft on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_ndavidson__cisco-iNAM-phi-sft.Cisco_CCNAmerged_datasets
