sentence-transformers/sentence-compression
Dataset Card for Sentence Compression This dataset is a collection of text-simplified pairs from the Sentence Compression project. See Sentence Compression for additional information. This dataset can be used directly with Sentence Transformers to train embedding models. Dataset Subsets pair subset Columns: "text", "simplified" Column types: str, str Examples:{ 'text': "The USHL completed an expansion draft on Monday as 10 players who were on the… See the full description on the dataset page: https://huggingface.co/datasets/sentence-transformers/sentence-compression.
Dataset Card for Sentence Compression
This dataset is a collection of text-simplified pairs from the Sentence Compression project. See Sentence Compression for additional information. This dataset can be used directly with Sentence Transformers to train embedding models.
Dataset Subsets
pair subset
- Columns: "text", "simplified"
- Column types:
str,str - Examples:
{
'text': "The USHL completed an expansion draft on Monday as 10 players who were on the rosters of USHL teams during the 2009-10 season were selected by the League's two newest entries, the Muskegon Lumberjacks and Dubuque Fighting Saints.",
'simplified': 'USHL completes expansion draft',
}- Collection strategy: Reading the Sentence Compression dataset from embedding-training-data.
- Deduplified: No
