CoolFace
Datasetpublic

Intelligent-Internet/arxiv

arXiv This is a arXiv dataset for use with the II-Commons-Store project. Dataset Details Dataset Description This dataset comprises a curated arXiv dataset. We provide a series of pre-computed embedding vector datasets based on ArXiv paper data to help users quickly start and test the semantic search API. These datasets contain paper metadata, text from certain sections, and optimized embedding vectors. They can be downloaded and used directly… See the full description on the dataset page: https://huggingface.co/datasets/Intelligent-Internet/arxiv.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
8likes274downloads
14 commits on main
238fbb61y ago

Update README.md

huoju
2ca0e781y ago

Update README.md

huoju
81179191y ago

Update README.md

Leask
26d48b91y ago

Update README.md

Leask
bc6beec1y ago

Update README.md

Leask
d4934d11y ago

Rename duckdb/arxiv.duckdb to duckdb/arxiv_snowflake2m_128_int8.duckdb

huoju
5db60f91y ago

Rename duckdb/arxiv.yaml to duckdb/arxiv_snowflake2m_128_int8.yaml

huoju
f3fd81e1y ago

Create arxiv.yaml

huoju
b935fa71y ago

Update README.md

Leask
93f22811y ago

Update README.md

Leask
7b83d5e1y ago

Update README.md

Leask
fb47ce11y ago

Update README.md

Leask
ab480281y ago

Upload folder using huggingface_hub

Leask
ad169251y ago

initial commit

Leask