CoolFace
20 results

vector

japanese-asr /whisper_transcriptions.reazon_speech_all.wer_10.0.vectorized1M<n<10M0 likes87k downloads2y agoHugging Facejapanese-asr /whisper_transcriptions.mls.wer_10.0.vectorized1M<n<10M1 likes36k downloads2y agoHugging Facephilippesaade /Wikidata_Vectors_0.2 Wikidata Entity Embeddings 0.2 Dataset Summary Wikidata Entity Embeddings is a dataset of embedding vectors for Wikidata entities. Each vector represents a Wikidata item (Q...) or property (P...) based on textual information extracted from Wikidata. The dataset is part of the Wikidata Embedding Project, an initiative led by Wikimedia Deutschland in collaboration with Jina AI and IBM DataStax. The project provides a publicly accessible Wikidata Vector Database to… See the full description on the dataset page: https://huggingface.co/datasets/philippesaade/Wikidata_Vectors_0.2.textfeature-extraction10M<n<100M3 likes5.7k downloads28d agoHugging Faceabotresol /emotion-vectors-gemma-4-31b-it-postfix Emotion vectors, google/gemma-4-31b-it (corrected extraction) Residual-stream activations for google/gemma-4-31b-it, pooled per story and averaged per emotion. Each emotion ends up as one direction in the model's activation space. Read LINEAGE.md before using this. This set supersedes abotresol/emotion-vectors-gemma-4-31b-it. The earlier extraction ran while the tokenizer padded on the left, so the step that skips a story's first 50 tokens skipped padding instead. This set… See the full description on the dataset page: https://huggingface.co/datasets/abotresol/emotion-vectors-gemma-4-31b-it-postfix.feature-extraction0 likes5.6k downloads2mo agoHugging Facevector-index-bench /vibeThis repository contains the datasets presented in VIBE: Vector Index Benchmark for Embeddings: https://github.com/vector-index-bench/vibe The datasets can be downloaded manually from this repository, but the benchmark framework also downloads them automatically. Datasets In-distribution datasets Name Type n d Distance agnews-mxbai-1024-euclidean Text 769,382 1024 euclidean arxiv-nomic-768-normalized Text 1,344,643 768 any dpr-jina-768-normalized… See the full description on the dataset page: https://huggingface.co/datasets/vector-index-bench/vibe.sentence-similarity2 likes3.8k downloads1mo agoHugging Facelu-christina /assistant-axis-vectors The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models This repository contains pre-computed axes and persona vectors for Gemma 2 27B, Qwen 3 32B, and Llama 3.3 70B, as described in the paper The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models. Paper | Code | Demo The Assistant Axis is a direction in activation space that captures how "Assistant-like" a model's current persona is. It can be used to: Monitor persona… See the full description on the dataset page: https://huggingface.co/datasets/lu-christina/assistant-axis-vectors.other12 likes2.4k downloads8mo agoHugging Face