CoolFace
Modelpublic

jinaai/jina-embeddings-v5-omni-nano-classification

sourceHugging Facecc-by-nc-4.0updated 1mo agoView on Hugging Face
3likes303downloads
27 commits on main
494a0cb1mo ago

readme: a video embeds its frames; drop the stale imageio claim

florian-hoenicke
db6fe381mo ago

sync custom_st.py from harness/src/hf_sources/nano_variant

florian-hoenicke
29fba331mo ago

Run the audio feature extractor whenever a request carries audio (#4)

florian-hoenicke, ziniuyu
b9d3b152mo ago

sync custom_st.py from harness/src/hf_sources/nano_variant

florian-hoenicke
77370462mo ago

sync vllm_llava_eurobert_audio.py from harness/src/hf_sources/nano_variant

florian-hoenicke
b1418092mo ago

audio: size the audio run from the real Whisper frame mask

florian-hoenicke
35f7ec12mo ago

audio: size the audio run from the real Whisper frame mask

florian-hoenicke
ca311504mo ago

Support attn_implementation="flash_attention_2"

koukandre
3f9341e4mo ago

readme: point users without torchcodec/torchvision (Windows, some Colab) to the av-only model.encode("clip.mp4") path

florian-hoenicke
3c56dcd4mo ago

readme: install instructions — add torchcodec for proc(videos=path) (transformers' video processor selects torchcodec by default; av/imageio are not consulted on that path)

florian-hoenicke
246308a4mo ago

docs: media query/document via encode_query/encode_document; nano requires transformers>=5.0 for multimodal

florian-hoenicke
a59cf994mo ago

fix(processor): count images/videos from grid_thw so a single PIL.Image works (was TypeError on len(Image)); byte-identical for list inputs

florian-hoenicke
48d90834mo ago

README.md: mirror retrieval paradigm — add 'Document: ' prefix to Quickstart text input, switch SBERT to encode_document, prefix the vLLM prompt, add prefix-requirement note.

florian-hoenicke
621203e4mo ago

config_sentence_transformers.json: mirror retrieval — set prompts={'document': 'Document: '} and default_prompt_name=null so SBERT users use encode_document(...), AutoModel/vLLM users prepend 'Document: ' manually (same paradigm as retrieval, no query side).

florian-hoenicke
fe4bf054mo ago

Revert config.json to pre-2026-05-17 content (parent 84442a148e)

florian-hoenicke
5bb6ccb4mo ago

Revert config_sentence_transformers.json to pre-2026-05-17 content (parent 84442a148e)

florian-hoenicke
3c2ab704mo ago

Revert vllm_llava_eurobert_audio.py to pre-2026-05-17 content (parent 84442a148e)

florian-hoenicke
9c8a5cb4mo ago

Revert modeling_llava_eurobert_audio.py to pre-2026-05-17 content (parent 84442a148e)

florian-hoenicke
eb0e0ea4mo ago

sync vllm_llava_eurobert_audio.py: auto-prepend default_text_prefix for text-only inputs (no-op on retrieval)

florian-hoenicke
fa915874mo ago

sync modeling_llava_eurobert_audio.py: auto-prepend default_text_prefix for text-only inputs (no-op on retrieval)

florian-hoenicke
bc8fc774mo ago

config.json: add default_text_prefix="Document: " so the modeling code and vLLM plugin auto-prepend this on text-only inputs (matching the text-* twin's SBERT-default behavior across all three code paths).

florian-hoenicke
775dc904mo ago

Set default_prompt_name="document" and prompts.document="Document: " to mirror the matching text-* twin; SBERT-default text vectors are now bit-identical to text-{nano,small}-{classification,clustering,text-matching}.

florian-hoenicke
84442a15mo ago

readme: add frontier plot + EIS section

florian-hoenicke
84e66275mo ago

add omni_frontier.png: omni model frontier plot from paper

florian-hoenicke
283ccf65mo ago

readme: add logo, ArXiv/Blog links, sibling-size link, broaden tags

florian-hoenicke
bb4333d5mo ago

fix(processor): expand image+video without placeholder collision

florian-hoenicke
77d31a25mo ago

Initial public release

Jina AI