CoolFace
Apppublic

prakhya15/trimodal-bind

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes
App README

TriModal-Bind

Retrieve the most semantically similar images from natural language using a shared multimodal embedding space learned through contrastive learning.

Features

  • —Text-to-image retrieval
  • —Contrastive multimodal embeddings
  • —DistilBERT text encoder
  • —PyTorch image encoder
  • —Gradio interface