CoolFace
Apppublic

drizzymedia/synapse-vision

sourceHugging Faceapache-2.0updated 2mo agoView on Hugging Face
1likes
App README

SynapseVLM Caption Evaluator

Drop multiple images, select a model, and compare generated captions to determine which Synapse vision model is best suited for captioning your image datasets.

Models (7–9B range, one resident at a time)

LabelModel
SynapseVLM-9B (official)Synapse/SynapseVLM-9B
SynapseVLM-9B UncensoredSynapse/SynapseVLM-9B-Unlocked
SynapseVLM-4B (fast)Synapse/SynapseVLM-4B
SynapseVision-8B-Instruct (previous generation)Synapse/SynapseVision-8B-Instruct

SynapseVLM is a native multimodal vision-language foundation model designed for image understanding, visual reasoning, and high-quality image caption generation.

Thinking mode is disabled by default for clean, consistent captions.

Hardware

Requires SynapseGPU (48 GB large inference slice).

Models are loaded on first use or when switching models (~1 minute). Caption generation typically completes within seconds per image after the model is loaded.