uv-scripts/vlm-object-detection
VLM Object Detection Instruction-prompted object detection with vision-language models via vLLM. Designed as a VLM-as-labeller primitive for bootstrapping object-detection datasets — give it a free-form prompt ("detect every photograph and illustration", "detect all PPE items", "detect every electronic component and identify its reference designator") and it returns bbox JSON ready for downstream labelling tools (Label Studio, FiftyOne, COCO conversion). Sibling: uv-scripts/sam3… See the full description on the dataset page: https://huggingface.co/datasets/uv-scripts/vlm-object-detection.
Sync from GitHub via hub-sync
Delete files example-ena24-bear.png with huggingface_hub
Upload README.md with huggingface_hub
Upload example-ena24-fox.png with huggingface_hub
Upload README.md with huggingface_hub
Upload example-ena24-bear.png with huggingface_hub
Upload README.md with huggingface_hub
Upload inspect-detections.py with huggingface_hub
Upload README.md with huggingface_hub
Upload README.md with huggingface_hub
Upload qwen3vl-detect-fewshot.py with huggingface_hub
Upload qwen3vl-detect.py with huggingface_hub
initial commit
