NVEagle/LocateAnything-Data
LocateAnything-Data 中文 · Paper · Model · Code Overview LocateAnything-Data is the public training-data release for LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding. LocateAnything formulates detection and visual grounding as a unified vision-language task. Given an image and a category, phrase, text string, or action-oriented instruction, the model predicts the corresponding bounding boxes or points. The data spans natural… See the full description on the dataset page: https://huggingface.co/datasets/NVEagle/LocateAnything-Data.
Add Object365 and OpenImages grounding views
Add manifest-driven subset downloader
Document Unsplash media retrieval TODO
Document upstream media hydration workflow
Link every dataset in data coverage
Simplify spatial supervision table
Add upstream license guidance and acknowledgements
Rewrite dataset card around data domains and usage
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add final English and Chinese dataset documentation
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
Add files using upload-large-folder tool
initial commit
