MiniCPM-V
minicpm-v46-strict-cycle-runpod-serverless
MiniCPM Strict Caption Harness
Strict 9-field cycling harness for MiniCPM-V-4.6 v2 captions. The harness asks the model for one field at a time, cleans each field, and assembles the final caption with deterministic v2 headers.
Smoke test without loading the model:
cd /Users/dustinpainter/datasets/image-datasets
PYTHONPATH=tools/mini-cap-harness python3 tools/mini-cap-harness/run_strict_cycle.py \
--backend mock \
--input… See the full description on the dataset page: https://huggingface.co/datasets/jaddai/minicpm-v46-strict-cycle-runpod-serverless.minicpmv_overfit_lora
Model Card for Model ID
Model Details
Model Description
Developed by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Model type: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Finetuned from model [optional]: [More Information Needed]
Model Sources [optional]
Repository: [More Information Needed]
Paper… See the full description on the dataset page: https://huggingface.co/datasets/cjfcsjt/minicpmv_overfit_lora.FLAME-ReCap-CC3M-MiniCPM-Llama3-V-2_5
Dataset description
Recaptioned CC3M by MiniCPM-Llama3-V-2_5.
Uses
The images are equivalent to https://huggingface.co/datasets/pixparse/cc3m-wds. Use data keys to index the original CC3M.
See https://github.com/MIV-XJTU/FLAME.
Citation
@article{cao2024flame,
title={FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training},
author={Cao, Anjia and Wei, Xing and Ma, Zhiheng},
journal={arXiv preprint arXiv:2411.11927}… See the full description on the dataset page: https://huggingface.co/datasets/caj/FLAME-ReCap-CC3M-MiniCPM-Llama3-V-2_5.FLAME-ReCap-YFCC15M-MiniCPM-Llama3-V-2_5
Dataset description
Recaptioned YFCC15M by MiniCPM-Llama3-V-2_5.
Uses
See https://github.com/MIV-XJTU/FLAME.
Citation
@article{cao2024flame,
title={FLAME: Frozen Large Language Models Enable Data-Efficient Language-Image Pre-training},
author={Cao, Anjia and Wei, Xing and Ma, Zhiheng},
journal={arXiv preprint arXiv:2411.11927},
year={2024}
}
@article{yao2024minicpmv,
title={MiniCPM-V: A GPT-4V Level MLLM on Your Phone},
author={Yao, Yuan… See the full description on the dataset page: https://huggingface.co/datasets/caj/FLAME-ReCap-YFCC15M-MiniCPM-Llama3-V-2_5.conceptbench_path_vqa_result_2_minicpm_v_8bconceptbench_path_vqa_result_2_minicpm_v_8b_evaluated
