streamline the fine-tuning process for multimodal models: PaliGemma 2, Florence-2, and Qwen2.5-VL
项目介绍
有效提示策略,适用于大型多模态模型,如GPT-4 Vision、LLaVA或CogVLM。🔥【此简介由AI生成】
Apache-2.0 Python534提交数captioningfine-tuningflorence-2multimodalobjectdetectionpaligemmaphi-3-visionqwen2-vltransformersvision-and-languagevqa
定制我的领域