Multimodal parameter-efficient fine-tuning
LoRA, QLoRA and P-Tuning v2 support multimodal image, text and video tasks while adapting industry knowledge and task capability with minimal impact on general foundation-model capability.
LoRA, QLoRA and P-Tuning v2
Adapters, prompt parameters, cross-modal projection layers and selected model parameters
Single GPU, multi-GPU on one machine and distributed multi-node training