MIMIC-IT: Multi-Modal In-Context Instruction Tuning paper page: https://
huggingface.co/papers/2306.05
425
… High-quality instructions and responses are essential for the zero-shot performance of large language models on interactive natural language tasks. For interactive vision-language tasks
MIMIC-IT: Multi-Modal In-Context Instruction Tuning for Vision-Language Tasks
By
–
