ViP-LLaVA: Making Large Multimodal Models Understand Arbitrary Visual Prompts
2024 · 42 citations
https://doi.org/10.1109/cvpr52733.2024.01227
Claim or correct this author profile
Claim this profile or suggest corrections to the name, affiliation, bio, photo, or paper titles. Approved changes appear as verified CitedEvidence overlays.
37
Papers
480
Citations
10
h-index
10
i10-index
Mu Cai is an academic researcher. The author has contributed to research in topics: Multimodal Machine Learning Applications & Natural Language Processing Techniques & Domain Adaptation and Few-Shot Learning. The author has an h-index of 10, co-authored 32 publications.
ORCID: 0009-0008-7967-9752https://doi.org/10.1109/cvpr52733.2024.01227
https://doi.org/10.18653/v1/2024.findings-acl.915
Click to start Chat