ExpVG: Investigating the Design Space of Visual Grounding in Multimodal Large Language Model

Open in new window