Distilling Vision-Language Foundation Models: A Data-Free Approach via Prompt Diversification

Open in new window