FaD-VLP: Fashion Vision-and-Language Pre-training towards Unified Retrieval and Captioning

Open in new window