Computer Vision Models That Learn From Language
Any typical successful computer vision model first undergoes pre-training on ImageNet and then proceeds to do the tasks such as classification or captioning of the image. But can the vision models learn more from language? To explore this, two researchers from the University Of Michigan introduced "VirTex", a pretraining approach to learn visual features via language using fewer images. The aim of this work is to demonstrate that natural language can provide supervision for learning transferable visual representations with better data-efficiency than other approaches. Introducing "VirTex": a pretraining approach to learn visual features via language using fewer images.
Jun-30-2020, 01:46:31 GMT
- Country:
- North America > United States > Michigan (0.26)
- Technology:
- Information Technology > Artificial Intelligence
- Vision > Image Understanding (0.63)
- Games > Go (0.40)
- Information Technology > Artificial Intelligence