Google Trains An AI Vision Model With Two Billion Parameter
Google Brain researchers announced a two-billion-parameter deep-learning computer vision (CV) model. The model was trained on three billion pictures and obtained a new state-of-the-art record of 90.45 percent top-1 accuracy on ImageNet. The ViT-G/14 model is based on Google's latest Vision Transformers development (ViT). On numerous benchmarks, including ImageNet, ImageNet-v2, and VTAB-1k, ViT-G/14 beat prior state-of-the-art systems. For example, the accuracy gain on the few-shot picture identification challenge was more than five percentage points.
Jun-28-2021, 16:35:10 GMT
- Technology: