Transfer Learning with Self-Supervised Vision Transformers for Snake Identification

Miyaguchi, Anthony, Gustineli, Murilo, Fischer, Austin, Lundqvist, Ryan

Jul-8-2024–arXiv.org Artificial Intelligence

We present our approach for the SnakeCLEF 2024 competition to predict snake species from images. We explore and use Meta's DINOv2 vision transformer model for feature extraction to tackle species' high variability and visual similarity in a dataset of 182,261 images. We perform exploratory analysis on embeddings to understand their structure, and train a linear classifier on the embeddings to predict species. Despite achieving a score of 39.69, our results show promise for DINOv2 embeddings in snake identification.

dataset, prediction, snake, (13 more...)

arXiv.org Artificial Intelligence

Jul-8-2024

arXiv.org PDF

Add feedback

Country:
- Africa > Sub-Saharan Africa (0.04)
- North America > United States
  - Georgia > Fulton County > Atlanta (0.04)
- Europe > France
  - Auvergne-Rhône-Alpes > Isère > Grenoble (0.04)

Genre:
- Research Report > New Finding (0.68)

Industry:
- Information Technology (0.69)

Technology:
- Information Technology > Artificial Intelligence
  - Vision (1.00)
  - Machine Learning
    - Statistical Learning (0.67)
    - Neural Networks > Deep Learning (0.47)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found