Transfer Learning with Pre-trained Conditional Generative Models

Yamaguchi, Shin'ya, Kanai, Sekitoshi, Kumagai, Atsutoshi, Chijiwa, Daiki, Kashima, Hisashi

Sep-29-2022–arXiv.org Artificial Intelligence

Transfer learning is crucial in training deep neural networks on new target tasks. Current transfer learning methods always assume at least one of (i) source and target task label spaces overlap, (ii) source datasets are available, and (iii) target network architectures are consistent with source ones. However, holding these assumptions is difficult in practical settings because the target task rarely has the same labels as the source task, the source dataset access is restricted due to storage costs and privacy, and the target architecture is often specialized to each task. To transfer source knowledge without these assumptions, we propose a transfer learning method that uses deep generative models and is composed of the following two stages: pseudo pre-training (PP) and pseudo semi-supervised learning (P-SSL). PP trains a target architecture with an artificial dataset synthesized by using conditional source generative models. P-SSL applies SSL algorithms to labeled target data and unlabeled pseudo samples, which are generated by cascading the source classifier and generative models to condition them with target samples. Our experimental results indicate that our method can outperform the baselines of scratch training and knowledge distillation. For training deep neural networks on new tasks, transfer learning is essential, which leverages the knowledge of related (source) tasks to the new (target) tasks via the joint-or pre-training of source models. There are many transfer learning methods for deep models under various conditions (Pan & Yang, 2010; Wang & Deng, 2018). For instance, domain adaptation leverages source knowledge to the target task by minimizing the domain gaps (Ganin et al., 2016), and fine-tuning uses the pre-trained weights on source tasks as the initial weights of the target models (Yosinski et al., 2014).

artificial intelligence, deep learning, machine learning, (19 more...)

arXiv.org Artificial Intelligence

Sep-29-2022

arXiv.org PDF

Add feedback

Country:
- Oceania
  - New Zealand > South Island
    - Marlborough District > Blenheim (0.04)
  - Australia > New South Wales
    - Sydney (0.04)
- North America
  - United States
    - Virginia (0.04)
    - Maryland (0.04)
    - Massachusetts > Middlesex County
      - Cambridge (0.04)
    - California > Santa Clara County
      - Palo Alto (0.04)
  - Canada > Newfoundland and Labrador
    - Newfoundland (0.04)
- Europe > United Kingdom
  - England > Staffordshire (0.04)
- Atlantic Ocean > North Atlantic Ocean
  - Chesapeake Bay (0.04)
- Asia
  - China (0.04)
  - Japan > Honshū
    - Kansai > Kyoto Prefecture > Kyoto (0.04)

Genre:
- Research Report > New Finding (0.46)

Industry:
- Automobiles & Trucks (1.00)
- Leisure & Entertainment > Sports (0.93)
- Consumer Products & Services (0.68)
- Health & Medicine (0.67)
- Transportation
  - Passenger (1.00)
  - Ground > Road (1.00)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning
  - Transfer Learning (1.00)
  - Neural Networks > Deep Learning
    - Generative AI (0.34)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found