A Scaling Law for Synthetic-to-Real Transfer: A Measure of Pre-Training

2021-08-25 02:29:28

Hiroaki Mikami, Kenji Fukumizu, Shogo Murai, Shuji Suzuki, Yuta Kikuchi, Taiji Suzuki, Shin-ichi Maeda, Kohei Hayashi

arXiv_CV

arXiv_CV Transfer_Learning Transformer

Abstract
Abstract (translated)
URL
PDF

Abstract

Synthetic-to-real transfer learning is a framework in which we pre-train models with synthetically generated images and ground-truth annotations for real tasks. Although synthetic images overcome the data scarcity issue, it remains unclear how the fine-tuning performance scales with pre-trained models, especially in terms of pre-training data size. In this study, we collect a number of empirical observations and uncover the secret. Through experiments, we observe a simple and general scaling law that consistently describes learning curves in various tasks, models, and complexities of synthesized pre-training data. Further, we develop a theory of transfer learning for a simplified scenario and confirm that the derived generalization bound is consistent with our empirical findings.

Abstract (translated)

URL

https://arxiv.org/abs/2108.11018

PDF

https://arxiv.org/pdf/2108.11018.pdf