Transfer Learning for Speech and Language Processing
\authorblockNDong Wang and Thomas Fang Zheng \authorblockA1. Center for Speech and Language Technologies (CSLT) Research Institute of Information Technology, Tsinghua University 2. Tsinghua National Lab for Information Science and Technology Beijing, 100084, P.R.China
Abstract
Transfer learning is a vital technique that generalizes models trained for one setting or task to other settings or tasks. For example in speech recognition, an acoustic model trained for one language can be used to recognize speech in another language, with little or no re-training data. Transfer learning is closely related to multi-task learning (cross-lingual vs. multilingual), and is traditionally studied in the name of ‘model adaptation’. Recent advance in deep learning shows that transfer learning becomes much easier and more effective with high-level abstract features learned by deep models, and the ‘transfer’ can be conducted not only between data distributions and data types, but also between model structures (e.g., shallow nets and deep nets) or even model types (e.g., Bayesian models and neural models). This review paper summarizes some recent prominent research towards this direction, particularly for speech and language processing. We also report some results from our group and highlight the potential of this very interesting research field111This survey will be continuously updated online () to reflect the recent progress on transfer learning. .
原文 arXiv:1511.06066;中英对照 + 大白话阅读 https://aha.fim.ai/paper/1511.06066v1