Scholay

学术搜索 · AI 审稿 · LaTeX 协作

Introducing shared-hidden-layer autoencoders for transfer learning and their application in acoustic emotion recognition

作者:Jun Deng, Rui Xia, Zixing Zhang, Yang Liu, Björn W. Schuller · 年份:2014 · DOI:10.1109/icassp.2014.6854517 · 被引用次数:75 · 研究领域:Speech Recognition and Synthesis、Music and Audio Processing、Speech and Audio Processing

This study addresses a situation in practice where training and test samples come from different corpora - here in acoustic emotion recognition. In this situation, a model is trained on one database while tested on another disjoint one. The typical inherent mismatch between the corpora and by that between test and training set usually leads to significant performance degradation. To cope with this problem when no training data from the target domain exists, we propose a `shared-hidden-layer autoencoder' (SHLA) approach for learning common feature representations shared across the training and test set in order to reduce the discrepancy in them. To exemplify effectiveness of our approach, we select the Interspeech Emotion Challenge's FAU Aibo Emotion Corpus as test database and two other publicly available databases as training set for extensive evaluation. The experimental results show that our SHLA method significantly improves over the baseline performance and outperforms today's state-of-the-art domain adaptation methods.