Recognition and retrieval of sound events using sparse coding convolutional neural network

Chien Yao Wang, Andri Santoso, Seksan Mathulaprangsan, Chin Chin Chiang, Chung Hsien Wu, Jia Ching Wang

研究成果: 書貢獻/報告類型會議論文篇章同行評審

9 引文 斯高帕斯(Scopus)

摘要

This paper proposes a novel deep convolutional neural network (CNN), called sparse coding convolutional neural network (SC-CNN), to address the problem of sound event recognition and retrieval task. Unlike the general framework of a CNN, in which feature learning process is performed hierarchically, the proposed framework models the whole memorizing procedures in the human brain, including encoding, storage, and recollection. Sound data from the RWCP sound scene dataset with added noise from NOISEX-92 noise dataset are used to compare the performance of the proposed system with the state-of-the-art baselines. The experimental results indicated that the proposed SC-CNN outperformed the state-of-the-art systems in sound event recognition and retrieval. In the sound event recognition task, the proposed system achieved an accuracy of 94.6%, 100% and 100% under 0db, 10db and clean noise conditions, respectively. In the retrieval task, the proposed system improves the mAP rate of the general CNN by approximately 6%.

原文???core.languages.en_GB???
主出版物標題2017 IEEE International Conference on Multimedia and Expo, ICME 2017
發行者IEEE Computer Society
頁面589-594
頁數6
ISBN(電子)9781509060672
DOIs
出版狀態已出版 - 28 8月 2017
事件2017 IEEE International Conference on Multimedia and Expo, ICME 2017 - Hong Kong, Hong Kong
持續時間: 10 7月 201714 7月 2017

出版系列

名字Proceedings - IEEE International Conference on Multimedia and Expo
ISSN(列印)1945-7871
ISSN(電子)1945-788X

???event.eventtypes.event.conference???

???event.eventtypes.event.conference???2017 IEEE International Conference on Multimedia and Expo, ICME 2017
國家/地區Hong Kong
城市Hong Kong
期間10/07/1714/07/17

指紋

深入研究「Recognition and retrieval of sound events using sparse coding convolutional neural network」主題。共同形成了獨特的指紋。

引用此