A survey on deep reinforcement learning for audio-based applications-Reference-Cited by-同舟云学术

A survey on deep reinforcement learning for audio-based applications

Published:2022-07-02 Issue:3 Volume:56 Page:2193-2240
ISSN:0269-2821
Container-title:Artificial Intelligence Review
language:en
Short-container-title:Artif Intell Rev

Author:

Latif Siddique^ORCID,Cuayáhuitl Heriberto,Pervez Farrukh,Shamshad Fahad,Ali Hafiz Shehbaz,Cambria Erik

Abstract

AbstractDeep reinforcement learning (DRL) is poised to revolutionise the field of artificial intelligence (AI) by endowing autonomous systems with high levels of understanding of the real world. Currently, deep learning (DL) is enabling DRL to effectively solve various intractable problems in various fields including computer vision, natural language processing, healthcare, robotics, to name a few. Most importantly, DRL algorithms are also being employed in audio signal processing to learn directly from speech, music and other sound signals in order to create audio-based autonomous systems that have many promising applications in the real world. In this article, we conduct a comprehensive survey on the progress of DRL in the audio domain by bringing together research studies across different but related areas in speech and music. We begin with an introduction to the general field of DL and reinforcement learning (RL), then progress to the main DRL methods and their applications in the audio domain. We conclude by presenting important challenges faced by audio-based DRL agents and by highlighting open areas for future research and investigation. The findings of this paper will guide researchers interested in DRL for the audio domain.

Funder

University of Southern Queensland

Publisher

Springer Science and Business Media LLC

Subject

Artificial Intelligence,Linguistics and Language,Language and Linguistics

Link

https://link.springer.com/content/pdf/10.1007/s10462-022-10224-2.pdf

Reference299 articles.

1. Abbeel P, Ng AY (2004) Apprenticeship learning via inverse reinforcement learning. In: Proceedings of the twenty-first international conference on Machine learning, p 1

2. Abdel-Hamid O, Mohamed Ar, Jiang H, Deng L, Penn G, Yu D (2014) Convolutional neural networks for speech recognition. IEEE/ACM Trans Audio Speech Lang Process 22(10)

3. Alamdari N, Lobarinas E, Kehtarnavaz N (2020) Personalization of hearing aid compression by human-in-the-loop deep reinforcement learning. IEEE Access 8:203503–203515. https://doi.org/10.1109/ACCESS.2020.3035728

4. Alfredo C, Humberto C, Arjun C (2017) Efficient parallel methods for deep reinforcement learning. In: The Multi-disciplinary Conference on Reinforcement Learning and Decision Making (RLDM)

5. Ali HS, ul Hassan F, Latif S, Manzoor HU, Qadir J (2021) Privacy enhanced speech emotion communication using deep learning aided edge computing. In: 2021 IEEE International Conference on Communications Workshops (ICC Workshops), pp. 1–5. IEEE

Cited by 32 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Practical Implementation of Automated Next Generation Audio Production for Live Sports;Journal of the Audio Engineering Society;2024-07-18

2. Procedurally generated AI compound media for expanding audial creations, broadening immersion and perception experience;International Journal of Electronics and Telecommunications;2024-06-25

3. The Integration of Artificial Intelligence and Video Production Skills in Workplace Development: A Study from the Perspective of Vocational Training;The Review of Socionetwork Strategies;2024-06-25

4. Artificial intelligence and learning environment: Human considerations;Journal of Computer Assisted Learning;2024-05-21

5. Identifying intentions in conversational tools: a systematic mapping;Proceedings of the 20th Brazilian Symposium on Information Systems;2024-05-20