An Audio-Based SLAM for Indoor Environments: A Robotic Mixed Reality Presentation

Lahemer, Elfituri S. F.; Rad, Ahmad

doi:10.3390/s24092796

Sensors

2024

DOI: 10.3390/s24092796

|View full text |Cite

An Audio-Based SLAM for Indoor Environments: A Robotic Mixed Reality Presentation

Elfituri S. F. Lahemer,

Ahmad Rad

Abstract: In this paper, we present a novel approach referred to as the audio-based virtual landmark-based HoloSLAM. This innovative method leverages a single sound source and microphone arrays to estimate the voice-printed speaker’s direction. The system allows an autonomous robot equipped with a single microphone array to navigate within indoor environments, interact with specific sound sources, and simultaneously determine its own location while mapping the environment. The proposed method does not require multiple a… Show more

Help me understand this report

Search citation statements

Order By: Relevance

Paper Sections

Select...

Citation Types

Supporting

Mentioning

Contrasting

Year Published

2024

Publication Types

Select...

Article1

Relationship

Self Cite0

Independent1

Authors

Journals

Cited by 1 publication

References 76 publications

Supporting

Mentioning

Contrasting

Order By: Relevance

Exploring deep learning frameworks for multi-track music synthesis

Liu

2024

Applied Mathematics and Nonlinear Sciences

View full text Add to dashboard Cite

It has been found that the existing methods for generating multi-track music fail to meet the market requirements in terms of melody, rhythm and harmony, and most of the generated music does not conform to the basic music theory knowledge. This paper proposes a multi-track music synthesis model that uses the improved WGAN-GP and is guided by music theory rules to generate music works with high musicality to solve the problems mentioned above. Through the improvement of the adversarial loss function and the introduction of the self-attention mechanism, the improved WGANGP is obtained, which is applied to multi-track music synthesis, and both subjective and objective aspects evaluate the performance of the model. The score of multi-track music synthesized by this paper’s model is 8.22, higher than that of real human works, which is 8.04, and the average scores of the four indexes of rhythm, melody, emotion, and harmony are 8.15, 8.27, 7.61, and 8.22, respectively, which are higher than that of the three models of MuseGAN, MTMG, and HRNN, except for the emotion index. The data processing accuracy and error rate of this paper’s model, as well as the training loss value and track matching, are 94.47%, 0.15%, 0.91, and 0.84, respectively, which are better than WGANGP and MuseGAN. The gap between synthesized multi-track music and the music theory rules of real music using the model in this paper is very small, which can fully meet practical needs. The deep learning model constructed in this paper provides a new path for the generation of multi-track music.

show abstract

Exploring deep learning frameworks for multi-track music synthesis

Liu

2024

Applied Mathematics and Nonlinear Sciences

View full text Add to dashboard Cite

show abstract

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

customersupport@researchsolutions.com

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

An Audio-Based SLAM for Indoor Environments: A Robotic Mixed Reality Presentation

Cited by 1 publication

References 76 publications

Exploring deep learning frameworks for multi-track music synthesis

Exploring deep learning frameworks for multi-track music synthesis

Contact Info

Product

Resources

About