Word Embeddings for Automatic Equalization in Audio Mixing

ORCID

Eduardo Reck Miranda: 0000-0002-8306-9585

Abstract

In recent years, machine learning has been widely adopted to automate the audio mixing process. Automatic mixing systems have been applied to various audio effects such as gain-adjustment, equalization, and reverberation. These systems can be controlled through visual interfaces, providing audio examples, using knobs, and semantic descriptors. Using semantic descriptors or textual information to control these systems is an effective way for artists to communicate their creative goals. In this paper, we explore the novel idea of using word embeddings to represent semantic descriptors. Word embeddings are generally obtained by training neural networks on large corpora of written text. These embeddings serve as the input layer of the neural network to create a translation from words to EQ settings. Using this technique, the machine learning model can also generate EQ settings for semantic descriptors that it has not seen before. We compare the EQ settings of humans with the predictions of the neural network to evaluate the quality of predictions. The results showed that the embedding layer enables the neural network to understand semantic descriptors. We observed that the models with embedding layers perform better than those without embedding layers, but still not as good as human labels.

DOI Link

10.17743/jaes.2022.0047

Publication Date

2022-09-12

Publication Title

Journal of the Audio Engineering Society

Volume

70

Issue

9

ISSN

1549-4950

Acceptance Date

2022-07-25

Embargo Period

2022-10-01

Keywords

Audio Mixing, Automatic Mixing, Equalization, Semantic Word Vectors

First Page

753

Last Page

763

Recommended Citation

Venkatesh, S., Moffat, D., & Miranda, E. (2022) 'Word Embeddings for Automatic Equalization in Audio Mixing', Journal of the Audio Engineering Society, 70(9), pp. 753-763. Available at: 10.17743/jaes.2022.0047

School of Art, Design and Architecture

Word Embeddings for Automatic Equalization in Audio Mixing

ORCID

Abstract

DOI Link

Publication Date

Publication Title

Volume

Issue

ISSN

Acceptance Date

Embargo Period

Keywords

First Page

Last Page

Recommended Citation

Search

Browse

About

Links

School of Art, Design and Architecture

Word Embeddings for Automatic Equalization in Audio Mixing

Authors

ORCID

Abstract

DOI Link

Publication Date

Publication Title

Volume

Issue

ISSN

Acceptance Date

Embargo Period

Keywords

First Page

Last Page

Recommended Citation

Share

Search

Browse

About

Links