Schedule as of Oct 11, 2022 - subject to change

Default Time Zone is EDT - Eastern Daylight Time

Back To Schedule
Thursday, October 27 • 1:45pm - 2:15pm
Comparison of Audio Spectral Features in a Convolutional Neural Network

Sign up or log in to save this to your schedule, view media, leave feedback and see who's attending!

Time-Frequency transformation and spectral representations of audio signals are commonly used in various machine learning applications. Typically the Mel-Spectrogram is used to create the input features to the network justified by the mel scale’s human auditory system basis. In this paper, we compare several spectral features in a gender detection speech model comparing their performance and showing that the Mel-Spectrogram is not always the best choice for input features.

avatar for Elias Nemer

Elias Nemer

Audio Engineer
Elias Nemer is involved in various aspects of audio signal processing related to AR and VR.He is currently an audio engineer at Meta Platforms.He previously worked at Cirrus Logic in areas related to ML for audio classification. At Broadcom, he developed audio algorithms for mobile... Read More →

Thursday October 27, 2022 1:45pm - 2:15pm EDT
Online Papers
  Applications in Audio
  • badge type: ALL ACCESS or ONLINE