AES Store

Journal Forum

Virtual Localization by Blind Persons - July 2012
1 comment

Effect of Spatial Location and Presentation Rate on the Reaction to Auditory Displays - July 2012
1 comment

Watermark-Aided Pre-Echo Reduction in Low Bit-Rate Audio Coding - June 2012
1 comment

Access Journal Forum

AES E-Library

Automated Speech/Other Discrimination for Loudness Monitoring

Inter-program audio level discrepancies continue to plague the content creation and broadcast industries. In particular, end users often are compelled to make adjustments to audio playback levels. It has previously been established that leveling programs based on dialogue loudness can improve listener satisfaction; however, it is difficult to measure as it requires a human to continuously monitor an audio stream and measure only the loudness during speech portions of the content. This paper presents an automated speech discrimination system that provides a means to detect portions of audio content that contain primarily speech. The speech/other discrimination system takes advantage of well known speech characteristics to achieve a total error of 3% despite delay and computational limitations.

Authors:
Affiliation:
AES Convention: Paper Number:
Subject:

Click to purchase paper or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member and would like to subscribe to the E-Library then Join the AES!

This paper costs $20 for non-members, $5 for AES members and is free for E-Library subscribers.

Learn more about the AES E-Library

E-Library Location:

Start a discussion about this paper!


 
Facebook   Twitter   LinkedIn   Google+   YouTube   RSS News Feeds  
AES - Audio Engineering Society