An Unsupervised Hybrid Approach for Online Detection of Sound Scene Changes in Broadcast Content

Sevkin, Gökhan; Craciun, Alexandra; Bäckström, Tom

AES E-Library

An Unsupervised Hybrid Approach for Online Detection of Sound Scene Changes in Broadcast Content

In this paper we describe an online system for broadcast content, which can detect sound scene changes with high accuracy. The system is unsupervised and does not require prior information on the segment classes. A scene change probability score is computed for each frame of the signal using a hybrid approach combining a model-based (Gaussian Mixture Model) with a distance-based (Hotelling’s T2-Statistic) segmentation method. The mixture model parameters are adapted online using the previous frames of the signal. Experiments on real recordings show that we can achieve more than 85% correct segment change detection with only 16% false detections.

Authors: Sevkin, Gökhan; Craciun, Alexandra; Bäckström, Tom
Affiliations: International Audio Laboratories, Erlangen, Friedrich- Alexander-Universität (FAU), Erlangen, Germany; Aalto University, Aalto, Finland(See document for exact affiliation information.)
AES Conference: 2017 AES International Conference on Semantic Audio (June 2017)
Paper Number: P1-2
Publication Date: June 13, 2017 Import into BibTeX
Subject: Semantic Audio
Permalink: https://www.aes.org/e-lib/browse.cfm?elib=18765

Click to purchase paper as a non-member or login as an AES member. If your company or school subscribes to the E-Library then switch to the institutional version. If you are not an AES member and would like to subscribe to the E-Library then Join the AES!

This paper costs $33 for non-members and is free for AES members and E-Library subscribers.

Learn more about the AES E-Library

E-Library Location: /conf/2017/semantic/semantic_audio_2017_paper_4.pdf

Start a discussion about this paper!

AES E-Library

An Unsupervised Hybrid Approach for Online Detection of Sound Scene Changes in Broadcast Content

ABOUT AES

Contact Us