• Home
  • Search
  • Nonparametric Bayesian sparse factor analysis for frequency domain blind source separation without permutation ambiguity
  • Open Access IconOpen Access
  • Cite Icon2
  • https://doi.org/10.1186/1687-4722-2013-4Copy DOI Icon

Nonparametric Bayesian sparse factor analysis for frequency domain blind source separation without permutation ambiguity

Show More
  • Abstract
  • Highlights & Summary
  • PDF
  • Literature Map
  • References
  • Citations
  • Similar Papers
Abstract

Blind source separation (BSS) and sound activity detection (SAD) from a sound source mixture with minimum prior information are two major requirements for computational auditory scene analysis that recognizes auditory events in many environments. In daily environments, BSS suffers from many problems such as reverberation, a permutation problem in frequency-domain processing, and uncertainty about the number of sources in the observed mixture. While many conventional BSS methods resort to a cascaded combination of subprocesses, e.g., frequency-wise separation and permutation resolution, to overcome these problems, their outcomes may be affected by the worst subprocess. Our aim is to develop a unified framework to cope with these problems. Our method, called permutation-free infinite sparse factor analysis (PF-ISFA), is based on a nonparametric Bayesian framework that enables inference without a pre-determined number of sources. It solves BSS, SAD and the permutation problem at the same time. Our method has two key ideas: unified source activities for all the frequency bins and the activation probabilities of all the frequency bins of all the sources. Experiments were carried out to evaluate the separation performance and the SAD performance under four reverberant conditions. For separation performance in the BSS_EVAL criteria, our method outperformed conventional complex ISFA under all conditions. For SAD performance, our method outperformed the conventional method by 5.9–0.5% in F-measure under the condition RT20 = 30–600 [ms], respectively.

Loading PDF

Similar Papers
  • Research Article
  • Citations2

Joint algorithm for permutation problem in frequency-domain blind speech source separation

  • Aug 20, 2008
  • Journal of Computer Applications
  • Jian-Jian Zhu
  • Research Article
  • Citations5

Single-channel blind source separation based on attentional generative adversarial network

  • Nov 18, 2020
  • Journal of Ambient Intelligence and Humanized Computing
  • Xiao Sun +4
  • Conference Article
  • Citations16

Acoustic event source localization for surveillance in reverberant environments supported by an event onset detection

  • Apr 01, 2015
  • Peter Transfeld +4
  • Conference Article
  • Citations10

IVA algorithms using a multivariate Student's t source prior for speech source separation in real room environments

  • Apr 01, 2015
  • Waqas Rafique +3
  • Conference Article
  • Citations3

Evaluation of fast convergence algorithm for blind source separation of real convolutive mixture

  • Jan 01, 2002
  • H Saruwatari +5
  • Conference Article
  • Citations9

Blind source separation based on fast-convergence algorithm using ICA and beamforming for real convolutive mixture

  • May 01, 2002
  • Hiroshi Saruwatari +4
  • PDF
  • Research Article
  • Citations17

Independent vector analysis based on overlapped cliques of variable width for frequency-domain blind signal separation

  • May 23, 2012
  • EURASIP Journal on Advances in Signal Processing
  • Intae Lee +1
  • Research Article
  • Citations5

Associative Memory Model-Based Linear Filtering and Its Application to Tandem Connectionist Blind Source Separation

  • Mar 01, 2017
  • IEEE/ACM Transactions on Audio, Speech, and Language Processing
  • Motoi Omachi +2
  • Conference Article
  • Citations14

Independent Positive Semidefinite Tensor Analysis in Blind Source Separation

  • Sep 01, 2018
  • Rintaro Ikeshita
  • Research Article
  • Citations12

Blind separation of non-stationary sources using continuous density hidden Markov models

  • Mar 19, 2013
  • Digital Signal Processing
  • Fanglin Gu +2
  • Conference Article
  • Citations4

Weakly Supervised Sound Activity Detection and Event Classification in Acoustic Sensor Networks

  • Dec 01, 2019
  • Janek Ebbers +4
  • Conference Article
  • Citations1

Frequency Domain Blind Source Separation Exploiting Higher-Order Dependencies

  • May 14, 2006
  • Taesu Kim +3
  • PDF
  • Research Article
  • Citations7

Evaluations on underdetermined blind source separation in adverse environments using time-frequency masking

  • Oct 23, 2013
  • EURASIP Journal on Advances in Signal Processing
  • Ingrid Jafari +3
  • Conference Article

Proposal of BSS method to separate the respiratory sound and the heart sound

  • Dec 01, 2019
  • Yuki Kubota +2
  • Conference Article
  • Citations1

Stereo Source Separation in the Frequency Domain: Solving the Permutation Problem by a Sliding K-means Method

  • May 01, 2019
  • Bang-Yin Chen +2
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.