Endre søk
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Blind Subband Beamforming With Time-Delay Constraints for Moving Source Speech Enhancement
Ansvarlig organisasjon
2007 (engelsk)Inngår i: IEEE Transactions on Audio, Speech, and Language Processing, ISSN 1558-7916, E-ISSN 1558-7924, Vol. 15, nr 8, s. 2360-2372Artikkel i tidsskrift (Fagfellevurdert) Published
Abstract [en]

A new robust microphone array method to enhance speech signals generated by a moving person in a noisy environment is presented. This blind approach is based on a two-stage scheme. First, a subband time-delay estimation method is used to localize the dominant speech source. The second stage involves speech enhancement, based on the acquired spatial information, by means of a soft-constrained subband beamformer. The novelty of the proposed method involves considering the spatial spreading of the sound source as equivalent to a time-delay spreading, thus, allowing for the estimated intersensor time-delays to be directly used in the beamforming operations. In comparison to previous approaches, this new method requires no special array geometry, knowledge of the array manifold, or acquisition of calibration data to adapt the array weights. Furthermore, such a scheme allows for the beamformer to efficiently adapt to speaker movement. The robustness of the time-delay estimation of speech signals in high noise levels is improved by making use of the non-Gaussian nature of speech trough a subband Kurtosis-weighted structure. Evaluation in a real environment with a moving speaker shows promising results, with suppression levels of up to 16 dB for background noise and interfering (speech) signals, associated to a relatively small effect of speech distortion.

sted, utgiver, år, opplag, sider
IEEE , 2007. Vol. 15, nr 8, s. 2360-2372
HSV kategori
Identifikatorer
URN: urn:nbn:se:bth-8479DOI: 10.1109/TASL.2007.903309Lokal ID: oai:bth.se:forskinfo978C238BCD0BA042C12574A500544EE6OAI: oai:DiVA.org:bth-8479DiVA, id: diva2:836203
Tilgjengelig fra: 2012-09-18 Laget: 2008-08-14 Sist oppdatert: 2025-09-30bibliografisk kontrollert

Open Access i DiVA

fulltekst(1200 kB)483 nedlastinger
Filinformasjon
Fil FULLTEXT01.pdfFilstørrelse 1200 kBChecksum SHA-512
c2be519f3b4017cecd218e906fa80b0ed4f6909d9071b4a816e66b398d3c6ad9c5ded6c1ce8b33e482f84825b47c0177dbdb0626646191c7ba97ee76c325d7be
Type fulltextMimetype application/pdf

Andre lenker

Forlagets fulltekst

Person

Claesson, Ingvar

Søk i DiVA

Av forfatter/redaktør
Claesson, Ingvar
I samme tidsskrift
IEEE Transactions on Audio, Speech, and Language Processing

Søk utenfor DiVA

GoogleGoogle Scholar
Totalt: 484 nedlastinger
Antall nedlastinger er summen av alle nedlastinger av alle fulltekster. Det kan for eksempel være tidligere versjoner som er ikke lenger tilgjengelige

doi
urn-nbn

Altmetric

doi
urn-nbn
Totalt: 346 treff
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf