REPORT kumatani:rr08-06/IDIAP Adaptive Beamforming with a Maximum Negentropy Criterion Kumatani, Kenichi McDonough, John Klakow, Dietrich Garner, Philip N. Li, Weifeng EXTERNAL http://publications.idiap.ch/attachments/reports/2008/kumatani-idiap-rr-08-06.pdf PUBLIC Idiap-RR-06-2008 2008 IDIAP \begin{abstract} In this paper, we address an adaptive beamforming application in realistic acoustic conditions. After the position of a speaker is estimated by a speaker tracking system, we construct a subband-domain beamformer in \emph{generalized sidelobe canceller} (GSC) configuration. In contrast to conventional practice, we then optimize the \emph{active weight vectors} of the GSC so as to obtain an output signal with \emph{maximum negentropy} (MN). This implies the beamformer output should be as non-Gaussian as possible. For calculating negentropy, we consider the $\Gamma$ and the generalized Gaussian (GG) pdfs. After MN beamforming, Zelinski post-filtering is performed to further enhance the speech by removing residual noise. Our beamforming algorithm can suppress noise and reverberation without the signal cancellation problems encountered in the conventional adaptive beamforming algorithms. We demonstrate the effectiveness of our proposed technique through a series of far-field automatic speech recognition experiments on the \emph{Multi-Channel Wall Street Journal Audio Visual Corpus} (MC-WSJ-AV). On the MC-WSJ-AV evaluation data, the delay-and-sum beamformer with post-filtering achieved a word error rate (WER) of 16.5\%. MN beamforming with the $\Gamma$ pdf achieved a 15.8\% WER, which was further reduced to 13.2\% with the GG pdf, whereas the simple delay-and-sum beamformer provided a WER of 17.8\%. \end{abstract}

<subfield code="a">REPORT</subfield>

</datafield>

<subfield code="a">kumatani:rr08-06/IDIAP</subfield>

</datafield>

<subfield code="a">Adaptive Beamforming with a Maximum Negentropy Criterion</subfield>

</datafield>

<subfield code="a">Kumatani, Kenichi</subfield>

</datafield>

<subfield code="a">McDonough, John</subfield>

</datafield>

<subfield code="a">Klakow, Dietrich</subfield>

</datafield>

<subfield code="a">Garner, Philip N.</subfield>

</datafield>

<subfield code="a">Li, Weifeng</subfield>

</datafield>

<subfield code="i">EXTERNAL</subfield>

<subfield code="u">http://publications.idiap.ch/attachments/reports/2008/kumatani-idiap-rr-08-06.pdf</subfield>

<subfield code="x">PUBLIC</subfield>

</datafield>

<subfield code="a">Idiap-RR-06-2008</subfield>

</datafield>

<subfield code="b">IDIAP</subfield>

</datafield>

<subfield code="a">\begin{abstract} In this paper, we address an adaptive beamforming application in realistic acoustic conditions. After the position of a speaker is estimated by a speaker tracking system, we construct a subband-domain beamformer in \emph{generalized sidelobe canceller} (GSC) configuration. In contrast to conventional practice, we then optimize the \emph{active weight vectors} of the GSC so as to obtain an output signal with \emph{maximum negentropy} (MN). This implies the beamformer output should be as non-Gaussian as possible. For calculating negentropy, we consider the $\Gamma$ and the generalized Gaussian (GG) pdfs. After MN beamforming, Zelinski post-filtering is performed to further enhance the speech by removing residual noise. Our beamforming algorithm can suppress noise and reverberation without the signal cancellation problems encountered in the conventional adaptive beamforming algorithms. We demonstrate the effectiveness of our proposed technique through a series of far-field automatic speech recognition experiments on the \emph{Multi-Channel Wall Street Journal Audio Visual Corpus} (MC-WSJ-AV). On the MC-WSJ-AV evaluation data, the delay-and-sum beamformer with post-filtering achieved a word error rate (WER) of 16.5\%. MN beamforming with the $\Gamma$ pdf achieved a 15.8\% WER, which was further reduced to 13.2\% with the GG pdf, whereas the simple delay-and-sum beamformer provided a WER of 17.8\%. \end{abstract}</subfield>

</datafield>

</record>

</collection>