Skip to main navigation Skip to search Skip to main content

Combining Novel Acoustic Features using SVM to Detect Speaker Changing Points

Haishan Zhong, David Cho, Vladimir Pervouchine, Graham Leedham

Research output: Contribution to conferencePaper

Abstract

Automatic speaker change point detection segments different speakers from continuous speech according to speaker characteristics. This is often a necessary step before applying speaker verification or identification systems. Among the features to represent a speaker in the speaker change point detection systems acoustic features are commonly used. Commonly used features are Mel Frequency Cepstral Coefficients (MFCC) and Linear Prediction Cepstral Coefficients (LPCC). However, the features are affected by speech content, environment, type of recording device, etc. So far, no features have been discovered, which values depend only on the speaker. In this paper four novel feature types proposed in recent major journals and conference papers for speaker verification problem, are applied to the problem of speaker change point detection. The features are also used to form a combination scheme via SVM classifier. The results shows that the proposed scheme improves the performance of speaker changing point detection as compared to the system that uses MFCC features. It was also found that some of the novel features of low dimensionality give comparable speaker change point detection accuracy to the high-dimensional MFCC features.
Original languageEnglish
Pages224-227
Publication statusPublished - 2008
EventBIOSIGNALS 2008: International Conference on Bio-inspired Systems and Signal Processing - Funchal, Portugal
Duration: 28 Jan 200831 Jan 2008

Conference

ConferenceBIOSIGNALS 2008: International Conference on Bio-inspired Systems and Signal Processing
CityFunchal, Portugal
Period28/01/0831/01/08

Keywords

  • Image Processing
  • Pattern Recognition and Data Mining
  • Computer Vision

Fingerprint

Dive into the research topics of 'Combining Novel Acoustic Features using SVM to Detect Speaker Changing Points'. Together they form a unique fingerprint.

Cite this