Speech Activity Detection (SAD)

Long-Term Spectral Pseudo-Entropy (LTSPE) Feature

Speech detection systems are known as a type of audio classifier systems which are used to recognize, detect or mark parts of audio signal including human speech. Here, a novel robust feature named Long-Term Spectral Pseudo-Entropy (LTSPE) is proposed to detect speech and its purpose is to improve performance in combination with other features, increase accuracy and to have acceptable performance. Experimental results show that if LTSPE is combined with other features, performance of the detector is improved.

Categories:: Signal Processing

534 Views

Long-Term Multi-Band Frequency-Domain Mean-Crossing Rate (FDMCR) Feature

In order to discriminate and mark audio signal segments which include normal human speech and discriminate segments which do not include speech (like silence, music and noise), Speech/Music Discrimination (SMD) systems are used. Using this definition, SMD systems can be considered as a specific or accurate type of speech activity detection system.

Categories:: Signal Processing

405 Views