Skip to content

MFCC

Definition

Mel-Frequency Cepstral Coefficients. The dominant acoustic feature set for automatic speaker recognition.

Related terms

FASR
Forensic Automatic Speaker Recognition. Likelihood-ratio framework using statistical speaker models (GMM, i-vectors, x-vectors).
Formants (F1, F2, F3, F4)
Resonant frequencies of the vocal tract that shape vowel quality. F1 and F2 together distinguish vowels (for example /i/ has low F1...
Fundamental Frequency (F0)
Rate of vocal-fold vibration during voiced speech, perceived as pitch. Typical adult male 80 to 180 Hz, adult female 165 to 255...
Likelihood Ratio (LR)
The ratio of two conditional probabilities: the probability of the observed evidence given the prosecution's hypothesis (same source), divided by the probability...
Source-Filter Model
Standard model of voice production: the glottal source (vocal-fold vibration) is filtered by the vocal-tract cavity to produce the speech signal.
Voice Spectrogram (Sonogram)
Time-frequency plot of speech. Wideband (about 300 Hz analysis bandwidth) shows formants; narrowband (about 45 Hz) shows harmonics.
Voiceprint
Lawrence Kersta's 1962 term for the claim that the spectrographic pattern of an individual's speech is unique enough to serve as a...

Explained in

Your journey to becoming a forensic professional starts here.

Practice with mock tests, learn from structured notes, and get your questions answered by a global forensic community, all in one place.