Formant Frequencies
Definition
Resonance frequencies of the vocal tract that shape vowel quality. F1 and F2 (the first and second formants) are the most informative for speaker comparison because they reflect both vocal tract anatomy and learned articulation habits.
- Origin
- Vocal tract resonance, not vocal fold vibration
- Most useful pair
- F1 and F2
- Reflects
- Vocal tract anatomy and articulation habits
- Measured via
- Spectrographic or LPC analysis
Common questions
Why are F1 and F2 more useful for speaker comparison than fundamental frequency (F0)?+
F0 is set by vocal fold vibration and varies with mood, effort, and disguise attempts. F1 and F2 depend on the fixed dimensions of the speaker's vocal tract as well as learned articulatory habits, so they tend to be more stable across a recording and harder for a speaker to voluntarily disguise.
Can formant values alone identify a speaker with certainty?+
No. Formant values overlap substantially across different speakers with similar vocal tract sizes, so they function as one class of feature within a broader acoustic-phonetic comparison, not as a standalone identifier comparable to a fingerprint or DNA match.
How does recording quality affect formant measurement?+
Low bandwidth, background noise, compression artefacts, and telephone filtering can distort or obscure formant tracks, particularly higher formants, which is why examiners check recording quality before relying on formant measurements in a comparison.
Related terms
- Forensic Speaker Comparison
- A systematic examination comparing acoustic and phonetic features of a questioned voice recording against known reference recordings of a named individual, expressed...
- I-Vector / X-Vector
- Fixed-length mathematical representations of a speech utterance used in automatic speaker recognition. I-vectors are derived from Gaussian mixture model statistics; x-vectors are...
- IAFPA
- International Association for Forensic Phonetics and Acoustics: the primary professional body for forensic phoneticians and audio analysts, which publishes guidelines on speaker...
- Likelihood Ratio (LR)
- The ratio of two conditional probabilities: the probability of the observed evidence given the prosecution's hypothesis (same source), divided by the probability...
- PLDA (Probabilistic Linear Discriminant Analysis)
- A statistical back-end model used with i-vector and x-vector systems to compute a similarity score between two utterance representations, normalised for within-speaker...