Skip to content

Formant Frequencies

Definition

Resonance frequencies of the vocal tract that shape vowel quality. F1 and F2 (the first and second formants) are the most informative for speaker comparison because they reflect both vocal tract anatomy and learned articulation habits.

Origin
Vocal tract resonance, not vocal fold vibration
Most useful pair
F1 and F2
Reflects
Vocal tract anatomy and articulation habits
Measured via
Spectrographic or LPC analysis

Common questions

Why are F1 and F2 more useful for speaker comparison than fundamental frequency (F0)?+

F0 is set by vocal fold vibration and varies with mood, effort, and disguise attempts. F1 and F2 depend on the fixed dimensions of the speaker's vocal tract as well as learned articulatory habits, so they tend to be more stable across a recording and harder for a speaker to voluntarily disguise.

Can formant values alone identify a speaker with certainty?+

No. Formant values overlap substantially across different speakers with similar vocal tract sizes, so they function as one class of feature within a broader acoustic-phonetic comparison, not as a standalone identifier comparable to a fingerprint or DNA match.

How does recording quality affect formant measurement?+

Low bandwidth, background noise, compression artefacts, and telephone filtering can distort or obscure formant tracks, particularly higher formants, which is why examiners check recording quality before relying on formant measurements in a comparison.

Related terms

Forensic Speaker Comparison
A systematic examination comparing acoustic and phonetic features of a questioned voice recording against known reference recordings of a named individual, expressed...
I-Vector / X-Vector
Fixed-length mathematical representations of a speech utterance used in automatic speaker recognition. I-vectors are derived from Gaussian mixture model statistics; x-vectors are...
IAFPA
International Association for Forensic Phonetics and Acoustics: the primary professional body for forensic phoneticians and audio analysts, which publishes guidelines on speaker...
Likelihood Ratio (LR)
The ratio of two conditional probabilities: the probability of the observed evidence given the prosecution's hypothesis (same source), divided by the probability...
PLDA (Probabilistic Linear Discriminant Analysis)
A statistical back-end model used with i-vector and x-vector systems to compute a similarity score between two utterance representations, normalised for within-speaker...

Explained in

Your journey to becoming a forensic professional starts here.

Practice with mock tests, learn from structured notes, and get your questions answered by a global forensic community, all in one place.