Overview: Vocal fry, also known as creaky voice, is a phonation type characterized by low, irregular vocal fold vibration, producing a rattling or popping sound at the lower end of the speaker's pitch range. In phonetic terms, it involves a shortened and thickened vocal fold configuration with reduced airflow, resulting in a distinct auditory quality. While often a natural feature of many languages (e.g., Danish, Vietnamese), vocal fry has also gained attention in sociolinguistic studies for its perceived associations with certain speech communities and contemporary media discourse.

Vocal fry is defined by a specific mode of vocal fold vibration that differs from modal voice in both articulatory mechanics and acoustic output. This section describes the underlying physiological adjustments and the resulting sound properties.

1.1 Articulatory mechanism

The production of vocal fry involves a tightly adducted but lax vocal fold posture, combined with low subglottal pressure. The arytenoid cartridges are compressed, bringing the vocal processes together while the body of the folds remains relatively slack. Air passes through the glottis in discrete pulses rather than in a continuous stream.

1.1.1 Vocal fold adduction and tension

During vocal fry, the vocal folds are adducted with moderate to high medial compression, but longitudinal tension is reduced. This combination shortens and thickens the folds, increasing their mass per unit length. The thyroarytenoid muscle is often more active than in modal phonation, contributing to the lax, bulky configuration.

1.1.2 Subglottal pressure and airflow

Subglottal pressure in vocal fry is typically lower than in modal voice, ranging from 3 to 8 cm H₂O. Airflow is also reduced—often below 50 mL/s—and is pulsatile rather than steady. Each glottal pulse is followed by a relatively long closed phase, producing the characteristic clicking or rattling quality.

1.2 Acoustic characteristics

The acoustic signal of vocal fry is distinguished by its low fundamental frequency, high cycle-to-cycle variability, and unique spectral shape. These features allow listeners to identify it perceptually.

1.2.1 Fundamental frequency range

The fundamental frequency (F0) of vocal fry generally falls below 70 Hz in adult speakers, often reaching as low as 20–30 Hz. Because the vibration is aperiodic, the F0 may fluctuate widely within a single utterance, sometimes dropping below the threshold of pitch perception.

1.2.2 Spectral properties (jitter, shimmer)

Vocal fry exhibits elevated jitter (frequency perturbation) and shimmer (amplitude perturbation) compared to modal voice. The harmonic spectrum is often irregular, with weak or absent upper harmonics above 1–2 kHz. The dominant energy concentrates in the lower frequencies, giving the sound a rough, creaky timbre.

1.3 Perception and auditory cues

Listeners typically perceive vocal fry as a low-pitched, rattling, or popping sound that occurs at phrase boundaries or on stressed syllables. The perceptual salience arises from the irregular pulse rate and the sharp onset of each glottal burst. In many languages, vocal fry is interpreted as a prosodic or phonemic cue rather than a voice quality error.

Vocal fry serves a range of linguistic roles across the world’s languages. It can differentiate lexical meanings, signal prosodic boundaries, and interact with tonal systems.

2.1 Phonemic contrast in languages

In some languages, vocal fry is a distinctive feature that contrasts with other phonation types, such as modal or breathy voice. These contrasts are often part of the segmental or tonal inventory.

2.1.1 Danish stød

Danish stød is a prosodic feature phonetically realized as a form of laryngealization, often involving vocal fry on a syllable. It functions as a phonemic contrast between minimal pairs (e.g., *hun* [hun] “she” vs. *hund* [hunˀ] “dog”). The presence of stød is marked by creaky voice on the vowel or following sonorant.

2.1.2 Jalapa Mazatec

Jalapa Mazatec, an Otomanguean language of Mexico, distinguishes four phonation types on vowels: modal, breathy, creaky (vocal fry), and a combination of breathy-creaky. Vocal fry in Jalapa Mazatec is a true phonemic contrast, as in *[tʃa̰a̰]* “he scratches” (creaky) vs. *[tʃaa]* “he says” (modal).

2.1.3 Other tonal and non-tonal languages

Vocal fry appears as a phonemic feature in many other languages. In Vietnamese, the *nặng* tone is produced with creaky voice. In some Burmese dialects, creaky voice distinguishes lexical tones. Non-tonal languages like Swedish also use creaky voice to mark certain accents, such as the Accent II in Central Swedish.

2.2 Prosodic marking

Beyond phonemic use, vocal fry frequently marks prosodic boundaries and discourse functions in spontaneous speech.

2.2.1 Boundary signaling

In many languages, vocal fry occurs at the end of intonational phrases, signaling a termination or a continuation of a turn. This is observed in English, German, and Mandarin, where a drop into creaky voice often accompanies a falling intonation at phrase boundaries.

2.2.2 Emphasis and discourse functions

Vocal fry can be used to mark emphasis, emotional affect, or to indicate certainty. In some dialects, speakers employ it parenthetically—for example, on discourse markers like “like” or “so”—to manage turn-taking or claim epistemic authority.

2.3 Interaction with tone and intonation

In tonal languages, vocal fry may be integrated into the tonal system as a register or a contour feature. For instance, in Cantonese, the low falling tone is sometimes realized with creaky voice, especially in older male speakers. In intonation languages, vocal fry can coincide with low pitch targets or phrase-final lowering, enhancing the perceptual contrast between high and low registers.

Vocal fry has attracted considerable attention in sociolinguistics due to its association with specific demographic groups and its variable social evaluation. This section examines its distribution and social meaning.

3.1 Demographic patterns

The occurrence of vocal fry in natural speech is not uniform across populations. Variation by age, gender, region, and social dialect has been documented.

3.1.1 Age and gender distribution

Studies have found that young women in several English-speaking communities (e.g., the United States, United Kingdom, Australia) use vocal fry more frequently than men of the same age. Older speakers, by contrast, tend to use it less often. However, in languages where vocal fry is phonemic (e.g., Danish), gender differences are less pronounced and more tied to dialect.

3.1.2 Regional and social dialects

Vocal fry is reported to be more common in certain regional dialects, such as the “Valley Girl” style in California and the “Aussie” accent. In some British dialects, like Estuary English, creaky voice appears regularly on final syllables. Social groups, such as certain online communities or subcultures, may adopt vocal fry as an identity marker.

3.2 Perceptions and stereotypes

Listener attitudes toward vocal fry vary widely, ranging from neutral acceptance to strong negative evaluations, especially in professional contexts.

3.2.1 Listener attitudes

Experimental studies show that when listeners hear vocal fry in a neutral phrase, they often rate the speaker as less competent, less trustworthy, or less employable—particularly if the speaker is female. However, these judgments are influenced by the listener’s own demographic background and the speaking context.

3.2.2 Vocational and media portrayals

In media and pop culture, vocal fry is frequently associated with young women and with speech styles stereotyped as “lazy” or “overly casual.” Some voice coaches advise against its use in professional settings, while others argue it is a legitimate stylistic choice. Internet memes and comedy sketches have lampooned vocal fry, often focusing on the speech of celebrities or social media influencers.

3.3 Historical and cross-cultural variation

Vocal fry is not a recent phenomenon. Historical recordings from the early 20th century show that creaky voice was present in some dialects of English and other languages. Cross-cultural comparisons indicate that what is considered vocal fry may differ phonetically: some dialects use a longer closure phase, while others use a more irregular pulse. The social meaning also varies—in some cultures, creaky voice is associated with authority (e.g., in certain East Asian traditions), while in others it is stigmatized.

In voice science, vocal fry is studied both as a normal phonatory option and as a sign of potential voice pathology. This section covers its role in voice disorders, assessment, and practical applications.

4.1 Voice disorders involving vocal fry

When vocal fry becomes persistent or occurs involuntarily, it may indicate a voice disorder. Two common conditions are hypofunctional dysphonia and muscle tension dysphonia.

4.1.1 Hypofunctional dysphonia

Hypofunctional dysphonia involves insufficient glottal closure and reduced vocal fold tension, often resulting in a weak, breathy voice that may alternate with creaky voice. Vocal fry in this context is a compensatory mechanism to maintain voicing during low airflow.

4.1.2 Muscle tension dysphonia

In muscle tension dysphonia, excessive laryngeal tension can lead to an adducted, hyperfunctional phonation that sometimes manifests as harsh, pressed vocal fry. The irregular vibration is a byproduct of high medial compression and poor damping.

4.2 Assessment and treatment

Clinicians use acoustic analysis and voice therapy to address vocal fry when it is part of a disorder or when a speaker wishes to modify it.

4.2.1 Acoustic analysis in clinics

In voice clinics, acoustic measures such as jitter, shimmer, and harmonic-to-noise ratio are used to quantify the irregularity typical of vocal fry. The presence of subharmonic frequencies and a low F0 can help distinguish pathological from stylistic vocal fry. Laryngostroboscopy often reveals a diadochokinetic pattern of slow, aperiodic fold vibration.

4.2.2 Voice therapy approaches

Voice therapy for maladaptive vocal fry typically involves breathing exercises, pitch modulation, and relaxation of the laryngeal musculature. Techniques such as resonant voice therapy and the “yawn-sigh” method help clients transition out of creaky phonation. For singers, specialized training may focus on blending vocal fry into modal voice for stylistic effect.

4.3 Phonetic training and forensic use

In phonetics laboratories, vocal fry is used to illustrate phonatory mode distinctions. Forensic linguists analyze vocal fry patterns in speaker identification, as the unique acoustic signature can aid in speaker recognition. Additionally, automatic speech recognition systems must account for vocal fry to avoid misclassification of low-pitched or aperiodic segments.

Vocal fry is a versatile technique in vocal music, ranging from subtle ornamentation to the core of certain vocal styles. Its use raises both artistic and health considerations.

5.1 Contemporary vocal styles

Many modern music genres incorporate vocal fry as a deliberate expressive tool. It adds texture, grit, or emotional weight to a performance.

5.1.1 Pop and R&B vocal fry

I n pop and R&B, vocal fry is often used as a stylistic ornament on sustained notes, especially at the end of phrases. Artists such as Britney Spears, Kesha, and SZA have popularized a “creaky” delivery that conveys intimacy, coolness, or vulnerability. The technique is sometimes termed “vocal fry singing” and is frequently imitated in amateur contexts.

5.1.2 Heavy metal and growl techniques

In heavy metal and other extreme genres, vocal fry forms the basis of “growling” and “screaming” techniques. By engaging the false vocal folds (ventricular folds) along with true vocal fold fry, singers achieve a low, distorted, powerful sound. This distorted vocal fry requires careful control to avoid vocal damage.

5.2 Voice pedagogy and health considerations

Voice teachers and speech-language pathologists debate the safety of voluntary vocal fry in singing. When used sparingly and with proper breath support, vocal fry is considered a low-risk technique. Overuse, however, can lead to vocal fatigue, hoarseness, or nodules if the laryngeal muscles are strained. Contemporary pedagogy increasingly includes exercises to safely access vocal fry without forcing, and many programs now teach “clean” fry as a healthy part of the singer’s palette.