movie

A screen accent works only when the actor can stop performing it

6 sources 1 primary source September 8, 2026

Text
A black director's chair labeled Jerome Butler in red on the set of Repo Men, with other production chairs behind it.

Dialect coach Jerome Butler's chair on the Toronto set of Repo Men. The empty seat suits a craft designed to become invisible: the coach's research and notes matter most when the actor no longer appears to be demonstrating them. Photograph via Butler's official site.[6]

Video mode

This article includes 2 embedded videos.

  1. 1 Jerome Butler demonstrates how dialect coaching turns linguistic detail into an actor's usable character voice YouTube embed
  2. 2 Erik Singer analyzes how complete accent systems succeed or break across 32 film performances YouTube embed

A screen accent is often judged like a party trick. A familiar star opens their mouth; the audience waits for a vowel it recognizes; a verdict—convincing or terrible—arrives before the scene has gathered any force. That game mistakes the loudest evidence for the whole craft. An accent is not a bag of unusual pronunciations. It is a working system of sound, rhythm, pitch, resonance, vocabulary, and physical habit, all of which must survive emotion, movement, interruption, and another take.

The better test is dramatic. Does the constructed voice place a character in a particular social world? Can the actor listen and react through it, vary it under pressure, and keep it coherent without sanding away every human inconsistency? Above all, can the technique stop announcing itself? The Los Angeles Times describes the coach's role as deliberately invisible: early preparation should make the accent feel effortless and second nature instead of another task competing with the acting.[3]

The two videos below approach that boundary from opposite sides. The National Endowment for the Arts follows dialect coach Jerome Butler into the process of making a voice playable; WIRED's Erik Singer listens to finished movie performances and takes their machinery apart.[1][2] Together they reveal a craft whose success depends on a paradox: enormous attention must be paid to speech so that, once the camera rolls, the actor can direct attention elsewhere.

First build a voice the character can inhabit

The NEA's 2017 portrait presents Butler not as a vendor of ready-made accents but as a translator between linguistic evidence and an individual actor's imagination.[1] The target is never simply “Boston,” “Dallas,” or “British.” Each label contains generations, neighborhoods, ethnicities, classes, occupations, and private histories. Before correcting a line, the coach and production have to decide whose voice could plausibly belong to this character—and which features matter enough to carry through a whole performance.

That decision begins with listening. Real speakers provide models; the script supplies the character's circumstances; the actor supplies a body and a way of learning. The coach turns those materials into a limited, usable plan rather than an encyclopedia. Butler's stated aim is to establish enough command of the new sound that imagination can take over.[1] Accuracy is therefore not an ornamental finish applied to performance. It is a foundation that must eventually bear improvisation, fatigue, speed, and feeling.

Watch how often Butler's explanation returns from an abstract accent to the particular performer in front of him. A sound can be taught through phonetic notation, a recording, imitation, or a physical prompt; no route is intrinsically the serious method. The serious method is the one the actor can retrieve during a scene. The Los Angeles Times notes that techniques must adapt to how an individual learns, and that video models let performers study the mouth, body language, and head movements as well as sound.[3] The voice is not separable from the person producing it.

This also explains why an ensemble accent cannot be created by handing everyone the same recording. A family or community needs shared features, but identical voices would sound as artificial as unrelated ones. On The Lord of the Rings, coach Roisin Carty and the production sometimes used one cast member's existing speech as a reference for others in the same family group; the invented languages also drew on Tolkien's descriptions and specialist recordings.[4] Coherence came from a common sound world, not mass-produced pronunciation.

The cover photograph adds the missing production reality. Butler's named chair sits among other department chairs on the Toronto set of Repo Men.[6] Dialect coaching is part of the set's labor network, subject to the same shrinking daylight, setup changes, and finite number of takes as the camera or costume departments. Preparation must therefore compress into quick, legible interventions. Carty describes giving discreet prompts from outside the frame and choosing the moment for a correction so that an actor does not repeat an error—or become so conscious of the correction that the scene stiffens.[4] Knowledge alone is insufficient; timing is part of the craft.

Then listen past the souvenir sounds

If Butler's portrait moves from evidence toward performance, Singer's WIRED video moves backward from performance toward evidence. Across 32 movie accents, he listens for systems rather than isolated “got it” words.[2] That scale is useful. Once examples sit beside one another, the ear begins to notice that a plausible accent is distributed across ordinary connective speech: the shape of a vowel, whether an r is sounded, where energy sits in the mouth, how a phrase rises or settles, how rapidly syllables collide.

The format can look like a ranking exercise, but its deeper value is diagnostic. Singer frequently separates what a performance sustains from what it merely signals.[2] A conspicuous consonant may identify a region for a second; rhythm and vocal posture have to keep identifying a person between the conspicuous moments. The distinction matters because audiences are especially alert to famous “marker” sounds. An actor can hit those markers and still leave the surrounding language in their habitual voice.

The most productive way to watch is not to memorize Singer's approvals and objections. Listen for the levels at which he diagnoses a choice. Pronunciation is one level, but melody, tempo, resonance, and the baseline arrangement of the mouth are others.[2] A strong performance coordinates them; a partial one may reproduce a few surface sounds while leaving its deeper timing unchanged. The lesson for viewers is generosity as much as precision: accent work is continuous behavior, and a single imperfect syllable need not erase an otherwise coherent vocal life.

Carey Mulligan's preparation to play Felicia Montealegre in Maestro clarifies why a real person raises the burden. She listened repeatedly to long recorded interviews and worked with Tim Monich, a dialect coach who had once met Montealegre.[5] Those sources offered more than a regional label: the record of a particular multilingual life and an eyewitness memory of the woman's vocal presence. A fictional character can be built from a plausible composite; a public figure must negotiate with an archive audiences may know. Both require selection. Neither is achieved by finding one supposedly pure regional sample and copying it without remainder.

That selection is also social. The Los Angeles Times reports that coaches may become cultural consultants, gathering speech models and tracing the cultural and anthropological reasons people speak as they do.[3] Carty's work on Middle-earth makes the same point through invention: even a language that has never had native speakers needs relationships, hierarchies, and group distinctions before it can feel inhabited.[4] A screen accent is credible when it describes a life, not merely when it names a place.

Continuity without a cage

Cinema adds a peculiar demand: spontaneous behavior must remain editable. A scene may be filmed in fragments over hours or days, with emotional peaks captured before their beginnings. The constructed voice has to carry across wide shots, close-ups, pickups, and later dialogue replacement. Yet rigid sameness would be another failure. People accelerate, retreat, code-switch, swallow words, and lose polish under stress. The coach protects a field of plausible variation rather than enforcing one immaculate reading.

This is why the pre-production model and the on-set relationship are inseparable. The Los Angeles Times describes a process that can begin with phonetics and recordings months before filming, continue on set, and extend into re-recorded dialogue in postproduction.[3] Carty emphasizes diplomacy: the right note delivered at the wrong moment can make the performer monitor their mouth instead of their scene partner.[4] Mulligan's repeated listening and Monich's personal memory of Montealegre show how archival evidence can be translated into one actor's preparation rather than copied mechanically.[5] The common principle is not softness. It is precision about where precision belongs.

The videos ultimately shift the standard for watching accents. “Could I detect the actor's original voice?” is a narrow question; traces may remain without damaging the character. Better questions are whether the choices are specific, related, sustainable, and dramatically responsive. Singer's close listening helps reveal the system in the finished work.[2] Butler's process shows why that system must become light enough to carry.[1]

An accent succeeds, then, not when every line sounds like a flawless audition sample, but when the voice can think. It must be prepared enough to stay rooted and flexible enough to be surprised. The dialect coach builds that freedom from recordings, phonetic decisions, bodily experiments, repetition, and carefully timed silence. When the work is finished, the audience should not hear a lesson being recited. It should hear a character with somewhere to come from and something urgent to say.

Sources

  1. National Endowment for the Arts, “Reading Between the Lines,” YouTube video, January 19, 2017 — Jerome Butler on making dialect work specific, internalized, and usable in performance.
  2. WIRED, “Movie Accent Expert Breaks Down 32 Actors' Accents,” YouTube video, November 16, 2016 — Erik Singer's comparative analysis of vocal posture, regional detail, consistency, and character-specific speech.
  3. Ada Tseng, “Explaining Hollywood: How to get a job as a dialect coach.” Los Angeles Times, February 6, 2023 — preparation methods, actor-specific teaching, cultural context, and the coach's role across production.
  4. Anna Tims, “How do I become … a dialect coach.” The Guardian, October 15, 2015 — Roisin Carty on The Lord of the Rings, ensemble voices, physical prompts, and discreet on-set correction.
  5. Lindsey Bahr, “How Carey Mulligan became Felicia Montealegre in ‘Maestro.’” Associated Press, December 20, 2023 — recorded interviews, multilingual biography, and collaboration with dialect coach Tim Monich.
  6. Jerome Butler, official site — on-set photographic archive, including Butler's production chair during the Toronto shoot of Repo Men.
Previous Dry Summer makes ownership look like a blockage Next The Louma crane moved the camera by leaving its operator on the ground

Recommended In movie

Matched by subject and format