A finished dub can make a conversation sound as though the actors simply performed the scene together in another language. Inside the recording session, the process can look very different. Performers may work alone, record dialogue in short sections and repeat individual lines several times while watching a character whose movements and expressions were completed long before the localized performance began.
Dubbing combines acting with an unusual amount of technical precision. The actor has an adapted script and a completed scene to work from, while the director and recording engineer help align the new performance with the existing picture. Depending on the studio and country, performers may follow audible beeps, visual timing systems or other cues that tell them exactly when dialogue needs to begin. The methods vary, but they are all designed to solve the same problem: creating a new performance without changing the performance already visible on screen.
A Scene Is Often Recorded in Small Pieces
Dubbing actors do not necessarily perform an entire scene from beginning to end. Dialogue is commonly divided into manageable recording units often referred to as cues or loops, allowing the performer and production team to concentrate on a specific line or short passage before moving forward. The actor watches the corresponding section of the picture, hears or studies the source performance, reviews the adapted dialogue and then records a new version against it.
The terminology can vary between dubbing traditions and studios, but the basic process gives the team much tighter control than attempting to record an entire episode continuously. If a character speaks for only a few seconds, that small section can be repeated until the acting and synchronization work together. Longer exchanges can be divided into several pieces, and individual sections can later be assembled into what sounds like one uninterrupted conversation. Caitlin Glass, who has worked extensively as both an actor and ADR director, has described dubbing as a process in which performers are simultaneously dealing with the character, the adapted words and the physical restrictions created by the animation.
The director provides context that may not be obvious from an isolated cue. A character could be responding to something that happened several minutes earlier or concealing information that becomes important later in the story. Because actors frequently record individually, they may not have another localized performer in the room delivering the opposite side of the conversation. The completed picture and original performance provide important references, while the director helps connect each short recording to the larger scene.
Three Beeps Tell the Actor When to Begin
One of the most recognizable techniques in English-language dubbing is the three-beep cue. The performer hears three evenly spaced electronic beeps before the beginning of a line. A fourth beep is not actually played. Instead, the actor learns to enter on the beat where that fourth beep would have occurred. The result is essentially a short count-in that gives the performer a precise starting point without placing another sound over the dialogue being recorded.
Experienced dubbing actors become accustomed to feeling that missing fourth beat while continuing to watch the screen and prepare the performance. The technique makes rhythm an important part of the job. The performer needs to enter at the correct moment but cannot become so preoccupied with the count that the resulting dialogue sounds mechanically timed. Once the line begins, the actor still has to follow the character’s emotion, pacing and visible movements while reaching the end of the dialogue within the available space.
Three beeps are not a universal language of dubbing. Different countries developed their own approaches, and technology has introduced additional visual cueing systems. France has long been associated with the bande rythmo, or rhythm band, in which adapted dialogue and timing information move in synchronization with the picture so performers can follow the text as the scene progresses. Other modern systems can display dialogue, timing markers or visual countdowns on the recording screen. The specific interface may change, but each method is intended to give actors information about when and how their dialogue fits the picture.
Dubbing Actors Perform While Watching Another Performance
Once the cue arrives, the technical exercise becomes an acting exercise. The performer may be watching an animated character or a live-action actor whose facial expressions, gestures and physical reactions cannot be altered to accommodate the new dialogue. The localized actor has to make a performance fit those choices rather than expecting the picture to follow their own timing.
The original-language performance is an important source of information. Actors and directors can listen for emotional intensity, pauses, changes in energy and the relationship between the voice and the physical action. That does not mean the dubbing performer must reproduce every vocal characteristic. As with casting, the objective is generally to preserve the character and intention while creating a performance that sounds believable in the target language. An emotional reaction that works naturally in Japanese, Korean, Spanish or another source language may require slightly different vocal timing when performed in English.
Spoken dialogue is only part of the work. Characters breathe, laugh, cry, gasp, struggle, shout and react physically, and many of those sounds have to be recreated for the localized track. Actors may record efforts associated with fights, falls, running or other physical action, sometimes moving their own bodies to generate a more convincing vocal response. Those moments may contain no conventional words at all, but they still form part of the performance audiences associate with the character.
A Perfectly Timed Take Can Still Be the Wrong Take
Synchronization is one of dubbing’s defining restrictions, but hitting the correct timing does not automatically produce usable acting. A performer can begin and end at exactly the right moments while sounding rushed, flat or emotionally disconnected from the scene. Conversely, the strongest dramatic reading may extend beyond the character’s visible speech. Much of the work inside the session involves finding a version that satisfies both requirements.
Actors can change the speed of a phrase, shift pauses or place emphasis differently to make adapted dialogue fit more naturally. The director may request another take because an emotional beat needs more weight, while the engineer can identify technical timing problems. If the line continues to resist the available space, the wording itself may sometimes be adjusted. Glass has described this kind of collaboration while directing Solo Leveling, where adapted dialogue could continue evolving in the session when the written version did not fit comfortably or did not sound right for the character.
The picture also determines how strict synchronization needs to be at any particular moment. A close-up of a live-action actor speaking directly toward the camera exposes mouth movements clearly, leaving less room for dialogue that visibly conflicts with them. A character shown from behind, at a distance or temporarily off screen can provide greater flexibility. Animation varies as well, from relatively simple mouth cycles to detailed facial animation. Dubbing teams therefore make synchronization decisions cue by cue rather than applying exactly the same standard to every line.
Actors May Perform Entire Conversations Alone
One of the strangest parts of dubbing for someone unfamiliar with the process is that an emotional conversation between several characters may never have been performed by those localized actors together. Individual recording is common in dubbing, allowing productions to schedule performers independently and concentrate on synchronizing one character at a time.
That arrangement places additional responsibility on the actor and director. If the other character’s localized dialogue has not yet been recorded, the performer cannot simply respond to their castmate’s delivery. The source-language track provides one reference, and the director can explain the emotional level, pace and intended relationship between the performances. If another localized actor recorded earlier, their completed material may sometimes be available for playback, giving subsequent performers an additional reference.
The order in which actors record can consequently shape what information is available during a session. Christopher Collet, who has directed English-language dubbing, has described early performers as effectively working in a vacuum because none of the other localized voices exist yet. Later actors can potentially hear more of the developing English track. The finished scene nevertheless needs to sound like a genuine exchange, regardless of whether its participants recorded together, days apart or in completely different sessions.
Pickups and Efforts Fill Out the Performance
Completing the main dialogue does not necessarily complete an actor’s work. A production can request pickups when a recorded line needs to be corrected or replaced. The reason might be an inaccurate pronunciation, a synchronization issue, a technical problem, a change to the approved script or simply a performance that no longer fits once the surrounding scene has been assembled. A pickup might replace an entire sentence or only a small portion of one.
Actors can also spend significant time recording efforts, including exertion sounds, impacts, breaths, cries and other nonverbal reactions. Background voices present another layer. Crowd conversations, indistinct chatter and other ambient vocal performances are commonly known as walla or group ADR in English-language production. Depending on how localization materials are supplied and how the dub is being produced, some background dialogue may need to be recreated so that the audible world surrounding the principal characters also belongs to the target-language version.
Alternate takes can provide editors and directors with additional choices, particularly when a scene can support more than one interpretation. All of these recordings contribute to a soundtrack that audiences eventually experience as a continuous whole. A few seconds of finished audio can therefore represent several takes, timing adjustments and performance decisions that are invisible once the scene is complete.
Individual Recordings Eventually Become a Finished Dub
After the recording session, the localized performances still need to be integrated into the production. Selected takes are edited and aligned with the picture, while dialogue can undergo further technical preparation before the final mix. Problems discovered during review can send specific lines back for pickups, particularly when synchronization, pronunciation or creative intent needs correction.
The localized voices ultimately have to coexist with the production’s music, effects and other soundtrack elements. Mixing determines how dialogue sits within that environment so the new voices do not feel detached from the world on screen. Professional dubbing workflows can also include creative and technical reviews intended to catch inconsistencies before the localized version is delivered. The dubbing director may participate in parts of that review, while engineers, mixers and other post-production specialists handle their respective technical responsibilities.
By that stage, a single conversation may have been assembled from actors who never met during recording, working through dozens of separate cues and performing against voices in another language. Pickups might have been recorded later, while breaths, reactions and background dialogue came from still other sessions. The finished scene hides those divisions.
That is one of the peculiar achievements of dubbing. Recording can involve beeps, timing markers, repeated loops, scripts, screens and constant attention to fractions of a performance that last only a few seconds. The audience is not supposed to notice any of them. All of that technical structure exists so that, when the finished dub plays, viewers can stop thinking about synchronization and simply listen to the characters speak.
