Category: Re-recording

  • How Do You Make ADR Sound Like It Was Never Replaced? Paul Carden and Chris Navarro on Performance, Technology, and the Art of Dialogue Replacement

    Paul Carden and Chris Navarro

    How do you make ADR sound like it was never replaced?

    A line of dialogue may last only a few seconds, yet replacing it convincingly can require an extraordinary combination of preparation, performance, technology and judgement. The original production recording must first be identified as unusable, the replacement carefully documented and prepared, the actor returned to the emotional and physical circumstances of a performance recorded months earlier, and the new dialogue captured so that it matches the timing, vocal quality, microphone perspective and acoustic character of the original scene. If the process succeeds, the audience should never know that any of this work happened. During their joint online guest lecture for Edinburgh Napier University, ADR supervisor Paul Carden and ADR mixer Chris Navarro demonstrated this complete process from beginning to end. Rather than discussing Automated Dialogue Replacement only in theory, they created a deliberately compromised line of production dialogue, prepared it for replacement, recorded it on an ADR stage and evaluated the result. Their demonstration revealed a process in which meticulous preparation and technical fluency serve a deceptively simple objective: allowing everyone involved to concentrate upon the performance.

    Carden began not in a recording studio, but outside beside a car. His objective was to demonstrate how an apparently simple piece of dialogue could become unusable during production. Wearing a lavalier microphone, he performed the five-word line “I’m late for work” while getting into the vehicle. The exercise immediately revealed how vulnerable production dialogue can be. Clothing could obscure the microphone, a zipped jacket could change the sound, movement could complicate recording and the closing car door could mask the line itself. The example was deliberately simple. Carden was standing almost still, knew exactly what he was going to say and had constructed the situation specifically for the demonstration. On a real production, actors may be walking, running, interacting with objects and performing emotionally demanding scenes while the production sound team manages multiple radio microphones and one or more booms. From this perspective, the surprise is often not that some lines require replacement, but that so much production dialogue survives at all.

    The demonstration also established one of the central tensions within ADR. Production dialogue is not simply speech recorded on location. It contains a performance created at a particular moment, within a particular physical environment, as part of an interaction with other performers. Months later, an actor may arrive on an ADR stage while already immersed in an entirely different project and be asked to recreate a few seconds from that earlier performance. The recording environment offers little of the original context. The actor stands in a comparatively neutral room, watches an image on a screen, listens for cues and attempts to reproduce not merely the words, but the emotional and physical conditions under which those words were originally spoken. Carden emphasised that this is one reason actors can find ADR difficult. Matching a line involves returning to a performance that may no longer feel immediate or familiar.

    Before any actor reaches the stage, however, the line must become an ADR cue. Carden demonstrated this preparation process by locating the damaged dialogue within the picture, defining the cue and documenting the reason for replacement. The cue was assigned a unique identifier, linked to its scene information and accompanied by notes explaining that the car door had closed over the line. He also demonstrated the value of professional flexibility. Where a production line was imperfect but potentially usable, he might identify the replacement as optional rather than forcing an unnecessary argument over whether it must be replaced. The cue sheet then carried the information required by the recording stage, including project details, version information, character and actor names, cue number, timecode and recording information. A five-word performance therefore arrived at the stage supported by an extensive information system designed to ensure that everybody was working on the correct material.

    Version control was particularly important. Carden described production teams as living and dying by version dates, reflecting the practical reality that picture changes can quickly make carefully prepared cues inaccurate. Cue numbering performs a similarly essential role. Each take is voice-slated so that the recording itself retains its identity even if paperwork becomes separated from the audio. These details may appear administrative when viewed from outside professional production, yet they protect the continuity of the entire process. An ADR session may contain hundreds of replacement lines recorded across multiple days, actors and facilities. The actor should not have to think about whether the cue is correctly identified or whether the stage is working from the right picture version. Preparation creates the stability within which performance can happen.

    Carden also offered one deceptively simple instruction for aspiring ADR mixers: keep recording until somebody clearly indicates that the take has ended. An actor may continue the performance, a director may decide to record another version immediately or the session may move into wild recording without a formal interruption. Stopping too early can lose useful material for no meaningful benefit. Storage is cheap; an unrecoverable performance is not. The advice reflected a wider principle that would continue through the lecture. Good ADR practice depends upon remaining attentive to what is happening in the room rather than allowing the machinery of recording to dictate the session.

    The second stage of the demonstration moved into Navarro’s ADR facility. Even though the project consisted of a single line created for the lecture, he established the session as though it were a conventional production. Folder structures, documents, Pro Tools sessions and media were organised according to the same principles he would use for a project involving one session or hundreds. His ADR template was already configured for the ordinary demands of recording, while remaining adaptable when unusual situations arose, such as multiple actors performing together with several microphones each. A good template does not prescribe every session in advance. It removes predictable technical work so that attention remains available for situations that cannot be predicted.

    Even the imported picture introduced a practical lesson. Navarro noticed that the system was responding sluggishly and identified the compressed H.264 picture as a likely cause, since the computer had to decode it continuously during playback. The moment was minor, but revealing. Professional technical fluency often appears not through dramatic troubleshooting, but through the ability to recognise quickly why a system is behaving unexpectedly and continue without allowing the problem to dominate the room. Earlier, Carden had offered similarly pragmatic advice about computer failure: save the work, restart the system and continue rather than allowing panic to consume valuable time. Both speakers treated technical expertise as calm familiarity rather than technological display.

    Microphone selection then revealed how ADR matching differs from conventional voice recording. The objective is not to capture the most beautiful possible version of the actor’s voice. It is to create a recording that can inhabit the existing production soundtrack without attracting attention. Navarro showed that the stage had previously been configured for voice-over work, with microphones positioned closely and directly to produce the clear, present sound required for narration. ADR demanded a different approach. Carden had recorded the original line using a lavalier, so the replacement needed to reproduce the qualities of that production perspective rather than simply offering a technically superior recording.

    Carden explained that ADR sessions commonly record both boom and lavalier microphones, ideally using the same models employed during production. Even when the original line appears to come predominantly from one microphone, alternatives can prove unexpectedly useful. Navarro demonstrated this by positioning a shotgun microphone close to Carden but substantially off-axis. Pointing it directly at the performer would have produced a cleaner and more extended sound, but that was not necessarily desirable. The off-axis position rejected particular frequencies and created a tonal character that could potentially sit closer to the lavalier recording. The result was counterintuitive: the boom microphone could, under those conditions, sound more like the production lavalier than the replacement lavalier itself. Matching therefore began not with assumptions about microphone categories, but with listening.

    This distinction between technical quality and contextual suitability runs throughout professional ADR. A beautifully recorded line can fail if it sounds too clean, too close, too rich or too controlled for the image and surrounding production dialogue. Conversely, a microphone position that would appear unconventional in another recording context may provide exactly the spectral character required for a convincing match. The mixer must understand microphones well enough to use their imperfections deliberately. The question is never simply which microphone sounds best. It is which recording can become part of the scene without revealing the process that created it.

    Once Carden stepped in front of the microphone, the demonstration moved from equipment towards performance. The familiar system of three audible beeps established the timing, with the actor beginning the line where an imaginary fourth beep would occur. In principle, the task appeared straightforward. In practice, Carden repeatedly entered early, became self-conscious about the timing and discovered how quickly an apparently trivial five-word line could become difficult once performance, synchronisation and technical awareness competed for attention. His reaction gave the students an unusually useful demonstration of the psychological demands placed upon actors during ADR. Knowing the line is not enough. Understanding the timing is not enough. Even achieving synchronisation is not enough. The replacement still has to sound as though it belongs to the original performance.

    At one point, Carden produced a take that was synchronised correctly but immediately recognised that the performance itself was wrong. Navarro’s response was subtle. He did not ask for shouting or a dramatically louder delivery. He suggested only slightly more projection. The difference was not simply level. A small change in the physical production of the voice altered its tone and gave the line the additional presence heard in the original recording. When the replacement was compared directly with production dialogue, the improvement became obvious. The exercise demonstrated that ADR matching cannot be reduced to waveform alignment, pitch correction or equalisation. Performance changes the spectrum of the voice before any microphone or processor becomes involved.

    Navarro later developed this point in greater depth. Two performances can have similar apparent loudness and pitch while differing substantially in vocal quality. An actor performing naturally on set may have a relaxed throat and a particular physical relationship with the surrounding scene. On the ADR stage, tension, self-consciousness or the effort to satisfy technical instructions can change the voice. A performer may become tighter, brighter or less natural even while reproducing the words and timing accurately. Matching therefore requires attention to qualities that are difficult to describe numerically. Projection, pitch, volume and rhythm matter, but so do muscular relaxation, breath and the physical origin of the voice.

    This presents the ADR mixer with a delicate problem. The mixer may hear precisely what is preventing a line from matching, yet communicating every technical observation to the actor may make the performance worse. Navarro warned that performers can absorb only so many notes before they begin thinking about the mechanics of speech rather than the character. A useful intervention must therefore translate technical listening into language that supports performance. Sometimes the correct decision is to offer a suggestion. Sometimes it is to communicate through the director. Sometimes it is to recognise that an imperfection is not important enough to justify disturbing the creative balance of the room. Hearing a problem and knowing whether to mention it are separate professional skills.

    The lecture also demonstrated an older approach to ADR recording that remains remarkably useful. Navarro sampled the original production line and played it repeatedly into Carden’s headphones. Instead of concentrating simultaneously upon picture, beeps and performance, Carden could hear the original line and immediately reproduce it, repeating the process several times. Navarro then edited the resulting takes into position and compared them with the production recording. The method offered a direct reference for timing, intonation, volume and vocal character, allowing the performer to respond to the sound of the original performance rather than attempting to reconstruct every detail intellectually.

    Navarro’s explanation of this technique revealed an important insight into perception. Picture can sometimes become a distraction. An actor watching for a mouth movement may wait until it becomes visually apparent, by which time the correct moment to begin has already passed. If the original audio is correctly synchronised to picture, matching its rhythm and timing can naturally reproduce picture sync. The performer can therefore concentrate upon hearing and responding rather than continually monitoring several streams of information at once. Despite the age of the sampling technique, Navarro regarded it as one of the most useful tools available on the stage. Its continued value comes not from technological sophistication, but from the way it simplifies the performer’s task.

    The same objective shaped Navarro’s control room. His system contained extensive routing, multiple microphone inputs, separate monitoring paths and a heavily customised control surface, yet the purpose of this complexity was to make the session itself feel simple. He might need to manage separate mixes for the control room, recording stage, actor headphones, supervisor headphones and remote participants, with each requiring different material during rehearsal, recording and playback. Attempting to reconfigure every route manually between passes would slow the session and create repeated opportunities for error. Navarro therefore designed systems that allowed complex changes to happen immediately.

    His customised keypad provided a particularly revealing example. Individual buttons could trigger sequences of macros that armed tracks, initiated recording and performed repetitive editing operations once a take had finished. Navarro had spent months developing the core system and continued refining it whenever he noticed himself repeating unnecessary keyboard operations. He had begun his career as an ADR recordist and understood the division of labour on a two-person stage, where one person could manage recordings and files while the mixer concentrated upon the performers. Working alone, he used automation to reproduce some of that support, delegating repetitive technical operations to macros so that his attention could remain directed towards the stage.

    This led to one of Navarro’s most important ideas: the ADR mixer must work at the speed of creativity. An actor or director may suddenly discover a new approach to a line and want to record it immediately. If the mixer responds by asking everyone to wait while tracks are configured, routes changed or files prepared, the idea may lose its immediacy. A delay of only a few seconds can alter the atmosphere of a performance. Technical speed therefore has a human purpose. The mixer learns the system thoroughly enough that machinery does not interrupt thought.

    Navarro connected this principle to advice he had received about respected ADR mixer Tommy O’Connell. Asked what distinguished O’Connell’s work, a sound editor gave a simple answer: he anticipates. Useful actions are completed before anyone needs to request them. Navarro interpreted this not as a mysterious talent, but as the result of attention. If the mixer is watching the stage, listening to conversations and understanding the direction in which a session is moving, preparation for the next action can begin before a formal instruction arrives. Anticipation therefore depends upon both technical readiness and social awareness. A mixer whose attention is buried in the workstation may complete every requested task correctly while still remaining one step behind the session.

    Carden’s preparation of the cue and Navarro’s management of the stage reveal two sides of the same professional process. Carden reduces uncertainty before the session begins through accurate cueing, documentation, version control and communication. Navarro reduces friction during the session through templates, routing, automation and anticipation. Neither form of preparation is intended to make the process more rigid. Both preserve the possibility that an actor, director or mixer can respond immediately when something unexpected and valuable occurs.

    Microphone signal flow offered another example of this balance. Navarro described a deliberately clean recording path from microphone through preamplifier into Pro Tools, avoiding unnecessary outboard processing. Within the workstation, compression and equalisation could be used sparingly, but he warned against making irreversible decisions without good reason. Recording completely flat preserves maximum flexibility for the re-recording mixer, while careful decisions made during the session can still produce useful, committed tracks. Heavy processing may be difficult to undo later. A technically impressive recording decision is not valuable if it reduces the options available to the production.

    Navarro’s session was configured to make microphone comparison immediate. Several microphone inputs could remain available, feeding dedicated record tracks at appropriate levels. A boom and lavalier could be recorded simultaneously as discrete channels, preserving both perspectives for later evaluation. The workflow reflected the uncertainty inherent in matching. The microphone expected to work best may not produce the most convincing result once the line is placed against production. Recording alternatives gives later editors and mixers material from which to construct the most believable transition.

    His discussion of recording format was equally pragmatic. For conventional ADR, he worked at the established production standard of 48 kHz, 24-bit audio. Higher sample rates could be valuable for sound effects intended for extensive manipulation, where recordings might later be slowed or processed heavily. ADR serves a different purpose. Recording at unnecessarily high rates would increase storage requirements before the material was eventually converted to the format used by the production. The appropriate technical choice depends upon what will happen to the sound. More data is not automatically more useful.

    As the lecture progressed, it became increasingly clear that the most difficult parts of ADR were not contained within any equipment specification. Navarro estimated that learning the technical fundamentals and becoming comfortable with Pro Tools took years, yet he distinguished that competence from the broader ability to run a session. Recording dialogue in sync with picture is ultimately a technical process that can be learned. Managing actors, directors, supervisors, performance, uncertainty and the emotional energy of the room represents a different level of expertise.

    Actors may arrive nervous. Directors may be highly collaborative, completely self-sufficient or resistant to suggestions. A performer may struggle with a line while becoming increasingly aware of the difficulty. The mixer must understand how much intervention the situation can support. Navarro described the importance of creating a calm and comfortable environment, particularly when clients are unfamiliar with the process. Confidence can be communicated without dominance. If someone is uncertain, the mixer can explain what will happen, answer questions and demonstrate that the session is under control. Technical authority becomes useful when it reduces anxiety rather than displaying superiority.

    Carden and Navarro also acknowledged the professional judgement involved in deciding when not to intervene. A mixer may recognise an aspect of a performance that could be improved, yet nobody has asked for technical input and the issue may not be important enough to disrupt the session. Navarro framed this as a balancing act. Would the intervention materially improve the line? Is the director open to suggestions? Will another note distract the actor from a performance that is already working? Expertise includes recognising problems, while professional maturity requires distinguishing consequential problems from imperfections that do not matter.

    The ADR stage therefore brings technical precision and emotional sensitivity into the same process. The mixer must hear minute changes in vocal quality, understand microphone behaviour, maintain synchronisation, manage several monitoring environments and operate the recording system almost instinctively. At the same time, attention must remain on body language, conversation, confidence, frustration and creative momentum. Mastery of the technology is essential precisely so that the technology no longer consumes the attention needed elsewhere.

    Carden’s demonstration also placed the scale of professional ADR into perspective. Their five-word line required production recording, cue preparation, documentation, stage setup, microphone selection, rehearsal, multiple takes, alternative recording methods, editing and comparison. A skilled actor might complete around ten or twelve conventional cues in an hour, while a feature film may contain 150 or 200 ADR lines. Complicated performances naturally take longer. What appears to an audience as a few moments of seamless dialogue can therefore represent days of concentrated work.

    Yet the success of that work is measured largely through its invisibility. When Navarro compared Carden’s replacement line with the production recording, the evaluation concerned relationships: volume, tone, vocal quality, perspective and the way the replacement interacted with the surrounding scene. The door slam that had originally damaged the line was restored around the new performance, returning the replacement to the event from which it had temporarily been separated. Heard alone, the ADR recording was merely a voice on a stage. Placed back into the scene, it became part of an action.

    This transformation captures the philosophy shared by both speakers. ADR is often described as the replacement of unusable dialogue, but their demonstration showed that replacement is only the beginning of the problem. The objective is reconstruction. The actor reconstructs a performance. The mixer reconstructs a microphone perspective. Editors reconstruct timing and continuity. The final soundtrack reconstructs the relationship between voice, action and environment so successfully that the audience experiences a single uninterrupted moment.

    The joint nature of the lecture made this especially clear. Carden approached ADR through the complete production process, showing how a line travels from location problem to documented cue and finally to the recording stage. Navarro approached the same process from inside the control room, revealing the technical systems, listening skills and interpersonal judgement required to capture a convincing replacement. Their perspectives met at the point where professional preparation serves human performance.

    By the end of the session, the deliberately damaged line had become something much larger than a technical demonstration. It revealed why ADR demands more than synchronisation, why the cleanest microphone is not always the correct microphone, why a performer can match timing while missing the voice, why an old sampler can remain useful in a modern digital workflow and why a highly automated control room can make a session feel more human rather than less. Above all, Carden and Navarro showed that successful ADR depends upon where attention is directed. The actor should be thinking about performance rather than machinery. The director should be thinking about the scene rather than routing. The mixer should be watching and listening to the room rather than fighting the workstation. The audience, finally, should be thinking about none of these things. When every part of the process works together, they simply hear a character speak.

  • How Do You Mix Television Sound Under Pressure? Frank Morrone on Dialogue, Workflow, and the Art of Re-Recording

    Frank Morrone

    How do you mix television sound under pressure?

    A television soundtrack may contain hundreds of dialogue recordings, sound effects, Foley performances, backgrounds, ADR takes and music stems, all competing for space within a mix that must remain clear, emotionally convincing and technically suitable for broadcast. The audience should never become aware of that complexity. They should simply understand every line, believe every environment and remain absorbed in the story. During his online guest lecture for Edinburgh Napier University, re-recording mixer Frank Morrone explored the craft behind achieving that apparent simplicity. Drawing upon a career in film and television that began in 1979, and projects including Lost, The Strain, Sleepy Hollow and Criminal Minds: Beyond Borders, he revealed a discipline shaped equally by technology, organisation, collaboration and judgement. Throughout the session, one principle emerged repeatedly. The most effective mixing workflows allow enormous technical complexity to disappear behind the story.

    Morrone began by tracing a career that developed across several different areas of professional audio. His earliest work took place in music studios, recording jazz and orchestral film scores before following those recordings into the dubbing theatre and becoming increasingly interested in the way complete soundtracks were assembled. A move into post-production allowed him to work across dialogue editing, music editing and Foley recording before concentrating upon re-recording mixing. That breadth of experience shaped the collaborative philosophy running throughout the lecture. Unlike a music recording session, where one engineer may remain closely involved from recording through to the final mix, film and television sound brings together work created by many different specialists. The dub stage is where those contributions finally meet. Successful mixing therefore depends upon understanding not only the material itself, but the people, processes and decisions that produced it.

    Television makes this collaboration particularly demanding. Morrone described an industry in which track counts have continued to increase while schedules and budgets have become progressively tighter. Sophisticated surround mixes must be created rapidly, and emerging formats add further complexity without removing the need to support conventional playback systems. His response is not simply to work faster. It is to design workflows that remove unnecessary decisions from the mixing stage. As soon as he joins a project, he communicates with the supervising sound editor about track requirements and provides a starting template so that incoming material already fits an established structure. Organisation begins before the mixer enters the room. Under severe time pressure, the ability to find, control and compare material immediately becomes part of the creative process itself.

    This reveals something deeper about Morrone’s understanding of expertise. His templates, early conversations with sound editors, knowledge of production microphones, preparation of alternative takes and habit of printing completed passes all anticipate problems before they are allowed to interrupt the mix. The same thinking extends beyond the dubbing theatre. He considers how broadcast processing will react to dynamics, how a mix will translate to domestic systems and how material created for one format will behave when heard through another. Professional experience, in this sense, is not simply the ability to solve problems quickly. It is the ability to recognise where problems are likely to emerge and construct a workflow in which many of them have already been addressed before they become urgent.

    The scale of Lost provided a striking illustration. Morrone showed the students sessions containing extraordinary numbers of elements, including dozens of tracks dedicated to the Smoke Monster alone. Its identity emerged from a deliberately ambiguous combination of animal voices, pneumatic machinery, roller-coaster wheels and other contrasting sources, creating something that resisted being understood as either entirely organic or entirely mechanical. Those effects existed alongside hard effects, backgrounds, Foley, production dialogue, ADR, group recordings and substantial music deliveries. The technology available at the time imposed strict limits upon voices and processing, requiring careful decisions about resource allocation as well as creative balance. Complexity could not simply be solved by adding more processing. The session itself had to be organised so that the mixers could navigate it instinctively.

    Custom fader layouts and VCA groups became essential to that process. Morrone described arranging controls so that principal dialogue could immediately be balanced against ADR, group recordings and music, while different categories of material remained independently accessible. Music sources could be separated from score, while dialogue in different languages could be isolated for international deliverables. Every layer of organisation reduced the time between hearing a problem and solving it. This became especially important during pilot season, when mixers might receive unusually elaborate material without first having time to develop a workflow around the programme. A sufficiently flexible template must already be capable of accommodating whatever arrives. Preparation, in this context, creates the conditions in which creative decisions can still be made under pressure.

    The production of Lost also demonstrated the tension between creative ambition and delivery requirements. Morrone recalled the exceptional resources devoted to the programme, including a large pilot budget and Michael Giacchino’s insistence upon recording a live orchestra for each episode. Yet the soundtrack still had to survive the restrictions of television broadcast. The team therefore created a more dynamic version for DVD before producing a contained broadcast mix designed to survive transmission processing. A similar approach was later adopted on The Strain. The distinction mattered. A mix can remain technically within specification and still behave poorly when subsequent broadcast processing responds to excessive dynamics. Re-recording therefore requires mixers to think beyond the dubbing theatre. They are mixing for every system through which the programme will eventually reach its audience.

    Dialogue occupied the centre of Morrone’s approach. His first objective is always to preserve the production performance wherever possible. When ADR has been recorded, he wants to know why. A line replaced for performance reasons presents a different problem from one replaced to solve a technical fault. If the director wanted a different performance, Morrone respects that decision while keeping the original available as an alternative. If the problem was technical, he first explores whether the production recording can be repaired. Modern restoration tools have dramatically expanded what can be rescued, reducing the need to replace performances that may possess subtleties difficult to recreate months later in an ADR studio.

    When ADR is necessary, matching involves far more than applying equalisation and reverb. Morrone keeps production dialogue available alongside the replacement so that he can compare transitions directly, matching tone, acoustic environment, pacing and performance. He prefers to receive several strong takes and recordings from both boom and lavalier microphones, recognising that the microphone apparently closest to the production perspective is not always the easiest to integrate. Understanding which microphones and wireless systems were used during production can also provide valuable clues, particularly where transmission systems have imparted their own sonic characteristics. ADR matching consequently becomes a process of reconstructing relationships rather than searching for a single corrective setting.

    Technology has transformed that work, though Morrone repeatedly warned against allowing powerful restoration tools to encourage excessive processing. Noise reduction, spectral repair, ambience matching, EQ matching and dereverberation can rescue material that would previously have required replacement. Yet he deliberately removes less noise than might appear necessary when dialogue is heard in isolation. Once backgrounds, effects and music return, much of the remaining noise may be perceptually masked. Processing that sounds impressively clean in solo can leave dialogue lifeless and constricted in the finished mix. Morrone therefore keeps copies of original material and sometimes returns to less processed versions during the final mix. The objective is not the cleanest possible dialogue track. It is dialogue that remains natural and convincing within the complete soundtrack.

    His attitude towards restoration reveals a broader philosophy of technology. Morrone began working with magnetic tape, a Cat 43 noise reduction unit and a notch filter, and he clearly values the extraordinary capabilities available to contemporary mixers. Yet greater technical power has not removed the need for judgement. In many respects, it has increased it. The ability to remove more noise does not mean that more noise should be removed, just as the ability to place sound almost anywhere within an immersive field does not mean that every available position should be used. New tools expand the range of possible decisions. They do not determine which decisions are appropriate. Throughout Morrone’s lecture, technical capability remained subordinate to perception.

    This distinction between isolation and context became one of the lecture’s most important ideas. Morrone described television mixing as beginning with dialogue and music, establishing the foundation against which the effects mixer can develop backgrounds and action. Once those elements come together, the mixers decide what should drive the scene. Some moments belong to music, others to effects, while still others require a more subjective perspective or a deliberate reduction of material. Experienced mixing partners develop an almost instinctive understanding of one another’s decisions. If an effect obscures a line during an early pass, Morrone knows that a trusted colleague will create space for it during refinement. Collaboration is therefore more than the division of tracks between two people. It is a shared process of deciding where the audience should listen.

    The idea that sounds must be judged in context reaches far beyond balancing dialogue against effects. Dialogue that appears too noisy when soloed may become entirely convincing once the world of the scene surrounds it. ADR that attracts attention after twenty repeated comparisons may pass unnoticed when the audience encounters it once within a continuous performance. Group recordings that sound absurd under isolated scrutiny may perform their role perfectly when placed at the correct distance behind principal dialogue. Morrone’s examples repeatedly challenged the assumption that individual elements should be perfected independently. The meaningful unit of evaluation is ultimately the audience’s experience of the scene. A soundtrack succeeds through relationships between sounds rather than the isolated perfection of its components.

    Normally, the priority within those relationships remains dialogue. Morrone described it as both the foundation of the mix and the principal carrier of storytelling. His concern with intelligibility, however, extends beyond maintaining a technical hierarchy between dialogue, music and effects. It is fundamentally about preserving attention. His most revealing test is not a meter reading but a listener asking what somebody has just said. At that moment, the audience member has been pulled out of the story and required to think about the failure of the soundtrack. Clear dialogue therefore supports immersion precisely by avoiding attention to itself. The better the mix communicates, the less the audience needs to think about the process of communication.

    Yet the rule is not absolute. For one sequence in The Family, a child was being used as bait to attract a kidnapper in a crowded shopping centre. The environment needed to feel genuinely busy, yet the available collection of separate crowd recordings and reverberant elements never created a convincing whole. Morrone took a six-channel recorder into a real shopping centre, captured the food court from different perspectives and brought the recordings back into the mix. The result allowed the environment to crowd the dialogue slightly, which was precisely what the scene required. Clarity remains fundamental, but realism sometimes depends upon controlled difficulty.

    The shopping-centre sequence illustrates an important tension within Morrone’s approach. His practice is built upon strong principles, but those principles do not become inflexible rules. Dialogue normally takes priority, except when allowing the environment to interfere with it makes the dramatic situation more believable. Restoration should preserve intelligibility, except when excessive cleaning destroys naturalness. Acoustic simulation is valuable, except when recording the real physical relationship between people and space produces a more convincing result. Expertise therefore involves knowing the rules well enough to understand what they protect, then recognising the moments when the needs of the scene justify bending them.

    This willingness to leave the dubbing theatre and record real spaces appeared repeatedly throughout the session. Morrone discussed impulse responses, convolution reverbs and carefully developed presets for rooms, vehicles and other environments, all of which provide useful starting points for worldizing sound. Yet he remained pragmatic about their limitations. A real location does not necessarily sound like the space suggested by the finished image, and even a carefully captured impulse response may require additional reflection, delay or reverberation before it feels convincing. Cars present an especially difficult problem, combining strong early reflections from glass with highly absorptive surfaces elsewhere. Experience gradually produces a library of useful starting points, though listening still determines the final result.

    Sometimes no simulation is as convincing as returning to the physical situation itself. Morrone described a scene in which a group of foster children were supposed to be creating chaos upstairs while a conversation took place below. Studio-recorded group voices did not reproduce the peculiar combination of footfalls, structural transmission and reflections travelling down a staircase into another room. His solution was direct. He gathered children on the upper floor of a house, recorded from downstairs and captured the complete acoustic event as it occurred. No increasingly elaborate chain of processing was required. The physical relationship between performers, building and microphone provided what the scene needed.

    The pressures of television production make such judgement particularly important. Morrone compared television mixing to boot camp. Schedules leave little room for hesitation, and mixers must develop workflows capable of producing strong results quickly. Once a scene has been successfully mixed, he prefers to print it rather than trusting that automation will remain untouched throughout later work. Accidental writes, changed sends and technical errors can occur even in experienced hands. Printing completed work provides security and allows later changes to be punched into established stems. Efficiency does not mean rushing blindly. It means reducing the opportunities for avoidable problems to consume the limited time available.

    Long sessions also introduce a more human limitation: hearing fatigue. Morrone described working on the highly dynamic soundtrack of The Strain, where sustained exposure to loud material forced him to think deliberately about auditory recovery. His solution was simple but important. He left the room for short periods, walked and allowed his ears to recover while his mixing partner continued working. The two mixers could alternate demanding passes, giving each other opportunities for rest without stopping the session. After decades of mixing, one of the useful discoveries was not another plug-in or processor, but the value of leaving the chair for ten minutes. Professional listening depends upon recognising the limits of the listener.

    Morrone’s discussion of client relationships revealed another dimension of the re-recording mix that students may rarely encounter in technical demonstrations. Mixers are working with directors, producers and other clients who may have strong preferences that differ from their own. Morrone described situations in which clients wanted music loud enough to compete with dialogue. His responsibility was to explain the likely consequences, demonstrate how the mix translated at lower levels and on smaller monitors, and search for a compromise that preserved the client’s intention while protecting intelligibility as far as possible. The mixer offers expertise, but does not own the programme. Knowing which decisions are worth challenging and which require accommodation is part of the craft.

    His description of deciding whether an issue represents a “hill to die on” reveals a sophisticated understanding of professional authority. Expertise does not give the mixer unlimited control over the work, nor does collaboration require the abandonment of professional judgement. The mixer must advocate for the audience, explain likely consequences and make alternatives audible, while recognising that the final creative intention belongs to the client. Professional judgement therefore includes negotiation. Sometimes expertise means defending a decision. Sometimes it means finding a compromise that neither side initially imagined. Occasionally, it means implementing a choice that remains contrary to personal taste while ensuring that it works as successfully as possible.

    Technical fluency plays an important role in maintaining those relationships. When a client requests a change, Morrone wants to make it immediately, play it once and continue. Searching through tracks or repeatedly troubleshooting a familiar process changes the atmosphere of the room and interrupts attention to the programme itself. A well-designed session keeps the conversation focused upon storytelling and communication. The deeper the mixer’s command of the tools, routing and session layout, the less those systems intrude upon the creative discussion.

    This is another form of transparency. Morrone repeatedly returned to the idea that technology should disappear from the client’s experience, yet transparency does not mean that technology has become unimportant. The opposite is closer to the truth. Considerable technical knowledge is required to make complex systems feel immediate. Templates, custom layouts, routing, monitoring, printed stems and intimate familiarity with the workstation create an environment in which a creative request can become an audible result without breaking the flow of the session. Mastery becomes visible through the absence of friction.

    The growth of immersive sound has expanded this challenge further. Morrone described mixers as working between extremes, from Dolby Atmos and sophisticated home theatres to stereo playback and mobile devices with earbuds. His philosophy was to begin with the best mix possible in the most capable format, then ensure that it translates successfully into simpler ones. Immersive technology may offer extraordinary spatial possibilities, but Morrone remained cautious about using novelty without considering perception. In particular, he defended dialogue remaining anchored to the centre. Moving voices between speakers can alter timbre and create distracting changes that audiences may notice without understanding their source. New formats create possibilities, though they do not invalidate principles developed through decades of listening.

    His argument about dialogue placement is particularly revealing. Audiences do not need to identify the technical source of a problem in order to experience discomfort. Morrone described viewers sensing that something was wrong when dialogue moved around an immersive field, even when they could not explain precisely what disturbed them. This places an unusual responsibility upon the mixer. Professional listening must sometimes diagnose experiences that ordinary listeners can feel but cannot name. The purpose of expertise is not to dismiss those responses as technically uninformed, but to understand the perceptual conditions that produced them.

    The same concern for translation shaped his approach to low frequencies. Subwoofers vary enormously between listening environments, and domestic listeners frequently adjust them far beyond calibrated levels. Morrone therefore warned against depending entirely upon the LFE channel for the weight of a soundtrack. Low-frequency energy can also be carried through the main channels, creating a result that remains powerful across a wider range of playback systems. His experience of hearing a domestic subwoofer struggle with low-frequency material from Sleepy Hollow reinforced the point. A soundtrack must survive real listening environments, not merely sound impressive on a perfectly calibrated dubbing stage.

    As the discussion widened, Morrone considered the future possibility of mixes that adapt more intelligently to different devices and contexts. Streaming, immersive audio, virtual reality and personalised playback were already creating pressure for soundtracks to function across radically different systems. Yet his underlying philosophy remained remarkably consistent. Whatever the format, begin with the strongest possible mix, preserve the storytelling hierarchy and understand how human perception responds to the result. Technology changes quickly. The responsibility to guide attention and communicate narrative does not.

    Questions from the students returned the discussion to practical preparation. Morrone strongly supported editors delivering material that is already sensibly balanced before it reaches the stage. Dialogue editors who use clip gain to create consistent levels save valuable mixing time, while backgrounds and effects arriving close to useful operating levels allow the mixer to begin creatively rather than first correcting avoidable problems. He described requesting particular editors for demanding productions precisely for this reason. Good preparation is noticed. In a professional environment where a pilot may need to be mixed in only a few days, the person who consistently delivers well-organised, intelligently balanced material becomes someone mixers actively want on the next project.

    His discussion of ADR offered another deceptively simple lesson about perception. Morrone sometimes works on difficult replacement lines privately through headphones while the effects mixer is making a pass. The client then hears the finished line only once in context rather than listening to it repeated dozens of times during adjustment. Repetition directs attention towards the repair and teaches the listener exactly where to expect it. The same awareness shaped his humorous rule about never soloing loop group in front of a client. Background conversations that work perfectly as part of a scene may sound absurd when isolated and scrutinised. What matters is not whether every element survives examination on its own, but whether it performs its intended role in the scene.

    These examples reveal that Morrone is not simply mixing sound. He is managing attention, expectation and knowledge. Once listeners have been taught where an edit exists, they may hear it differently. Once an element has been isolated, they may judge it according to criteria that have little relevance to its actual purpose. Perception is shaped not only by acoustic information, but by what listeners have been encouraged to notice. Part of the mixer’s craft therefore lies in protecting the audience’s experience from unnecessary awareness of the mechanisms used to construct it.

    Yet group recording could also become a powerful storytelling tool. On Criminal Minds: Beyond Borders, episodes moved between international locations while much of the production remained based in Los Angeles. Carefully performed local-language group recordings, combined with music and other environmental elements, became essential to establishing each location convincingly. Here, group material could be brought forward rather than hidden. There was no universal rule governing how loudly an element should be mixed. Its appropriate level depended upon what the scene needed to communicate.

    By the end of the session, Morrone’s account of re-recording mixing had moved far beyond faders, plug-ins and delivery specifications. The technology matters enormously, as do templates, routing, restoration tools, monitoring and control surfaces, but those things serve a larger process. A mixer must understand performance, storytelling, perception, collaboration, translation and the subtle politics of working with clients under pressure. The session may contain hundreds of tracks, yet the audience should hear a coherent world rather than the complexity required to construct it. Morrone’s lecture revealed a craft built upon anticipation, contextual judgement and the careful management of attention. Preparation preserves the possibility of creativity under pressure. Technical knowledge allows technology to disappear from the conversation. Rules provide essential foundations, while experience reveals when the needs of a scene require them to bend. Great television sound is not created by making every element impressive or every recording perfect in isolation. It emerges from understanding what the audience needs to hear, recognising what they should never need to notice, and making hundreds of individual decisions feel like one continuous experience.

  • How Do You Make an Orchestra Fit Inside a Television Show? Phil McGowan on Recording, Mixing, and the Sound of Star Trek: Picard

    Phil McGowan

    How do you make an orchestra fit inside a television show?

    At first glance, the answer appears straightforward. Musicians gather in a studio, microphones are placed around the room, a conductor raises a baton, and the music is recorded. Yet during his online guest lecture for Edinburgh Napier University, recording and mixing engineer Phil McGowan revealed a process that is considerably more complex. Drawing upon his work on Star Trek: Picard, McGowan described a world of orchestral recording that combines musical performance, engineering, editing, production management, and problem-solving. By the end of the lecture, it became clear that recording an orchestra is only one small part of a much larger process. Throughout the lecture, McGowan repeatedly returned to the importance of preparation, organisation, and communication. Although microphones, software, and recording techniques played important roles, many of the challenges he described ultimately concerned coordinating people, decisions, and workflows across an unusually complex production process.

    McGowan began by introducing the recording sessions for the third season of Star Trek: Picard. Across ten episodes, the score was recorded using large orchestral forces, with most episodes featuring a sixty-five-piece ensemble recorded at Warner Brothers Studios in Burbank. For the majority of the season, the orchestra was divided across separate recording sessions. Strings and woodwinds were recorded together, while brass was recorded later. Only the final episode brought the entire eighty-piece orchestra into the room simultaneously. Although audiences often imagine a film score as a single orchestra performing together, McGowan explained that modern production frequently relies upon these layered recording approaches. Recording sections separately provides greater flexibility during mixing while allowing music editors and dubbing mixers more control later in the production process.

    Yet even before a note is recorded, a surprising number of decisions have already been made. The placement of every section within the room affects both the recording and the eventual mix. Strings, woodwinds, brass, piano, harp, and other instruments each occupy carefully chosen positions. Microphone placement becomes equally important. Looking at the recording diagrams shown during the lecture, it was difficult not to be struck by the sheer number of microphones involved. Individual sections receive dedicated spot microphones, larger groups receive overhead microphones, and the entire orchestra is captured by an array of room microphones positioned high above the ensemble.

    What was particularly interesting, however, was McGowan’s repeated emphasis that the most important microphones are often not the closest ones. In a well-designed scoring stage, much of the orchestra’s character emerges from a relatively small number of carefully positioned room microphones. Spot microphones provide detail, definition, and control, though the overall impression of the orchestra often comes from the way the ensemble interacts with the acoustic space itself. Rather than constructing an orchestral sound entirely from individual instruments, the recording process begins with capturing the orchestra as a unified musical body.

    This relationship between detail and cohesion appeared repeatedly throughout the lecture. Modern recording technology allows engineers to place microphones extremely close to instruments. Individual players can be isolated with remarkable precision. Yet McGowan’s approach demonstrates considerable restraint. Spot microphones are available when needed, though many remain relatively low in the final mix. The objective is not to maximise separation. Instead, it is to preserve the sense that listeners are hearing a single orchestra performing together within a shared acoustic environment.

    Recording the orchestra is only the beginning. Once the sessions finish, the material enters a complex process of editing and mixing. Here, McGowan’s role becomes particularly interesting. The raw recordings arrive alongside extensive collections of programmed material supplied by the composer. Modern television scores often combine live orchestral recordings with sampled instruments, synthesizers, percussion libraries, pads, textures, and electronic elements. One of the mixer’s responsibilities is deciding how these different layers should coexist.

    What emerged from the lecture was a strong preference for using the live recordings whenever possible. Sampled instruments often provide useful support, additional weight, or subtle reinforcement, though McGowan repeatedly emphasised that the live orchestra remains the foundation of the sound. The samples are rarely intended to replace the musicians. Instead, they are carefully blended into the mix where appropriate.

    Organisation becomes essential at this stage. Large orchestral sessions generate enormous numbers of tracks. Strings, brass, woodwinds, percussion, piano, harp, synthesizers, effects, and auxiliary elements all require separate management. McGowan demonstrated how sessions are organised into stems, allowing different components of the score to be adjusted independently later in the production process. These stems become particularly important when the music eventually reaches the dubbing stage, where it must coexist with dialogue, sound effects, Foley, ambience, and every other element of the soundtrack.

    This relationship between music and the rest of the soundtrack formed one of the most revealing parts of the discussion. Audiences often imagine that a score reaches the screen in essentially the same form in which it leaves the recording studio. McGowan demonstrated that the reality is considerably more complicated. The music mixer occupies a position between composition and final dubbing, shaping material that must eventually coexist with dialogue, Foley, ambience, sound effects, and every other component of the soundtrack.

    This creates an unusual challenge. During the mixing process, the final soundtrack often does not yet exist. Dialogue may still be evolving. Effects tracks may be incomplete. Editorial changes may continue arriving. The mixer therefore works partly with the present version of the programme and partly with an anticipated future version. Decisions must account not only for what is currently on screen but also for what will eventually happen when the material reaches the dubbing stage.

    In this sense, music mixing becomes an act of translation. The composer’s intentions need to remain intact, though they must also survive the practical realities of television production. A passage that sounds spectacular in isolation may compete with dialogue once the final soundtrack is assembled. A delicate orchestral texture may disappear beneath effects. A dramatic crescendo may need flexibility if the editorial structure changes. The mixer therefore balances musical priorities with narrative requirements, ensuring that the score remains expressive while still serving the larger needs of the programme.

    McGowan described the importance of communication throughout this process. Conversations with composers, music editors, producers, and re-recording mixers help establish how the material will ultimately be used. Stem structures become especially valuable here. By separating different orchestral and electronic elements into organised groups, later stages of production retain the flexibility needed to support storytelling decisions. What appears to be a purely technical workflow is therefore deeply connected to narrative concerns.

    Seen in this light, the music mixer occupies a remarkably important position within the production chain. The role involves much more than balancing levels or applying plug-ins. It requires understanding composition, orchestration, recording, editing, post-production, and storytelling simultaneously. The objective is not simply to make the music sound good. The objective is to ensure that the music can fulfil its dramatic function once every other element of the soundtrack is finally assembled.

    Questions of storytelling therefore remain central throughout the process. Although the lecture contained detailed discussions of microphones, reverbs, routing structures, and plug-ins, these technical topics were rarely presented as ends in themselves. Instead, they were framed as tools supporting dramatic communication. Reverb is not merely an acoustic effect. It helps create scale, atmosphere, and emotional character. Stem structures are not simply organisational devices. They provide flexibility for storytelling. Even microphone choices ultimately serve narrative goals.

    A particularly striking example emerged in McGowan’s discussion of reverberation. For Star Trek: Picard, the production deliberately embraced a more expansive orchestral sound inspired by earlier generations of science-fiction scoring. Rather than pursuing absolute clarity or dryness, the score was allowed to inhabit larger acoustic spaces. The resulting sound connects contemporary production practices with earlier traditions of science-fiction scoring associated with composers such as Jerry Goldsmith and James Horner. Listening to McGowan describe these decisions, it became clear that technical choices often carry historical and aesthetic significance as well.

    The lecture also offered a fascinating glimpse into the practical realities of large-scale media production. Television schedules are rarely generous. Recording sessions must fit within union regulations, musicians’ availability, studio bookings, editorial deadlines, and dubbing schedules. Scores are often recorded while other parts of the production remain unfinished. Picture edits may continue evolving. Visual effects may still be in development. Deadlines continue approaching regardless.

    Under such conditions, consistency becomes invaluable. McGowan described how recording setups, templates, routing structures, and mixing approaches are designed to remain stable across multiple episodes. Establishing reliable systems allows creative decisions to happen more efficiently. Rather than reinventing workflows repeatedly, engineers can focus their attention on the musical and dramatic needs of each project.

    Another recurring theme throughout the lecture was collaboration. Large orchestral productions depend upon extensive networks of expertise. Composers, orchestrators, contractors, recording engineers, Pro Tools operators, music editors, re-recording mixers, musicians, producers, and showrunners all contribute to the final result. No individual controls every aspect of the process. Instead, successful productions emerge through coordination between specialists whose work overlaps at crucial moments.

    Listening to McGowan describe recording sessions, one gains a strong sense of the trust involved. Musicians are trusted to perform complex scores with remarkable efficiency. Engineers are trusted to capture those performances accurately. Music editors are trusted to manage revisions and conforming. Dubbing mixers are trusted to integrate the score into the larger soundtrack. The finished music reflects not only technical skill but also a highly collaborative production culture.

    Perhaps the most interesting aspect of the lecture was the way it challenged romantic ideas about orchestral recording. Popular accounts often focus on dramatic moments: the orchestra enters the room, the conductor raises a baton, and the music comes to life. Those moments certainly exist. Yet McGowan’s account suggests that the real craft often lies elsewhere. It lies in preparation, organisation, consistency, communication, editing, and the countless small decisions that allow large productions to function successfully.

    Looking back across the lecture, what emerges most clearly is not simply a story about recording orchestras. It is a story about connecting different stages of a creative process. Recording sessions, editing workflows, stem preparation, music mixing, and final dubbing all form part of a chain in which every decision influences what follows. Managing that chain requires technical expertise, though it also requires communication, anticipation, and an understanding of how music functions within narrative storytelling. Every stage of the process involves balancing competing demands. Technical precision must coexist with musical expression. Flexibility must coexist with consistency. Individual details must support larger dramatic goals. The orchestra must sound impressive in its own right while still serving the needs of the programme.

    For students interested in recording, mixing, or film music production, this may be the lecture’s most valuable lesson. Technology remains important. Microphones matter. Software matters. Recording techniques matter. Yet none of these elements exist in isolation. They are part of a larger system whose purpose is ultimately narrative. The audience does not hear microphone placements, stem structures, or routing templates. They hear music supporting a story.

    For Phil McGowan, the challenge is not simply recording an orchestra. The challenge is shaping hundreds of performances, thousands of audio tracks, and countless technical decisions into something that helps bring a fictional world to life. By the time audiences sit down to watch Star Trek: Picard, most of that work has become invisible. The orchestra feels as though it simply belongs there. Achieving that illusion, however, requires an extraordinary amount of craft.

  • What Did They Say? Gary Bourgeois on Dialogue, Attention, and the Art of Film Mixing

    Gary Bourgeois

    What happens when an audience misses a line of dialogue?

    At first glance, the consequences seem relatively minor. A viewer leans towards a friend. Someone quietly asks for clarification. A sentence is repeated. Yet during his online guest lecture for Edinburgh Napier University, veteran re-recording mixer Gary Bourgeois suggested that this moment reveals something important about the relationship between sound and storytelling. The audience has stopped following the narrative and started thinking about the soundtrack. For Bourgeois, whose career spans more than five decades across film, television, music, and streaming media, preventing that moment has remained one of the central responsibilities of a mixer.

    This might appear surprising. Popular discussions of film sound often focus on spectacle. We talk about explosive action sequences, immersive surround sound systems, powerful musical scores, and increasingly sophisticated technologies. Yet Bourgeois repeatedly returned to a much simpler idea. Sound exists to support communication. Every creative and technical decision ultimately serves the story. If audiences cannot understand what matters at the moment it matters, even the most technically impressive soundtrack has failed in its primary task.

    Throughout the lecture, Bourgeois described film mixing as a process of guiding attention. A finished soundtrack may contain dialogue, Foley, ambience, music, effects, backgrounds, transitions, and countless other elements. These sounds do not all demand equal attention simultaneously. Their relationships are constantly shifting. During a conversation, dialogue may occupy the foreground while music retreats slightly into the background. During a dramatic reveal, music may briefly become the dominant element. An action sequence may allow effects to take centre stage before returning attention to character and narrative. Mixing therefore involves much more than balancing levels. It involves shaping the audience’s experience of a story.

    This perspective helps explain why Bourgeois places such importance on dialogue. Writers spend months or years developing scripts. Actors devote enormous effort to performance. Directors construct scenes around the communication of information, emotion, and character. If a crucial line becomes unintelligible, the audience loses access to part of that work. More importantly, they momentarily leave the fictional world. Instead of thinking about the characters, they begin thinking about the soundtrack. The illusion is interrupted.

    One of the most interesting aspects of the lecture concerned the relationship between film mixing and human perception. During the discussion, we explored the idea that many mixing decisions effectively replicate forms of selective attention that listeners perform naturally. In everyday life, people can focus on a particular voice within a crowded room, follow a conversation in a noisy taxi, or attend to one sound source while ignoring dozens of others. The auditory system constantly prioritises information. Bourgeois agreed that much of professional mixing involves recreating these perceptual priorities for audiences. The mixer helps listeners focus on what matters without drawing attention to the process itself.

    Seen in this light, many familiar audio tools acquire a different significance. Equalisation is not simply a way of adjusting frequencies. Compression is not merely a method of controlling dynamics. Reverb is not only about creating a sense of space. These processes become valuable insofar as they help establish relationships between sounds. A dialogue track may require subtle equalisation to distinguish it from surrounding ambience. A sound effect may need certain frequencies reduced so that speech remains intelligible. A reverberant environment may need careful shaping to preserve clarity. The technical operations matter, though their ultimate purpose remains perceptual. Ultimately, they help prevent the audience from asking the question that opened the lecture. What did they say?

    Several examples from Bourgeois’ career illustrated this philosophy particularly well. Large-scale productions such as Transformers are often associated with spectacle, scale, and sonic intensity. Audiences remember giant robots, enormous impacts, and dense layers of sound. Yet Bourgeois described how even the most elaborate action sequences depend upon careful control of attention. One memorable example involved introducing a single frame of silence immediately before an explosion. The audience never consciously notices this interruption. Nevertheless, the brief absence of sound creates a perceptual contrast that makes the subsequent impact feel considerably larger. The effect depends not on additional volume but on the way listeners perceive change.

    Examples such as this reveal a recurring principle running throughout the lecture. Effective sound design often depends less upon adding material than upon managing relationships between existing elements. A soundtrack filled continuously with dramatic gestures eventually loses its ability to surprise. Contrast becomes difficult. Emphasis becomes impossible. Restraint therefore plays an important role within the mixer’s craft. Sometimes the most effective decision is deciding what not to hear.

    This concern with attention also shapes Bourgeois’ attitude towards immersive audio formats such as Dolby Atmos. The technology provides extraordinary creative possibilities. Sounds can move through three-dimensional space with remarkable precision. Environments can become more detailed and immersive than ever before. Yet Bourgeois consistently framed these capabilities in relation to storytelling rather than technology. An Atmos mix succeeds when it helps audiences engage more deeply with a scene. It fails when the technology becomes the focus of attention itself. More speakers do not automatically produce better storytelling. The same principles still apply. Audiences need to understand what matters and why it matters.

    A particularly revealing section of the lecture explored Bourgeois’ lifelong curiosity about listening. Long before spatial audio became a major industry topic, he was conducting informal experiments with binaural recording, environmental acoustics, and perceptual phenomena. Rather than treating recording purely as a professional necessity, he approached it as an opportunity to investigate how sound behaves.

    One story involved recording a stream in rural Canada. Expecting to capture clear differences between close, medium, and distant perspectives, he recorded the same source from multiple locations. When he returned to the studio, however, the recordings sounded remarkably similar. What initially appeared disappointing became an important lesson. Distance is often communicated less by direct sound than by reflections, environmental interactions, and contextual cues. The stream itself had changed very little. The surrounding environment had provided most of the information listeners normally use to judge distance. Stories such as this reveal another dimension of Bourgeois’ approach. Technical expertise emerges not only from formal training but also from observation. Throughout the lecture, he repeatedly emphasised the importance of listening carefully to the world. Many of the insights that shaped his professional practice originated in moments of curiosity rather than commercial necessity. A recording experiment, an unusual acoustic environment, or an unexpected perceptual effect could become the foundation for future creative decisions.

    His reflections on Canada extended this theme further. Bourgeois noted that a surprisingly large number of Hollywood film mixers originate from Canada. While partly humorous, the observation led into a broader discussion about listening environments. Growing up in quieter surroundings encouraged attention to subtle acoustic details, spatial relationships, and environmental sounds. Whether or not this fully explains the phenomenon, the anecdote reinforced a larger point. Listening is not a passive activity. It is a skill developed through experience, practice, and sustained attention.

    The conversation eventually turned towards emerging technologies, particularly artificial intelligence. Here again, Bourgeois adopted a perspective shaped by decades of professional experience. Throughout his career he has witnessed repeated technological transformations. Analogue workflows gave way to digital systems. New recording formats emerged. Distribution platforms changed. Entire production processes evolved. Each transition created uncertainty alongside opportunity.

    Rather than treating AI as fundamentally different from earlier technological developments, Bourgeois viewed it as another stage within a continuing process of change. New tools will inevitably alter professional practice. Some tasks may become easier. Others may disappear entirely. Yet the underlying challenge remains remarkably consistent. Practitioners must learn how new technologies work, understand their limitations, and identify meaningful ways of applying them. Avoiding change rarely proves productive. Understanding it usually does.

    Looking back across the lecture, what emerges most clearly is a conception of mixing rooted in attention. Compressors, equalisers, reverbs, Atmos systems, loudness standards, recording technologies, and AI tools all matter. Yet they matter only insofar as they help audiences remain connected to a story. Bourgeois repeatedly returned to the same fundamental question. Can the audience understand what matters at the moment it matters?

    Many discussions of sound focus primarily on technology. Gary Bourgeois offered a useful reminder that technology is ultimately a means rather than an end. The purpose of a soundtrack is not to demonstrate technical sophistication. Its purpose is to support communication, emotion, and narrative understanding. The most successful mixes often pass unnoticed precisely because they allow audiences to remain fully absorbed in the world unfolding before them.

    Perhaps that is why the simple question that opened the lecture remains so revealing. What happens when an audience misses a line of dialogue? For Bourgeois, the answer extends far beyond a few misunderstood words. It represents a brief fracture in the relationship between story and listener. Much of the mixer’s craft is devoted to preventing that fracture from occurring. Every adjustment, every balance decision, every technical process ultimately serves the same goal: helping audiences hear not merely the sounds of a film, but the story those sounds are trying to tell.

  • Dubbed to Perfection: Graham Hartstone’s Guide to Enhancing Storytelling Through Sound

    Graham Hartstone, a highly respected dubbing mixer and former head of post-production at Pinewood Studios, shared his expertise in an online guest lecture. Drawing on his extensive career in film sound, which spans decades and includes work on major productions, he offered a wealth of insights into the art and technical precision of rerecording sound for film.

    Graham Hartstone

    The Evolution of Sound and Its Role in Storytelling

    Hartstone’s career began in 1961 as a cable operator, progressing through various roles in sound before ultimately leading the dubbing team at Pinewood. His experience includes working on iconic films such as the James Bond series and collaborations with directors like Stanley Kubrick and Ridley Scott. He reflected on the shift from analogue mixing techniques to the expansive digital tools available today, discussing how technological advancements have changed the sound mixing process.

    Throughout his career, Hartstone emphasised that sound must serve the narrative, with careful attention to dialogue clarity, atmospheric cohesion, and the interplay between sound effects and music. He discussed the importance of premixing, highlighting how dialogue, effects, and Foley must be balanced to create a seamless final mix. Foley, he stressed, should blend naturally rather than draw attention to itself. Using Aliens as an example, he described how even background movements were carefully crafted to maintain immersion without overwhelming the primary action.

    Collaborations, Challenges, and International Versions

    Hartstone shared experiences working with directors who had strong opinions on sound, such as James Cameron and Stanley Kubrick. Kubrick was known for personally directing foreign language dubs to maintain creative control, often insisting that his own team handle translations to ensure consistency across different languages. Hartstone recalled how Kubrick’s meticulous nature extended to every aspect of post-production, with dialogue edits often requiring multiple iterations to match the director’s high standards. Kubrick even insisted on making foreign dubs sound as close to the original English version as possible, ensuring that voice tone and performance retained the same impact.

    James Cameron was similarly demanding, particularly about technical precision in sound. Hartstone shared an example from Aliens, where Cameron required the sound of motion trackers to be carefully crafted to enhance suspense. He recalled how Cameron would repeatedly review sound effects, adjusting subtle details to make sure they perfectly complemented the tension of each scene. This attention to detail extended to mixing explosions and gunfire, where Cameron wanted the audience to feel every impact without overwhelming the dialogue.

    The challenges of working on large-scale productions also included meeting tight deadlines and working with evolving edits. Hartstone noted that in films like Blade Runner, changes were often made up to the last minute. He shared how the iconic ambient soundscape of Los Angeles in Blade Runner was built from unused Alien sound elements, giving the city a layered, futuristic atmosphere. He also recounted how Ridley Scott requested late-stage changes to music and sound effects after test screenings, requiring the mixing team to make quick adjustments to balance the soundtrack effectively.

    For international versions, Hartstone explained that dialogue premixes had to be prepared well in advance of final mixes to allow time for translation and dubbing. On GoldenEye, special care was taken to ensure the foreign dubs matched the English version’s intensity, particularly during action sequences. His team provided detailed mixing notes, ensuring that foreign versions retained the same dynamic range and impact. He also explained the additional complexities of preparing mixes for different distribution formats, including airline and television edits, which required removing or replacing strong language while maintaining natural speech flow.

    Practical Techniques for Mixing

    Hartstone provided a wealth of practical advice for sound mixers, focusing on achieving clarity, balance, and impact.

    Dialogue Mixing and Clarity

    He advised using high-pass and low-pass filters to enhance dialogue clarity, suggesting a high-pass filter at around 80Hz to eliminate unwanted low-end rumble and a low-pass filter at around 9kHz to reduce sibilance. He explained that dialogue should be prioritised in the mix, ensuring that off-screen lines remain intelligible by adjusting levels and adding subtle reverb to match distance perception.

    Hartstone also discussed the importance of perspective in dialogue mixing. He emphasised that the audio should match the framing of the shot—voices should not shift unnaturally in relation to the camera’s viewpoint. For example, close-up dialogue should be crisp and intimate, while wide shots should have a more open sound, reflecting the environment. When working with ADR (Automated Dialogue Replacement), he recommended blending it with the original production sound by matching room acoustics and microphone placement to avoid inconsistencies.

    Balancing Sound Elements and Surround Mixing

    Hartstone stressed the importance of dynamic balance between different sound elements. He warned against overusing compression, explaining that while it can help smooth out levels, excessive compression can make a mix sound unnatural. Instead, he recommended using automation and manual level adjustments to retain natural dynamics, especially for dialogue-driven scenes.

    For surround mixing, Hartstone advised positioning ambient sounds carefully to avoid distracting the audience. Dialogue and primary sound effects should remain anchored in the front channels, while environmental sounds and subtle atmospheric elements should be spread across the surround channels. He suggested that surround effects should be used sparingly in dialogue-heavy scenes but can be more pronounced in action sequences to enhance immersion.

    Layering Explosions and Action Sequences

    Hartstone shared techniques for mixing action-heavy films, particularly regarding explosions and gunfire. He explained that layering sound elements helps create depth and realism. For an explosion, he suggested layering three key components: a bass-heavy thump for impact, a mid-range crack for texture, and high-end debris for detail. He recommended ensuring that these layers are carefully mixed so that the low end does not overpower dialogue and other important sounds.

    He also discussed the importance of spatial placement for action scenes. For instance, gunfire should have directional placement in the mix to match the on-screen perspective. He recalled how, on James Bond films, the team carefully panned gunfire and bullet ricochets to follow the action, adding realism and depth to chase and fight sequences.

    Checking Mixes Across Different Playback Systems

    To ensure consistency, Hartstone recommended testing mixes on multiple playback systems, from large cinema screens to nearfield monitors. He suggested switching between full surround and stereo playback to detect phase issues or missing elements. He also noted that checking the mix at lower volumes can help identify problems with clarity, as important dialogue or sound effects may get lost when played at lower levels.

    Additionally, he highlighted the importance of attending final screenings to verify the mix in the intended playback environment. He recalled how, during a Blade Runner premiere screening, last-minute mix adjustments were needed to correct sound balance issues, reinforcing the importance of checking the final product under real-world conditions.

    Final Thoughts

    Graham Hartstone’s lecture provided a detailed exploration of film sound design, offering valuable lessons for professionals and enthusiasts alike. His expertise underscored how vital a well-crafted soundtrack is in shaping the audience’s experience, blending technical precision with creative storytelling.