The Theory of Editing7 min readLessonBy · Production Lead at Studio432

J-Cuts and L-Cuts Explained for Video Editors

The two most powerful cuts in dialogue editing are named after letters — and once you understand them, you will use them on every project.

Most cuts happen when picture and sound change at exactly the same moment. J-cuts and L-cuts break that rule deliberately — the audio from one shot bleeds into another before or after the picture switches, and the result feels smoother and more cinematic than a straight cut almost every time. These split edits are a core technique in narrative film, documentary, interviews, and podcast video, and learning them is one of the fastest ways to raise the quality of your editing. This lesson explains what each cut is, why it works, and how to apply it in real-world projects.

Where the Names Come From

The names J-cut and L-cut come from the shape the edit makes on a traditional two-track timeline. Picture the video track sitting above the audio track. In a straight cut, both tracks end and begin at exactly the same point — the edit looks like a clean vertical line.

In a J-cut, the audio from the incoming shot starts before its picture arrives. On the timeline this means the audio clip on the lower track extends to the left, ahead of the video clip above it — creating a shape that looks like the letter J rotated on its side. The sound leads, the picture follows.

In an L-cut, the opposite happens. The audio from the outgoing shot continues playing after its picture has already cut away to a new image. The audio clip extends to the right, past the end of its video clip — forming the shape of an upside-down L. The picture moves on while the sound lingers.

Both edits are sometimes called split edits or audio splits, and you will see those terms used interchangeably in software documentation and editing communities.

Why Split Edits Feel So Natural

Human attention works asymmetrically: we often hear something before we turn to look at it, and we keep listening after our gaze has moved on. A J-cut mirrors the first experience — we hear a new voice or sound, then our visual attention follows. An L-cut mirrors the second — we look somewhere new while still processing what was just said.

A straight cut that changes picture and sound simultaneously draws attention to the edit itself. It can feel abrupt, mechanical, or jarring, especially in dialogue-heavy content where the rhythm of conversation is supposed to feel unbroken. Split edits absorb that seam by keeping at least one element continuous across the cut, so the viewer's brain registers a transition rather than a hard break.

This is why experienced editors reach for J-cuts and L-cuts by default in interview and dialogue scenes, reserving hard cuts for moments where an abrupt break is intentional — a shock cut, a joke punchline, a sudden reveal.

The J-Cut: Audio Leads the Picture

In a J-cut, you hear the next shot before you see it. The most common version is the reaction setup: you are watching Shot A, and you start to hear the voice or ambient sound of Shot B while the picture is still on Shot A. When the picture finally cuts to Shot B, the viewer has already been primed for it — the transition feels earned rather than imposed.

J-cuts work especially well at the start of a new scene. Bringing in the sound of the new environment — music, crowd noise, traffic, a voice — a second or two before the picture cuts there gives the viewer an audio cue that something is about to change. It replaces a jarring teleport with a gentle pull.

In interviews and talking-head videos, a J-cut lets you cut from B-roll back to your speaker without it feeling like a hard slap back to the head cam. You hear the speaker begin their next thought, then the picture reveals them — the edit breathes.

  • Introduce a new scene with its sound before the picture arrives
  • Cut from B-roll back to a speaker by leading with their voice
  • Pull the viewer's attention toward the incoming shot with an audio cue
  • Soften the return to a talking head after cutaway footage

The L-Cut: Audio Trails the Picture

In an L-cut, the audio from Shot A continues after the picture has already moved to Shot B. This is one of the most common edits in documentary and narrative film, and it happens so naturally that most viewers never consciously notice it.

The classic application is dialogue: Person A finishes speaking, and the picture cuts to Person B's face while Person A's last few words are still playing. The viewer reads Person B's reaction while still hearing the tail of what was said — which is exactly how attention works in a real conversation. You look at the listener to gauge their response while the speaker is still wrapping up.

L-cuts also carry emotional weight in storytelling. A character hears bad news; the picture cuts to a landscape or an empty room while their voice or the ambient sound of the scene continues. The audio holds the emotional context while the picture has already moved. This technique is common in documentary narration, where a subject's voice continues over archival footage or environmental shots.

In interview editing, L-cuts let you extend a speaker's audio over B-roll without losing the thread of what they are saying. The words keep going; the picture illustrates them.

  • Show a listener's reaction while the speaker finishes their thought
  • Hold ambient or emotional sound while the picture moves to a new image
  • Extend interview audio across B-roll without breaking the narrative
  • Create a contemplative pause by letting a character's voice linger over an empty frame

Applying Split Edits in Dialogue and Interviews

For interview and podcast video work, split edits are less a stylistic choice and more a baseline expectation. A competent edit of a talking-head interview will almost always use L-cuts when cutting to B-roll — the interviewee's voice continues under the footage, maintaining the logic of what they are saying. Cutting the audio at the same point as the picture creates a choppy experience that pulls the viewer out of the content.

In multi-camera dialogue scenes, J-cuts and L-cuts let you control where emphasis falls. Cutting the picture to the listener before the speaker has finished their line tells the audience that the listener's reaction is what matters in this moment. Holding on the speaker after the cut has happened, with the new shot playing over the tail of the dialogue, invites the audience to sit with the words before moving on.

The overlap length matters. A split of fifteen to thirty frames (half a second to a full second at 30fps) is usually enough to smooth a cut without the overlap becoming noticeable. Longer overlaps — two to four seconds — work well when the trailing audio carries specific emotional or narrative content that the incoming picture should illustrate. Experiment with the length to feel what serves the scene.

Executing Split Edits in Your NLE

Every major non-linear editor supports split edits, though the mechanics vary. In Adobe Premiere Pro, you can hold Alt (Option on Mac) while dragging the edge of a clip to trim audio and video independently. The rolling edit tool, combined with linked selection off, is another approach — slide just the audio or just the video edge of a cut to create the split.

In DaVinci Resolve, the blade or trim tools allow you to cut audio and video tracks separately. The timeline viewer lets you see the linked tracks clearly, so the J or L shape of your edit is visually obvious as you work.

Final Cut Pro handles this with the precision editor, which expands a cut to show the overlap between clips. Trimming in the precision editor is one of the most intuitive ways to set J-cuts and L-cuts for editors new to the technique.

Regardless of software, the workflow is similar: make your primary cut, then unlink audio from video and drag the audio edge of one clip to extend it across the cut point. Listen back with the picture playing and adjust the overlap until it feels right — trust your ear over the waveform.

Common Mistakes and How to Avoid Them

The most common mistake is overlapping audio that conflicts with the incoming or outgoing sound. If the speaker's voice overlaps with music, a sound effect, or another voice that starts at the cut point, the result is a muddy audio mix rather than a smooth transition. Always listen to the full audio mix when evaluating a split edit — what looks right on the timeline may sound wrong in context.

Another frequent error is making the overlap too short or too long by accident. An overlap of just two or three frames is often imperceptible and does not do the work of softening the cut. An overlap of several seconds can confuse the viewer about which shot they are currently in. Start with half a second to a full second and adjust from there.

Finally, avoid using split edits so consistently that they become formulaic. In a fast-paced action sequence or a comic scene where timing is everything, a hard cut at the precise moment of impact is the right choice. Split edits are a tool for smooth transitions — not a rule that applies to every cut in every sequence.

FAQ
What is the difference between a J-cut and an L-cut?

In a J-cut, the audio from the incoming shot begins before its picture appears — sound leads the image. In an L-cut, the audio from the outgoing shot continues after the picture has already cut away to a new shot — sound trails the image. Both are called split edits because audio and video change at different points rather than simultaneously.

Why are they called J-cuts and L-cuts?

The names come from the shape each edit makes on a two-track timeline where video sits above audio. A J-cut has the audio clip extending to the left ahead of its video clip, creating a J-like shape. An L-cut has the audio clip extending to the right past the end of its video clip, creating an L-like shape. The visual description stuck and became the standard terminology.

When should I use a J-cut versus an L-cut?

Use a J-cut when you want to introduce the sound of a new shot before revealing it visually — for scene transitions, returning from B-roll, or pulling the viewer's attention toward what is coming. Use an L-cut when you want to hold on a sound or voice after the picture has moved on — for showing listener reactions, carrying narration over imagery, or extending interview audio under B-roll footage.

How long should the audio overlap be in a split edit?

There is no fixed rule, but a practical starting range is fifteen to thirty frames, which equals roughly half a second to one second at standard frame rates. This is usually enough to smooth a cut without making the overlap obvious to the viewer. For emotionally significant moments where the trailing audio carries narrative weight, two to four seconds can work well. Always judge by listening with the picture playing, not by looking at the timeline alone.

Do J-cuts and L-cuts work only in dialogue scenes?

No — they are most common in dialogue and interviews, but split edits apply anywhere audio and visual timing can be separated for effect. Documentary scene transitions, music video sequences, corporate videos cutting between interview and B-roll, and even scripted narrative scenes in non-dialogue moments all benefit from thoughtful use of J and L-cuts. The principle — keeping at least one element continuous across a cut to reduce abruptness — is universally useful.

F
Written & reviewed by
Faran@432
Production Lead & Consultant · Studio432

Faran is the production lead at Studio432 — the studio arm of Club432, the Karachi collective behind 100+ filmed live sessions and concert films. He plans and consults on shoots worldwide, and owns the gear, settings and craft standards behind everything published here.

Don’t want to learn it — want it done?

Studio432 edits concert footage, music videos, gaming content, vlogs and short-form for creators worldwide. Send the footage, get a quote by email.

Start a project →