Editing & Sound

J Cut: When the Sound Arrives First

Also called J-cut, audio lead, pre-lap, split edit

A J cut is an edit where the audio from the next shot starts before the picture changes, so you hear the new scene while still looking at the old one. It is the standard way to make a transition feel led rather than abrupt, and it is why most dialogue scenes never cut on picture and sound together.

What the letter means

Line up two clips on a timeline. If you drag the second clip's audio to the left so it starts before its picture does, the pair makes an L rotated: video on top starting later, audio underneath starting earlier. The shape is a J, and the name has stuck since tape editing.

Mechanically it is a split edit: picture and sound are cut at different points instead of together. That is the norm rather than the exception. Cutting picture and sound at the same frame on every transition is what makes an edit feel like a slideshow.

Why it works

Hearing comes before seeing in an edit. When a new sound arrives while you are still watching the old shot, you start preparing for the change, and by the time the picture arrives you were already going there. The transition feels caused rather than imposed.

The other effect is on attention. Sound arriving early tells the audience where to look next, which is why a J cut is so useful in exposition-heavy scenes: a line of dialogue from the next room pulls the audience forward before the location has even been established.

Three common shapes:

  • Dialogue lead. The next scene's first line starts over the end of the current shot. This is the pre-lap, and it is the most common transition in television drama.
  • Ambience lead. Traffic, a school playground, or a ward's beeping starts a second early, so the new location registers before you see it.
  • Music lead. A cue starts before the cut, which makes the whole following sequence feel like it began earlier than it did.

How to place one

The choice is where the audio starts, and the reliable approach is to let the outgoing picture tell you. Bring the sound in on a natural pause: after a line has landed, at the end of a movement, on a breath. Coming in over the middle of a spoken line puts two competing voices in the same moment and reads as a mistake.

Two habits worth keeping. Fade the incoming audio in rather than hard cutting it, unless the shock is the point. And do not stack it: a J cut on top of a music swell on top of a whip pan means nothing survives, because three signals arriving at once cancel each other.

Building one from generated material

Nothing here depends on the picture, which makes this device unusually easy with AI footage. Generated clips normally arrive either silent or with unusable audio, so you are building the sound layer anyway, and a split edit is just a decision about where each element starts.

Practical notes:

  • Generate audio and picture separately, on purpose. Whether the sound comes from a text to speech read of the next line, a generated ambience bed, or a music cue, keep it as its own element so its start point is yours to choose.
  • Leave handles. Ask for a slightly longer clip than you need at the head of the incoming shot. A J cut consumes the tail of the outgoing picture, and it is the audio head you need room in.
  • Keep the incoming shot silent at its start. If a generated clip has baked-in audio, the lead you laid in earlier will collide with it at the cut.
  • Use one continuous ambience under both shots when the two locations are meant to be near each other. That, plus an early line, is enough to convince an audience that two independently generated shots are in the same building.

The same logic runs in reverse for an L cut, where the outgoing sound carries past the picture. Most sequences use both, alternating so no transition lands the same way twice.

The prompt for this

A starting point that reliably produces the effect. Adjust the subject and setting; keep the technical clauses.

A woman opens a heavy office door and steps into a bright corridor, natural motion, static camera, cool daylight, 35mm

Try J-Cut yourself

Open the generator with a starting point already filled in.

Frequently asked questions

Why is it called a J cut?
From the shape it makes on a timeline. The audio track extends to the left of the video track, below and ahead of it, so the two clips together look like the letter J. An L cut is the mirror image, with the audio trailing to the right.
What is the difference between a J cut and an L cut?
Direction. A J cut brings the next scene's sound in early, so sound leads picture. An L cut holds the previous scene's sound after the picture has changed, so sound trails picture. Both are split edits, and most scene transitions use one or the other.
What is a pre-lap?
The same thing, named from the other side. A pre-lap is a line of dialogue from the next scene laid over the tail of the current one. Editors tend to say pre-lap when it is specifically dialogue and J cut when it is any audio, including room tone or music.
How early should the audio come in?
Usually a beat or two, from half a second up to a couple of seconds. Long enough that the audience registers the new sound before the picture confirms it, short enough that they are not left wondering what they are hearing. Longer leads work when the sound is unambiguous, like traffic or a crowd.
Does a J cut work in a dialogue scene?
It is where the device lives. Bringing a reply in slightly before the cut to the speaker makes a conversation feel like it is being driven rather than assembled, and it stops each exchange landing in the same rhythm. Straight cuts on every line are what makes dialogue coverage feel mechanical.

Related terms