Skip to content

Live Action and Animation in One Commercial: Shared World or Separate Viewpoint?

Advertising

Live Action and Animation in One Commercial: Shared World or Separate Viewpoint?

A treatment can survive a clumsy transition. It usually cannot survive an exchange in which nobody has decided who can see whom.

The transition is the part everyone describes. The wall dissolves into ink, the cup becomes a drawing, a line draws itself across the actor's coat. Meanwhile the treatment never says whether the animated thing is a participant in the room, a picture of somebody's thought, or a joke told to the audience over everyone's heads. Those are three different commercials sharing one paragraph of description.

Fix the relationship first and the transition gets easier, because now you know what it has to accomplish. Skip the relationship and the transition becomes the whole idea — which is not enough weight to carry a passage, and is why a lot of mixed media reads as a change of surface rather than an event.

Choose an exchange, not a transition

Pick one place where a live-action person and an animated element do something to each other, and describe it as an exchange: someone initiates, someone responds, something is different afterwards. Two beats minimum. A single moment — the frame where the live action becomes animation — is a transition, and transitions have their own separate question to settle, about what a cut or a continuous move claims about time and identity. That question sits downstream of this one. You can answer it correctly and still have no idea what the bird is doing in the room.

To keep the comparison honest, hold the action constant and vary only the media relationship. Here is an invented exchange, built for this comparison and not drawn from any production.

A worktable. A lamp at the far-left corner, a mug at the far-right corner, a stack of printed label sheets, and one trimmed label lying loose in front of the first person, Nadia. Ben sits across from her. The label is small, well under a gram, and light enough that lifting it is not the hard part — which matters, because it means the difficulty of this scene is never weight. It is contact, depth and who is watching.

Nadia wants Ben to look at that label. In the first version, a drawn bird lifts it and carries it to him. In the second, the bird is her idea of the handover and her hand does the work.

Version one: the drawn bird is in the room

Shared world means the animated element is a physical participant with the same standing as the mug. Both people can see it. It obeys the lamp. It can pass in front of something.

Four shots:

One. Camera over Nadia's shoulder, looking across the table. The label sits low in the foreground; Ben's hands are on the far side; the lamp is at frame-left, the mug at frame-right. The drawn bird alights on the near-left edge of the table. Nadia's head turns with it — she is watching it, not watching Ben.

Two. Reverse, over Ben's shoulder. The lamp is now at frame-right, the mug at frame-left. The bird takes the label's near corner in its beak. Ben's eyes lead it. Nadia, in the background, is still tracking it.

Three. Close on the label, swinging from the beak as the bird crosses in front of the lamp.

Four. Ben's open palm, the release, the fingers closing.

Now the questions. The bird's feet touch the table edge — does it land, or hover unsupported? If it never lands, standing still in midair is a rule you have chosen, and the audience will notice it before you explain it. The beak has one contact point on one corner of the label, so the label hangs, swings, and rotates; its orientation is continuity, and shot three to shot four is where it usually breaks. If it swings past ninety degrees, it does so in every intervening frame. Ben's receiving hand has to be somewhere the label can actually arrive, and the label's resting position on his palm in shot four has to match where shot five would show it.

Then depth. Decide the flight path through the room once, and the occlusion in each shot follows from it. If the bird travels along the left side of the table, it passes between the shot-one camera and the lamp, and therefore behind the lamp when the camera reverses. That inversion is correct, not a mistake — but only because the path was fixed first. What you cannot do is choose occlusion per shot for compositional convenience. Viewers read which object is nearer, and with no change of camera, a bird that is nearer the mug in one shot and farther from it in the next has moved, whether or not you intended it to.

The lamp is also a light source, so the bird either casts a shadow or has a reason not to. If you keep the shadow, its direction is set by the lamp; if you drop it, be aware that you have told the audience something about what the bird is made of.

And eyelines. Ben's look leads the bird, so one of them has to give. Framestore's page on Paddington in Peru, inspected 18 September 2026, describes coordinating physical space, co-star eyelines, and practical and digital prop scale to keep an animated character's interaction inside the actors' world. That is a first-party account of a feature production, not a commercial, and it supports the category of questions rather than any particular answer; I did not examine the film or any clip. What it usefully confirms is that space, sightlines and prop scale belong on the list of things to investigate — not that they are solved. Whether the bird's path is authored first and the actor's look directed to a reference, or the performance is captured first and the animation fitted to it, is a production decision with consequences in both directions. Present it as a decision. Do not present it as a method you have already validated.

Two more rules that people forget to state. First, persistence: when does the bird exist? If it is absent from a shot, is that because it is out of frame, behind the mug, or gone? An animated participant that quietly stops appearing will read as a conjuring trick. Second, Nadia's own performance. She is a person standing next to a bird, so she may look at it, and her lack of surprise is itself information about the world.

One subtlety worth protecting: in this version, Ben's look is what proves the world is shared. If only Nadia ever tracks the bird, the passage can still read as her private image even though the bird lands on the table and casts a shadow. The second person's attention is the evidence.

Version two: whose shot contains the bird

Separate viewpoint means the bird is a representation — of Nadia's idea, her anticipation, her way of seeing the handover. Her real hand picks up the label and passes it. The bird is not in the room.

Two versions of that, and they behave very differently.

Substitution. The bird performs exactly what the hand performs, in the same place, at the same time. The animation is a re-skin of real contact. Nothing about the label's actual path changes; the drawn figure is doing the same work in the same beat. This is the tidier of the two, because the real action and the animated action are the same event described twice, and the rules are easy to state.

Commentary. The bird does something the hand does not — carries the label in a loop, drops and catches it, takes a route the hand never took. Now the animation is expressing an attitude rather than depicting a contact. That is legitimate, but it has a hard consequence: no shot may show both the real hand and the bird handling the label, because the audience will see two agents moving one object. The label cannot rise from the table under a beak and also travel along a hand in the same frame.

Which shots may contain the drawn element? This is the rule that actually defines the separate viewpoint, more than any visual style. If the bird is Nadia's, it belongs in shots that belong to her — her point of view, or a shot composed around her attention. Put it in an objective two-shot from across the table, with no one perceiving it, and you have made a third thing: an unattributed commentary layer addressed to the audience. That is also a real option, and some commercials live there. What they do not do is alternate among all three without saying so, because the audience builds a rule from the first two appearances and then spends the rest of the spot noticing violations.

The reactions are where a private version most often breaks. Ben responds to the handover, not to a bird he cannot perceive. If he glances at it, ducks it, or waits for it to move, the private rule dies in that shot and does not come back. He may look at Nadia's face — that is a reaction to her, and it is fine — but it needs its own direction, because an actor given nothing to look at will often glance past the other person's shoulder at the empty air, and the shot will look like someone noticing a bird he is not supposed to see.

Watch the two hands and the object, too. If Nadia's hand is visibly on the label while the bird is also carrying it, the substitution reads as an error rather than a subjective image. One route is to keep the hand out of frame whenever the bird has the label, and let the hand appear only in shots owned by nobody's viewpoint. That is a choice about what each shot shows, and it costs you the two-shot at the moment you might want it most.

Finally, continuity of the representation itself. A drawn figure that is somebody's thought can change — simplify, brighten, lose detail as she stops paying attention. But the change has to be doing work, because if it drifts between the first reveal and the last shot for no reason, the audience will assume it means something and will be wrong.

What each version lets you claim

Compare the two on capability. The shared world gives you a third participant: a drawn figure with its own agency, physical comedy, an action neither person could perform, and a surprise both of them get to share. Its cost is that everything about it is now accountable to the room — light, depth order, contact points, scale, eyelines, and how long it stays.

The separate viewpoint gives you privacy. It lets the animation carry a feeling the action cannot carry, and it leaves the live action completely real. Its cost is that it cannot move an object. If the drawn bird lifts the label, either a hand lifted it too, or the label never moved.

So what does this particular exchange need? A brief about an idea becoming a joint game — the label passing from something only Nadia is thinking into something two people are handling together — matches the shared version, because the moment Ben's hand closes on it, the idea is no longer hers alone. The animation should be out there in the room where he can see it, which means accepting the interaction questions in full: contact, scale, occlusion, eyelines, and the bird's persistence across every shot of the passage. If instead the intended relationship is that the animation never leaves her head, the separate version is right, and then it should stay separate in every shot, including the wide one. Choosing the shared version and shooting the two-shot as if the bird were a private thought gives you neither.

There is one more thing to keep separate from all of it. If the product is a sheet of labels, a drawn bird lifting, peeling, or placing one demonstrates nothing about the product. Animation makes no measurement. Even in the shared-world version, when the drawn bird handles the label, the audience is watching an illustrated action, and a sequence that looks like a demonstration can be taken as a claim about adhesion, repositioning, or print quality that no one has supported. Put the drawn sequence beside the demonstration, or make it plainly a story about two people. Do not let a drawn gesture stand in for evidence about how a real material behaves.

Where the passage has to end

Decide the relationship, then write down its rules in a form you can check: who sees the animated element, what it may touch, what it may pass in front of, how long it exists, and whose shots it can appear in. Then read every reaction and every contact in the passage against those rules. Most mixed-media treatments that collapse do not collapse at the join. They collapse four shots later, when someone reacts to something they were never supposed to see, or a label moves without a hand.

The consequential question for the people who will build it is not which transition style is prettier. It is what the animated element is, in this room, for these two people, and therefore what every collaborator has to make true on set and in the composite. Naming a technique does not repair an inconsistent world. Clarifying the exchange first is the part that makes the technique possible.

Frequently asked questions

What must be decided before designing the transition between live action and animation?

Decide the relationship: is the animated element a physical participant in the room, a picture of someone's thought, or a joke addressed to the audience? That determines who can see it and what the transition has to accomplish. Without it, the transition becomes the whole idea.

What makes an exchange rather than a transition?

An exchange has at least two beats: someone initiates, someone responds, and something is different afterward. A single moment where live action becomes animation is a transition, and its own question about time and identity comes downstream.

In the shared-world version, what should be written down and checked?

Rules for landing or hovering, beak contact and label orientation, the receiving hand's position, a fixed flight path and its occlusion, whether the lamp casts a shadow, eyelines, and persistence across shots. Ben's look, not only Nadia's, is the evidence that the world is shared.

What breaks a separate-viewpoint version?

Showing both the real hand and the animated element handling the same object, or letting a live-action character react to something they cannot perceive. Also, placing the animated element in objective shots no one perceives turns it into an unattributed commentary layer. The private rule has to hold in every shot.

Can a drawn action demonstrate a product claim?

No. Animation makes no measurement. A bird lifting, peeling, or placing a label demonstrates nothing about adhesion, repositioning, or print quality, and a sequence that looks like a demonstration may be taken as an unsupported claim. Keep it beside the demonstration or make it plainly a story.

More in Advertising Browse all articles