Pan is among the most misused camera terms in AI video prompting — not because Seedance cannot pan, but because the word describes at least five completely different camera behaviors. An action-follow pan is a horizontal sweep that chases kinetic subject movement through space; a reactive pan is a lateral pivot that fires after an event rather than during it, re-centering attention on a secondary subject; a documentary observational pan is a slow scanning movement that builds environmental authenticity through its unhurried quality; a vertical tilt-pan moves the camera upward or downward to reveal a subject's scale; and a multi-type pan sequence explicitly labels several named pan types (whip-pan, lateral pan, tilt) as distinct grammar units within a single shot. Writing "pan left" without specifying which of these you mean leaves the model with an underdetermined camera instruction, and the result is typically a generic lateral drift that neither follows action nor builds rhythm.
The five Seedance pan prompts below cover each discipline. The salon fight sequence shows how establishing the pan as a camera persona produces better action coverage than specifying individual pan moves at individual timestamps. The Moses VHS pool scene shows how a reactive pan fires after the main event — not during it — to deliver a comedic punchline. The Jallianwala Bagh documentary demonstrates how a slow observational pan functions as an environmental authenticity signal whose baseline quality shifts automatically when the scene changes. The giant character low-angle tilt-pan shows how specifying a pan's start point, direction, and destination produces a scale reveal that a wide establishing shot cannot achieve. And the volcanic warrior sequence shows how labeling distinct pan types in numbered shot notes — "Whip-pan" versus "Fast lateral pan" — produces genuinely different camera behaviors within the same action sequence.
1. The action-follow pan — camera persona as kinetic tracking mode
See the full prompt on scenic.sh →
"The intense fight features fast-paced martial arts choreography: she dodges attacks, delivers powerful kicks and punches, spins with flowing blonde hair, shatters large mirrors and glass panels in dramatic slow-motion explosions of shards... The camera follows the action with dynamic handheld shots, pans, and dramatic angles."
Why this works: At 410 likes, this prompt's defining camera technique is the final sentence — "The camera follows the action with dynamic handheld shots, pans, and dramatic angles" — which is not a specific movement instruction but a camera persona declaration. It tells the model that panning is its tracking mode throughout the sequence, not a move to execute at a particular second. The action-follow pan is established globally, and the model applies it wherever the action demands it: chasing a spin, sweeping to a new attacker, tilting for an incoming kick.
The "handheld" qualifier does essential work. A clean action-follow pan implies fixed hardware (tripod, slider) following a choreographed path; a handheld action-follow pan implies an operator trying to keep pace with unexpected movement. These produce different motion qualities: a handheld pan has micro-drift, slight over-correction, and a brief lag before re-acquiring the subject after a strike. These imperfections are what make action panning read as live rather than directed. Without "handheld," the pan would be technically competent but cinematically distant.
The phrase "dramatic angles" alongside "pans" gives the model permission to break from pure horizontal movement when an upward or diagonal pan better serves the action — a kick might warrant a low-angle lateral pan, an aerial move a tilt-down sweep. The three terms ("handheld shots, pans, and dramatic angles") function as a camera vocabulary the model fills in according to which move serves each moment.
The takeaway: specify the action-follow pan as a camera persona at the end of the prompt, not as individual move instructions at specific timestamps. "The camera follows the action with dynamic handheld shots, pans, and dramatic angles" is a complete action-pan specification. The "handheld" qualifier produces realistic operator lag. "Dramatic angles" allows tilt and bank variations when the action demands them.
2. The reactive pan — lateral pivot as comedic punctuation after the event
See the full prompt on scenic.sh →
"The shaky camera then quickly pans over to Moses, who is now standing normally in the water. He looks toward the camera, winks, and takes a shot of tequila from a small glass he was holding."
Why this works: At 74 likes, this prompt's defining technique is placement: the pan fires after the main action, not during it. A man has fallen from a diving board into water that Moses parted then re-closed; the camera's first job was to capture the fall. The pan to Moses happens only after the splash resolves. This sequencing is the reactive pan — the camera is a surprised observer re-centering on the entity responsible for what just occurred.
The pan does not reveal Moses. Moses is established at the opening of the prompt standing in the parted water. The pan re-frames him: moving from the splash to the calm figure who caused it. This re-framing is the punchline delivery mechanism. The wink and the tequila are the joke; the pan is the setup beat that physically carries the viewer there. A cut to Moses would be faster but would feel editorial; the pan feels reactive, as if the operator spotted him and swung the camera over.
The "shaky" qualifier encodes surprise. A smooth pan to Moses would read as a planned camera position — as if a director told an operator to land on Moses after the splash. A shaky pan reads as an operator scrambling to catch a reaction they did not anticipate. The camera's shakiness throughout ("extremely shaky and reactive the entire time") is a consequence of the VHS-camcorder persona established in the opening sentence — once that persona is set, the model maintains its reactive behavior, including the pan, without needing individual re-instruction at each moment.
The takeaway: place the reactive pan after the main event, not during it. The pan re-frames onto the secondary subject after the primary action resolves. The "shaky" or camera-persona qualifier encodes the operator's surprise. Establish a camera persona (VHS camcorder, documentary observer) at the opening — the model will maintain that persona's reaction behavior throughout the entire clip.
3. The documentary observational pan — slow crowd scan as baseline that changes with the scene
See the full prompt on scenic.sh →
"The camera casually records families sitting together during the Baisakhi gathering. Children run across the open ground while elders talk peacefully. The operator slowly pans across the crowd."
Why this works: At 40 likes, this prompt's defining technique is using the pan as an environmental authenticity signal rather than a narrative device. The pan at 00:00–00:02 does not advance the plot — it simply scans a peaceful crowd. The scanning motion establishes two things simultaneously: that the camera operator is a casual participant (not a filmmaker directing the scene), and that the gathering has enough spatial depth to warrant scanning — this is a large public space, not a private moment.
The transition from slow observational pan (00:00–00:02) to reactive handheld shake (00:04–00:08 as confusion spreads) is where the technique earns its structural value. The initial slow pan is the baseline: calm, curious, unhurried. When the panic begins, the camera behavior changes without any explicit new camera instruction — "confusion spreads through the crowd" is an action description, not a camera note, but the model extrapolates from it to increased handheld shake and drifting framing. The contrast between the calm pan and the later frantic movement is what makes the panic feel sudden.
The "DV camcorder aesthetic" instruction determines the pan's character in the same way "handheld" determines the action-follow pan. A DV camcorder pan is slow and soft, with autofocus hunting and slight exposure drift as the lens sweeps across faces in different lighting conditions. These qualities are all implied by the camera persona — you do not need to specify "soft pan with autofocus hunting" after establishing the camera as a DV camcorder operated by a casual participant. The persona carries the pan's quality.
The takeaway: use the documentary observational pan as a baseline camera behavior established early, then let the scene change the camera rather than re-instructing it. A slow crowd scan at the opening sets both the spatial scale and the operator's passive observer role. The model will extrapolate from scene events (confusion, panic, retreat) to changed camera behavior without additional camera notes. Establish the camera persona first — the pan quality follows automatically.
4. The vertical tilt-pan — upward pan along a body to reveal colossal scale
See the full prompt on scenic.sh →
"The camera is positioned at a dramatic low angle, looking up from the ground, starting close to the tiny figure and panning up alongside the character to emphasize their immense, colossal scale."
Why this works: At 7 likes but technically precise, this prompt demonstrates the vertical tilt-pan — where the camera rotates upward along a subject's body to reveal a scale relationship that a single wide framing cannot encode. The pan begins at the ground (the tiny figure climbing) and rises to the face (the colossal character's expression), encoding the scale contrast through the length of the tilt rather than through perspective tricks or an establishing shot.
The phrase "starting close to the tiny figure and panning up alongside the character" is the complete technical specification. It gives the model the starting point (ground level, close to the small figure), the camera action (pans up), and the axis of travel (alongside the giant's body). The camera is not tilting at empty air; it is traveling the body of a subject, which gives the pan a built-in destination — the face — rather than an undirected upward drift.
The "dramatic low angle" qualifier establishes the starting perspective before the pan begins. A vertical tilt-pan from a standard eye-level angle would show the character growing progressively larger as the camera tilts up; a vertical tilt-pan from a low angle starts with maximum lens distortion (the tiny figure's size exaggerated by the foreshortened perspective) and maintains that dramatic compression throughout the upward movement. The low angle is not purely aesthetic — it is the structural basis for the scale emphasis the pan is meant to convey.
The emotional destination ("wide-eyed expression and a slight blush as they notice the tiny figure climbing them") confirms that the tilt-pan must reach the face. The pan is a journey with a purpose: the scale is revealed over the course of the movement, and the face at the top is the payoff moment. Without a destination, a "pan upward" instruction becomes an undifferentiated vertical drift.
The takeaway: specify the vertical tilt-pan with three elements: start point, travel direction, and destination. "Starting close to [small element] and panning up alongside [large subject]" is a complete instruction. Add a low-angle qualifier to begin with maximum distortion that emphasizes scale contrast. Give the pan a named destination — the face, the top of a structure — rather than "pan upward" alone.
5. The multi-type pan sequence — shot-labeling pan types as explicit camera grammar
See the full prompt on scenic.sh →
"Shot 2 (2.0-4.0s) - Whip-pan; he extends his arm; a glowing hammer streaks in with blue electricity and lands perfectly in his hand. Shot 4 (6.5-8.5s) - Fast lateral pan; he leaps nimbly between the figures, channeling blue lightning through graceful kinetic movements."
Why this works: At 11 likes, this prompt demonstrates that naming different pan types within a single sequence produces genuinely distinct camera behaviors for each shot rather than a generic "action camera" result. Shot 2 is a whip-pan — an abrupt, blur-heavy rotation that fires at a specific action beat (the hammer returning to the warrior's hand). Shot 4 is a fast lateral pan — a horizontal sweep that tracks the character moving across space between figures. Both are pans; neither is the same camera move.
The whip-pan in Shot 2 is tied to an action cue: "he extends his arm; a glowing hammer streaks in." The camera move fires when something happens, not at an arbitrary timestamp. The action event is the trigger. This is the beat-triggering principle: the model treats the action description as the timing cue and executes the whip at the moment of arm extension and hammer arrival, not at 2.0 seconds flat.
The numbered shot structure with a camera label as the first item in each shot note is the structural template. "Shot N — Whip-pan; [action]" and "Shot N — Fast lateral pan; [action]" are read by the model as: camera first, action second. Listing the camera label first signals that the movement is the primary constraint and the action should be fitted into that movement, not the reverse. If the action description came first, the model would tend to serve the action and let the camera follow incidentally.
The five-shot architecture (tilt-down, whip-pan, tracking, lateral pan, push-in) covers five different camera behaviors in twelve seconds — establishing that the model can maintain distinct shot grammar across a sequence when each shot is individually specified. The alternation between different movements also serves the pacing: tilt-down to establish scale, whip-pan for the kinetic reveal, tracking for the chase, lateral pan for the crossing, push-in for the energy release.
The takeaway: when you need multiple distinct pan types in a single sequence, label each type explicitly as the first item in a numbered shot note. "Shot N — Whip-pan; [action]" produces a different camera behavior from "Shot N — Fast lateral pan; [action]." List the camera label before the action description. Tie each pan to a specific action cue rather than only a timestamp — "he extends his arm" is a better trigger than "2.0-4.0s" alone.
Seedance pan prompt cheat sheet
Across all five, the structural principles that make Seedance pan prompts work:
- Action-follow pan — declare the pan as a camera persona at the end of the prompt ("The camera follows the action with dynamic handheld shots, pans, and dramatic angles"), not as individual move instructions at specific timestamps. "Handheld" produces operator lag and micro-correction. "Dramatic angles" allows tilt and bank when the action demands them.
- Reactive pan — fire after the main event resolves, not during it. The pan re-frames onto the secondary subject (the reaction, the cause). A camera persona (VHS camcorder, documentary observer) encodes the reactive quality automatically — set it once at the opening and the model maintains it.
- Documentary observational pan — use as a baseline that changes when the scene changes. A slow crowd scan establishes spatial scale and the observer's passive role. Let scene events ("confusion spreads") change the camera behavior — the model extrapolates from situation to camera response without additional camera notes.
- Vertical tilt-pan — specify start point, direction, and destination: "starting close to [small element] and panning up alongside [large subject]." A low-angle qualifier maximizes scale distortion from the first frame. A named destination (the face, the structure top) tells the model where the pan ends.
- Multi-type pan shot grammar — label each pan type as the first item in its numbered shot note. "Shot N — Whip-pan; [action]" and "Shot N — Fast lateral pan; [action]" produce different behaviors. Tie each pan to an action cue, not only a timestamp. Different labels in the same sequence produce genuinely different camera movements.
Browse the Scenic cinematic gallery for more camera movement examples, or see Seedance zoom prompts and Seedance tracking shot prompts for complementary camera techniques. Read how to write Seedance 2 prompts for the complete prompting guide.