Seedance 2.5 셀카 한 장으로 30초 원테이크 영상 만들기: 브라질 파벨라 풋살 프롬프트 테크닉
단 한 장의 인물 사진 레퍼런스로 얼굴과 체형을 고정하고, 9:16 모바일 핸드헬드 원테이크 및 초 단위 타임라인 지시어로 30초 스트리트 풋살 영상을 사실감 있게 연출하는 Seedance 2.5 실전 프롬프트 팁입니다.
AI 크리에이터 TechHalla(@techhalla)가 바이트댄스의 차세대 비디오 파운데이션 모델인 Seedance 2.5 환경에서 본인의 셀카 한 장만을 레퍼런스로 입력하여, 30초 동안 얼굴과 체형의 일관성을 유지하면서도 역동적인 브라질 파벨라 스트리트 풋살 경기를 담아낸 원테이크 비디오 연출 프롬프트를 공개했습니다.

이미지 출처: @techhalla / X
기존 비디오 생성 모델 환경에서 15초 이상의 장기 샷을 생성할 경우, 주인공의 이목구비가 점진적으로 변형되거나 주변 엑스트라의 얼굴과 동기화되는 페이스 몰핑(Face Morphing) 현상이 빈번하게 발생했습니다. 특히 여러 명의 인물이 빠르게 뒤엉키는 스포츠 경기 씬에서는 카메라가 흔들리거나 각도가 바뀔 때마다 주인공의 정체성이 훼손되기 일쑤였습니다. TechHalla는 인물 사진 단 한 장을 엄격한 앵커로 잠그고, 모바일 스마트폰 특유의 질감과 초 단위 타임라인을 결합하여 이 문제를 극복하는 실전 프롬프트 템플릿을 체계화했습니다.
단일 인물 사진 레퍼런스 락과 9:16 핸드헬드 원테이크의 결합 원리
이번 프롬프트 기법의 핵심 돌파구는 Seedance 2.5의 이미지 레퍼런스 처리 메커니즘을 극대화하면서, 인위적인 'AI 질감(AI Gloss)'을 역으로 상쇄시키는 모바일 카메라 문법을 도입한 점에 있습니다.
TechHalla는 주인공의 외형을 묘사할 때 대머리 두상, 짙은 흑백 혼합 수염, 동일한 눈매와 골격 구조를 단 한 장의 사진에 엄격히 구속(Strict Lock)하도록 명령했습니다. 여기에 상대 수비수를 주인공의 복제 캐릭터가 아닌 별개의 20대 초반 청년으로 완전히 분리 정의함으로써, 다인물 상호작용에서 발생하는 모델의 신원 혼선(Identity Bleed)을 사전에 차단했습니다.
더불어 30초 전체 길이를 컷 전환 없이 단 하나의 9:16 세로형 핸드헬드 롱테이크로 고정했습니다. 시네마틱 카메라의 과도하게 매끄러운 트래킹을 배제하고 다음과 같은 스마트폰 고유의 아티팩트를 의도적으로 주입했습니다.
- 모바일 원시 영상 질감(Raw mobile footage): 흔들리는 카메라 앵글, 거친 노이즈 그레인, 압축 블록 노이즈, 해질녘 하늘의 미세한 노출 과다.
- 렌즈 물리적 결함 유도: 렌즈 먼지, 오토포커스 탐색 현상(Autofocus hunting), 빠른 발놀림과 축구공에 묻어나는 자연스러운 모션 블러.
- 인위적 후가공 배제: 영화적인 컬러 그레이딩, 슬로모션, 인위적 컷 편집, 텍스트 자막, 소셜 스티커의 전면 배제.
- 입체적 현장 음향 동기화: 바닥에 튕기는 볼의 둔탁한 타격음, 콘크리트에 미끄러지는 스니커즈 마찰음, 포르투갈어 함성, 펜스가 덜컹거리는 진동음.
이러한 모바일 현장감 설정은 모델이 과도하게 다듬어진 CG 스타일로 수렴하는 것을 방지하고, 실제 누군가가 현장에서 스마트폰을 들고 촬영한 듯한 사실성을 부여합니다.
모듈형 프롬프트 구조: 스타일·레퍼런스·공간 설정의 분리 설계
TechHalla의 프롬프트는 Seedance 2.5가 지시문의 계층 구조를 명확히 해석할 수 있도록 5개의 독립된 모듈 블록([STYLE + CAMERA + ATMOSPHERE], [IMAGE REFERENCES], [PLACE], [TIMELINE], [STYLE & QUALITY BOOSTERS])으로 분리 설계되었습니다.
첫 번째 스타일 모듈은 브라질 파벨라의 거친 콘크리트 코트(Quadra)와 해질녘의 주황빛 하늘, 웅웅거리기 시작하는 저가형 투광등 조명, 철조망을 둘러싸고 스마트폰을 든 동네 주민들의 분위기를 규정합니다.
두 번째 인물 레퍼런스 모듈은 입력된 사진을 '주인공의 얼굴과 체형에 대한 유일하고 엄격한 잠금 장치'로 선언합니다. 주인공의 복장을 단순한 짙은 색 티셔츠와 반바지, 운동화로 제한하여 불필요한 의상 디테일 변형을 막고, 상대 수비수를 날렵한 체형의 로컬 선수로 대조시켜 상호작용의 시각적 명확성을 확보했습니다.
세 번째 공간 모듈은 양쪽 끝이 찌그러진 금속 프레임과 기워진 그물로 만들어진 임시 골대, 콘크리트 바닥의 물웅덩이 자국, 분필로 그어진 X 표시, 언덕 위로 빽빽하게 늘어선 벽돌집과 빨랫줄 등 구체적인 배경 앵커를 배치하여 30초 내내 공간적 일관성을 유지하도록 유도합니다.
초 단위 타임라인 지시문과 30초 연속 연출 프롬프트 전문
이 프롬프트의 가장 독창적인 기술적 장치는 30초를 총 18개 구간으로 세분화한 초 단위(Second-by-Second) 연속 타임라인 지시문입니다. 단순한 상황 설명에 그치지 않고 시간의 경과에 따른 액션, 카메라 동선, 인물의 반응을 순차적으로 통제합니다.
- 0-3초 (오프닝 및 1대1 대치): 펜스에서 스마트폰이 이미 돌아가고 있는 상태로 시작하여 어깨 페이크와 발놀림으로 긴장감을 조성합니다.
- 3-5초 (알까기 기술과 카메라 윕): 상대의 가랑이 사이로 볼을 빼내는 전격적인 '카네타(Caneta, Nutmeg)' 순간과 수비수의 충격, 카메라의 빠른 패닝 및 초점 회복을 지정합니다.
- 5-10초 (가속·슈팅·득점): 헐거운 임시 골대를 향한 전력 질주, 투광등 플레어, 강력한 인스텝 슈팅과 골망이 출렁이는 순간 및 즉각적인 함성을 연출합니다.
- 10-24초 (관중 난입과 축하 세리머니 카오스): 청소년들이 철조망 문을 넘어 콘크리트 코트로 쏟아져 들어오고, 카메라를 든 촬영자가 어깨를 부딪치며 뒤엉키는 물리적 상호작용을 구현합니다.
- 24-30초 (클로즈업 세리머니와 자연스러운 종료): 주인공이 군중 사이를 빠져나와 렌즈를 향해 조깅하며 손가락을 가리키고, 컷 없이 실제 소셜 미디어 스토리 영상처럼 생생한 혼돈 속에서 자연스럽게 마무리됩니다.
아래는 AI 크리에이터 TechHalla가 공유한 Seedance 2.5 실전 프롬프트 전문입니다. 사용자의 외형 사진과 원하는 환경 조건에 맞추어 수정하여 활용할 수 있습니다.
[STYLE + CAMERA + ATMOSPHERE]
Raw mobile phone footage, vertical 9:16, one unbroken continuous handheld take for the full 30 seconds. Shaky cam, grain, compression blocks, slight overexposure on the dusk sky, lens dirt, autofocus hunting, motion blur on the ball and feet. No cinematic grade, no slow-mo, no cuts, no text, no stickers. Gritty concrete football court in a Brazilian favela at dusk: painted lines half worn off, chain-link fence, brick houses stacked uphill, laundry lines, scooters, warm orange sky, cheap floodlight starting to buzz. Crowd of teenagers and locals ring the cage, phones already half-up. Audio: ball thuds, sneakers squeak, Portuguese shouts, fence rattle, then full scream when it goes in.
[IMAGE REFERENCES]
Use the provided photo as the single strict lock for the protagonist’s face: bald head, thick dark salt-and-pepper beard, same eyes and bone structure for the entire take. He wears a simple dark t-shirt, shorts, and sneakers. Body type locked. He is the older bald man in the 1-on-1. The young athletic street footballer facing him is lean, early 20s, jersey and low socks, quick feet — a real kid from the court, not a clone of the protagonist.
[PLACE]
Enclosed concrete pick-up court / “quadra”: small makeshift goal with a bent metal frame and patched net at the far end, puddle stains, chalk X, spectators pressed to the fence. Dusk. Favela hillside behind.
[TIMELINE — ONE CONTINUOUS VERTICAL TAKE, SECOND BY SECOND]
0-1s: Phone already rolling, vertical, held by someone at the fence. Shaky. Bald bearded man and the young footballer face off center-court, ball at the bald man’s feet. Crowd murmur. Hook: he taps the ball forward and commits.
1-2s: He feints left with the shoulder. The kid mirrors. Concrete grit under sneakers. Camera jolts trying to keep both in frame.
2-3s: He rolls the ball onto his right foot, body open. The kid crouches, arms out, ready to block the lane.
3-4s: Lightning caneta — the bald man pushes the ball through the kid’s open legs in one clean nutmeg. Ball pops out behind the defender. Kid’s head snaps down in shock.
4-5s: The bald man is already past him, accelerating onto the loose ball. Vertical cam whips, soft focus for a frame, then snaps sharp on his back and bald head.
5-6s: First touch settles the ball ahead of him toward the small makeshift goal. The kid turns late, chasing. Crowd noise spikes.
6-7s: Two strides. He shapes for a strike. Floodlight flare hits the lens. Fence and heads blur at the edges.
7-8s: Explosive shot — full laces through the ball. Motion blur on the leg and ball. The ball rockets toward the patched net.
8-9s: Ball smashes the top corner of the tiny goal, net bulges, frame rattles. Goal. Instant scream from the fence.
9-10s: The kid freezes mid-run, hands on head. The bald man opens his arms, still moving forward, beard readable, face locked to the photo.
10-11s: Vertical cam shakes hard as the first teenagers vault or squeeze through the gate onto the concrete.
11-12s: Bodies flood the pitch from the near side. Dust and sneakers. Someone bangs the chain-link. Audio clips.
12-13s: The bald man gets swarmed — backs, arms, a leap from a kid trying to jump on him. Camera shoved by a shoulder, still rolling, still continuous.
13-15s: Chaotic celebration fill: phones up in frame filming the phones, shirts pulled, someone kicks the ball away into the fence. The young defender laughs despite himself at the edge.
15-18s: The bald man is lifted half a second by two teens, sneakers off the ground, then set down hard. He points at the goal, yelling with the crowd. Grain heavy. Autofocus hunts his face and locks.
18-21s: Camera operator gets pushed closer — tighter vertical on his bald head and beard in the middle of the pile, kids screaming into the mic. A scooter horn from the street outside the court.
21-24s: Wide-ish again as the pile breaks: people running in circles, one kid sliding on his knees on the concrete, another hanging on the fence. The makeshift goal leans. The ball sits in the net.
24-27s: The bald man breaks free enough to do a short cocky jog toward the phone, pointing at the lens, breathing hard, same locked face. Crowd still pouring behind him.
27-30s: Final hold: vertical frame packed with bodies, the bald man center, arms out, dusk flare, grain, fence rattle, celebration still peak. No cut. Recording ends mid-chaos like a real story upload.
[STYLE & QUALITY BOOSTERS]
Raw vertical continuous phone take, shaky grainy favela dusk court, face locked to the still, lightning caneta nutmeg then an explosive strike into the small makeshift goal, crowd storms the pitch, authentic street light and lens flare, motion blur on ball and feet, viral social footage energy, no cinema camera, no slow-motion, no text.
Seedance 2.5를 활용해 본인의 얼굴을 고정한 상태로 사실적인 원테이크 영상을 제작하고자 하는 크리에이터라면, 이러한 타임라인 분할과 물리적 카메라 결함 유도 지시문을 참고하여 다양한 스포츠나 일상 액션 씬에 응용해 볼 수 있습니다.
원문 출처
- TechHalla (@techhalla) 공식 X 포스트: Seedance 2.5 단일 인물 사진 레퍼런스 30초 원테이크 풋살 프롬프트 스레드 — AI 크리에이터 TechHalla가 실제 30초 생성 비디오 결과물과 함께 공개한 원문 스레드입니다. 단일 셀카 인물 락 기법 및 초 단위 모바일 연출 지시문 전문을 확인할 수 있습니다.
- ByteDance Seedance 2.5 모델 정보: 틱톡 및 캡컷(CapCut)의 모회사 바이트댄스가 개발한 차세대 비디오 파운데이션 모델로, 단일 레퍼런스 이미지 기반 장기 일관성 유지와 네이티브 오디오 동기화 역량을 지원합니다.