본문 바로가기
성공지식백과 로고성공지식백과
가이드

그림 못 그려도 4K 애니메이션 만들기: 시드댄스 2.0 프롬프트 전문 공개

그림 실력 없이 캐릭터가 안 바뀌고 컷이 안 잘려 보이는 4K 애니메이션 60초를 만드는 전 과정입니다. 스토리 시트 생성 프롬프트, 레퍼런스 시트 4종 프롬프트, 실제로 쓴 영상 프롬프트 4컷 전문을 하나도 자르지 않고 공개합니다.

Seedance 2.0HiggsfieldGPT Image 2AI영상AI애니메이션
12분 읽기
공유:

그림을 한 장도 못 그려도 애니메이션이 나옵니다

AI 영상 툴에 프롬프트를 넣는 것 자체는 어렵지 않습니다. 문제는 컷을 두 개, 세 개 이어 붙일 때 생깁니다. 첫 컷에 뽑아놓은 캐릭터가 두 번째 컷에서 다른 사람이 되고, 15초짜리를 뽑았는데 15초 내내 카메라가 한 앵글로 굳어 있습니다. 이 두 가지 때문에 대부분 여기서 그만둡니다.

이 글은 성공지식백과 유튜브 채널에서 4K 애니메이션 60초를 만든 영상의 배포 자료입니다. 이미지는 GPT Image 2로, 영상은 힉스필드의 시드댄스 2.0 4K로 뽑았습니다. 실제로 쓴 프롬프트를 하나도 빼지 않고 전문 그대로 담았습니다. 화면에서 복사해 붙여넣은 그 텍스트입니다.

이 가이드에 담긴 것
  1. 스토리 시트 생성 프롬프트
    한 줄 아이디어를 넣으면 로그라인·색 규칙·컷 구성표까지 나오는 프롬프트
  2. 레퍼런스 시트 4종 프롬프트
    주인공 · 동료 · 몬스터 · 배경 로케이션 보드 — GPT Image 2 · 2K · high
  3. 영상 프롬프트 4컷 전문
    시드댄스 2.0 · 15초 · 4K — 하드컷 7비트와 한국어 대사 블록 포함
  4. 실제로 걸린 함정들
    4K가 720p로 나옴, 앵글이 안 바뀜, 한국어 대사가 뭉개짐, 시트가 그림째로 움직임
  5. 크레딧을 아끼는 순서
    720p·fast로 구도를 잡고 확정된 컷만 4K·std로 다시 뽑는 방식

순서를 바꾸면 망합니다

네 단계입니다. 이 순서를 지키는 것 자체가 절반입니다.

1. 스토리 시트: 뭘 찍을지 글로 먼저 확정
2. 레퍼런스 시트: 주인공, 동료, 몬스터, 배경 네 장
3. 영상 프롬프트: 카메라 비트, 대사 블록, 오디오 3층
4. 4K 생성: 시트를 물려서 뽑고 이어 붙이기

바로 4번부터 하면 컷마다 얼굴이 바뀌고, 세계관이 새로 지어지고, 15초 내내 한 앵글로 굳습니다.

주의
주인공 시트만 만들면 절반만 해결됩니다

처음에 주인공 캐릭터시트만 한 장 만들어 물렸습니다. 주인공은 유지가 됐는데 몬스터와 배경이 컷마다 새로 지어졌습니다. 골목이 갑자기 다른 도시가 되고, 그림자 괴물의 생김새가 컷마다 달라집니다. 그래서 세계관이 붙지 않습니다. 동료, 몬스터, 배경까지 네 장을 다 만들어야 컷이 한 편으로 이어집니다.

1단계: 스토리 시트

여섯 가지를 적습니다. 로그라인 한 줄, 세계관, 색 규칙, 캐릭터 외형 고정문, 로케이션, 컷 구성표입니다.

이 중에 색 규칙이 제일 중요합니다. 아군 색과 적 색을 보색에 가깝게 잡아 놓고, 이 세 줄을 모든 컷 프롬프트 끝에 반복해서 넣습니다. 이것만으로 컷 사이 톤이 잡힙니다.

이번 작업에서는 이렇게 잡았습니다. 시안은 우리 편(주인공의 검, 강아지의 문양), 앰버와 주황은 적(그림자의 눈, 흩어지는 불티), 배경은 마젠타와 시안 네온입니다.

스토리 시트를 AI로 뽑는 프롬프트

한 줄 아이디어만 있으면 시트 전체가 나옵니다. 맨 위 한 줄만 바꿔서 쓰시면 됩니다.

스토리 시트 생성 프롬프트
내 아이디어: (여기에 한 줄. 예 — 밤의 도시에서 그림자 괴물을 베는 여전사와 흰 진돗개)

이 아이디어로 AI 영상 생성용 스토리 시트를 만들어줘.
결과물은 15초짜리 컷 4개, 총 60초 분량이야.

아래 7개를 순서대로, 표 형태로 채워서 줘.

1. 로그라인 한 줄
2. 세계관 규칙 3줄 — 무슨 일이 벌어지는지, 왜 주인공만 할 수 있는지, 적이 어떻게 소멸하는지
3. 색 규칙 3줄 — 아군 색 / 적 색 / 배경 색. 아군과 적은 보색에 가깝게 잡아서 화면에서 즉시 구분되게 할 것
4. 캐릭터 외형 고정문 — 주인공, 동료, 적 졸개, 적 보스. 각각 머리·의상·액세서리·무기를 색까지 지정해서 한 문단으로. 영상 프롬프트에 그대로 복붙할 문장이니까 형용사 말고 구체적인 재질과 색으로 써줘
5. 로케이션 3곳 — 전부 같은 시간대, 같은 날씨, 같은 조명 규칙을 공유해야 함
6. 컷 구성표 — 컷 4개. 각 컷에 담는 사건, 대사, 감정을 표로. 감정은 컷마다 달라야 하고 전체가 기승전결이 되게
7. 컷 사이 카메라 핸드오프 — 각 컷이 끝나는 카메라 위치와, 다음 컷이 그 위치에서 시작하는 방식

제약:
- 기존 애니메이션·게임·영화의 캐릭터나 설정을 닮게 하지 말 것. 전부 오리지널
- 캐릭터 이름은 받침이 겹치지 않는 열린 음절 두 개짜리로 (AI가 발음을 못 만든다)
- 대사는 컷당 한 줄 이하, 짧게
- 배경에 읽히는 글자나 실제 브랜드가 없어야 함
전체 보기

뽑은 뒤 반드시 손볼 것 세 가지

AI가 준 시트를 그대로 쓰면 안 됩니다. 세 군데는 사람이 봐야 합니다.

항목확인할 것
색 규칙아군 색과 적 색이 화면에서 실제로 구분되는가. 둘 다 어두우면 다시 잡습니다
캐릭터 이름받침이 겹치는 이름이면 바꿉니다. 백구는 발음이 안 나왔고 두부는 나왔습니다
컷마다 감정네 컷이 전부 멋있음 하나면 60초가 지루해집니다

직접 채우실 분들을 위한 빈 템플릿

스토리 시트 빈 템플릿
# 스토리 시트 — 「(제목)」

## 1. 로그라인
(누가 / 어디서 / 무엇을 하는 이야기인지 한 줄)

## 2. 세계관 규칙
- (무슨 일이 벌어지는가)
- (왜 주인공만 할 수 있는가)
- (적은 어떻게 죽는가 — 시각적으로)

## 3. 색 규칙
아군 =        (예: 시안)
적   =        (예: 앰버·주황)
배경 =        (예: 마젠타·시안 네온)
※ 이 세 줄을 모든 컷 프롬프트 끝에 복붙한다

## 4. 캐릭터 외형 고정문
주인공 — (머리 / 상의 / 하의 / 액세서리 / 무기, 색까지)
동료   — (종류 / 특징 / 표식 / 소지품)
적 졸개 — (크기 / 재질 / 눈 / 실루엣)
적 보스 — (졸개 대비 몇 배 / 추가 요소)

## 5. 로케이션
(장소1) → (장소2) → (장소3)
※ 전부 같은 시간대·같은 날씨·같은 조명

## 6. 컷 구성표
| # | 씬 | 담는 사건 | 대사 | 감정 |
|---|---|---|---|---|
| 1 |  |  |  |  |
| 2 |  |  |  |  |
| 3 |  |  |  |  |
| 4 |  |  |  |  |

## 7. 컷 사이 카메라 핸드오프
| 컷 | 나가는 카메라 | 다음 컷이 받는 위치 |
|---|---|---|
| 1 |  |  |
전체 보기

이번 영상의 컷 구성표

감정이 컷마다 하나씩 배정돼 있는지가 핵심입니다.

#담는 사건대사감정
1발견과 각성옥상 부감 → 강아지가 짖음 → 발도·점화 → 낙하가자, 두부야!고요 → 경보 → 결의
2골목 교전착지 → 나란히 질주 → 졸개 습격 → 협공 → 소멸두부, 왼쪽!속도·합격
3우두머리와 위기안개에서 보스 등장 → 밀림 → 강아지가 막아섬 → 클로 일격물러서!공포·위기
4피니시등을 딛고 도약 → 일섬 → 붕괴 → 착지 → 애교잘했어, 두부야.해소·보상

컷 사이 핸드오프

앞 컷이 끝나는 카메라 위치를 다음 컷 프롬프트 맨 앞에 적어줍니다. 이게 없으면 컷 사이가 툭 끊어집니다.

나가는 카메라다음 컷이 받는 위치
1하이 와이드 부감, 골목으로 낙하그 부감에서 같이 낙하
2바닥 가까이 낮게, 안개가 뭉침그 낮은 위치에서 안개 시작
3강아지 등 높이, 웅크린 자세강아지 등 높이에서 출발
4스카이라인 풀백

프롬프트 맨 앞에는 이렇게 한 줄 넣습니다.

text복사
Continues directly from a high looking-down angle as the two fall into the alley.

2단계: 레퍼런스 시트 네 장

이미지는 전부 GPT Image 2로 뽑았습니다. 해상도 2K, 퀄리티 high, 비율 16:9입니다. 같은 프롬프트를 나노 바나나에도 넣어봤는데 서양식 화풍에 가까운, 평면적인 2D 카툰으로 빠졌습니다. 그 그림체를 좋아한다면 나노 바나나를 골라도 됩니다. 요즘 극장에서 나오는 3D 애니메이션 질감을 원한다면 GPT Image 2 쪽이 맞습니다.

주인공 캐릭터시트

정면 한 장만 물리면 모델이 옆과 뒤를 마음대로 지어냅니다. 정면, 3/4, 측면, 후면을 한 장에 박아두면 그 문제가 사라집니다.

프롬프트는 순서가 중요합니다. 이미지 모델은 앞쪽에 쓴 단어를 더 세게 반영하기 때문에 구도와 정체성이 앞, 품질이 뒤로 갑니다. 실제로 쓴 슬롯 순서는 이렇습니다.

text복사
① 구도        4뷰 턴어라운드, 같은 지면선
② 동일인 선언  네 장 다 같은 인물 + 같은 헤어스타일
③ 배경        순백 심리스 스튜디오
④ 나이·피부
⑤ 얼굴        얼굴형 · 턱선 · 광대 · 코 · 입
⑥ 눈          모양 · 색 · 하이라이트
⑦ 눈썹
⑧ 머리        색 · 길이 · 스타일 · 마감
⑨ 렌더 방식   셀 셰이딩 / 3D 애니 질감
⑩ 체형
⑪ 의상        위 → 아래 순서로 빠짐없이
⑫ 조명
⑬ 품질
⑭ 네거티브

의상은 위에서 아래로 하나도 빼지 않고 적어야 합니다. 한 군데라도 비워두면 컷마다 그 부분이 달라집니다.

주인공 캐릭터시트 — GPT Image 2 · 2K · high · 16:9
Character turnaround model sheet, four consistent full-body views of the same character in a row, evenly spaced and aligned on the same ground line — front view, three-quarter view, side profile, and back view, identical original female character on all four views with the exact same hairstyle in every view, pure white seamless studio background, professional animation studio character sheet presentation, strikingly beautiful young Korean woman in her early twenties with warm light-tan skin, oval face with a defined tapered jawline, high sharp cheekbones, straight slim nose, full lips with a natural rose tint, mature adult bone structure and adult facial proportions, large upturned almond eyes with deep amber-gold irises and crisp stylized iris highlights, confident calm expression, straight dark eyebrows with a slight arch, long jet-black hair pulled into a high ponytail that hangs down past the shoulder blades, electric-cyan streaks in the under-layer of the ponytail, loose face-framing strands at the temples, glossy stylized CG strand grouping, the identical high ponytail clearly visible in the front view as well as the other three views, stylized 3D anime feature-film render, cel-shaded with clean hard shadow terminators and soft gradient fill light, subtle subsurface skin shading, crisp graphic linework accents, athletic lean build with balanced heroic proportions, wearing a cropped structured black bomber jacket with an asymmetric diagonal wrapped collar and glowing cyan piping along the seams, a fitted charcoal high-neck sleeveless top underneath, tapered black tactical trousers, a wrapped magenta sash belt knotted at the left hip with the tail hanging down, layered matte black knee guards, chunky black and white combat sneakers, a slim glowing cyan cuff on the right forearm, small silver hoop earrings, a slender curved energy blade with a faintly glowing cyan edge sheathed diagonally across the lower back, even neutral character sheet lighting consistent across all four views with a cool cyan rim light and warm key light, high-end 3D animation studio quality, modern action-anime feature film aesthetic, clean neutral background, sharp, 4K, single character design only, no other characters, no duplicate alternate designs, no mannequin, no furniture, no background objects, empty seamless studio, no text, no watermark, no logos, no frame borders, no babyface, no overly youthful rounded proportions, fully original character that does not resemble any real person or any existing copyrighted character

동료 캐릭터시트

동물은 4뷰에 얼굴 클로즈업을 한 장 더 넣습니다. 클로즈업 컷을 쓸 일이 많기 때문입니다.

동료(흰 진돗개) 캐릭터시트
Animal companion character turnaround model sheet, four consistent full-body views of the same dog in a row aligned on the same ground line — front view, three-quarter view, side profile, and back view — with one small head close-up at the far right, identical original white Korean Jindo dog in every view, pure white seamless studio background, professional animation studio character sheet presentation, a medium-sized white Korean Jindo dog with a thick plush double coat, erect triangular ears, a tightly curled tail carried over the back, broad wedge-shaped head, dark almond eyes with warm amber irises, black nose, sturdy athletic build, faint glowing electric-cyan spirit markings tracing along the shoulders and haunches, a slim woven charcoal collar with a small glowing cyan talisman tag, alert and appealing expression that reads cute but capable, stylized 3D anime feature-film render, cel-shaded with clean hard shadow terminators and soft gradient fill light, detailed stylized fur strand grouping, even neutral character sheet lighting consistent across all views with a cool cyan rim light and warm key light, high-end 3D animation studio quality, modern action-anime feature film aesthetic, clean neutral background, sharp, 4K, single animal design only, no other animals, no humans, no duplicate alternate designs, no furniture, no background objects, empty seamless studio, no text, no watermark, no logos, no frame borders, fully original design that does not resemble any existing copyrighted character

몬스터 시트

졸개와 우두머리를 한 장에 같이 넣습니다. 같은 종족이라는 걸 모델이 알아야 생김새를 같은 결로 맞춰줍니다. 적의 색은 주인공과 반대로 잡습니다. 주인공이 시안이면 적은 앰버와 주황이고, 그러면 화면에서 즉시 구분됩니다.

졸개와 보스의 차이는 수치로 적어야 합니다. 3배 크기, 눈 2개와 세로 4개, 장갑판 추가처럼 말입니다. 그리고 입을 없애면 표정 연기 부담이 사라지고 실루엣이 강해집니다.

몬스터 시트 — 졸개 3뷰 + 우두머리
Creature model sheet on a pure white seamless studio background, professional animation studio presentation, split into two zones. Left zone: three consistent full-body views of the same grunt monster in a row — front view, side profile, and back view — an original shadow creature about human height, a gaunt humanoid silhouette made of dense ink-black smoke with a semi-solid charcoal shell, elongated limbs ending in tapered claws, no visible mouth, two glowing amber eyes set deep in a smooth featureless head, ragged smoke trailing off the shoulders and heels, faint ember flecks drifting from the body. Right zone: one towering boss variant of the same species at roughly three times the height, heavier armored charcoal plating across the chest and shoulders, four glowing amber eyes in a vertical cluster, long ragged smoke mantle, cracked molten-orange fissures running through the torso, hunched powerful posture. Consistent design language between grunt and boss, stylized 3D anime feature-film render, cel-shaded with clean hard shadow terminators and volumetric smoke shading, even neutral model sheet lighting with a cool rim light and warm amber accent, high-end 3D animation studio quality, modern action-anime feature film aesthetic, sharp, 4K, no humans, no other creature species, no furniture, no background objects, empty seamless studio, no text, no watermark, no logos, no frame borders, fully original creature design that does not resemble any existing copyrighted character

배경 로케이션 보드: 효과가 제일 큽니다

장소마다 배경을 따로 만들면 컷 사이 톤이 튑니다. 여러 장소를 한 장에 3분할로 만들고, 세 패널이 같은 시간대와 같은 조명 규칙을 공유한다고 문장으로 못 박습니다. 이 한 문장이 전부입니다.

사람, 동물, 차량은 전부 뺍니다. 배경만 순수하게 뽑아야 여러 컷에서 재활용됩니다.

배경 로케이션 보드 — 옥상 / 골목 / 광장 3분할
Environment concept board for an animated film, one wide image evenly divided into three vertical panels with thin neutral gutters between them, all three panels sharing the exact same colour script and time of day — a rain-soaked night, deep blue-black sky, wet reflective surfaces, magenta and cyan neon as the only strong light sources, warm amber window pools as accent. Left panel: a high rooftop of an old low-rise building, wet concrete, a low parapet, rusted water tanks and antenna masts, looking out over a dense neon skyline receding into rain haze. Centre panel: a narrow back alley at street level, cluttered with stacked crates and hanging cables, glowing signage boards running up both walls, standing puddles mirroring the neon, steam rising from a floor vent. Right panel: a wide open plaza intersection with a huge blank billboard tower, empty crosswalk lines, scattered street lamps, low mist across the wet ground. No characters, no people, no animals, no vehicles. Stylized 3D anime feature-film background painting, cel-shaded with clean graphic shapes, volumetric neon haze, cinematic depth, high-end 3D animation studio quality, modern action-anime feature film aesthetic, sharp, 4K, no readable text, no real brand names, no logos, no watermark, no frame borders, fully original environment design

영상 프롬프트에서는 컷마다 어느 패널인지 지정합니다. Location matches the rooftop panel of the environment board처럼 씁니다.

네 장에 공통으로 넣는 네거티브

시트에 글자나 워터마크가 구워지면 그 컷은 버려야 합니다. 아래 문장을 프롬프트 끝에 항상 붙입니다.

시트 공통 네거티브
single character design only, no other characters, no duplicate alternate designs, no mannequin, no furniture, no background objects, empty seamless studio, no text, no watermark, no logos, no frame borders

(성인 캐릭터면 추가) no babyface, no overly youthful rounded proportions

(배경 보드면 추가) no readable text, no real brand names, no logos

3단계: 카메라를 초 단위로 쪼갭니다

여기가 결과를 가장 크게 바꾼 항목입니다.

처음에는 카메라를 이렇게 썼습니다. 시네마틱하게, 크레인으로 올라가면서, 주변을 도는 식으로요. 그렇게 뽑았더니 15초 내내 화면 구성이 거의 안 바뀌었습니다. 카메라만 부드럽게 움직이고 앵글은 그대로였기 때문입니다.

INFO
연속 무빙과 하드컷은 다릅니다

크레인, 틸트, 오빗은 카메라가 계속 이어져 있는 움직임입니다. 화면 구성이 비슷하게 유지되기 때문에 앵글이 안 바뀐 것처럼 보입니다. 편집된 애니메이션처럼 보이려면 셋업 자체를 갈아엎는 하드컷이 섞여야 합니다. 프롬프트에 HARD CUT이라는 단어를 글자 그대로 써야 하고, 그냥 앵글만 나열하면 모델이 전부 부드럽게 이어버립니다.

하드컷 4원칙

#원칙프롬프트에 쓰는 말
1리버스 앵글을 강제한다뒤에서 찍었으면 다음은 정면. `HARD CUT to the REVERSE angle`
2프레임 크기를 건너뛴다와이드에서 미디엄, 클로즈업으로 순하게 가지 말고 와이드 다음에 바로 익스트림 클로즈업
3대사는 화자 단독 컷으로 끊는다`the woman ALONE in frame` / `the dog ALONE in frame`
4인서트를 끼운다칼날, 눈, 목걸이 태그 같은 익스트림 클로즈업 한 컷이 리듬을 만든다

비트를 3~4초로 잡으면 죽습니다

비트를 길게 잡으면 대사 한 줄 하고 2~3초를 그냥 서 있는 화면이 나옵니다. 저도 처음엔 이걸로 컷을 여러 번 버렸습니다.

항목이렇게 하면 죽음이렇게
15초 안의 컷 수4~5개6~8개
비트 하나 길이3~4초1.5~2.5초
가장 긴 비트제한 없음3초 이하, 마지막 여운 컷만 예외
빈 동작얼굴을 잡고 유지금지
카메라 비트 + 밀도 선언 템플릿
PACING: every beat contains motion or a reaction. No idle frames, no static holds. Dialogue is spoken ON TOP of movement, never during a pause. Cut away from each beat while its action is still in progress.

CAMERA BEATS — SEVEN HARD CUTS:
0.0 to 2.0 s — (셋업. 동작이 이미 시작된 상태로)
2.0 to 4.0 s — HARD CUT to the REVERSE angle: (정면 클로즈업, 화자 단독. 대사는 여기)
4.0 to 5.5 s — HARD CUT to an EXTREME CLOSE-UP of (칼날·눈·손 같은 인서트)
5.5 to 7.5 s — HARD CUT to a REVERSE low front-on close-up of (대답하는 쪽 단독)
7.5 to 9.5 s — HARD CUT to a wide side view: (액션 임팩트. 슬로모는 여기만)
9.5 to 11.5 s — HARD CUT to (반대편 각도에서 결과)
11.5 to 15.0 s — HARD CUT to a HIGH WIDE (마지막 여운. 3초 넘겨도 되는 유일한 비트)

한국어 대사가 제대로 나오게 쓰는 법

영어 대사는 크게 손댈 게 없습니다. 프롬프트 자체가 영어로 학습돼 있어서 웬만하면 그대로 나옵니다. 문제는 한국어입니다. 문장 속에 그냥 섞어 넣으면 영어로 말해버리거나 발음이 뭉갭니다. 대사만 따로 블록으로 빼야 합니다.

다섯 가지가 들어갑니다.

#요소빠뜨리면
1`the language spoken in this video is KOREAN`영어로 말합니다
2한글 대사를 독립된 줄에설명문에 묻혀서 무시됩니다
3로마자 발음을 따로 (`Pronounced: "..."`)발음이 뭉갭니다
4`Do not speak this line in English`영어로 말합니다
5`no subtitles, no captions, no on-screen text`한글을 화면에 글자로 굽습니다

그리고 말하는 사람을 화면에 잡으라고 같이 지시해야 합니다. 화자가 화면에 없으면 모델이 대사를 그냥 버립니다.

한국어 대사 블록
SPOKEN DIALOGUE — the language spoken in this video is KOREAN.
The woman speaks one line aloud in fluent natural Korean, lip-synced,
clearly audible over the music, in the low steady voice of a young woman.
Her line, spoken in Korean, is:
가자, 두부야!
Pronounced: "ga-ja, doo-boo-ya" (Dubu is the dog's name).
Frame her face as she says it, while she is already moving.
Do not speak this line in English. Do not whisper it.
Do not render any subtitles, captions, or on-screen text anywhere in the frame.
The dog's sound is two short sharp dog barks, not speech.
주의
캐릭터 이름부터 발음하기 쉬운 걸로 지으세요

강아지 이름을 처음에 백구로 잡았습니다. 흰 진돗개니까 그렇게 갔는데, 발음이 제대로 안 나왔습니다. 받침이 겹치니까 뭉개집니다. 두부로 바꾸니 해결됐습니다. 받침 없이 열린 음절 두 개짜리 이름이 훨씬 안전합니다. 개 짖는 소리는 대사가 아니라 소리로 지시합니다. a short sharp dog bark sound처럼 쓰면 됩니다.

오디오는 세 층으로 나눠서 씁니다

audio: rain 한 줄만 쓰면 밋밋한 환경음만 나옵니다. 환경 베드, 효과음, 음악을 따로 지시합니다.

규칙
환경 베드컷 전체에 깔리는 배경음. 장소마다 반향을 다르게
SFX컷마다 최소 세 개. 비트에 붙는 소리를 하나씩
Music컷마다 다르게. 공포 구간은 음악을 비워야 무섭습니다

우두머리가 나오는 공포 구간에서는 음악을 거의 지웠습니다. 소리를 빼니까 오히려 더 무서워집니다.

text복사
Audio bed: steady rain, distant city hum, wet concrete reverb.
SFX: rain hammering metal, two sharp dog barks, metal blade friction on the draw,
     deep energy ignition hum, the rain shockwave, boots scraping wet concrete.
Music: near-silent low synth pad under the opening wide, a low bass stinger on the bark,
       hard percussion entering on the draw, strings climbing to a peak as they leap.

4단계: 실제 영상 프롬프트 4컷 전문

완성본에 그대로 쓴 프롬프트입니다. 시드댄스 2.0, 15초, 4K, mode std, 16:9입니다. 시트 네 장은 전부 image_references에 물렸습니다.

컷 1 — 발견과 각성 (옥상)
SHOT 1 of 4 — DISCOVERY AND AWAKENING. A dense 15-second sequence cut from SEVEN HARD CUTS between completely different camera setups. This is edited coverage, NOT one continuous camera move.

PACING: every beat contains motion or a reaction. No idle frames, no static holds. Dialogue is spoken ON TOP of movement, never during a pause. Cut away from each beat while its action is still in progress.

Designs must match the reference sheets exactly: the young woman with a high black ponytail with electric-cyan under-streaks, cropped black bomber jacket with cyan piping, magenta sash, curved blade across her lower back; the white Korean Jindo dog with erect triangular ears, tightly curled tail, glowing cyan spirit markings, charcoal collar with a glowing cyan talisman tag. Location: the rain-soaked night rooftop from the environment board, magenta and cyan neon skyline behind.

CAMERA BEATS — SEVEN HARD CUTS:
0.0 to 2.0 s — EXTREME WIDE aerial high above the rooftop, rain falling through frame, two small silhouettes at the parapet against the neon skyline.
2.0 to 3.5 s — HARD CUT to an EXTREME CLOSE-UP of rain striking the glowing cyan talisman tag swinging on the dog's collar.
3.5 to 5.5 s — HARD CUT to a LOW front-on close-up of the white Jindo dog ALONE in frame: its ears snap upright and it barks twice down at the alley, cyan spirit markings blazing alight.
5.5 to 7.5 s — HARD CUT to the REVERSE angle: tight front-on close-up of the woman ALONE in frame, head snapping down toward the dog, her hand already rising to her shoulder for the blade as she speaks her line.
7.5 to 9.0 s — HARD CUT to an EXTREME CLOSE-UP of the blade edge tearing free of the sheath, cyan energy racing along the steel.
9.0 to 11.0 s — HARD CUT to a WIDE side view: the ignition shockwave blasts the rain outward in a ring, her ponytail and jacket snapping back, the dog braced beside her.
11.0 to 15.0 s — HARD CUT to a HIGH WIDE looking straight down as both leap off the parapet and fall away toward the neon alley far below.

SPOKEN DIALOGUE — the language spoken in this video is KOREAN. The woman speaks one line aloud in fluent natural Korean, lip-synced, clearly audible over the music, in the low steady voice of a young woman. Her line, spoken in Korean, is:
가자, 두부야!
Pronounced: "ga-ja, doo-boo-ya" (Dubu is the dog's name). Frame her face as she says it, while she is already moving.
Do not speak this line in English. Do not whisper it. Do not render any subtitles, captions, or on-screen text anywhere in the frame.
The dog's sound is two short sharp dog barks, not speech.

Audio bed: steady rain, distant city hum, wet concrete reverb. SFX: rain hammering metal, two sharp dog barks, metal blade friction on the draw, deep energy ignition hum, the rain shockwave, boots scraping wet concrete. Music: near-silent low synth pad under the opening wide, a low bass stinger on the bark, hard percussion entering on the draw, strings climbing to a peak as they leap.
Cinematic anime feature-film lighting, cyan energy against magenta neon, volumetric rain haze, modern action-anime CG render, sharp 4K, no subtitles, no text overlay.
전체 보기
컷 2 — 골목 교전
SHOT 2 of 4 — THE ALLEY FIGHT. A dense 15-second sequence cut from SEVEN HARD CUTS between completely different camera setups. This is edited coverage, NOT one continuous camera move.

PACING: every beat contains motion or a reaction. No idle frames, no static holds. Dialogue is spoken ON TOP of movement, never during a pause. Cut away from each beat while its action is still in progress.

Designs must match the reference sheets exactly: the young woman with a high black ponytail with electric-cyan under-streaks, cropped black bomber jacket with cyan piping, magenta sash, wielding a glowing cyan curved blade; the white Korean Jindo dog with erect ears, curled tail, glowing cyan spirit markings; the human-sized grunt shadow creatures with ink-black smoke bodies, cracked charcoal shells, tapered claws and two glowing amber eyes each. Location: the narrow rain-soaked alley from the environment board — stacked crates, hanging cables, glowing signage up both walls, neon puddles, steam from a floor vent.

CAMERA BEATS — SEVEN HARD CUTS:
0.0 to 2.0 s — HIGH looking straight down as the two fall into the alley, the camera dropping with them and rotating.
2.0 to 3.5 s — HARD CUT to GROUND LEVEL, extreme low: their feet slam into a puddle and water explodes outward across the lens.
3.5 to 5.5 s — HARD CUT to a fast LATERAL TRACKING shot running alongside the pair at shoulder height as they sprint down the alley together, neon signage streaking past, slight dutch tilt.
5.5 to 7.5 s — HARD CUT to the REVERSE angle: front-on tracking retreating ahead of the woman ALONE in frame as she charges forward and shouts her line.
7.5 to 9.0 s — HARD CUT to a LOW angle on the white Jindo dog ALONE as it launches sideways and slams a shadow creature into the wall.
9.0 to 11.0 s — HARD CUT to an EXTREME CLOSE-UP of the cyan blade carving through a creature's torso, extreme slow motion, sparks and water suspended in the air.
11.0 to 15.0 s — HARD CUT to a WIDE from the alley floor as all three creatures burst into drifting orange embers, the camera settling low where mist begins to gather.

SPOKEN DIALOGUE — the language spoken in this video is KOREAN. The woman shouts one line aloud in fluent natural Korean, lip-synced, breathless and urgent, clearly audible over the music. Her line, spoken in Korean, is:
두부, 왼쪽!
Pronounced: "doo-boo, wen-jjok" (Dubu is the dog's name; the line means: Dubu, left). Frame her face as she shouts it, while she is already running.
Do not speak this line in English. Do not whisper it. Do not render any subtitles, captions, or on-screen text anywhere in the frame.
The dog's sound is one short sharp bark, not speech.

Audio bed: rain, alley reverb, dripping pipes. SFX: bodies landing in water, running footsteps and paw splashes, cloth snapping, a heavy body slamming into a wall, blade slash, one sharp bark, embers crackling. Music: fast percussion loop with driving bass; the music drops out almost completely during the slow-motion slash, then slams back in.
Cinematic anime feature-film lighting, cyan blade energy against amber creature eyes and magenta neon, modern action-anime CG render, sharp 4K, no subtitles, no text overlay.
전체 보기
컷 3 — 우두머리와 위기 (광장)
SHOT 3 of 4 — THE BOSS AND THE CRISIS. A dense 15-second sequence cut from SEVEN HARD CUTS between completely different camera setups. This is edited coverage, NOT one continuous camera move.

PACING: every beat contains motion or a reaction. No idle frames, no static holds. Dialogue is spoken ON TOP of movement, never during a pause. Cut away from each beat while its action is still in progress.

Designs must match the reference sheets exactly: the young woman with a high black ponytail with electric-cyan under-streaks, cropped black bomber jacket with cyan piping, magenta sash, holding a glowing cyan curved blade; the white Korean Jindo dog with erect ears, curled tail, glowing cyan spirit markings; the towering boss shadow creature at three times human height with heavy armored charcoal plating, a vertical cluster of glowing amber eyes, cracked molten-orange fissures and a ragged smoke mantle. Location: the wide misty neon plaza from the environment board — huge blank billboard tower, crosswalk lines, low mist across wet ground.

CAMERA BEATS — SEVEN HARD CUTS:
0.0 to 2.0 s — GROUND LEVEL in the churning mist as something massive displaces it, only shifting shapes and amber glow through the fog.
2.0 to 4.0 s — HARD CUT to an EXTREME LOW ANGLE craning fast up the boss's entire body as it rises out of the mist, ending on its cluster of amber eyes igniting.
4.0 to 5.5 s — HARD CUT to the REVERSE angle: front-on close-up of the woman ALONE in frame, driven back a step, blade snapping up as she shouts her line.
5.5 to 7.0 s — HARD CUT to an EXTREME CLOSE-UP of the molten-orange fissures pulsing across the boss's chest plating as it draws breath.
7.0 to 9.0 s — HARD CUT to a LOW front-on shot of the white Jindo dog ALONE in frame, stepping forward between her and the boss, hackles up, cyan markings blazing, growling then barking twice.
9.0 to 11.5 s — HARD CUT to a huge clawed arm sweeping down and smashing the wet pavement beside them, water and debris exploding, both thrown sideways.
11.5 to 15.0 s — HARD CUT to a low side two-shot: the woman plants her feet and resets her stance while the dog drops into a braced crouch in front of her, the camera lowering to the dog's back height.

SPOKEN DIALOGUE — the language spoken in this video is KOREAN. The woman shouts one line aloud in fluent natural Korean, lip-synced, sharp and urgent, clearly audible over the music. Her line, spoken in Korean, is:
물러서!
Pronounced: "mul-leo-seo" (meaning: get back). Frame her face as she shouts it.
Do not speak this line in English. Do not whisper it. Do not render any subtitles, captions, or on-screen text anywhere in the frame.
The dog's sounds are a low growl and two sharp barks, not speech.

Audio bed: thick mist, wet plaza reverb, distant rain. SFX: deep sub-bass rumble as the boss rises, stone and metal groaning, a low dog growl, two sharp barks, a massive claw impact cratering wet pavement, debris raining down. Music: almost nothing — hold a single sub-bass drone and let the silence do the work; a sharp string stab only on the claw impact.
Cinematic anime feature-film lighting, amber creature glow against cyan blade light, volumetric mist, modern action-anime CG render, sharp 4K, no subtitles, no text overlay.
전체 보기
컷 4 — 피니시
SHOT 4 of 4 — THE FINISH. A dense 15-second sequence cut from SEVEN HARD CUTS between completely different camera setups. This is edited coverage, NOT one continuous camera move.

PACING: every beat contains motion or a reaction. No idle frames, no static holds. Dialogue is spoken ON TOP of movement, never during a pause. Cut away from each beat while its action is still in progress. Only the final beat may hold.

Designs must match the reference sheets exactly: the young woman with a high black ponytail with electric-cyan under-streaks, cropped black bomber jacket with cyan piping, magenta sash, wielding a glowing cyan curved blade; the white Korean Jindo dog with erect ears, curled tail, glowing cyan spirit markings; the towering boss shadow creature with armored charcoal plating, a vertical cluster of glowing amber eyes and cracked molten-orange fissures. Location: the misty neon plaza from the environment board.

CAMERA BEATS — SEVEN HARD CUTS:
0.0 to 2.0 s — LOW from BEHIND the braced white Jindo dog: the woman's boot plants hard on its back and she launches upward out of the top of frame.
2.0 to 4.0 s — HARD CUT to the REVERSE angle high in the air: front-on medium shot of the woman at the apex, blade drawn back overhead, neon city far below behind her.
4.0 to 5.5 s — HARD CUT to an EXTREME CLOSE-UP of her eyes narrowing, cyan light from the blade washing across her face.
5.5 to 8.0 s — HARD CUT to a WIDE ground-level side view as the cyan crescent tears clean through the towering boss. Extreme slow motion, embers and rain suspended in the air.
8.0 to 9.5 s — HARD CUT to an EXTREME CLOSE-UP of the boss's cluster of amber eyes going dark one by one as its face crumbles into drifting embers.
9.5 to 11.5 s — HARD CUT to a low wide as the entire boss collapses into a rising storm of orange embers and she lands in a hard crouch in the foreground, blade trailing cyan light.
11.5 to 15.0 s — HARD CUT to an intimate low two-shot: the dog trots into frame and pushes its nose into her free hand, she smiles and rubs its head; the camera pulls back and cranes up past them to reveal the full neon skyline as embers rain down.

SPOKEN DIALOGUE — the language spoken in this video is KOREAN. The woman speaks one line aloud in fluent natural Korean, lip-synced, warm and slightly out of breath, clearly audible over the music. Her line, spoken in Korean, is:
잘했어, 두부야.
Pronounced: "jal-hae-sseo, doo-boo-ya" (Dubu is the dog's name; the line means: well done, Dubu). Frame her face as she says it in the final two-shot.
Do not speak this line in English. Do not whisper it. Do not render any subtitles, captions, or on-screen text anywhere in the frame.
The dog's sound is one short bright happy bark, not speech.

Audio bed: rain easing off, wide plaza reverb. SFX: a huge energy sweep, deep collapsing roar, ember crackle and sparkle, hard landing impact on wet ground, dog snuffling breath, one bright bark. Music: a full orchestral hit on the slash, then everything drops away to a long reverb tail and resolves into warm piano and strings over the final two-shot.
Cinematic anime feature-film lighting, cyan energy and warm orange embers against magenta neon, modern action-anime CG render, sharp 4K, no subtitles, no text overlay.
전체 보기

설정값과 실제로 걸린 함정들

설정은 이렇게 뒀습니다.

항목주의
모델`seedance_2_0`
해상도480p / 720p / 1080p / 4K4K와 1080p는 std에서만 열립니다
mode`std`fast로 두면 720p까지만 나옵니다
길이4~15초15초가 최대
비율16:9컷마다 같아야 이어붙일 때 안 깨집니다
레퍼런스`image_references``start_image` 아님
오디오기본 켜짐끄면 무음으로 나옵니다
genreaction / horror / drama / epic컷 성격에 맞춰 고릅니다
주의
4K를 눌렀는데 720p로 나온다면 mode를 보세요

시드댄스 2.0은 mode가 fast로 되어 있으면 480p와 720p까지만 지원합니다. 4K와 1080p는 mode를 std로 두어야 열립니다. 4K로 갈 거라면 std를 기본값으로 두고 시작하는 편이 낫습니다. 그리고 시트는 반드시 image_references 칸에 넣어야 합니다. start_image에 넣으면 시트 그림 자체에서 영상이 출발해서, 네 방향으로 나란히 서 있는 상태로 움직이기 시작합니다.

컷마다 필요한 시트만 골라 물려도 됩니다.

물리는 시트
옥상 설정주인공 · 동료 · 로케이션
골목 교전주인공 · 동료 · 몬스터 · 로케이션
보스 등장몬스터 · 주인공 · 동료 · 로케이션
피니시주인공 · 동료 · 몬스터 · 로케이션

프리셋 추천이 뜨면 그대로 진행하지 마세요

힉스필드가 "이 프롬프트는 프리셋 XX 같다"고 추천을 띄우면서 생성을 멈추는 경우가 있습니다. 프리셋을 쓰면 직접 짠 카메라 비트와 캐릭터 고정문이 무시됩니다. 추천을 받지 말고 원래 프롬프트 그대로 다시 요청합니다.

한 번에 여러 개를 밀어넣지 마세요

동시에 여러 잡을 제출하면 rate_limit_reached가 뜹니다. 두세 개씩 끊어서 넣고, 앞의 것이 끝나면 다음을 넣습니다. 15초 4K는 큐가 밀리면 편당 5분에서 10분까지 걸립니다.

크레딧을 아끼는 순서

4K는 한 컷 값이 큽니다. 실측값은 이렇습니다.

조건크레딧
4K · std · 8초176
4K · std · 15초약 330
60초 완성 (15초 × 4컷)약 1,300

한 번에 마음에 들게 나오는 경우가 거의 없습니다. 비트를 잘못 나눴거나, 대사가 뭉개졌거나, 캐릭터 옷이 달라져서 다시 뽑게 됩니다. 그 재시도 값을 4K로 내면 60초 만드는 데 크레딧이 몇 배로 붑니다.

그래서 해상도를 두 단계로 나눕니다.

1. 구도와 비트를 잡는 단계는 720p·fast로 돌립니다. 값이 싸니까 같은 컷을 다섯 번, 열 번 돌려도 부담이 없습니다. 여기서 카메라 비트가 실제로 먹었는지, 대사가 한국어로 나오는지, 캐릭터가 시트대로 나오는지를 확인합니다.
2. 컷이 확정되면 그때만 4K·std로 같은 프롬프트를 다시 넣습니다.

실패 컷 값을 안 내는 것이 핵심입니다. 프롬프트가 확정된 뒤에 4K로 넘어가면 컷당 한 번으로 끝나는 경우가 많습니다.

TIP
무제한이 포함된 플랜이라면 토글을 켜세요

힉스필드 플랜에 따라 특정 모델을 무제한으로 쓸 수 있는 경우가 있습니다. 이때 생성 화면의 무제한 토글 버튼을 켜고 만들어야 크레딧이 차감되지 않습니다. 토글을 안 켜면 무제한 대상 모델이어도 평소처럼 크레딧이 빠집니다. 4K는 한 컷 값이 큰 편이라 첫 생성 전에 토글 상태부터 확인하는 편이 낫습니다. 참고로 저는 울트라 연간 플랜을 쓰고 있습니다.

이어붙이기

컷을 전부 같은 설정으로 뽑으면 재인코딩 없이 붙습니다. 해상도, 비율, 오디오 설정이 하나라도 다르면 붙일 때 깨집니다.

bash복사
# concat.txt 에 절대경로로 파일 목록을 만든 뒤
ffmpeg -f concat -safe 0 -i concat.txt -c copy 00-full.mp4

파일 목록은 절대경로로 씁니다. 상대경로는 concat.txt 위치 기준이라 자주 깨집니다.

뽑기 전 체크리스트

생성 버튼 누르기 전

0/21 완료

뽑은 뒤 확인할 것

생성이 끝난 뒤

0/13 완료

구도 확인은 눈으로 하면 놓칩니다. 프레임을 뽑아서 나란히 놓고 봅니다.

bash복사
# 15초 클립에서 6장 뽑아 셋업이 다 다른지 확인
for n in 20 70 130 190 260 330; do
  ffmpeg -y -i cut.mp4 -vf "select='eq(n\,$n)',scale=460:-1" -frames:v 1 "f$n.jpg"
done

여섯 장이 여섯 개 다른 구도면 카메라 비트가 먹은 것입니다. 두세 장이 비슷하면 비트를 더 쪼개고 HARD CUT 표기를 늘립니다.

정리

캐릭터가 컷마다 바뀌는 문제는 레퍼런스 시트 네 장으로 잡습니다. 컷이 잘려 보이는 문제는 카메라를 초 단위 하드컷으로 쪼개서 잡습니다. 이 둘만 해결하면 나머지는 설정 문제입니다. 크레딧은 720p로 구도를 잡고 확정된 컷만 4K로 올리는 것만 지켜도 크게 줄어듭니다.

오늘 하실 것은 하나입니다. 스토리 시트 한 장만 먼저 써보세요. 로그라인 한 줄과 색 규칙 세 줄만 정해도 나머지가 훨씬 쉬워집니다.

관련 글
GPT 이미지 2 + 시댄스 2 가이드 커버 이미지
가이드

GPT 이미지 2 + 시댄스 2 가이드

GPT 이미지 2로 실사급 이미지를 만들고, 힉스필드의 시댄스 2로 영상까지 연결하는 전체 워크플로우를 정리했습니다. 사고 모드 활성화, 4K 출력, 이미지 투 비디오 변환까지 한 번에 확인할 수 있습니다.

AI 영상 광고 샷·렌즈 프롬프트 가이드 (Seedance 2.0 기준) 커버 이미지
가이드

AI 영상 광고 샷·렌즈 프롬프트 가이드 (Seedance 2.0 기준)

같은 AI 영상 모델을 쓰는데 결과물 수준이 갈리는 이유는 프롬프트에 있습니다. 광고 영상에서 검증된 샷 5개와 렌즈 5개를 제품군별로 매칭하고, 상위 1% 프롬프트의 공통 구조까지 한 문서로 정리했습니다.