GPT Image 2.5 Flare & Sunburst가 EvoLink에 출시되었습니다GPT Image 2.5 체험하기

MiniMax H3 프롬프트와 영상 예제

실제 MiniMax H3 결과, 필요한 입력, 재사용 가능한 변수가 포함된 검증 프롬프트 40개를 확인하세요. 템플릿을 복사하거나 EvoLink Playground로 바로 보낼 수 있습니다.

MiniMax H3는 Hailuo 3, Hailuo 3.0 또는 Hailuo 03으로도 불립니다.

생성 모드

사용 사례

프롬프트 40개 중 40개

MiniMax H3 브랜드·제품 광고 프롬프트

MiniMax H3로 생성한 세로형 아이웨어 광고: 이음매 없는 흰 스튜디오에서 미래적인 랩어라운드 안경을 착용한 두 모델멀티모달 레퍼런스
15s9:16레퍼런스 3개

브랜드·제품 광고

미래적 아이웨어 캠페인

키 비주얼, 모델의 얼굴, 제품 디자인을 세 이미지가 각각 담당하는 패션 광고입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Generate a vertical screen 9:16 high-end fashion glasses commercial, taking overall reference to the storyboard rhythm, editing speed, white studio texture and cool fashion atmosphere of the given video. The picture is a minimalist white booth, a seamless white background, a strong sense of high-end advertising, and a clean, simple, handsome, avant-garde, international first-line fashion blockbuster texture. Key visual character reference picture 1, two full-body female models, one black female model and one European and American model, maintain their high-end clothing, body posture, white studio light and shadow, fashion show temperament and overall cool attitude. Both of them wear futuristic high-end glasses. The design of the glasses refers to Figure 3, emphasizing the covered curved surface, sharp geometric cat-eye/goggle hybrid outline, mirror reflection, streamlined temples, and the texture of high-end fashion accessories. Please refer to Figure 2 for the appearance details of the two characters.

필요한 입력

  • 이미지 1: 키 비주얼. 모델의 전신, 의상, 스튜디오 조명, 태도를 담습니다.
  • 이미지 2: 두 인물의 얼굴 디테일. 컷이 바뀌어도 같은 사람으로 유지하기 위해 사용합니다.
  • 이미지 3: 제품. 빠른 편집에서도 실루엣과 재질이 살아남을 만큼 선명하게 촬영합니다.
  • 영상 전체에 유지할 이음매 없는 스튜디오 배경.

작동하는 이유

  • 세 개의 레퍼런스에 세 가지 명확한 역할을 준 것은 참조 경로가 상정한 사용법 그대로입니다. 다중 이미지 프롬프트가 무너지는 원인은 모호함입니다.
  • 제품을 이름이 아니라 형태로 설명합니다. 감싸는 곡면, 캣아이와 고글을 섞은 날카로운 윤곽, 미러 반사, 유선형 템플.
  • 단일 재질의 흰 스튜디오는 배경 변수를 없애, 모델의 연산을 제품과 인물에 집중시킵니다.

교체 가능한 변수

  • 제품 카테고리
  • 모델 캐스팅
  • 스튜디오 색상
  • 편집 속도
  • 의상

제약 사항

  • 공개된 프롬프트는 참고 영상의 템포도 언급하지만, 이 사례로 공개된 자료는 이미지 세 장입니다. 해당 문장은 스타일 지시로 보거나 직접 만든 클립을 Video 1로 첨부하세요.
  • 참조 요청은 이미지 9장, 영상 3개, 오디오 3개까지 가능하지만 파일 총합은 12개까지이며, 오디오만 단독으로 참조 소재가 될 수는 없습니다.

설정

멀티모달 레퍼런스 · 15s · 9:16 · 레퍼런스 3개

MiniMax H3로 생성한 제품 영상: 검은 스튜디오의 반사되는 받침대 위에서 회전하는 프리미엄 오버이어 헤드폰텍스트-투-비디오
15s16:9레퍼런스 없음

브랜드·제품 광고

프리미엄 헤드폰 쇼케이스

15초를 네 개의 시간 블록으로 나누고 각 블록에 카메라 움직임과 역할을 부여한 제품 영상입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Create a 15-second luxury cinematic product showcase for premium wireless over-ear headphones. 0–4s: Begin with an extreme macro tracking shot moving across the soft memory-foam ear cushion, fine fabric texture, brushed-metal hinge and precision-machined controls. A narrow light band travels across the surface, revealing realistic materials against a deep black studio background. 4–8s: Pull back into a three-quarter hero view. The headphones rotate slowly above a glossy reflective pedestal. The ear cups pivot naturally while the adjustable headband extends slightly, demonstrating flexible construction and comfort. Maintain exact symmetry, stable geometry and consistent proportions. 8–12s: Transition into an elegant exploded-view reveal. The ear cushion, acoustic driver, internal sound chamber, control ring and outer shell separate smoothly in perfect alignment. Subtle luminous sound waves pulse outward from the driver while the camera performs a restrained side orbit. 12–15s: Every component reconnects seamlessly. The headphones settle into a centered front-facing hero composition as soft rim lighting defines the silhouette. Complete a gentle dolly-in toward the ear cups. Premium technology-commercial finish, controlled reflections, realistic shadows, shallow depth of field, crisp surface detail, stable product shape, no hands, no distortion, no onscreen text.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오는 프롬프트만으로 동작합니다.
  • 이름이 아니라 재질과 구조로 설명할 수 있는 제품.

작동하는 이유

  • 시간이 0~4초, 4~8초, 8~12초, 12~15초로 나뉘어 있어 모델이 희망 목록이 아니라 샷 리스트를 받습니다.
  • 블록마다 카메라 움직임이 다릅니다(매크로 트래킹, 물러나기, 측면 오빗, 돌리 인). 이것이 영상을 떠도는 화면이 아니라 편집된 화면으로 읽히게 합니다.
  • 마지막 문장이 금지 목록(손 금지, 왜곡 금지, 화면 문자 금지)이며 제품 렌더가 실패하는 세 가지 전형을 막습니다.

교체 가능한 변수

  • 제품
  • 재질과 마감
  • 각 블록의 시간 배분
  • 배경과 조명
  • 분해도 등장 여부

제약 사항

  • 길이는 4~15초의 정수이며 요청에서 지정합니다. 프롬프트의 시간 블록 합계가 이와 일치해야 합니다.
  • 시간 블록은 연출 지시이지 엄격한 타임라인이 아닙니다. 15초 영상이라면 네 개 이하로 유지하세요.
  • 이 클립은 작성자가 720p로 공개했습니다. EvoLink의 H3 경로는 2K를 출력합니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3로 생성한 음료 광고: 더위에 지쳐 걷던 인물이 차가운 주스를 한 모금 마시자 거리가 병 주위로 무성한 초록으로 피어나는 장면텍스트-투-비디오
15s16:9레퍼런스 없음

브랜드·제품 광고

여름 더위 음료 광고

문제-해소 구조의 정석적인 음료 광고입니다. 더위에 지친 인물이 한 모금을 마시는 순간 환경 전체가 변신하고, 물방울이 맺힌 제품 히어로 샷으로 마무리됩니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

A young person walks under the blazing summer sun, looking exhausted and sweating heavily. The road shimmers with heat waves, and everything appears dry and dull. Suddenly, they grab a chilled bottle of premium fruit juice from a cooler and take a refreshing sip. Instantly, the environment transforms—lush green trees bloom, vibrant flowers appear, a cool breeze flows, water splashes through the air, and glowing particles surround the scene. Ice cubes and fresh fruit slices (orange, mango, or according to the flavour) swirl around the bottle in cinematic slow motion. End with a stunning close-up of the juice bottle covered in cold water droplets against a bright, refreshing background. Ultra-realistic, premium commercial, 4K, cinematic lighting, high-detail, smooth camera movements, vibrant colours, luxury beverage advertisement. Tagline ideas: Beat the Heat. Taste the Freshness. Every Sip Brings Life. Refresh Your Day, Naturally. Stay Cool. Stay Fresh.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 광고가 메마른 더위와 싱그러운 신선함이라는 전후 상태 변화로 설계되어 있어, 모델이 분위기 목록이 아니라 실행할 변신 하나를 명확히 받습니다.
  • 제품이 늦게 등장해 클로즈업 히어로 샷으로 클립을 끝냅니다. 모델이 패턴으로 익힌 음료 광고의 표준 비트 순서 그대로입니다.
  • 풍미 요소(얼음, 과일 조각)가 추상적인 형용사가 아니라 슬로모션으로 소용돌이치는 물리적 사물로 연출됩니다.

교체 가능한 변수

  • 음료 종류와 풍미 단서
  • 지친 인물이 놓인 배경
  • 변신 후의 환경
  • 태그라인 문구

제약 사항

  • 끝에 붙은 태그라인 목록은 카피 아이디어이지 화면에 새겨지는 문자가 아닙니다. 화면에 렌더링하고 싶다면 태그라인 하나를 비주얼 지시부로 옮기세요.
  • 변신 한 번이 예산의 전부입니다. 두 번째 제품이나 장면 전환을 더하면 15초 창이 감당하지 못합니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3 UGC·크리에이터 광고 프롬프트

MiniMax H3가 다섯 레퍼런스로 생성한 방송 오프닝: 애니메풍 스트리머가 LIVE 배지와 팔로워 배너가 있는 Twitch풍 레이아웃에서 채팅 레일을 읽는 장면멀티모달 레퍼런스
15s16:9레퍼런스 5개

UGC·크리에이터 광고

VTuber 방송 오프닝

다섯 소재로 만드는 프로덕션입니다. 이미지 네 장이 정체성, 플랫폼 UI, 방, 오프닝 카드를 각각 분담하고, 오디오 레퍼런스가 무음 리드인까지 포함해 실제 립싱크 스트리머 연기를 이끕니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Use @Image4as the opening card only: circular Luna avatar, black background, cream “LUNALIVE”, rose “STREAM STARTING”, warm circles, tiny mint accent. Static, chime. Use @Image1 for Luna’s identity only: same face, long center-parted black hair, blue-gray eyes, pink anime hoodie, white headphones around neck, delicate necklace, pale nails. Do not copy the drink, pose, or background from @Image1 . Use @Image2 only for Twitch-like platform chrome: dark top bar with “LUNALIVE”, red “LIVE” badge, “2.4K viewers”, right “STREAM CHAT” rail, bottom title area, rounded “FOLLOW” and pink “SUBSCRIBE” buttons. Do not copy Luna, pose, drink, or room from @Image2 Use @Image3 only for the cozy room behind Luna inside the video area: desk, monitor, white PC, plush shelves, curtain fairy lights, soft pink/purple light. No empty-room showcase shot. Use @Audio1as Luna’s actual vocal performance and behavior reference. @Audio1has a silent lead-in: 0.0–2.4s must be treated as no speech. Preserve her voice identity, cadence, tone, breaths, pauses, emphasis, warmth, and streamer mannerisms. Do not replace the voice, do not generate a different influencer voice, and do not add extra spoken lines beyond @Audio1 Lip sync, mouth shapes, jaw movement, smiles, glances, nods, and small hand movements must follow the audio waveform after 2.4s. Create a 15-second 16:9 Twitch-like stream opening. Important performance direction: Luna is reading chat, not delivering a camera monologue. Place Luna slightly left of center in the video area with the right “STREAM CHAT” rail clearly visible. Whenever she speaks, her eyes angle screen-right toward the chat rail as if she is reading the messages out loud. She returns to camera only for brief reactions. Add constant small movement: eye darts to chat, eyebrow lifts, tiny nods, head tilts, shoulders shifting, one subtle hand gesture near the desk. No stiff talking-head pose. One cursor only, no cursor trail, no duplicate panels, no duplicate buttons. Render only large UI text cleanly. Chat feels alive with typing dots and soft short blurred lines, but only these chat lines are readable: “hi Luna”, “welcome back”, “so cozy”, “gugugaga?”, “RAID INCOMING!”. Do not invent usernames. Chat pops stay quieter than Luna’s voice. [0–2 seconds] Open on @Image4 “LUNALIVE / STREAM STARTING”. Absolute no-speech zone: @Audio1is silent here, Luna is not shown, no mouth movement, no voice on the card. One soft chime only. No movement. [2–7.4 seconds] Hard cut to Luna live at 2.0s, slightly left of center in the cozy room from @Image3 with @Image2 chrome active: “LUNALIVE”, red “LIVE”, “2.4K viewers”, right “STREAM CHAT” rail, bottom title “COZY NEON” and “Just Chatting”. She settles for a beat, eyes already moving toward the chat rail. At about 2.4s when speech begins in @Audio1 , match lip sync exactly while she reads toward the chat rail, not into camera. Chat shows typing dots and soft blurred lines. [7.4–7.9 seconds] First audio pause = chat beat. Typing dots, then readable messages pop in: “hi Luna”, “welcome back”. Luna’s eyes track the new messages on the right rail; small nod and smile follow the audio pause. [7.9–12.5 seconds] Continue matching @Audio1Keep her gaze mostly on the chat rail while speaking, like she is reading and reacting. Add one more readable message: “so cozy”. During any softer phrase, she leans slightly forward as if reading; during brighter phrases, eyebrows lift and shoulders react. No frozen face. [12.5–13.9 seconds] Bigger audio pause = bigger chat beat. “gugugaga?” appears, then “RAID INCOMING!”, and a clean “NEW FOLLOWER” banner slides in with a gentle pop. Luna reads the raid message from the chat rail, then reacts brighter as the audio resumes. [13.9–14.8 seconds] Finish @Audio1 with accurate lip sync. If the audio winds down, Luna stops talking, gives a small wave toward chat, and settles into a warm listening pose. End on the stable live frame: Luna slightly left of center, eyes toward the right chat rail, red “LIVE”, “2.4K viewers”, no end card. Audio mix: 0–2s card is silent except one soft chime. @Audio1voice begins only after the cut to Luna and remains primary. Tiny chat pops under the voice. Warm low room tone. No crowd noise, no music lyrics, no rain, no traffic.

필요한 입력

  • 이미지 1: 스트리머의 정체성 전용. 얼굴, 머리, 후디. 포즈와 배경은 복사하지 않는다고 명시됩니다.
  • 이미지 2: 스트리밍 플랫폼 UI. 채팅 레일, LIVE 배지, 버튼.
  • 이미지 3: 그녀 뒤의 아늑한 방.
  • 이미지 4: 정적인 "stream starting" 오프닝 카드.
  • Audio 1: 실제 보컬 테이크. 0~2.4초의 무음이 무발화 구간으로 대본화되어 있습니다.

작동하는 이유

  • 레퍼런스마다 역할 하나와 명시적인 "복사 금지" 두 개가 붙습니다. 라이브러리에서 가장 깔끔한 소재별 역할 배정입니다.
  • "독백이 아니라 채팅을 읽는 중"이라는 연기 지시가 시선과 미세 제스처를 바꿔 놓습니다. 클립이 라이브처럼 느껴지는 이유입니다.
  • 읽을 수 있는 채팅 메시지가 정확히 다섯 문자열로 화이트리스트되고 나머지는 블러로 남아 UI 텍스트가 깨끗하게 렌더링됩니다.

교체 가능한 변수

  • 스트리머 정체성 원판
  • 플랫폼 UI 스타일
  • 화이트리스트된 채팅 문구
  • 오디오 테이크와 그 쉼

제약 사항

  • 원본 게시물은 @Image1…@Audio1로 씁니다. EvoLink에서는 배열 위치로 지칭하고 오디오 옆에 이미지를 최소 한 장 두세요. 오디오는 결코 단독으로 전달될 수 없습니다.
  • 이 경로는 이미지 9장, 영상 3개, 오디오 3개까지 받으며 파일 총합 12개가 상한입니다.

설정

멀티모달 레퍼런스 · 15s · 16:9 · 레퍼런스 5개

MiniMax H3로 생성한 UGC 광고: 밝은 침실에서 여성이 스포이트 병의 헤어 세럼을 바르며 자신을 촬영하는 장면멀티모달 레퍼런스
15s3:4레퍼런스 1개

UGC·크리에이터 광고

UGC 헤어 세럼 광고

제품 레퍼런스 하나로 만드는 4장면 TikTok풍 세럼 체험담입니다. 셀피 인트로, 도포 클로즈업, 거울 결과, 세면대 제품 샷. 불완전함이 곧 스타일링입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Create a 15-second authentic UGC-style hair growth serum ad using the provided product image as the exact product reference. SCENE 1 — 0–3s A young woman films herself in a bright bedroom using a smartphone front camera. Natural lighting, handheld movement, casual appearance. She looks at the camera and says: “I’ve been trying this hair growth serum lately…” SCENE 2 — 3–7s Cut to a close handheld shot of the woman holding the serum bottle. She removes the dropper, applies a few drops directly to her scalp, and gently massages it in. Keep the movement natural and slightly imperfect like real UGC content. SCENE 3 — 7–11s Mirror selfie shot. She runs her fingers through her hair, showing healthy-looking, fuller hair while casually talking to the camera: “And honestly, I love how easy it is to add to my routine.” SCENE 4 — 11–15s Close-up product shot on her bathroom counter. She picks up the bottle and smiles toward the camera. End with natural on-screen text: “Simple hair care. Every day.” Style: authentic TikTok/Reels UGC, smartphone camera, realistic skin texture, natural expressions, subtle handheld motion, imperfect framing, casual home environment, soft daylight, realistic audio, no cinematic commercial look, no excessive beauty filters. Preserve the exact product packaging, label, bottle shape, and branding from the reference image.

필요한 입력

  • 이미지 1: 제품 샷. 네 장면 전체에서 패키지, 라벨, 병 모양이 이 원판으로 고정됩니다.

작동하는 이유

  • 장면 순서가 실제 크리에이터 콘텐츠(훅 → 시연 → 결과 → 제품)를 그대로 따라가 피드에서 네이티브하게 읽힙니다.
  • "진짜 UGC처럼 살짝 불완전하게"와 시네마틱 금지 목록이, 구매자가 그냥 넘겨 버리는 매끈한 광고 룩을 무력화합니다.
  • 대사가 짧고 인용되어 있고 구어체라, 네이티브 오디오가 자연스러운 체험담 말투로 전달합니다.

교체 가능한 변수

  • 제품 원판
  • 두 마디 대사
  • 침실과 욕실 배경
  • 마무리 화면 문구

제약 사항

  • 제품 정체성은 전적으로 레퍼런스 이미지에 있습니다. 라벨을 텍스트로도 묘사하면 충돌을 부릅니다.
  • 공개된 클립의 실측값은 2:3으로, API가 받아들이지 않는 비율입니다. 피드에 어울리는 세로 화면은 3:4나 9:16으로 요청하세요.

설정

멀티모달 레퍼런스 · 15s · 3:4 · 레퍼런스 1개

MiniMax H3로 생성한 푸드 브이로그: 젊은 여성이 빠른 점프 컷 사이에서 웃으며 고메 치즈버거를 폰 카메라 쪽으로 들어 올리는 장면멀티모달 레퍼런스
15s21:9레퍼런스 1개

UGC·크리에이터 광고

버거 나이트 UGC 브이로그

버거 레퍼런스를 고정한 8숏 푸드 브이로그 템플릿입니다. 셀피 훅, 박스 오픈, 줌 펀치, 치즈 풀, 한입 리액션. 하드 점프 컷이 페이스를 책임집니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Duration: 15 seconds | Aspect Ratio: 16:9 | Style: Authentic UGC / iPhone selfie-vlog, handheld, natural light, TikTok/Reels aesthetic. Product Reference: Use the uploaded gourmet burger image as the only product reference. Preserve the bun shape, patty thickness, cheese melt, lettuce, tomato, sauces, and proportions exactly in every shot. Character Description Name: Hana A young Japanese woman in her early 20s with natural beauty, long dark hair in a loose ponytail, oversized cream sweatshirt, minimal makeup, bright smile, friendly lifestyle-vlogger personality. Shot Breakdown SHOT 1 (0–2s) — Selfie showing the burger box. Dialogue: "Burger night!" SHOT 2 (2–4s) — Opens the box. SHOT 3 (4–6s) — Quick zoom on the burger. SHOT 4 (6–8s) — Hands lifting the burger with cheese stretching naturally. SHOT 5 (8–10s) — Bite reaction. Dialogue: "Okay... that's incredible." SHOT 6 (10–12s) — Casual close-up b-roll while reaching for fries. SHOT 7 (12–14s) — Toasting the burger toward the camera. Dialogue: "You need this." SHOT 8 (14–15s) — Freeze frame with overlay: "burger cravings = solved 🍔" Look & Feel Warm apartment lighting, genuine phone footage, slight grain, natural autofocus breathing, handheld imperfections, fast jump cuts. Negative Prompt cinematic grading, commercial production, CGI burger, fake cheese, distorted hands, warped food, perfect stabilization, studio lighting, text glitches, logo distortion.

필요한 입력

  • 이미지 1: 음식 히어로 샷. 번 모양, 패티, 치즈 녹은 정도, 비율이 모든 컷에서 동일하게 유지됩니다.

작동하는 이유

  • 약 2초짜리 마이크로 숏 여덟 개가 실제 푸드 크리에이터의 편집 방식과 일치해 에너지가 포맷에 네이티브합니다.
  • 이름이 있는 캐릭터("Hana")와 짧게 인용된 대사가 과잉 명세 없이 안정적인 얼굴과 목소리를 만듭니다.
  • 음식 물리에 자체 지시가 있습니다(치즈가 자연스럽게 늘어나기). 머니 숏이 요행이 아니라 대본입니다.

교체 가능한 변수

  • 음식과 그 원판
  • 캐릭터 스타일링
  • 두 마디 대사
  • 프리즈 프레임 마무리 문구

제약 사항

  • 음식 묘사는 레퍼런스 이미지에만 두세요. 부정 목록(CGI 버거 금지, 가짜 치즈 금지)이 리얼리즘을 지킵니다.

설정

멀티모달 레퍼런스 · 15s · 21:9 · 레퍼런스 1개

MiniMax H3 타이포그래피·텍스트 모션 프롬프트

MiniMax H3로 생성한 패션 필름: 수채 리본이 흰 하이넥 드레스의 여성 주위로 카메라를 끌고 다니는 동안 LET SILENCE BLOOM 문구가 물리적 글자로 형성되는 장면텍스트-투-비디오
15s16:9레퍼런스 없음

타이포그래피·텍스트 모션

수채화 쿠튀르 원테이크

수채 리본이 단 하나의 연속 카메라를 공간 속으로 끌고 다니고, 타이포그래피가 물리적 사물로 존재하며, 모든 움직임이 음악에 맞아떨어지는 아방가르드 패션 필름입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

## Concept AQUARELLE No.7 — an avant-garde haute couture film where watercolor becomes a living medium. 15-second cinematic sequence. The rhythm controls the visual world: 0–4s: restrained silence and negative space 4–8s: gradual density buildup 8s: drop moment, expanding into fluid long-form motion 12–15s: transition into a final fashion poster composition ## Character Identity Lock Maintain the exact same female character throughout the entire video. Identity: - Young woman - Long black hair - Calm and refined facial features - White high-neck pigment dress - Black wide belt - Pigment heels Strict consistency: - Same face - Same hairstyle - Same age - Same body proportions - Same garment structure Any new colors must only appear through watercolor gradually absorbing into the fabric. Core concept: Her fingertips can extract transparent watercolor ribbons from the air. These watercolor ribbons can: - Pull the camera through space - Transform the environment - Shape physical typography - Interact with depth and materials The world follows real cinematic physics: water, paper, fabric, light, shadows, and depth of field must feel physically believable. ## Camera Direction Strict one-take shot. No cuts. No teleportation. No hidden transitions. Camera journey: Macro shot of a floating water droplet → watercolor ribbon emerges → camera pulls back to reveal the woman → camera circles around her → enters a paper art gallery → passes through dimensional typography → rises into a final overhead fashion poster. ## Typography Only allow these words: LET SILENCE BLOOM AQUARELLE No.7 WEAR THE UNSEEN Typography is not a flat overlay. Letters must have: - Physical depth - Wet watercolor reflections - Paper fiber edges - Shadows - Occlusion - Material interaction ## Visual Style Avant-garde fashion editorial. Inspired by: - Museum catalog composition - Handmade ivory paper - Sculptural negative space - Elegant Didone serif typography - Translucent watercolor calligraphy Color palette: - Pale cyan - Crimson lake - Smoky violet - Ink black ## Motion Language Every movement follows the music: Kick: Camera movement and paper folding. Wooden snare: Paper structures physically fold and transform. Sub-bass: Changes the perception of spatial scale. ## Storyboard 30 beats, 0.5 seconds each. 01 0.0–0.5 Macro shot: A transparent water droplet floats in the air, reflecting a blurred silhouette of a black-haired woman. 02 0.5–1.0 The droplet stretches with the breath-like vocal, becoming a pale cyan watercolor thread. 03 1.0–1.5 Camera travels backward along the thread as paper fibers slowly emerge into focus. 04 1.5–2.0 The thread wraps around the lens. Focus shifts to her raised fingertip. 05 2.0–3.0 Camera continues pulling back, revealing her face and white high-neck dress. 06 3.0–4.0 She moves her wrist. The watercolor thread guides a smooth camera arc. 07 4.0–8.0 Additional watercolor colors emerge from her movement. Paper folds, typography begins forming, and the phrase "LET SILENCE BLOOM" appears as a physical object. 08 8.0–12.0 The camera passes through a transparent paper flower structure. The environment expands into an endless ivory paper gallery. Her dress absorbs watercolor naturally. The ribbons create sculptural forms around her body. 09 12.0–15.0 Camera cranes upward. The composition transforms into a luxury fashion advertisement poster. Typography appears: AQUARELLE No.7 WEAR THE UNSEEN Final frame: A museum-level fashion editorial poster. The woman remains centered, calm, and elegant. A final watercolor droplet remains suspended in the air. ## Negative Prompt No: - Cuts - Scene changes - Identity change - Face swap - Extra limbs - Deformed hands - Random costume changes - Explosive paint effects without physical cause - Incorrect typography - Chinese characters - Extra subtitles - Extra logos - Watermarks

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 카메라 여정이 하나의 연속 체인(물방울 → 리본 → 인물 공개 → 갤러리 → 오버헤드 포스터)으로 쓰여 있고 "no cuts"가 강행 규칙으로 명시됩니다.
  • 타이포그래피가 화이트리스트로 제한됩니다. 세 문구만 등장할 수 있고 종이 섬유 가장자리, 젖은 수채 반사, 오클루전 같은 재질 속성이 부여됩니다. 글자가 깨끗하게 렌더링되는 이유입니다.
  • 0.5초 단위 30비트 스토리보드가 소리를 공간에 매핑합니다. 킥은 종이를 접고, 스네어는 구조물을 변형시키고, 서브베이스는 스케일 감각을 바꿉니다.

교체 가능한 변수

  • 허용된 세 문구
  • 안료 팔레트
  • 의상 구조
  • 갤러리 환경

제약 사항

  • 문구 화이트리스트가 타이포그래피 품질 장치입니다. 텍스트를 더 추가하면 이 프롬프트가 막으려던 철자 흔들림이 되돌아옵니다.
  • 원테이크 프롬프트는 요란하게 실패합니다. 어느 한 비트라도 컷을 암시하면 공간 체인 전체가 끊어집니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3로 생성한 교육 애니메이션: 둥근 글자 A가 부풀어 웃는 빨간 사과로 변하고 옆에 APPLE 단어가 표시되는 파스텔 장면텍스트-투-비디오
15s16:9레퍼런스 없음

타이포그래피·텍스트 모션

A-B-C-D 학습 애니메이션

글자, 소리, 사물, 동작, 단어라는 고정 학습 루프를 가진 유아 파닉스 애니메이션입니다. 각 글자가 해당 사물로 물리적으로 모핑하고 내레이션은 초 단위로 대본화되어 있습니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Create a 15-second animated educational video that teaches young children the letters A, B, C, and D. The learning pattern for every letter must be: LETTER → SOUND → OBJECT → PLAYFUL ACTION → OBJECT NAME Target audience: children ages 3 to 6. Visual style: Use adorable rounded 3D characters, soft pastel colors, gentle facial expressions, and simple recognizable objects. Combine this with a premium minimalist technology aesthetic featuring clean white space, elegant composition, soft studio lighting, subtle reflections, smooth gradients, rounded geometry, crisp typography, and extremely polished transitions. The animation should feel playful and child-friendly while remaining calm, uncluttered, and beautifully designed. Use a clean off-white background with a different soft color glow behind each letter. 0:00–0:01 | Introduction A small smiling star mascot bounces into the center of the screen. Colorful letters briefly float around it. Display the text: “Let’s learn!” The mascot taps the screen, creating a soft ripple that reveals the first letter. 0:01–0:04 | A is for Apple Show a large uppercase “A” and smaller lowercase “a” beside it. Use thick, rounded, highly readable typography. The narrator says: “A. A says ah. A is for Apple.” The uppercase A gently inflates and transforms into a shiny red apple. Its top point becomes the apple stem, and a small green leaf unfolds from the side. The apple gains a cute smiling face and performs one soft bounce. Display the word: “APPLE” Highlight the first letter A in red. Add a soft pop and a tiny crunchy sound. 0:04–0:07 | B is for Ball The apple rolls across the screen and leaves behind a curved red trail. The trail loops twice and forms a large uppercase “B,” with a lowercase “b” appearing beside it. The narrator says: “B. B says buh. B is for Ball.” The two rounded sections of the B expand and merge into a colorful striped ball. The ball bounces twice with playful squash-and-stretch animation. Display the word: “BALL” Highlight the first letter B in blue. Synchronize each bounce with a soft musical note. 0:07–0:10 | C is for Cat On its final bounce, the ball stretches into a curved shape and becomes a large uppercase “C.” A lowercase “c” slides gently into place beside it. The narrator says: “C. C says kuh. C is for Cat.” The C rotates and becomes the curled tail of a cute orange cat. The rest of the cat forms from soft rounded shapes. The cat stretches, blinks, and gives one gentle wave with its paw. Display the word: “CAT” Highlight the first letter C in orange. Add a quiet and friendly “meow.” 0:10–0:13 | D is for Duck The cat’s tail uncurls and transforms into the curved side of a large uppercase “D.” A lowercase “d” pops up beside it. The narrator says: “D. D says duh. D is for Duck.” The straight line of the D becomes the duck’s neck. The curved section becomes its round yellow body. A small orange beak and two tiny wings pop into place. The duck waddles forward, flaps its wings, and gives one cheerful quack. Display the word: “DUCK” Highlight the first letter D in yellow. Add tiny water ripples beneath its feet. 0:13–0:15 | Recap The apple, ball, cat, and duck slide into four clean rounded tiles. Place their letters above them: “A B C D” The mascot returns and points to each object as they bounce once in sequence. Narrator: “A, B, C, D. Great job!” Finish with the text: “Great job!” Use a small sparkle animation and a warm musical chime. Animation requirements: Keep each letter fully visible for a moment before it transforms. Show uppercase and lowercase versions clearly. Make every object instantly recognizable. Use smooth shape morphing so children can visually understand how the letter becomes the object. Maintain stable spelling, clean letterforms, accurate object shapes, and consistent character design. Use gentle squash-and-stretch, soft motion blur, subtle shadows, polished lighting, and precisely synchronized sound effects. Avoid fast camera movement, cluttered backgrounds, harsh colors, tiny text, warped letters, random symbols, duplicated objects, scary expressions, or overly complex transformations. The final video should feel cute, educational, memorable, calming, and exceptionally polished.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • LETTER → SOUND → OBJECT → ACTION → NAME 루프가 동일한 구조로 네 번 반복됩니다. 반복은 H3가 가진 가장 강력한 안정화 장치입니다.
  • 모든 모핑이 기하학적 설명입니다(A의 꼭짓점이 사과 꼭지가 되고, B의 두 볼록면이 공이 됩니다). 그래서 글자 형태가 변신을 버텨 냅니다.
  • 내레이션이 구간별 타임스탬프와 함께 정확히 인용되어 네이티브 오디오를 이끌고 자막 텍스트와의 싱크를 유지합니다.

교체 가능한 변수

  • 네 글자와 사물
  • 마스코트 디자인
  • 글자별 색 글로우
  • 내레이션 목소리

제약 사항

  • 철자 안정성은 "keep each letter fully visible before it transforms" 규칙에 달려 있습니다. 이를 삭제하면 일그러진 글리프가 되돌아옵니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3로 생성한 키네틱 타이포그래피: Every great change라는 단어들이 가는 세리프 글자로 어둠 속에서 떠오르고 금빛 입자가 흩날리는 장면텍스트-투-비디오
15s16:9레퍼런스 없음

타이포그래피·텍스트 모션

키네틱 명언 타이포그래피

순수 키네틱 타이포그래피입니다. 한 문장을 구절 단위로 드러내면서 구절마다 장면의 분위기, 모션 언어, 팔레트를 바꾸고, 마지막에 전체 문장이 오프화이트 엔드 카드에 고정됩니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Create a 15-second cinematic text-animation video built around the quote: “Every great change begins quietly, grows through courage, and becomes impossible to ignore.” The quote should appear gradually as a visual story. Each new phrase must transform the design, atmosphere, movement, and emotional intensity of the scene. Use elegant typography, accurate spelling, cinematic lighting, smooth transitions, and perfectly readable text. 0:00–0:03 | “Every great change” Begin with a completely black screen. A tiny point of warm light slowly appears in the center, like the first spark of an idea. The words “Every great change” emerge softly from the darkness, one word at a time. Use thin, elegant serif typography with wide letter spacing. “Every” fades in gently. “Great” grows slightly larger. “Change” forms from small drifting particles that gather into solid letters. Keep the scene quiet, minimal, and mysterious. 0:03–0:06 | “begins quietly,” The camera slowly moves closer to the text. The previous words shrink and reposition toward the upper-left corner as the phrase “begins quietly,” appears in delicate lowercase letters. Animate the phrase as though it is being written by an invisible hand. Each letter should create a subtle ripple in the darkness. Introduce faint textures, soft shadows, floating dust, and gentle light rays. The comma should appear last and create a small circular pulse. 0:06–0:09 | “grows through courage,” The pulse expands and transforms the scene from darkness into a rich sunrise gradient with deep orange, red, and golden tones. The words “grows through courage” rise upward from the bottom of the frame. Animate “grows” by gradually increasing its size and weight. Animate “through” along a curved path. Animate “courage” in bold uppercase letters that push through a translucent barrier, causing it to crack into geometric fragments. The movement should feel powerful but controlled. 0:09–0:12 | “and becomes” The fragments rotate in slow motion and reorganize into a clean editorial grid. The phrase “AND BECOMES” appears across the frame in condensed sans-serif typography. Animate the letters with fast tracking changes, vertical stretching, masking, and perspective movement. The camera accelerates forward through the center of the word “BECOMES.” The sound and visual energy should steadily build. 0:12–0:14 | “impossible to ignore.” Reveal a vast bright space filled with light, moving shapes, and large-scale typography. The words “IMPOSSIBLE TO IGNORE” appear one after another. “IMPOSSIBLE” expands beyond the edges of the screen. “TO” remains small and perfectly centered. “IGNORE” slams into place with strong visual impact, briefly shaking the surrounding grid and shapes. Use bold contrast, dramatic scale, sharp shadows, and synchronized motion. 0:14–0:15 | Final quote All movement stops instantly. The complete quote appears centered on a clean off-white background: “Every great change begins quietly, grows through courage, and becomes impossible to ignore.” Use refined black typography with “change,” “courage,” and “impossible” highlighted in deep red. Hold the final composition clearly for the last second. Maintain one continuous visual journey from darkness to light, silence to impact, and simplicity to complexity. Keep every phrase connected through visual transformations rather than hard cuts. Use realistic motion blur, precise kerning, clean masks, stable letterforms, smooth camera movement, subtle film grain, cinematic sound design, rising ambient music, soft particles, controlled color transitions, and a final deep impact sound. Avoid misspelled words, warped letters, duplicated characters, unreadable text, random symbols, excessive flickering, chaotic layouts, inconsistent fonts.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 각 구절이 자기만의 애니메이션 동사(떠오르기, 손글씨, 상승, 내리꽂기)를 가진 3초 블록을 소유해, 에너지 상승이 형용사가 아니라 구조로 만들어집니다.
  • 감정의 곡선이 디자인 언어에 매핑됩니다. 어둠에서 일출 그러데이션으로, 다시 에디토리얼 그리드로. 모델이 단어가 아니라 팔레트 대본을 받습니다.
  • 단호한 정지("all movement stops instantly")와 홀드되는 엔드 카드가 섬네일로 쓸 수 있는 읽기 좋은 마지막 프레임을 보장합니다.

교체 가능한 변수

  • 명언과 구절 분할
  • 강조 단어와 포인트 색
  • 구절별 애니메이션 동사
  • 엔드 카드 스타일

제약 사항

  • 구절은 6단어 안팎으로 유지하세요. H3는 문장 길이의 줄보다 짧은 디스플레이 텍스트를 훨씬 안정적으로 렌더링합니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3 숏드라마·서사 프롬프트

MiniMax H3로 생성한 세로형 숏드라마 예고편: 촛불이 켜진 성 내부의 뱀파이어 주연과 인간 여주인공멀티모달 레퍼런스
15s9:16레퍼런스 2개

숏드라마·서사

뱀파이어 로맨스 숏드라마

한 이미지로 두 주인공을, 다른 이미지로 장소를 고정하고 관계·템포·쇼트 크기에 집중합니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Generate a 15-second, 9:16 vertical trailer segment for an international live-action vampire romance short drama. Use Figure 1 as the appearance reference for the male and female leads, and Figure 2 as the scene reference. Keep both leads' identities consistent, with a realistic live-action look and premium short-drama production quality. Story: an innocent human heroine accidentally enters a forbidden area of an old castle and awakens a sleeping aristocratic vampire. He discovers that she carries an aura connected to an ancient war, which sparks a powerful urge to control her and a dangerous fascination with her. She fears him but does not completely submit, resisting his pressure. Overall style: an international ReelShort / DramaBox vampire-romance trailer. Dark romance, dangerous attraction, fate, intense control, brooding oppression, and a striking reversal. Keep the visuals premium, restrained, and tightly paced, like the opening 15-second hook of a hit short drama. No gore, cheap horror, Halloween aesthetic, or modern street feel. Format: 9:16 vertical composition for TikTok / ReelShort / DramaBox. Use primarily medium close-ups, close-ups, and extreme close-ups, emphasizing faces, eye contact, pressure, and relationship tension within the vertical frame.

필요한 입력

  • 이미지 1: 두 주연이 함께 담긴 참조. 모델이 두 사람을 하나의 캐스팅으로 읽게 합니다.
  • 이미지 2: 장소 참조. 성 내부의 조명과 재질을 담습니다.
  • 한 문장짜리 설정과 명시된 감정의 반전. 15초 훅이 감당할 수 있는 것은 반전 한 번이지 한 회 분량의 이야기가 아닙니다.

작동하는 이유

  • 장르 레퍼런스(ReelShort / DramaBox 예고편)를 지정하면 템포, 색보정, 구도 관습을 몇 단어로 전달할 수 있습니다.
  • 미디엄 클로즈업, 클로즈업, 익스트림 클로즈업으로 쇼트 어휘를 고정했습니다. 세로 화면이 답답하지 않고 고급스럽게 읽히는 이유가 여기에 있습니다.
  • 제외 목록(유혈, 값싼 공포, 핼러윈 느낌, 현대 거리 분위기)이 이 장르가 무너지는 네 가지 전형적인 방식을 차단합니다.

교체 가능한 변수

  • 주연 외모 참조
  • 장소
  • 설정과 반전
  • 목표 플랫폼의 톤
  • 쇼트 크기 구성

제약 사항

  • 세로 출력은 참조 경로의 aspect_ratio로 요청합니다. 프롬프트에 "9:16"이라고 쓰는 것만으로는 지정되지 않습니다.
  • 두 주연이 한 장에 함께 담긴 이미지가, 따로 잘라낸 인물 사진 두 장보다 동일성 유지에 훨씬 유리합니다.

설정

멀티모달 레퍼런스 · 15s · 9:16 · 레퍼런스 2개

MiniMax H3로 생성한 비주얼 노벨 인터페이스 전환: 고정된 시작 프레임과 마지막 프레임 사이의 변화첫 / 마지막 프레임
15s16:9레퍼런스 2개

숏드라마·서사

오토메 비주얼 노벨 전환

첫 프레임과 마지막 프레임을 고정하고, 그 사이의 여정과 감정 변화만 텍스트로 지시합니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Use the first image as the opening frame and the second image as the exact final frame to generate an otome visual-novel interface transition. Overall feel: a premium Chinese otome romance-interaction interface capturing an intimate moment before and after a performance. Transition naturally from "choose to watch his performance" to "Han Xu is drawn in by the heroine's words and reacts with intrigued interest." UI text, choices, and dialogue boxes should appear with refined otome-game presentation. Keep the transition silky smooth and the emotion suggestive yet restrained.

필요한 입력

  • 이미지 1: UI 상태를 포함한 시작 프레임.
  • 이미지 2: 정확히 도달하고자 하는 마지막 프레임.
  • 두 상태 사이의 감정 변화를 한 문장으로.

작동하는 이유

  • 양쪽 끝이 고정되어 있으므로 모델은 구도가 아니라 보간을 풉니다. 예측 가능한 샷을 얻는 가장 확실한 방법입니다.
  • 프레임을 나열하는 대신 감정의 전환("그의 무대를 보기로 선택" → "이끌리며 흥미를 보임")을 지목해, 연기가 컷을 지탱합니다.
  • UI 요소를 장르 고유의 방식으로 움직이라고 지정해 인터페이스가 다시 그려지는 것을 막습니다.

교체 가능한 변수

  • 시작 프레임과 종료 프레임
  • 둘 사이의 감정선
  • UI 표현 스타일
  • 전환 속도

제약 사항

  • 두 프레임은 이미지 투 비디오 경로에서 image_start와 image_end로 전달합니다. 참조 경로는 이 필드를 아예 받지 않습니다.
  • 두 프레임의 구도와 조명이 비슷할수록 보간이 매끄럽습니다. 서로 무관한 구도는 전환이 아니라 컷이 됩니다.

설정

첫 / 마지막 프레임 · 15s · 16:9 · 레퍼런스 2개

MiniMax H3로 생성한 시대극 시퀀스: 1940년대 군복의 병사들이 연기 자욱한 미국 시골 거리에서 클래식 자동차 뒤로 엄폐하는, 기록 필름처럼 촬영된 장면텍스트-투-비디오
15s16:9레퍼런스 없음

숏드라마·서사

1940년대 전쟁 뉴스릴 리얼리즘

시대 고증에 충실한 1947년 뉴스릴 시뮬레이션입니다. 리얼리즘이 세 겹의 시스템에서 나옵니다. 시대에 일치하는 미술, 물리 계약, 그리고 모든 현대 사물을 금지하는 부정 목록.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Create a 15-second ultra-photorealistic live-action war sequence set in the United States in 1947, designed to look like authentic historical footage captured on a 1940s film camera. The entire scene must feel grounded, documentary-like, raw, and physically realistic. Environment: A rural American town in 1947 with wooden houses, old brick buildings, telephone poles, dirt roads, vintage American cars from the 1940s, wooden fences, farmland, and period-accurate street details. Overcast afternoon light, light fog, drifting smoke, dust in the air, damaged buildings, scattered debris, and a tense wartime atmosphere. Characters: American soldiers wearing historically accurate late-1940s military uniforms, helmets, boots, and equipment. Civilians wear authentic 1940s American clothing. Natural faces, realistic skin texture, sweat, dirt, fatigue, and believable body movements. 0–3s — Establishing Shot: Wide handheld shot of a quiet rural American street suddenly filled with smoke and confusion. Vintage 1940s vehicles are parked along the road while soldiers move quickly between wooden buildings. Civilians rush toward safer areas. 3–6s — Tension: Camera moves through the street at shoulder height, following several soldiers as distant gunfire is heard. They immediately react and take cover behind a vintage vehicle and a brick wall. Their movements are cautious and realistic. 6–10s — Combat: Fast handheld tracking shot as the soldiers move between cover while distant gunfire impacts the environment. Small pieces of wood, dust, and debris fall naturally from nearby impacts. Weapon recoil, movement, and body weight must be physically accurate. Keep the violence realistic and restrained. 10–13s — Human Moment: Camera briefly focuses on a soldier helping an injured civilian move behind cover. Their breathing, facial expressions, body language, and movement should feel natural and unscripted. 13–15s — Final Shot: Camera pulls back into a wide shot of the American town as smoke slowly moves through the street. Soldiers remain behind cover while vintage vehicles and damaged buildings fill the background. The scene ends with an authentic, tense 1940s documentary feeling. Visual Style: Ultra-photorealistic live-action, authentic 1940s American environment, vintage 35mm film texture, subtle film grain, natural imperfections, realistic exposure, handheld documentary cinematography, muted historical color palette, realistic smoke and dust, natural shadows, accurate depth of field. Physics: Strictly obey real-world gravity, momentum, inertia, friction, recoil, weight, collision physics, and human biomechanics. No exaggerated explosions, impossible movements, superhero behavior, or choreographed-looking combat. Negative Prompt: modern buildings, modern cars, smartphones, modern clothing, modern weapons, futuristic technology, CGI appearance, video-game graphics, fantasy, superhero action, excessive explosions, excessive blood, gore, impossible physics, unrealistic recoil, slow-motion physics, distorted faces, extra limbs, floating objects, plastic skin, artificial-looking environments.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 시대 고증이 두 방향으로 강제됩니다. 긍정적으로(1940년대 자동차, 군복, 전신주), 그리고 부정적으로(스마트폰 금지, 현대 건물 금지). 실패할 수 있는 양방향을 모두 닫습니다.
  • 물리 문단("strictly obey gravity, momentum, recoil…")이 렌더 사양서처럼 읽히며 슈퍼히어로식 움직임을 눈에 띄게 억제합니다.
  • 조용한 인간적 비트(민간인을 돕는 병사)가 10~13초에 배치되어, 쉼 없는 액션 대신 다큐멘터리적 신뢰감을 만듭니다.

교체 가능한 변수

  • 시대와 장소
  • 인간적 순간
  • 필름 스톡 룩
  • 전투 비트의 강도

제약 사항

  • 절제 조항("violence realistic and restrained", 고어 금지)은 이 출력이 실제로 쓸 만한 이유의 일부입니다. 변형할 때도 유지하세요.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3 캐릭터·모션 프롬프트

MiniMax H3로 생성한 클레이 애니메이션: 용암 협곡을 뛰어넘는 클레이 여우와 그 아래를 통과하는 카메라첫 / 마지막 프레임
10s16:9레퍼런스 1개

캐릭터·모션

클레이 여우의 협곡 점프

시작 이미지를 하나의 결정적 동작으로 만들고 카메라 경로도 동일하게 정밀하게 지시합니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Claymation style. A sprinting fox reaches the edge of a cliff and launches without hesitation, making a dramatically tense, heroic slow-motion leap across a vast lava canyon. While the fox is airborne, the camera rushes at high speed beneath its belly in a sweeping dynamic move, fully revealing the terrifying depth of the chasm and the fox's clay body at maximum extension in midair.

필요한 입력

  • 스타일을 이미 담고 있는 시작 이미지. 여기서는 클레이 여우와 그 재질입니다.
  • 일어나길 원하는 동작 하나. 세 개가 아닙니다.

작동하는 이유

  • 프롬프트는 그림이 아니라 변화를 씁니다. 시작 프레임이 외형을 이미 담고 있으므로 모든 단어가 움직임을 사는 데 쓰입니다.
  • 카메라에 자체 지시(여우의 배 아래를 고속으로 통과하는 움직임)가 주어져 있고, 이것이 단순한 점프를 하나의 샷으로 만듭니다.
  • 정점의 순간("공중에서 최대한 뻗은 상태")을 지목해 모델에 타이밍을 구성할 목표 포즈를 제공합니다.

교체 가능한 변수

  • 캐릭터와 재질 스타일
  • 환경과 위험 요소
  • 카메라 경로
  • 슬로모션 강조점

제약 사항

  • 이미지 투 비디오 경로는 입력 이미지에서 출력 비율을 결정하며 aspect_ratio를 허용하지 않습니다.
  • 시작 프레임이 이미 보여 주는 것을 다시 설명하지 마세요. 정적인 외형을 반복하는 것이 이 경로의 프롬프트를 낭비하는 가장 흔한 방식입니다.

설정

첫 / 마지막 프레임 · 10s · 16:9 · 레퍼런스 1개

MiniMax H3로 생성한 모션 전이: 참조 영상에서 복제한 스트릿 댄스를 추는 두 참조 캐릭터멀티모달 레퍼런스
10s16:9레퍼런스 3개

캐릭터·모션

스트릿 댄스 모션 전이

참고 영상의 안무를 이미지로 제공한 두 캐릭터에게 옮기는 짧고 명확한 프롬프트입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Have the characters perform street dance following the movements in Video 1. Use Figure 1 and Figure 2 as the character references.

필요한 입력

  • 이미지 1과 이미지 2: 두 캐릭터의 깨끗한 전신 참조 각 한 장.
  • Video 1: 복제할 동작. 2~15초이며 연기자가 화면에 온전히 담겨 있어야 합니다.

작동하는 이유

  • 정보를 참조 소재가 담고 있기 때문에 프롬프트가 짧아도 됩니다. 영상이 동작을, 이미지가 정체성을 가집니다.
  • 각 소재를 배열 위치로 지칭하므로 어떤 참조가 무엇을 제공하는지 모호하지 않습니다.
  • 그 외에는 아무것도 요구하지 않습니다. 조명도 카메라도 스타일도 지정하지 않은 이유는, 추가 지시가 복제하려는 동작과 경쟁하기 때문입니다.

교체 가능한 변수

  • 캐릭터
  • 원본 안무
  • 환경
  • 연기자 수

제약 사항

  • 참조 영상의 총 길이는 15초 이내여야 하며, 각 클립은 2~15초·23.976~60 FPS여야 합니다.
  • 참조 영상은 MP4 또는 MOV에 H.264 또는 H.265를 사용하고 개당 최대 50 MB, 전체 JSON 본문은 64 MB 미만이어야 합니다.

설정

멀티모달 레퍼런스 · 10s · 16:9 · 레퍼런스 3개

MiniMax H3가 레퍼런스 시트에서 생성한 캐릭터 소개: 땋은 포니테일의 흑발 전사가 돌 폐허에서 부츠부터 전신 히어로 포즈까지 드러나는 장면멀티모달 레퍼런스
15s1:1레퍼런스 1개

캐릭터·모션

캐릭터 시트 히어로 인트로

라이브러리에서 가장 많은 좋아요를 받은 커뮤니티 템플릿입니다. 캐릭터 레퍼런스 시트 한 장이 부츠에서 얼굴, 전신으로 이어지는 시네마틱 등장 연출을 이끌며, 어떤 오리지널 캐릭터 디자인에도 통합니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Use @[char ref] as the sole character reference. Preserve the exact identity, face, body proportions, hairstyle, outfit, colors, materials and overall silhouette of the character throughout the entire video. Do not redesign, simplify or replace any defining visual features. Create a cinematic character introduction focused on presence, silhouette, attitude and controlled motion. 0–4s Begin with a close shot of a defining lower-body or detail element such as boots, shoes, feet, hands, clothing hem or an important accessory. The character enters frame or settles into position. The camera slowly tracks upward while hair, clothing and secondary elements move naturally in the wind or environment. 4–8s Reveal more of the body with a medium or medium-wide shot from the back, side or three-quarter angle. The character stands in a calm, composed way inside the environment. The camera makes a smooth orbit, arc or lateral move to gradually reveal the character’s face and silhouette. 8–12s Move into a tight cinematic portrait or upper-body shot. The character performs one subtle signature action that fits their personality, such as lifting the chin, turning the head, adjusting clothing, brushing hair aside, opening a hand, looking toward camera, or shifting posture. Keep the motion minimal and intentional. The expression should match the character’s vibe. 12–15s End with a strong full-body hero shot that clearly presents the entire design and silhouette. Use a low-angle, eye-level or slightly dramatic framing depending on the character’s personality. The character settles into a natural final pose and holds it confidently for a clean final reveal. VISUAL DIRECTION Premium cinematic presentation. Match the visual medium and rendering style of @[char ref]. Emphasize clean silhouette, elegant staging, subtle secondary motion, believable hair and cloth movement, strong composition, atmospheric depth and polished lighting. The scene should feel like a high-end anime, game or film character introduction. CAMERA Use a clear progression from detail reveal to partial reveal to face reveal to full-body hero reveal. Camera movement should be smooth, controlled and intentional. Avoid chaotic motion. ENVIRONMENT Place the character in a fitting environment that supports their identity and mood. The background should enhance the character without distracting from them.

필요한 입력

  • 이미지 1: 캐릭터 레퍼런스. 디자인 시트나 깨끗한 전신 렌더면 됩니다. 영상이 여기서 정체성, 의상, 재질, 렌더링 스타일을 물려받습니다.

작동하는 이유

  • 공개 사다리(디테일 → 부분 → 얼굴 → 전신 히어로)가 고정된 카메라 문법이라, 모델이 변주 여력을 숏 구성이 아니라 캐릭터에 씁니다.
  • "레퍼런스의 시각 매체와 렌더링 스타일을 따르라"는 지시 덕분에 프롬프트 하나가 애니메, 게임 렌더, 실사 캐릭터에 똑같이 통합니다.
  • 세 번째 블록의 "시그니처 동작" 슬롯 하나에 개성이 담깁니다. 의도적으로 작게 잡은 제스처 하나입니다.

교체 가능한 변수

  • 캐릭터 레퍼런스 시트
  • 시그니처 동작
  • 환경 무드
  • 마지막 포즈 프레이밍

제약 사항

  • 원본 게시물은 레퍼런스를 @[char ref]로 표기합니다. EvoLink 경로에서는 소재를 배열 위치로 지칭하세요. "Image 1"이라고 쓰면 image_urls[0]에 매핑됩니다.
  • 캐릭터 하나, 환경 하나. 이 템플릿은 의도적으로 장소를 바꾸지 않으며, 실루엣이 안정적으로 유지되는 이유가 그것입니다.

설정

멀티모달 레퍼런스 · 15s · 1:1 · 레퍼런스 1개

MiniMax H3 시네마틱·VFX 프롬프트

MiniMax H3로 생성한 SF 예고편: 거대한 원형 우주 관문 앞에 선 작은 인영과 어둠 속에서 떠오르는 타이틀멀티모달 레퍼런스
15s16:9레퍼런스 2개

시네마틱·VFX

SF 미스터리 예고편

두 개의 레퍼런스로 분위기와 주인공을 고정하고, 프롬프트로 푸시인·타이틀·사운드를 제어하는 시네마틱 예고편입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Realistic cinematic look, high-contrast lighting, and a tight pace. Use Figure 1 as the overall atmosphere and style reference, and Figure 2 as the protagonist reference. Shot 1 — Ultra-wide establishing shot. A huge circular cosmic gateway nearly fills the frame. The person is only a tiny figure seen from behind before the gateway, positioned toward the lower right. The ground is wet and reflective, and the center of the gateway is pitch black. The camera slowly pushes forward. A large title fades in from the edge of the darkness, blurred at first and then sharp: "THE STARS WERE LISTENING". Use an extremely condensed, heavy, all-caps typeface in dark red mixed with rust red, with subtle grain and misted edges. Audio: a deep low-frequency pulse, faint metallic vibrations in the distance, and a soft hit as the text becomes sharp. → Hard cut.

필요한 입력

  • 이미지 1: 분위기와 스타일 기준. 샷이 물려받을 환경, 색 팔레트, 색보정을 담습니다.
  • 이미지 2: 주인공 참조. 인물이 화면에서 작게 잡혀도 얼굴과 실루엣을 알아볼 수 있을 만큼 크게 촬영합니다.
  • 실제로 화면에 넣고 싶은 타이틀 문구. H3는 적은 그대로 글자를 렌더링하므로 짧게 쓰고 철자를 정확히 확인하세요.

작동하는 이유

  • 참조 소재마다 역할이 하나씩만 주어집니다. 이미지 1은 분위기, 이미지 2는 인물이므로 모델이 어느 쪽이 룩을 결정하는지 추측할 필요가 없습니다.
  • 샷이 한 번의 연속 푸시인과 하나의 사건(타이틀이 선명해지는 순간)으로만 구성되어, 15초가 실제로 담을 수 있는 분량에 맞습니다.
  • 사운드가 "영화음악" 같은 모호한 표현이 아니라 저역 펄스, 멀리서 들리는 금속성 진동, 글자가 선명해질 때의 타격음이라는 세 층으로 적혀 있습니다.

교체 가능한 변수

  • 타이틀 문구와 서체 처리
  • 설정 샷의 관문 또는 랜드마크
  • 타이틀 색상
  • 앰비언스 베드
  • 마지막 컷 처리 방식

제약 사항

  • 소재는 배열 위치로 지칭합니다(‘Figure 1’, ‘Figure 2’). image_urls 순서와 일치시키세요. @image1 표기는 이 API 계약에 포함되지 않습니다.
  • 공개된 프롬프트는 의도적으로 열어 둔 출발점입니다. 하드 컷으로 끝나므로 두 번째 샷을 직접 이어 붙일 수 있습니다.

설정

멀티모달 레퍼런스 · 15s · 16:9 · 레퍼런스 2개

MiniMax H3로 생성한 텍스트-투-비디오 클립: 해질 무렵 주방을 핸드헬드로 촬영하는 가운데 손으로 그린 빛나는 생명체가 소품 사이를 움직이는 장면텍스트-투-비디오
15s16:9레퍼런스 없음

시네마틱·VFX

주방의 빛나는 생명체

실사 주방과 손그림 발광 애니메이션을 섞고, 카메라의 불안정함과 제외 요소를 명확히 쓴 텍스트-투-비디오 템플릿입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

15-second, 16:9 landscape video. Blend live-action footage of a small kitchen at dusk with hand-drawn glowing animation. The last light of sunset lingers by the window. The lived-in kitchen contains an old wooden table, a half-washed mug, a slightly fogged glass bottle, and a hanging dishcloth. Give the footage subtle one-handed smartphone shake, hesitant close-range focusing, exposure fluctuations caused by backlight, and slightly coarse noise in the shadows. It should not look carefully arranged like an advertisement; instead, it should feel like someone hurriedly captured an unbelievable event at home. Do not show huge eyes, gaping mouths, fangs, threatening or lunging movements, sudden black frames, or jump scares. Use only kitchen room tone, cloth rubbing, the soft clink of a mug, water dripping from the faucet, the camera operator's footsteps and quiet breathing, plus gentle electronic sounds and tiny calls from the hand-drawn creature.

필요한 입력

  • 업로드할 자료가 없습니다. 이 경로는 프롬프트만으로 동작합니다.
  • 길이와 화면비는 프롬프트 문장이 아니라 요청 파라미터로 지정합니다.

작동하는 이유

  • 한 손 촬영의 흔들림, 머뭇거리는 초점, 역광에 따른 노출 변화, 어두운 부분의 거친 노이즈까지 카메라의 결함이 명시되어 있습니다. 이것이 "광고"가 아니라 "집에서 급히 찍은 영상"이라는 인상을 만듭니다.
  • 거대한 눈, 송곳니, 달려드는 동작, 점프 스케어를 금지하는 짧고 명확한 목록이 귀여운 생명체가 공포물로 흘러가는 것을 막습니다.
  • 오디오 지시가 소리마다 나뉘어 있습니다. 방의 공기음, 천, 머그컵, 수도꼭지, 발소리, 숨소리, 그리고 생명체 자신의 소리입니다.

교체 가능한 변수

  • 공간과 시간대
  • 생명체 디자인과 행동
  • 식탁 위 소품
  • 카메라의 어떤 결함을 보여줄지
  • 오디오 레이어

제약 사항

  • 프롬프트에 길이와 화면 구성이 적혀 있어도, API에는 duration과 aspect_ratio를 요청 필드로 따로 전달해야 합니다.
  • 부정 지시는 짧고 명확한 목록일 때 가장 잘 작동합니다. 금지 항목을 쌓을수록 동작 묘사에 쓸 예산이 줄어듭니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3가 첫 프레임에서 생성한 애니메 레이스: 두 호버 바이크가 시안과 크림슨 빛 궤적을 남기며 비에 젖은 산길 헤어핀에서 순위를 다투는 장면첫 / 마지막 프레임
15s16:9레퍼런스 1개

시네마틱·VFX

호버 바이크 애니메 레이스

CRITICAL ENTITY LOCK을 건 첫 프레임 기반 스포츠 애니메 레이스입니다. 완전히 명세된 바이크 두 대와 이름 붙은 레이서 두 명만 등장하며, 타임스탬프된 헤어핀, 드래프트, 포토 피니시로 이어집니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Cinematic Anime Video Scene Generate a 15-second horizontal 16:9 original high-speed hover-bike racing anime video from the provided first frame. CRITICAL ENTITY LOCK: There must be exactly 2 racers and 2 bikes in the entire video: RENJI on VALKYRIE-01 (cyan/black drift bike) and ELENA on AERO-X (crimson/white draft bike). Do not add extra racers, drone support vehicles, spectators, or traffic. Maintain total visual consistency for both bikes, helmet visors, suit patterns, repulsor spark colors, and bike liveries throughout the sequence. Entity identity: VALKYRIE-01: Matte-black and cyan angular hover-bike, exposed repulsor pads, lateral drift brakes, blue plasma exhaust trails, ridden by Renji (cyan trim suit). AERO-X: Pearl-white and neon-crimson aerodynamic hover-bike, enclosed canopy, crimson energy draft aura, white-hot central booster, ridden by Elena (crimson/gold visor suit). Video style: High-budget modern sports anime, sakuga-level velocity animation, crisp line art, vibrant neon lighting contrast, high-speed camera tracking, hyper-realistic friction and energy particle effects. Set on a wet downhill mountain pass at dawn. Camera and pacing: Continuous forward velocity, zero slow-motion interruptions: 0.0s - 3.0s: High-speed rear-tracking shot diving into the first downhill hairpin curve; instant drift initiation. 3.0s - 7.5s: Tight side-parallel tracking shot as bikes navigate rock debris and trade positions through S-curves. 7.5s - 11.5s: Close camera lock on the draft-slingshot maneuver; high-energy particle displacement as booster ignition occurs. 11.5s - 15.0s: Low-angle front-facing camera lock on the final sprint to the finish line bridge, ending on a hyper-speed photo-finish freeze. Action timing: 0.0s - 1.5s: Sequence begins at speed. VALKYRIE-01 leads downhill; AERO-X locks onto its rear bumper. Anti-gravity repulsors spray road water and blue sparks into the frame. 1.5s - 4.0s: First sharp hairpin. VALKYRIE-01 deploys lateral drift airbrakes with a burst of blue thruster fire, sliding sideways at 300 km/h. AERO-X stays glued inside its slipstream aura. 4.0s - 7.0s: Mountain debris hazard. VALKYRIE-01 hops over a boulder using a repulsor burst. AERO-X ducks under it, scraping the neon magenta guardrail in a cloud of friction sparks. 7.0s - 10.0s: S-Curve exchange. Bikes lean side-by-side; their repulsor fields collide, creating a bright electrical shockwave. ELENA pulls the overdrive lever; AERO-X's rear fins extend. 10.0s - 13.0s: Slingshot maneuver. AERO-X bursts out of VALKYRIE-01's draft, igniting its central white plasma booster. Both bikes roar down the final straightaway side-by-side. 13.0s - 15.0s: Final sprint toward the finish light gate. Water sprays violently behind them. Both nose cones cross the finish line simultaneously in a flash of light. Final freeze frame. Motion quality: Fluid 2D animation, extreme speed-line integration, stable bike geometry, flawless vehicle reflection rendering, zero limb or body clipping, high-frame-rate kinetic realism. Environment: Wet mountain pass asphalt, sheer cliff walls, neon cyan and magenta guardrail lights, early dawn sky with pink/purple clouds, water spray, floating spark particles. Final output: 15 seconds, horizontal 16:9, original high-budget sports racing anime, exactly 2 racers, relentless kinetic pacing, dynamic cinematography, no subtitles, no watermarks, no logos.

필요한 입력

  • 시작 프레임: 자신의 그림체로 그린 두 레이서와 바이크의 스틸. 시퀀스 전체가 이 이미지에서 확장됩니다.

작동하는 이유

  • 인원 조사식 개체 고정("exactly 2 racers and 2 bikes… no spectators, no traffic")이 레이싱 프롬프트가 흔히 겪는 군중 증식 실패를 차단합니다.
  • 두 머신 모두 도색, 물리 특성, 라이더 슈트가 이름 붙은 정체성으로 주어져 고속에서도 모델이 둘을 구분할 수 있습니다.
  • 카메라와 액션이 서로를 참조하는 별도의 타임라인에 살아, 한 번에 두 카메라 무브를 요구하지 않으면서도 쉼 없는 페이스를 유지합니다.

교체 가능한 변수

  • 시작 프레임 그림체
  • 바이크 정체성과 도색
  • 트랙 위험 요소
  • 결승선 연출

제약 사항

  • 이것은 첫 프레임 경로입니다. 시작 이미지가 그림체를 결정하며, 이미지 투 비디오는 aspect_ratio 파라미터를 받지 않습니다. 프레임이 화면비를 정의합니다.

설정

첫 / 마지막 프레임 · 15s · 16:9 · 레퍼런스 1개

MiniMax H3 네이티브 오디오·대사 프롬프트

MiniMax H3로 생성한 대사 교체: 등장인물의 대사가 참조 오디오의 새 대사로 바뀐 결과멀티모달 레퍼런스
10s16:9레퍼런스 2개

네이티브 오디오·대사

대사와 연기 교체

기존 대사와 새 대사를 그대로 인용하고, 연기가 변할 수 있는 범위를 제한합니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Replace the girl's line in Video 1, "We can't be together. It's not that we don't love each other; we truly can't make it to the end," with the line from Audio 1: "Don't go, okay? This time, let's not let go of each other." Slightly adjust the corresponding performance.

필요한 입력

  • Video 1: 교체할 대사가 들어 있는 클립.
  • Audio 1: 새 대사. WAV 또는 MP3, 최대 15 MB·15초.
  • 두 대사 모두 프롬프트에 한 글자도 빠짐없이 적기.

작동하는 이유

  • 기존 대사를 인용하면 "중간쯤의 대사"가 아니라 클립의 어느 구간을 다뤄야 하는지 정확히 지정됩니다.
  • 새 대사를 인용하면 립싱크가 오디오만으로 추론하지 않고 명확한 목표를 갖습니다.
  • "해당 연기를 살짝 조정한다"는 표현이 연기 변경에 상한을 두어, 나머지 테이크가 살아남게 합니다.

교체 가능한 변수

  • 원본 클립
  • 교체할 대사
  • 새 대사와 목소리
  • 연기를 얼마나 바꿀 수 있는지

제약 사항

  • 오디오만 단독으로 참조 소재가 될 수는 없습니다. 반드시 이미지나 영상과 함께 전달해야 합니다.
  • 참조 오디오와 영상은 요청당 각각 총 15초까지입니다.
  • 고정할 요소를 명시하세요. 대사 이야기만 쓰면 구도, 의상, 배경이 흔들립니다.

설정

멀티모달 레퍼런스 · 10s · 16:9 · 레퍼런스 2개

MiniMax H3로 생성한 보이스 레퍼런스 클립: 참조 오디오의 음색으로 작성된 대사를 말하는 등장인물멀티모달 레퍼런스
10s16:9레퍼런스 2개

네이티브 오디오·대사

‘바람을 따라’ 보이스 레퍼런스

말할 문장과 누구의 음색으로 말할지를 정하는 클립 하나로 구성된 최소 오디오 레퍼런스 프롬프트입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Character dialogue: "Follow the wind, live free. Leave worries behind, enjoy the moment." Use Audio 1 as the voice-timbre reference.

필요한 입력

  • Video 1: 대사를 말할 캐릭터.
  • Audio 1: 목표 음색의 깨끗한 샘플. 2~15초이며 배경 음악이 깔리지 않은 것이 좋습니다.
  • 말하게 할 대사를 그대로 적은 문장.

작동하는 이유

  • 대사를 요약이 아니라 인용으로 적었기 때문에 타이밍과 립싱크가 붙잡을 구체적인 목표가 생깁니다.
  • Audio 1에 "음색"이라는 좁은 역할 하나만 부여했고, 일반적인 배경음악처럼 넘기지 않았습니다.
  • 그 외에는 아무것도 지정하지 않아 구도와 연기는 참조 클립이 계속 통제합니다.

교체 가능한 변수

  • 대사
  • 보이스 레퍼런스
  • 캐릭터 클립
  • 말하는 속도

제약 사항

  • 보이스 레퍼런스가 옮기는 것은 음색이며 억양, 감정, 속도는 함께 넘어가지 않습니다. 중요하다면 프롬프트에 따로 쓰세요.
  • 참조 오디오에 음악이나 여러 화자가 섞이면 품질이 떨어집니다. 분리된 단일 음성을 제공하세요.

설정

멀티모달 레퍼런스 · 10s · 16:9 · 레퍼런스 2개

MiniMax H3가 캐릭터와 오디오 레퍼런스로 생성한 몽타주: 뿔 달린 백발 캐릭터가 비트에 맞춘 빠른 숏으로 다섯 환경을 오가는 장면멀티모달 레퍼런스
15s1:1레퍼런스 2개

네이티브 오디오·대사

오디오 싱크 환경 몽타주

캐릭터 이미지가 정체성을 고정하고 오디오 클립이 편집을 지휘하는 이중 레퍼런스 몽타주입니다. 다섯 환경, 환경마다 여섯 개의 버스트 컷, 모든 컷이 트랙의 악센트에 맞아떨어집니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Use @[char ref] as the strict character reference and @[audio ref] as the timing, rhythm and editing reference. Keep the character’s exact identity, proportions, hairstyle, outfit, colors and overall style consistent throughout. Create a 15-second cinematic burst-cut video showcasing the character across 5 different environments that naturally fit their design, vibe and world. AUDIO SYNC Synchronize the entire edit to @[audio ref]. Cuts, camera accents, transitions and environment changes should land precisely on strong beats, half-beats and musical accents. Let audio1 control the pacing and intensity of the montage. STRUCTURE - 5 environments total - 3 seconds per environment - 6 burst-cut shots per environment - 30 shots total Each environment must be clearly different in atmosphere, lighting, scale and visual language. Show each environment through rapid cinematic angles: wide establishing shots, aerials, low angles, side views, tracking shots, close environmental details, medium shots and hero frames. Every cut must reveal a new angle, distance, composition or spatial relationship. Avoid repeated framing. Mix static shots, push-ins, pull-backs, tracking, orbit and crane-like movement. Keep character movement subtle and natural. The focus is environmental variety, cinematic framing and tight synchronization with audio1. Hard constraints: - exactly 5 environments - exactly 6 shots per environment - exactly 30 shots total - environment changes must follow audio1’s musical phrasing - cuts and motion accents synchronized to audio1 - no outfit changes - no character duplication - no morphing - no text or UI - no blurry unreadable frames - maintain strict character consistency

필요한 입력

  • 이미지 1: 모든 숏이 유지해야 할 정체성의 캐릭터.
  • Audio 1: 페이스를 소유하는 트랙. 컷, 전환, 환경 변경이 그 비트를 따릅니다.

작동하는 이유

  • 모달리티 간 분업이 깔끔합니다. 이미지는 "누구", 오디오는 "언제"에 답하며 어느 쪽도 텍스트와 다투지 않습니다.
  • 정확한 산수(환경 5 × 숏 6 = 컷 30)가 강행 제약으로 명시되어 막연한 몽타주를 셀 수 있는 구조로 바꿉니다.
  • "캐릭터 움직임은 절제"가 에너지를 전부 카메라 다양성으로 밀어 넣습니다. 몽타주는 동작 다양성보다 카메라 다양성을 훨씬 잘 견딥니다.

교체 가능한 변수

  • 캐릭터 레퍼런스
  • 오디오 트랙과 그 프레이징
  • 다섯 환경
  • 숏 타입 구성

제약 사항

  • 원본 게시물은 @[char ref]와 @[audio ref]로 씁니다. EvoLink에서는 배열 위치(이미지 1, Audio 1)로 지칭하고, 오디오는 결코 단독 레퍼런스 타입이 될 수 없다는 점을 기억하세요.
  • 이 경로의 참조 오디오 클립은 2~15초여야 합니다.
  • 공개된 클립의 실측값은 8:9로, API가 받아들이지 않는 비율입니다. 같은 정사각형에 가까운 화면은 1:1로 요청하세요.

설정

멀티모달 레퍼런스 · 15s · 1:1 · 레퍼런스 2개

MiniMax H3 게임·인터페이스 프롬프트

MiniMax H3가 아홉 레퍼런스로 생성한 UI 애니메이션: 미니멀한 크리처 도감에서 커서가 초원의 복슬복슬한 뿔 달린 크리처 MEADOW CROWN을 선택하는 장면멀티모달 레퍼런스
15s16:9레퍼런스 9개

게임·인터페이스

크리처 도감 UI 데모

아홉 개의 레퍼런스를 명시적 역할에 매핑합니다. UI 레이아웃 원판 하나와 크리처 카드 여덟 장을 고정 카메라 도감 데모로 애니메이션화하고, 커서가 항목을 클릭해 나가다 마지막 크리처가 커서를 삼킵니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Use Image 1 as the exact UI/layout/style reference for the creature encyclopedia screen. Use Images 2–9 as the exact creature references. Map them like this: Image 2 = card A = LUMI HARE Image 3 = card B = CLOUD WISP Image 4 = card C = EMBER FENNEC Image 5 = card D = TIDE BEHEMOTH Image 6 = card E = PETAL VULPIN Image 7 = card F = ORCHARD EYE Image 8 = card G = MEADOW CROWN Image 9 = card H = FROST GLIDER Create a 15-second 16:9 video. Keep the camera locked. Keep the interface, layout, typography, panels, icons and overall composition stable, elegant and readable. The UI should feel like a modern minimal digital creature encyclopedia, similar to a sleek pokedex. No scene cuts, no extra text, no extra buttons, no UI distortion. Sequence: 0–2.5s: Cursor clicks card A. Main creature becomes LUMI HARE. Title changes to “LUMI HARE”. Creature blinks and rotates slightly. 2.5–5s: Cursor clicks card C. Main creature becomes EMBER FENNEC. Title changes to “EMBER FENNEC”. Cursor drags to rotate it left and right. 5–7.5s: Cursor clicks card E. Main creature becomes PETAL VULPIN. Title changes to “PETAL VULPIN”. Cursor pokes it a few times. It reacts, annoyed. 7.5–10s: Cursor clicks card G. Main creature becomes MEADOW CROWN. Title changes to “MEADOW CROWN”. Cursor taps near the face/horns. It recoils slightly. 10–12s: Cursor clicks card D. Main creature becomes TIDE BEHEMOTH. Title changes to “TIDE BEHEMOTH”. Cursor keeps poking it. 12–15s: TIDE BEHEMOTH gets angry, opens its mouth very wide, lunges forward, and swallows the cursor. Then it returns to idle. Title stays “TIDE BEHEMOTH”. Rules: - When a card is selected, both the main creature and the main title must update. - Only animate cursor, selection state, title change, and the selected creature. - Only one cursor. - Keep motion subtle and clean until the final swallow. - No cuts, no camera move, no UI distortion, no extra text. Audio: soft UI click sounds, subtle hover sounds, tiny creature reaction sounds, then a sharper aggressive creature sound and one comedic swallow gulp at the end.

필요한 입력

  • 이미지 1: 타이포그래피, 패널, 구도를 소유하는 도감 UI 레이아웃.
  • 이미지 2~9: 카드당 크리처 한 마리. 각각 프롬프트의 카드 맵에서 이름이 지정됩니다.

작동하는 이유

  • 이미지-카드 매핑 표(Image 2 = card A = LUMI HARE…)는 가능한 가장 직설적인 역할 배정입니다. 모델이 어느 소재가 무엇인지 결코 추측하지 않습니다.
  • 카메라가 고정되고 네 가지(커서, 선택 상태, 타이틀, 활성 크리처)만 움직일 수 있어 UI 모션의 실패 표면이 거의 0으로 줄어듭니다.
  • 코믹 비트(크리처가 커서를 삼키기)가 맨 끝에 배치되어 데모가 페이오프 전까지 깔끔하게 유지됩니다.

교체 가능한 변수

  • UI 스타일 원판
  • 여덟 크리처와 이름
  • 인터랙션 대본
  • 효과음 구성

제약 사항

  • 이미지 9장은 이 경로의 타입별 최대치이며, 파일 총합 12개 상한 때문에 그 위에 영상 3개와 오디오 3개를 더할 수는 없습니다.
  • UI 텍스트 안정성은 "no camera move, no UI distortion"에 의존합니다. 카메라를 풀면 일그러진 패널이 되돌아옵니다.

설정

멀티모달 레퍼런스 · 15s · 16:9 · 레퍼런스 9개

MiniMax H3로 생성한 게임플레이풍 클립: 제네릭 FPS HUD가 표시된 화면에서 소총을 조준한 1인칭 시점이 연기 자욱한 군 기지를 전진하는 장면텍스트-투-비디오
15s16:9레퍼런스 없음

게임·인터페이스

FPS 게임플레이 시뮬레이션

캡처된 게임플레이처럼 읽히는 1인칭 슈터 시퀀스입니다. 플레이어가 조작하는 카메라 문법, 완전히 명세된 제네릭 HUD, 그리고 안무가 아니라 전술로 쓰인 페이스.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Camera: First-person perspective at eye level with authentic handheld player movement, as if recorded directly from a modern AAA military shooter. The player carries a highly detailed assault rifle with realistic animations, visible hands, tactical gloves, dynamic reload mechanics, and weapon sway. **Opening Action:** The video immediately begins with the player already aiming down a roadway inside a modern military base. Multiple enemy soldiers are visible in the distance near sandbags, barricades, and military vehicles. The player carefully tracks one target, making small aim corrections while maintaining ADS (aim down sights). Fire several controlled bursts immediately at the visible enemies, producing realistic muzzle flashes, shell casings ejecting, smoke, recoil, hit reactions, and dust impacts around the targets. Continue firing in multiple short bursts while adjusting aim between enemies, simulating authentic FPS gameplay rather than scripted animation. **Movement:** After the opening firefight, lower slightly from ADS and begin advancing cautiously along the road beside concrete barriers, Hesco walls, and parked military vehicles. Frequently check left and right corners, briefly stop to reacquire targets, then raise the weapon and fire additional controlled bursts whenever enemies appear ahead. Continue pushing forward with deliberate player-controlled movement, using cover naturally and maintaining believable tactical pacing. **Environment:** Large modern military base with guard towers, armored vehicles, shipping containers, blast barriers, damaged buildings, smoke plumes, burning debris, scattered shell casings, dust clouds, and atmospheric battlefield haze. Cool natural daylight mixed with smoke and orange firelight creates a cinematic battlefield atmosphere. **Camera Motion:** Authentic player-controlled movement with subtle head bob, weapon sway, natural mouse-look adjustments, small left-right corrections while aiming, realistic recoil, smooth tracking of moving targets, brief pauses before shooting, and fluid forward progression. Avoid cinematic camera moves—everything should feel like genuine live gameplay captured by a skilled player. **Visual Quality:** Ultra-photorealistic, AAA game graphics with realistic PBR materials, detailed weapon models, physically accurate lighting, volumetric smoke, dynamic particle effects, crisp textures, realistic bullet impacts, muzzle flash illumination, motion blur only during rapid movement, and high-end military shooter presentation. **Gameplay UI:** Display a realistic modern FPS HUD inspired by games like PUBG, Battlefield, or Call of Duty (without copying exact copyrighted assets). Include: * Central dynamic crosshair or reticle * Ammo counter with magazine and reserve ammunition * Fire mode indicator * Compass at the top * Squad/team status panel * Mini-map in the upper corner * Health bar * Tactical equipment icons (grenades, medkit) * Hit markers when bullets connect * Directional damage indicators * Kill notification feed * Objective marker in the distance * Subtle interaction prompts and realistic HUD animations The HUD should feel polished, modern, and fully integrated into the gameplay, enhancing the illusion of authentic recorded footage from a contemporary military FPS. #MiniMaxH3

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 리얼리즘 목표가 "숙련된 플레이어가 녹화한 화면"이라, 헤드 밥, 조준 보정, 멈췄다 쏘는 페이스가 카메라 행동으로 명세됩니다.
  • HUD가 히트 마커와 킬 피드까지 항목화되면서도 명시적으로 제네릭을 유지합니다. 실제 게임으로 읽힐 만큼 촘촘하고, 게시해도 안전할 만큼 일반적입니다.
  • "Avoid cinematic camera moves"가 핵심 반전입니다. 대부분의 프롬프트가 원하는 바로 그것이 이 프롬프트를 망가뜨립니다.

교체 가능한 변수

  • 환경과 진영 스타일링
  • HUD 요소 구성
  • 교전 리듬
  • 날씨와 빛

제약 사항

  • 프롬프트는 HUD를 "카피하지 않고 영감만 받은" 상태로 유지합니다. 이 조항을 보존하세요. 특정 게임의 HUD를 복제하는 것은 다른(그리고 더 위험한) 작업입니다.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3 브이로그·셀피 카메라 프롬프트

MiniMax H3로 생성한 셀피 스타일 클립: 숲에서 자신을 찍던 여성이 휴대폰을 돌려 나무 사이에 추락해 연기를 내뿜는 UFO를 비추는 장면텍스트-투-비디오
15s1:1레퍼런스 없음

브이로그·셀피 카메라

셀피 캠 UFO 발견

폰 리얼리즘 쇼케이스입니다. 오토포커스 브리딩과 롤링 셔터가 살아 있는 핸드헬드 셀피 영상, 대본화된 일본어 대사, 얼굴에서 추락한 UFO로 뒤집히는 카메라 전환.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Ultra photorealistic live-action captured on an iPhone 17. Authentic handheld selfie footage with premium cinematic documentary color grading, realistic HDR, deep green foliage, warm sunlight, subtle teal shadows, natural skin tones, gentle filmic contrast, rolling shutter, autofocus breathing, slight motion blur, and natural handheld shake. A lush forest in daytime with dense trees, wild plants, an uneven dirt trail, scattered leaves, soft sunlight through the canopy, and a gentle breeze. The atmosphere is quiet and slightly unsettling. A cute Japanese woman in her early twenties wearing a stylish bikini walks through the forest while recording herself in selfie mode. She suddenly notices something ahead, looks shocked, turns the camera, and points into the distance. A large crashed UFO is partially embedded in the forest floor. Its metallic hull is badly damaged with broken panels, scorch marks, exposed internal structures, thick gray smoke, and occasional sparks. She says in Japanese: 「ちょっと待って! あそこ見て! UFOじゃない!? 完全に墜落してるんだけど! 煙まで出てる! やばい、本物かもしれない! ちょっと近づいてみる!」 She alternates between filming herself and the UFO while continuing to point at it. Continuous single take. Natural walking movement, realistic hand tremors, slight framing imperfections, quick pans, and autofocus shifts between her face and the UFO. Natural sunlight creates cinematic highlights, soft shadows, realistic reflections on the UFO, subtle volumetric light, and realistic smoke. Audio: footsteps on leaves, gentle wind, birds becoming quieter near the crash site, creaking branches, faint electrical crackling from the UFO, and distant eerie unidentified animal calls echoing through the forest. Negative: no blood, no visible aliens, no monsters, no horror creature reveal, no excessive explosions, no CGI, no cartoon style, no text, no subtitles, no watermark, no logo.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 폰 특유의 결함(오토포커스 헤매기, 롤링 셔터, 손 떨림, 어긋난 프레이밍)이 열거되어 있습니다. 리얼리즘은 "realistic"이라는 단어가 아니라 이름 붙인 결함에서 나옵니다.
  • 대사가 일본어 원문 그대로 인용되어 있어, H3의 네이티브 오디오가 웅얼거림이 아니라 립싱크가 맞는 실제 발화를 생성합니다.
  • 오디오가 다이제틱으로 겹겹이 쌓입니다. 발소리, 바람, 조용해지는 새소리, 전기 튀는 소리. 배경음악이 없어 발견된 폰 영상이라는 환상이 깨지지 않습니다.

교체 가능한 변수

  • 대사와 언어
  • 발견되는 물체
  • 숲 또는 도시 배경
  • 의상과 캐릭터 스타일링

제약 사항

  • 부정 목록(외계인 금지, 호러식 정체 공개 금지)이 하중을 받칩니다. 클립을 티저 영역에 붙들어 두고 모델이 장면을 증폭시키는 것을 막습니다.
  • 공개된 클립은 1:1로 렌더링되었습니다. 프레임은 화면비 파라미터로 지정하고, 자신과 물체를 오가는 프레이밍은 텍스트에 유지하세요.

설정

텍스트-투-비디오 · 15s · 1:1 · 레퍼런스 없음

MiniMax H3로 생성한 다큐멘터리 클립: 거리 사진가가 카페 앞 노신사와 테리어를 프레이밍한 뒤 찍힌 사진을 시청자에게 보여 주는 장면텍스트-투-비디오
15s16:9레퍼런스 없음

브이로그·셀피 카메라

거리 사진가의 순간

카메라 안의 카메라가 있는 압축 다큐멘터리 비트입니다. 사진가가 캔디드한 거리 장면을 프레이밍해 셔터를 누른 뒤, 방금 찍은 사진을 카메라를 돌려 시청자에게 보여 줍니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

A young Western female street photographer walks through a lively downtown street and notices an elderly man sitting outside a café with his small dog. She carefully composes the candid moment through her camera, captures the photo, then turns the camera toward the viewer to proudly show the shot she just took. She smiles, says “Look at that,” then continues walking through the city. Ultra-photorealistic visuals, natural handheld documentary movement, realistic camera interaction, authentic facial expressions, accurate hand movements, realistic dog behavior, natural daylight, cinematic depth of field, continuous character consistency, immersive city ambience, premium documentary realism.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 사진 보여 주기 비트가 H3에 영상 속 정합적인 스틸 이미지 렌더링을 강제합니다. 자연스러운 제스처 하나로 읽히는 조용한 능력 시연입니다.
  • 인터랙션 체인이 완전히 명세되어(구도 → 촬영 → 뒤집기 → "Look at that" → 계속 걷기) 클립이 한 테이크 안에 완결된 아크를 가집니다.
  • 피사체가 정체성이 아니라 역할로 묘사됩니다. 노신사, 작은 개, 카페. 거리 장면이 일반적이고 안전하게 유지됩니다.

교체 가능한 변수

  • 캔디드 피사체
  • 도시와 빛
  • 한 마디 대사
  • 카메라 소품 종류

제약 사항

  • 카메라 속 사진은 그것이 찍힌 장면과 일치해야 합니다. 피사체를 바꾸면 두 묘사를 함께 바꾸세요.

설정

텍스트-투-비디오 · 15s · 16:9 · 레퍼런스 없음

MiniMax H3로 생성한 여행 브이로그: 여성이 커튼을 열어 해변 테라스를 드러낸 뒤 셀피 모드로 전환해 바다를 등지고 아침 인사를 하는 장면텍스트-투-비디오
15s9:16레퍼런스 없음

브이로그·셀피 카메라

해변 아침 브이로그 아크

의도적인 모드 전환이 있는 초 단위 아침 브이로그입니다. 처음 5초는 시네마틱한 3인칭이고, 그 뒤에야 인물이 자신을 찍기 시작합니다. "브이로그"가 시작되는 순간입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Create a 15-second ultra-realistic cinematic lifestyle vlog video, vertical 9:16, featuring the same young woman throughout the entire video. Preserve her facial identity, facial proportions, hairstyle, skin tone and overall appearance consistently in every shot. She wears the same outfit throughout: fitted white V-neck T-shirt with a small subtle logo, blue denim jeans, natural makeup, long softly wavy brown hair. 0:00–0:01 — Wake-up: Close-up inside a beautiful bright bedroom. The woman is lying comfortably on the bed, slowly wakes up, stretches naturally and opens her eyes. She is NOT filming a vlog yet and does not hold a phone or camera. Soft morning sunlight enters through the curtains. 0:01–0:02 — Gets up: Medium shot. She sits up on the bed, smiles softly, fixes her hair and gets ready to start her morning. Natural, effortless movement. 0:02–0:03 — Walks to window: She walks toward the large glass balcony door/window. Camera follows her naturally from behind/side. 0:03–0:04 — Seaside reveal: She opens the curtains/door and looks outside. Reveal a breathtaking blue ocean, coastal hills, flowers, balcony and beautiful morning sunlight. She smiles happily while taking in the view. 0:04–0:05 — Steps outside: She walks out onto the seaside terrace. Gentle ocean breeze moves her hair naturally. Wide cinematic shot showing the beautiful surroundings. 0:05–0:06 — VLOG START: Only now she starts filming herself in handheld selfie-vlog style. She looks into the camera with a bright natural smile and says: “Good morning!” 0:06–0:07 — Show the view: She turns the camera away from herself and slowly pans across the stunning ocean, coastal mountains, flowers and terrace. Smooth handheld vlog movement. 0:07–0:08 — Back to selfie: Selfie shot. She looks into the camera and happily says: “This place is just perfect!” 0:08–0:09 — Location reveal: Wide cinematic shot of the cozy seaside terrace with wooden table, chairs, plants and flowers overlooking the ocean. 0:09–0:10 — Walk to table: Medium tracking shot as she walks toward the table, enjoying the view. Her hair and T-shirt move gently in the sea breeze. 0:10–0:11 — Sit and relax: She sits at the seaside table, smiling peacefully and enjoying the ocean view. A refreshing orange-colored juice is placed on the table. 0:11–0:12 — Juice close-up: Cinematic close-up of her hand picking up the glass of fresh orange juice. Beautiful ocean bokeh in the background, natural sunlight reflecting through the glass. 0:12–0:13 — Vlog toast: Selfie shot. She raises the juice toward the camera with a cheerful smile and says: “Cheers to good days!” 0:13–0:14 — Happy close-up: Beautiful close-up of her smiling naturally at the camera, ocean and warm sunlight softly blurred behind her. 0:14–0:15 — Ending: Camera moves from her toward the sparkling ocean and peaceful coastal landscape. Warm sunlight, gentle waves and a relaxing cinematic ending. Overall Style Ultra-realistic, cinematic travel vlog, natural handheld camera movement, realistic human motion, smooth transitions, soft morning sunlight, realistic ocean waves, gentle wind in hair and clothes, beautiful coastal atmosphere, premium lifestyle aesthetic, natural expressions, authentic vlog feeling, shallow depth of field, cinematic composition, realistic skin texture, high detail, 4K quality.

필요한 입력

  • 업로드할 자료가 없습니다. 텍스트 투 비디오 경로는 프롬프트만으로 동작합니다.

작동하는 이유

  • 기상 비트의 "she is NOT filming yet" 노트가 가장 흔한 브이로그 아티팩트, 즉 브이로그 시작 전에 폰이 등장하는 문제를 예방합니다.
  • 1초짜리 비트 열다섯 개가 셀피 캠과 풍경 컷어웨이를 실제 여행 브이로그 문법 그대로 번갈아 배치합니다.
  • 짧게 인용된 세 마디("Good morning!")가 지속해야 할 독백 없이 네이티브 오디오에 자연스러운 체크포인트를 줍니다.

교체 가능한 변수

  • 공개되는 장소
  • 세 마디 대사
  • 의상과 정체성 고정
  • 음료 소품

제약 사항

  • 프롬프트는 9:16을 요구하며 요청에도 그 비율을 넣으면 됩니다. 공개된 클립은 대략 3:2로 재인코딩되었는데, 방향은 산문이 아니라 요청에서 지정해야 한다는 또 하나의 이유입니다.

설정

텍스트-투-비디오 · 15s · 9:16 · 레퍼런스 없음

MiniMax H3 편집·변환 프롬프트

MiniMax H3로 생성한 영상 편집: 원본 클립의 신문, 의자, 선글라스, 불타는 자동차가 교체되거나 제거된 결과멀티모달 레퍼런스
10s16:9레퍼런스 1개

편집·변환

여러 요소를 바꾸는 장면 편집

원본 클립에 여섯 개의 변경을 대상 → 결과 목록으로 쓰고 장면은 다시 설명하지 않습니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Replace the newspaper in the reference video with a green-covered book; change the chair the character is sitting on to a red sofa; remove the sunglasses worn by the character to retain a clear face; remove the car burning effect to keep the vehicle in a normal state; change the photo the character takes out of his arms to a small black book; and add a tree on the left side of the screen

필요한 입력

  • Video 1: 편집할 클립. 2~15초, MP4 또는 MOV, H.264 또는 H.265, 최대 50 MB.
  • 편집 목록. 각 항목에 무엇을 어떻게 바꿀지 명시합니다.

작동하는 이유

  • 모든 지시가 대상과 결과의 쌍이라 해석의 여지가 없습니다("신문"이 "초록색 표지의 책"이 됩니다).
  • 삭제 지시에는 원하는 최종 상태도 함께 적혀 있어("선글라스를 없애 얼굴이 드러나도록") 물건이 있던 자리에 구멍이 남지 않습니다.
  • 장면 묘사가 전혀 없기 때문에 모델은 원본 클립을 사실로 받아들이고 차이만 적용합니다.

교체 가능한 변수

  • 편집 항목 수
  • 교체할 사물
  • 제거할 효과
  • 추가할 요소

제약 사항

  • EvoLink가 제공하는 H3 경로는 세 가지이며, 편집 작업은 원본 클립을 Video 1로 넣는 reference-to-video로 수행합니다.
  • 참조 영상의 길이는 과금 대상이므로 업로드 전에 필요한 구간만 남기고 잘라내세요.
  • 변경 대상 옆에 지켜야 할 것이 있다면 무엇이 그대로여야 하는지 명시하세요. 적히지 않은 영역은 모델이 마음대로 다룰 수 있습니다.

설정

멀티모달 레퍼런스 · 10s · 16:9 · 레퍼런스 1개

MiniMax H3로 생성한 무대 영상: 연기 속에서 정장 색을 맞바꾸는 두 마술사와 빨강에서 파랑으로 바뀌는 커튼첫 / 마지막 프레임
7s16:9레퍼런스 1개

편집·변환

마술사 의상 교환

두 의상을 바꾸고, 한 세부 요소는 유지하며, 배경은 지정한 타이밍에 색상을 전환하는 지시 수행 테스트입니다.

세부정보 보기

전체 프롬프트(검증된 영문 원문)

Two magicians stand onstage facing the audience and perform a "swap" trick. They wave their wands at the same time, and a cloud of smoke rises. When it clears, their suit colors have switched: the person on the left wears a white suit, and the person on the right now wears a black suit, while both magicians' glove colors remain unchanged. They bow to thank the audience. The red curtain behind them closes, transitioning from deep red to deep blue.

필요한 입력

  • 두 연기자와 의상, 무대가 담긴 시작 이미지.
  • 교환되는 대상에 대한 명확한 전후 상태.

작동하는 이유

  • 교환을 연기(煙)라는 사건으로 가려서, 모델이 화면에서 모핑하지 않고 변화를 수행할 정당한 순간을 얻습니다.
  • 바뀌면 안 되는 것(장갑 색)이 바뀌어야 하는 것 바로 옆에 적혀 있습니다. 편집이 번지지 않게 하는 방법이 바로 이것입니다.
  • 샷이 정해진 최종 상태로 끝납니다. 인사, 닫히는 커튼, 짙은 파랑으로 안착하는 색.

교체 가능한 변수

  • 연기자와 의상
  • 그대로 유지할 세부
  • 교환을 가리는 사건
  • 마지막 색 전환

제약 사항

  • 이미지 투 비디오 경로는 입력 이미지에서 화면비를 결정하며 aspect_ratio 필드를 허용하지 않습니다.
  • 적히지 않은 속성은 모델이 마음대로 다룰 수 있습니다. 변화 속에서 지켜야 할 세부가 있다면 반드시 적으세요.

설정

첫 / 마지막 프레임 · 7s · 16:9 · 레퍼런스 1개

MiniMax H3 API를 찾고 계신가요?

초 단위 요금, 3가지 생성 라우트, 연동 문서가 있는 모델 페이지로 이동하세요.

API 페이지 열기

MiniMax H3 프롬프트 프레임워크

H3는 순서가 있는 레퍼런스를 이해하고 오디오도 생성합니다. 일반적인 공식보다 실제 워크플로우에 맞춰 작성하세요.

H3 프롬프트 기본 구조

목표 → 순서가 있는 레퍼런스 → 주체와 정체성 → 시간 순서의 동작 → 카메라 경로 → 오디오 또는 대사 → 유지할 요소 → 최종 상태 순으로 작성하면 소리와 결말을 놓치지 않습니다.

텍스트-투-비디오

클립 안에서 완료할 수 있는 타임라인을 작성하세요. 주체, 소수의 동작, 하나의 연속 카메라 경로, 대사·환경음·음악을 따로 지정합니다.

첫 / 마지막 프레임

이미지가 외형을 이미 결정합니다. 사이의 변화, 카메라 움직임, 유지할 내용, 최종 상태만 설명하세요.

멀티모달 레퍼런스

각 자산에 하나의 역할만 주고 배열 순서로 지칭합니다. Image 1은 캐릭터, Image 2는 장소, Video 1은 모션, Audio 1은 음색을 담당합니다. 요청당 파일은 총 12개까지입니다.

기존 클립 편집

대상 → 결과를 적는 ‘변경’ 목록과 그대로 둘 내용을 적는 ‘유지’ 목록으로 나누세요. 필요한 구간만 Video 1로 전송합니다.

오디오와 대사

원하는 대사를 그대로 인용합니다. 보이스 레퍼런스는 음색에만 사용하고, 억양·감정·속도는 프롬프트에 별도로 적습니다.

MiniMax H3 프롬프트 FAQ

MiniMax H3와 Hailuo 3는 같은 모델인가요?

네. MiniMax H3는 Hailuo 3, Hailuo 3.0, Hailuo 03으로도 불립니다. EvoLink는 MiniMax H3라는 제품명과 텍스트-투-비디오, 이미지-투-비디오, 레퍼런스-투-비디오 라우트를 제공합니다.

코드 없이 이 프롬프트를 사용할 수 있나요?

네. ‘이 프롬프트 사용’을 누르면 검증된 영문 원본과 맞는 워크플로가 MiniMax H3 Playground로 전달됩니다. 레퍼런스 파일은 그곳에서 추가하세요.

MiniMax H3 클립은 몇 초까지 만들 수 있나요?

세 라우트 모두 4〜15초의 정수 길이를 지원하며 현재 2K로 출력합니다. 카드의 길이는 공개된 예제 클립과 같습니다.

한 프롬프트에 레퍼런스를 몇 개 사용할 수 있나요?

레퍼런스 라우트는 최대 9개 이미지, 3개 비디오, 3개 오디오를 받지만 파일 총합은 12개까지입니다(9+3+3 전체는 거부됩니다). 이미지나 비디오가 최소 하나 필요하며 오디오만 보낼 수는 없습니다. 비디오와 오디오는 2〜15초여야 합니다.

프롬프트에 왜 ‘Image 1’이나 ‘Video 1’을 쓰나요?

API는 image_urls, video_urls, audio_urls 배열의 위치로 자산을 구분합니다. 검증된 프롬프트에 남겨 둔 영어 표기는 특정 자산에 역할을 부여합니다.

프롬프트와 클립의 출처는 어디인가요?

각 카드에 출처가 표시됩니다. MiniMax 공식 예제는 중국어 원문의 검증된 영문판을 사용하고, 커뮤니티 예제는 X의 원본 게시물로 연결됩니다.

API로 이 프롬프트를 사용할 수 있나요?

네. 같은 영문 프롬프트를 Playground와 API의 prompt 필드에 사용할 수 있습니다. API 가이드에서 요청, 비동기 태스크, 콜백, 오류 처리를 설명합니다.

프롬프트를 고르고 생성하세요

모든 프롬프트는 EvoLink의 실제 MiniMax H3 라우트에 연결됩니다. Playground로 보내거나 통합 API로 같은 요청을 구축하세요.