TIP

✈️ 15초 만에 파리를 여행하는 법: AI VLOG 제작 레시피

디플릭 Editorial

🗓️ 2026.05.19

파리 여행의 낭만을 15초에 압축하는 법: AI 스토리텔링 브이로그 레시피



파리 여행의 낭만을 단 15초의 고밀도 에너지로 압축한다면 어떤 느낌일까요?


우리가 살펴볼 사례는 '어떻게 하면 AI로 만든 영상이 보고 싶은 이야기가 되는가'에 대한 하나의 해답을 보여줍니다. 단순히 풍경을 나열하는 것이 아니라, 대사 한 마디로 시작해 고조되는 감정의 흐름을 담아낸 이 프로젝트의 디테일을 함께 짚어볼까요?



©X, @craftian_keskin



🖼️ 이미지 분석: '나'를 기록하는 가장 자연스러운 방법


"여행의 주인공은 풍경이 아니라, 그 풍경 속에 스며든 '나'의 순간들입니다.”


이미지들의 매력은 인위적인 포즈가 아니라, 친구와 걷다가 혹은 혼자 여행하며 툭툭 찍은 듯한 '시점의 자연스러움'에 있습니다.


  • 비주얼 포인트 01. 1인칭과 3인칭의 조화: 정면을 응시하는 셀피뿐만 아니라, 뒷모습이나 먼 곳을 응시하는 캔디드(Candid) 샷을 섞어 시청자가 마치 함께 여행하는 동행자가 된 듯한 느낌을 줍니다.
  • 비주얼 포인트 02. 빛과 공기의 서사: 쨍한 고화질보다는 해 질녘의 역광이나 살짝 과노출된 하늘을 활용해, 그날의 날씨와 분위기(Vibe)를 이미지에 연출 했습니다.
  • 비주얼 포인트 03. 여행자의 페르소나: 일관된 패션 아이템(베레모, 선글라스)을 활용해 여러 장의 사진이 하나의 일관된 스토리라인으로 연결되게 구성했습니다.



[💻복사해서 바로 쓰는 ‘자연스러운 여행자’ 이미지 프롬프트]


{ "Objective": "Generate a 3x3 grid collage of travel vlog-style photos featuring a Japanese female traveler exploring iconic locations in Paris, captured with a low-quality smartphone aesthetic.", "Format": { "Aspect Ratio": "1:1 square collage", "Type": "Photo grid", "Grid": { "Rows": 3, "Columns": 3, "Total Images": 9, "Spacing": "Equal spacing" } }, "Concept": { "Theme": "Paris travel vlog", "Style": "Casual, candid, influencer-style moments", "Aesthetic": "Low-quality smartphone photos with natural imperfections" }, "Persona Details": { "Character": "Japanese female traveler", "Appearance": "Slender, very attractive, natural beauty", "Wardrobe": "Chic Parisian summer outfits (light dresses, beret, sunglasses)", "Mood": "Romantic, curious, cheerful" }, "Visual Style": { "Camera": "Handheld smartphone", "Quality": "Low resolution, slight grain/noise", "Lighting": "Mixed natural lighting with occasional overexposure", "Vibe": "Authentic, unpolished, travel diary aesthetic" } }



디렉터의 팁:
이미지를 생성할 때 looking at camera 보다는 caught in a moment나 walking away 같은 키워드를 사용해 보세요. 카메라를 의식하지 않는 찰나의 순간이 이미지를 훨씬 더 '진짜'처럼 만듭니다.





🎥 영상 콘텐츠 분석: 대사가 숨을 불어넣는 숏폼 문법


"시선을 붙드는 건 0.5초 단위의 컷 편집과 사운드 동기화입니다."


이 영상이 감각적으로 느껴지는 이유는 숏츠(Shorts)와 릴스(Reels) 플랫폼의 특성을 완벽히 이해했기 때문입니다.


  • 포인트 01. 대사(Dialogue)의 힘: 영상 시작과 끝에 배치된 일본어 대사는 시청자의 청각을 즉각적으로 자극합니다. "15초 만에 파리를 다 보여줄게"라는 선언적 오프닝과 "파리, 너무 좋아"라는 여운 있는 엔딩은 영상의 완결성을 부여하며 감성적 몰입을 돕습니다.
  • 포인트 02. 고밀도 압축 편집: 드럼 비트에 맞춰 0.5초 단위로 전환되는 컷 편집은 지루할 틈을 주지 않습니다. 특히 빠른 속도의 '줌 인/아웃'과 '스핀 전환'은 여행의 흥분을 시각적으로 치환한 효과를 줍니다.
  • 포인트 03. 스토리텔링의 기승전결: 단순한 장소 나열이 아닙니다. [호기심 유발(오프닝)] -> [에너지의 폭발(중반부 몽타주)] -> [감성적 마무리(엔딩)]라는 고전적인 서사 구조를 15초 안에 완벽히 녹여냈습니다.




[💻 복사해서 바로 쓰는 ‘15초 숏츠’ 비디오 제작 프롬프트]


Prompt:

Style: High-energy cinematic Paris travel vlog, ultra-vivid colors, fast-paced editing, dynamic handheld camera work, authentic influencer aesthetic, energetic motion, natural human movement, trendy TikTok/Reel editing, cinematic realism, bright summer atmosphere.

Duration: 15 seconds

Aspect Ratio: 9:16

[00:00-00:02] EXTREME CLOSE-UP selfie shot. The cute Japanese woman rushes toward the camera laughing breathlessly with the Eiffel Tower behind her. Fast handheld motion, hair blowing wildly in the wind. She whispers directly into the mic in playful ASMR Japanese: 「ねぇ、15秒でパリ全部見せるね。」 ("Hey, I’ll show you all of Paris in 15 seconds.") Suddenly the rock music DROPS HARD with explosive transition cuts.

[00:02-00:03] FAST MULTI-SHOT MONTAGE: — Whip-pan selfie at Eiffel Tower — Quick laugh close-up — Fast spin transition beside the Seine — Sunglasses flip toward camera Heavy rock beat syncs perfectly with cuts.

[00:03-00:05] Louvre Pyramid sequence. Hyper-dynamic moving camera circles around her as she grabs the camera and runs toward the glass pyramid. Rapid cuts between: — Looking back smiling — Close-up grin — Tourists rushing past — Low-angle fashion shot — Fast handheld vlog movement Electric guitar intensifies.

[00:05-00:06] Notre-Dame Cathedral. She suddenly turns toward camera while crowds blur behind her in motion blur timelapse. She laughs loudly in Japanese: 「ヤバい、映画みたい!!」 ("This is insane, it feels like a movie!!")

[00:06-00:08] Arc de Triomphe rapid montage: — Wide cinematic shot beneath the monument — Jump spin transition — Camera tilted upward dramatically — Walking directly toward lens — Quick smile close-up — Speed-ramped city traffic around her Fast energetic editing synced to drum hits.

[00:08-00:10] Inside the Louvre beside the Mona Lisa. Chaotic fun vlog energy: — Selfie grin with Mona Lisa behind — Tourists moving rapidly — Camera flash effect — Close-up eye contact with camera — Fast snap zoom transition Rock music briefly cuts for crowd ambience and camera shutter sounds.

[00:10-00:12] Versailles Gardens cinematic sequence. Golden sunlight floods the scene. Massive drone pullback while she runs through the gardens laughing. Dress and hair flow naturally in the wind. Rapid intercuts: — Sunglasses on — Twirl — Looking over shoulder — Running toward camera Music reaches emotional uplifting chorus.

[00:12-00:13] Massive clock tower interior shot. Fast cinematic orbit around her face as sunlight beams through the giant clock glass. She whispers softly in Japanese: 「パリ、大好き。」 ("I love Paris.") Brief ASMR pause before music explodes back in.

[00:13-00:15] FINAL ULTRA-FAST PARIS MONTAGE: — Champs-Élysées walking shot — Arc de Triomphe at sunset — Eiffel Tower sparkle — Louvre smile — Notre-Dame turn-back shot — Fast spinning selfie transition — Final freeze-frame smile directly into camera Camera flash freeze ending.

Audio: High-energy female-fronted pop rock soundtrack with explosive drums, electric guitar riffs, fast transitions synced to beat drops, authentic city ambience, camera clicks, crowd energy, soft ASMR whispers, cinematic travel energy.

Negative prompts: No slow pacing, no empty locations, no stiff movement, no robotic facial expressions, no blurry face, no low-energy scenes, no static camera, no dark moody lighting, no unrealistic physics, no awkward crowd behavior.



💡디플릭의 생각: "기술이 아닌 감성의 영역"


이번 사례가 우리에게 주는 진짜 인사이트는 'AI를 어떻게 부리는가'보다 '어떤 감각을 담아낼 것인가'가 훨씬 중요해졌다는 점이에요.


단순히 "파리에서 찍은 예쁜 사진"이 아니라, 주인공이 말을 걸고, 옷깃이 바람에 날리며, 친구의 시선으로 찍은 듯한 자연스러운 장면의 설계가 핵심이죠. 숏폼 플랫폼에서 통하는 콘텐츠를 만들고 싶다면, 화려한 기술 이전에 '대사 한 줄의 힘'과 '사람 냄새 나는 연출'을 먼저 고민해 보는 건 어떨까요?


AI는 우리가 상상한 그 '무드'를 현실로 옮겨주는 가장 똑똑한 조력자가 되어줄 테니까요.



원문 출처

다양한 아티클을 확인해보세요