Complex Multi-Action Street Photography
A complex stress-test prompt for image models, requiring multiple distinct character actions, spatial separation, and text rendering in a single frame.

프롬프트
RAW iPhone 사진 품질, 전혀 편집되지 않은 스마트폰 카메라 느낌, 자연스러운 노출, 사실적인 조명.
구도:
프레임의 맨 왼쪽에는 정장을 입은 남성이 바닥에 넘어져 있습니다. 녹색 셔츠를 입은 노부인이 몸을 숙여 그를 일으켜 세우려 합니다.
맨 뒤쪽 배경, 장면의 정반대 편에는 {argument name="animal" default="검은 고양이"} 한 마리가 창가에 앉아 밖을 응시하는 모습이 선명하게 보입니다.
앞쪽 중앙 영역에서는 두 쌍의 커플(총 4명의 성인)이 자연스러운 손동작과 풍부한 바디 랭귀지를 사용하며 서로 격렬하게 논쟁하고 있습니다.
맨 오른쪽에는 세 사람이 함께 춤을 추고 있습니다. 그들의 셔츠에는 {argument name="logo" default="OpenAI 로고"}가 선명하게 표시되어 있습니다. 각 사람은 23, 44, 98이라는 큰 숫자가 적힌 모자를 쓰고 있습니다.
장면의 높은 하늘 위로는 비행기 한 대가 커다란 현수막을 매달고 날아가고 있습니다. 현수막에는 다음과 같은 문구가 선명하게 적혀 있습니다:
“{argument name="banner text" default="@WolfRiccardo — NEW GPT-IMAGE MODEL COMING SOON!"}”
프레임의 맨 앞쪽, 카메라와 가장 가까운 곳에서는 {argument name="celebrity" default="Will Smith"}가 바닥에 놓인 스파게티 접시/더미에 발이 걸려 앞으로 넘어지는 즉흥적이고 코믹한 순간이 연출됩니다.
모든 요소가 동일한 사진 안에 동시에 나타나야 합니다.
모든 인물, 동작, 사물, 숫자, 로고 및 텍스트를 명확하게 구분하고 공간적으로 분리하십시오.
기계 번역 — 실제 결과물을 만든 것은 원문 프롬프트입니다.
원문 프롬프트
RAW iPhone photo quality, completely unedited smartphone-camera look, natural exposure, realistic lighting.
COMPOSITION:
On the FAR LEFT side of the frame, a man wearing a formal business suit has fallen onto the ground. An elderly woman wearing a green shirt is bending down and trying to help him stand up.
In the FAR BACKGROUND, directly across the scene, a {argument name="animal" default="black cat"} is clearly visible sitting by a window and staring outside.
In the FRONT-MIDDLE area, TWO DIFFERENT COUPLES — four adults total — are having a heated argument with each other, using natural hand gestures and expressive body language.
On the FAR RIGHT side, THREE PEOPLE are dancing together. Their shirts clearly display the {argument name="logo" default="OpenAI logo"}. Each person wears a hat with a large readable number: 23, 44, 98
High above the scene, an airplane is flying across the sky while towing a large banner behind it. The banner clearly reads:
“{argument name="banner text" default="@WolfRiccardo — NEW GPT-IMAGE MODEL COMING SOON!"}”
At the VERY FRONT of the frame, closest to the camera, {argument name="celebrity" default="Will Smith"} is accidentally tripping over a plate/pile of spaghetti on the ground and falling forward in a spontaneous, comedic moment.
Everything must appear inside the SAME photograph at the SAME time.
Make every character, action, object, number, logo and text clearly distinguishable and spatially separated.