
sheet-p3.png
01-kitchen.png
모델 gpt-image-2.5-sunburst(생산과 같은 모델, OpenAI 직접) · 크기 1008×1792 ·
품질 high — 현행 생산은 medium이고, 형이 high 로 바꾸라 하신 값입니다(12 K1) ·
참조 2장(시트 → 배경, 생산과 같은 순서) · 판마다 4장 · 만든 장수 44장(실패 0장) ·
실제 비용 $2.32
A 판은 생산 buildHostMasterPlatePrompt 전문 그대로입니다. 단 생산이 프롬프트에 박는
Character={요청서 JSON}만 최소 스펙으로 대신했습니다 —
네 판이 모두 같은 값이라 판끼리 비교는 됩니다. 다만 「생산과 완전히 같다」고는 못 합니다.
E 판의 빛 지시는 배경 이름을 박지 않고 「두 번째 이미지를 읽어라」로 썼습니다 —
배경이 바뀌어도 그대로 쓸 수 있어야 스펙이 되기 때문입니다.
D-1 은 처음 보낼 때 HTTP 520(1.3초)으로 끊겨 한 번 다시 보냈습니다.
요청이 안 간 것이지 결과가 나빠서 다시 뽑은 것이 아닙니다 — 나머지 15장은 모두 첫 번째 것입니다.
보실 곳 — 빛이 한쪽 어깨·뺨에 얹혔나 · 반대쪽 목덜미와 어깨가 그늘로 떨어졌나.
이 방의 창은 왼쪽에 있습니다. 위아래 한 쌍이 같은 번호입니다.
위 = 앞 판(발밑 그림자까지) · 아래 = 빛 방향 덩어리 하나만 더한 것.
그 밖은 프롬프트·배경·설정이 한 글자도 같고 발밑 그림자 문장도 그대로 있습니다.
🔴 숫자를 안 붙였습니다.








보실 곳 — ① 신발 밑창 닿는 자리에 짙은 선이 있나 ② 바닥의 빛 무늬가 인물 앞에서 끊기나.
위아래 한 쌍이 같은 번호입니다. 위 = 지금 쓰는 문장 · 아래 = 접지 문장 하나만 더한 것.
그 밖은 프롬프트·배경·설정이 한 글자도 같습니다.
🔴 숫자를 안 붙였습니다. 접지 그림자를 숫자로 걸어 보려고 자를 둘 만들었는데 둘 다 못 썼습니다 —
방이 내는 빛 무늬(밝기 30~90단계)가 인물 그림자(5~10단계)를 덮습니다. 이 자리는 눈이 자입니다.








흐린 배경 위에서는 인물이 뜹니다(실측: 01-kitchen 1.85 ↔ 04-dining 0.78).
그래서 흐린 3장만 「같은 방 그대로, 초점만 전부 또렷하게」로 다시 만들었습니다.
방·가구·창·빛은 바꾸지 말라고 못 박았습니다.
검수 자 — 인물이 설 자리 옆(어깨 높이) 잔결이 15 를 넘나. 셋 다 넘었습니다
(기준 04-dining = 21.89).
⚠️ 픽셀 단위로 같지는 않습니다 — 02-living 은 구도가 살짝 옮겨졌고
06-bedroom 은 커튼 주름이 다릅니다. 같은 방·같은 가구·같은 빛까지입니다.
🔴 원본은 안 건드렸습니다. 바꿔 넣을지는 형·main 판단입니다.






라플라시안 표준편차 = 그 조각의 잔결 양. 인물÷배경이 1 에 가까울수록 한 카메라로 찍힌 것입니다. 판마다 4장의 평균이고, 조각 자리는 여덟 판에 똑같이 댔습니다.
| 판 | 인물 | 배경 | 바닥 | 인물÷배경 |
|---|---|---|---|---|
| A | 13.5 | 5.0 | 7.7 | 2.74 ±0.40 |
| C | 13.6 | 5.7 | 6.8 | 2.38 ±0.10 |
| D | 13.9 | 5.6 | 7.4 | 2.47 ±0.14 |
| E | 13.2 | 5.6 | 8.1 | 2.37 ±0.06 |
| F | 12.9 | 6.0 | 8.8 | 2.16 ±0.21 |
| G | 13.1 | 6.5 | 9.3 | 2.02 ±0.08 |
| H | 12.9 | 6.0 | 9.0 | 2.17 ±0.21 |
| I | 11.6 | 6.3 | 8.6 | 1.85 ±0.12 |
| J | 12.5 | 16.1 | 10.4 | 0.78 ±0.04 |
| K | 13.1 | 15.5 | 10.5 | 0.85 ±0.07 |
| L | 11.2 | 5.0 | 8.8 | 2.25 ±0.12 |
🔴 앞서 제가 「여덟 판이 전부 같다, 안 움직인다」고 올린 것은 잘못이었습니다.
그때는 배경 조각을 왼쪽 위 한 곳만 썼고 판마다 1장씩만 봤습니다. 좌우 양쪽을 같은 높이에서 재고
판마다 4장을 평균 내니 A → F → G → I 로 내려갑니다.
🔸 그리고 내려간 이유가 인물만 흐려진 것이 아닙니다 — 인물은 내려가고(13.5→11.6)
배경과 바닥은 올라갔습니다(5.0→6.3 · 7.7→8.6). 두 쪽이 서로에게 다가갔습니다.
⚠️ 그래도 1 에는 한참 멉니다. 자가 말하는 것은 「방향이 맞다」까지고, 「됐다」가 아닙니다. 판정은 눈으로 하십니다.
위 넷 J = high · 아래 넷 K = medium. 그 밖은 한 글자도 같습니다.
얼굴 조각 잔결 — J 14.5 · K 15.6. medium 이 오히려 높습니다. 뭉개지지 않았습니다.
값 — 한 장 만드는 출력 토큰 1234 → 320, 걸린 시간 26초 → 17초, 값 $0.058 → $0.031.
⚠️ 다만 K-3 한 장이 구도가 더 당겨졌습니다(머리끝 J ±1.0%p · K ±1.8%p). 4장이라 단정은 못 합니다.









발·바닥
발·바닥
발·바닥
발·바닥더한 세 문장 (12 K8)
Match lighting, shadows and color temperature to the second image.
Use the same key light direction as the second image.
Render photorealistic contact shadows where both shoes meet the floor, and a soft cast shadow falling in the same direction and with the same softness as the shadows already present on that floor.

발·바닥
발·바닥
발·바닥
발·바닥











더한 여섯 덩어리
① 광원이 어디서 오나 🟢 공식 어휘
Read the light in the second image and relight the person with it: identify the dominant light source in that room and its direction, height and hardness, and light the person from that same direction, matching the angle and softness of the shadows already visible on the floor and walls.
② 얼굴과 어깨의 명암 🔴 형 지적을 옮긴 것
Put that light on the head and shoulders, not only on the floor: the side of the face, neck and shoulder that faces the light source is the brightest part of the person, the opposite side of the face and shoulder falls into soft shadow, and the transition between them is as gradual as the shading on the surfaces behind the person.
③ 색온도가 피부·옷에 🟢 공식 어휘
Match color temperature on the person to the second image: the lit side of the skin and the knitwear take the same warmth as the lit surfaces of that room, the shaded side takes the same cooler neutral as its shadows, and the person shares one white balance with the background instead of reading as a separately lit cutout.
④ 반사광(bounce) 🔴 형 지적을 옮긴 것
Add the light the room throws back onto the person: warm bounce from the floor lifting the underside of the chin, the neck and the lower arms, and a softer neutral bounce from the bright surfaces on the shaded side, so the person's edges pick up colour from the surfaces immediately around them.
⑤ 채도·대비·질감 🟢 공식 어휘
Match exposure, contrast, saturation and grain between the person and the room: the person is no sharper, no more saturated and no more contrasty than the background, with the same highlight roll-off and the same subtle film grain across the whole frame.
⑥ 하지 마라 🟢 공식 원문
The result should look like a real photograph someone could have taken in that room, not an overly enhanced or cinematic movie-poster image. Avoid studio lighting, flat even illumination, cinematic lighting, dramatic color grading, glamorization or heavy retouching, and avoid a person who is brighter, sharper or cleaner than the room they stand in.
🔴 ②④는 OpenAI 공식 예제에 없습니다. 형 지적을 문장으로 옮긴 것입니다. ①③⑤⑥은 공식 가이드의 어휘·원문에서 왔습니다.












🔴 F 는 D·E 와 한 변수 차이가 아닙니다. 주어·참조 역할·접촉·매체·질감이 한꺼번에 바뀝니다. 「접근 자체가 맞나」를 보는 판이라 그렇게 짰습니다. 이기면 그다음에 무엇이 이겼는지 쪼갭니다.
🔴 공식 문서에는 「인물 참조 + 배경 사진 참조 → 새로 찍은 사진」 예시가 없습니다. 미검증 영역입니다. 공식은 「참조 1장 + 글로 쓴 배경」이거나 「참조 2장 + 옮겨심기」 둘뿐입니다.
왜 바꿨나 — 같은 공식 문서가 문법을 둘로 갈라 놨습니다. Combine references는 "Place X into the setting of Y"이고 목표가 "the composite looks naturally captured"입니다. Insert a person into a scene은 "Generate a scene where this person is…"이고 목표가 "a real photograph someone could have taken"입니다. A~E 는 앞쪽 문법이었습니다.
일곱 덩어리
① 주어가 「사진」이다 🟠 레퍼런스 구조
A candid photograph taken in the room shown in the second image, showing the woman from the first image standing on its wooden floor.
② 첫 참조 = 정체성만 🟢 공식 + 🟡 역할 한정
Use the first image strictly as the identity reference: preserve her facial structure, hair, skin texture, natural asymmetry, body proportions, and her complete outfit and shoes. Do not copy its background or its multi-view layout.
③ 둘째 참조 = 장소로만 🟢 공식 + 🟡 역할 한정
Use the second image as the location where this photograph was taken: the same room, the same furniture and surfaces, the same window and the same daylight. Do not treat it as a flat backdrop placed behind her.
④ 접촉·하중 — 우리에게 없던 축 🟠 레퍼런스 구조
She stands centered on that floor facing the camera at eye level, both feet flat and parallel at shoulder width, her weight carried evenly through both soles so the shoes press into the floor, knees relaxed, arms hanging naturally with both hands empty.
⑤ 환경광과 표면 반사를 한 문장에 🟢 공식 어휘 + 🟠 레퍼런스 구조
The daylight coming through that room's window falls across her the same way it falls across the floor and the counters, creating the same direction and softness of shadow on her face, shoulders and clothing, and the pale floor and bright surfaces throw warm bounce back onto the underside of her chin, her neck and her forearms.
⑥ 매체·질감 🟢 공식 원문 어휘
Shot like a 35mm photograph at eye level using a 50mm lens, the whole frame at one exposure with natural color balance, subtle film grain, and the same softness and detail on her as on the room behind her.
⑦ 하지 마라 🟢 공식 원문
The image should look like a real photograph someone could have taken in that room, honest and unposed, with real skin texture and worn materials. No glamorization, no heavy retouching. Avoid studio lighting, flat even illumination, cinematic lighting, dramatic color grading, or a person who is brighter, sharper or cleaner than the room she stands in. One person, one continuous 9:16 image. No text, logo, sign, or badge.












🔴 G 는 F 에 셋을 더한 것입니다. F 문장은 한 글자도 안 바꿨습니다.
① 찍는 사람의 자리 — taken by someone standing on the same floor about three metres in front of her, holding the camera at their own eye level
② 🔴 바닥의 빛이 인물 위를 지나간다 — the patches of window light already lying across that floor continue across her shoes, her trousers and her torso … without a break at her outline
③ 불완전함 — visible pores and fine skin texture, a few stray hairs, soft creases and pilling in the knitwear, slightly uneven skin tone, and the same faint sensor grain over the person as over the room
🔸 ①②는 레퍼런스 구조를 옮긴 것, ③은 OpenAI 공식 §4.3 어휘입니다. 셋을 한꺼번에 더했습니다 — 이기면 그다음에 쪼갭니다.

다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나

🔴 G 와 낱말 하나 차이입니다. Shot like a 35mm photograph → Shot like a 35mm editorial photograph. 그 외는 한 글자도 다르지 않습니다(글자 수 차이가 정확히 " editorial" 인 것을 확인하고 돌렸습니다).
형 지시 — "editorial 이라는 키워드가 들어가면 조금 더 자연스러운가". 톤을 바꾼 것이 아니라 낱말 하나가 무엇을 하는지만 보는 판입니다.

다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나

🔴 I 는 G 에서 한 덩어리만 갈아 끼운 것입니다. 나머지는 한 글자도 안 바꿨습니다.
뺀 것 — visible pores and fine skin texture, a few stray hairs, soft creases and pilling…
넣은 것 — The camera missed focus very slightly, the way a real one does: a soft, slightly missed focus over the whole frame and a faint chromatic aberration at the high-contrast edges. Focus falls on her near eye; her far eye, her ears and the back of her shoulders soften away from it, and no part of her clothing is rendered sharper than the counters standing beside her at the same distance.
🔸 왜 바꿨나 — 초점을 「얼마나 흐리게」가 아니라 「무엇이 일어났나」로 쓴 것입니다. GPT Image 대상 실측 글들이 soft / slightly missed focus·chromatic aberration은 먹고 f/1.8 같은 수치는 안 먹는다고 적습니다. visible pores 는 뺐습니다 — AI 가 그리는 모공은 균일해서 오히려 가짜 신호를 더한다는 근거가 나왔습니다.

다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나

🔴 J 는 I 와 프롬프트가 한 글자도 다르지 않습니다. 배경 참조 한 장만 바꿨습니다.01-kitchen.png(배경이 아웃포커스 · 잔결 7.98) → 04-dining.png(같은 실내 · 같은 창빛인데 또렷 · 19.14)
🔸 main 요청은 「F 또는 G 를 바탕으로」였는데 I 를 바탕으로 썼습니다 — I 가 지금 제일 낮아서, 배경만 바꾼 한 변수 비교가 그대로 성립하고 결과도 바로 쓸 수 있기 때문입니다.
⚠️ 프롬프트를 한 글자도 안 바꾸느라 the counters standing beside her(조리대)가 그대로 남아 있습니다. 이 방에는 조리대가 없습니다 — 그래도 한 변수만 바꾸는 쪽을 택했습니다.

다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나

🔴 K 는 J 와 프롬프트도 배경도 한 글자 안 바꿨습니다. quality 하나만 high → medium 입니다.
왜 — OpenAI 공식 인물합성 예제 4개가 전부 medium 이고, 현행 생산 기본값도 medium 입니다(high 는 12 K1 형 지시).
🔴 볼 것은 비율이 아니라 얼굴입니다 — 아래 「얼굴」 칸에 J 와 나란히 확대해 두었습니다.

다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나

🔴 L 은 K 와 배경만 다릅니다. 04-dining → 03-studio.
🔴🔴 그런데 이 판에서는 「인물÷배경」이 뜻을 잃습니다. 03-studio 는 흐린 방이 아니라 결이 아예 없는 흰 사이클로라마입니다(잔결 2.35 — 흐려서가 아니라 그릴 것이 없어서). 분모가 거의 0 이면 비율은 무엇을 해도 높습니다. 2.25 는 「실패」가 아니라 「이 자가 안 맞는 자리」입니다.
🟢 눈으로는 안 떠 보입니다 — 발밑에 그림자가 붙고 창빛 조각이 바닥에 깔려 인물 위로 이어집니다.
⚠️ 접지 그림자를 숫자로 걸어 보려 했는데 자가 흔들려서(±16) 근거로 못 씁니다. 이 칸의 판정은 눈으로만 드립니다.

다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나


다리·바닥 — 바닥의 빛이 바지 위로 이어지나
