我发出的编辑提示非常简单:
分享链接 · https://openimages.ajiang.me/p/2074499009633239461/




camera_lens · deep-focuscomposition · multi-panelcomposition · wide-shotlighting · natural-daylightquality_flags · embedded-textquality_flags · multi-panel-layout
提示词(原文,逐字保留)
The edit prompt I asked was very simple:
"Replace the kid on the cliff with Kai (character sheet reference image) and keep same acting posture as he is looking towards the valley. Fix some of the small people scale who look bigger than their houses down in the valley."
Here are the results.
Image 1:
Original MidJourney render.
Image 2:
NanoBanana Flash.
Image 3:
NanoBanana Pro.
Same edit. Same prompt. Different model.
👉Flash is great for speed.
But if you're preparing keyframes for Veo, Seedance, Kling, Higgsfield, or any cinematic workflow where every frame matters...
...don't use Flash.
👉The Pro model preserves noticeably more of the original composition, lighting, textures, and material detail.
That said...
Even Pro still degrades a surprising amount of what MidJourney created. 😅
- Fine textures disappear.
- Materials become softer.
- Small design decisions get simplified.
- The image slowly drifts away from the original art direction.
👉GPT Image 2...well see for yourself...I'm not impressed.
Right now, I honestly don't think there's a model that can make significant edits while perfectly preserving everything that made the original image great.
It's currently the weakest link in the AI filmmaking pipeline.
Hopefully not for long.
中文译文
我发出的编辑提示非常简单:
"把悬崖上的小孩替换成 Kai(角色设定参考图),并保持他望向山谷时的同一表演姿态。修正山谷中一些看起来比自家房子还大的小人比例问题。"
以下是结果。
图像 1:
原始 MidJourney 渲染图。
图像 2:
NanoBanana Flash。
图像 3:
NanoBanana Pro。
同样的编辑。同样的提示。不同的模型。
👉Flash 胜在速度。
但如果你正在为 Veo、Seedance、Kling、Higgsfield 或任何电影级工作流准备关键帧,每一帧都至关重要……
……那就不要用 Flash。
👉Pro 模型在保留原始构图、光照、纹理和材质细节方面明显更出色。
话虽如此……
即便是 Pro,依然会出乎意料地损失大量 MidJourney 原本的细节。😅
- 精细纹理消失。
- 材质变得柔化。
- 一些微小的设计决策被简化。
- 图像慢慢偏离原本的艺术方向。
👉GPT Image 2……你自己看吧……我不满意。
老实说,目前我不认为有哪个模型能在进行重大编辑的同时,完美保留让原图出色的所有元素。
它目前是 AI 电影制作流程中最薄弱的一环。
希望不会持续太久。
来源与署名
原文由 @AndrewFromDO 发布在 X:查看原推文。 本页逐字保留原文并提供机器翻译的中文解读;版权归原作者所有。
同工具的更多提示词
- 凌乱的拼贴画,拥挤的葡萄园场景,照片级真实感的背景,色彩过多,酒瓶焦点不突出,千篇一律的葡萄酒广告,厚重的画框,幼稚的纸质工艺,细节不足,标签模糊不清,现代极…@ou_zhen599
- GPT image 2 on the Chatgpt 提示词 👇 以上传的照片作为唯一的面部参考,创建一张超现实主义的电影感冬季肖像,在保留该人物 100%…@iamsofiaijaz
- 随机涂鸦,装饰性边框,对称线条艺术边框,照片级写实风景背景,拥挤的插图,颜色过多,酒瓶焦点弱,通用的葡萄酒广告,幼稚的卡通风格,细节低,标签模糊,文字杂乱,扁…@ou_zhen599
- 这个场景看起来莫名熟悉。有人知道这是哪位老师吗? GPT- image 2 提示词 👇 一张用 35mm 胶片拍摄的抓拍照片,照片中一位二十出头的漂亮年轻东亚…@johnAGI168
- GPT Image 2 { "prompt": { "title": "Agency-Grade Brand Identity System Poster(代…@Nas_tech_AI
- 软袋包装、纸盒、塑料罐、玻璃罐、金属罐、大型家庭装、写实背景场景、暗淡的色彩、弱动态、杂乱无章的布局、低饱和度、通用零食广告、模糊的产品、低廉的卡通风格、浑浊…@ou_zhen599