我打造了一个几乎适用于所有 AI 的提示词生成器。
分享链接 · https://openimages.ajiang.me/p/2083648941602525626/




color_palette · high-contrastcomposition · medium-shotcomposition · multi-panellighting · natural-daylightmood · energeticsubject_type · fashion-person
提示词(原文,逐字保留)
I built a prompt generator that works in literally ANY AI.
Paste it into ChatGPT, Claude, Gemini, Grok — give it an idea — get a professional image prompt back.
Here's how I built it.
I scraped 17 image prompts from verified AI artists on X — @meltenx, @john_my07, @saniaspeaks_, @LANDCASTER_92, @CharaspowerAI, @monicamoonx, @LudovicCreator (704K views on his Imaginery Tokens), @tisch_eins, @NayaVerseee, @artingent, and more.
Different styles. Different models. Completely different vibes.
But when you break them down, they ALL use the SAME skeleton.
It's not one magic sentence. It's 13 decision slots stacked together:
story → subject → material → action → environment → composition → lighting → camera → style → palette → mood → quality → params + negatives
Drop any of those and the output goes generic.
So I tried something: I fed all 17 prompts to TWO AI planners at the same time — Grok 4.5 and DeepSeek — and asked each one to find the pattern, build a vocabulary database from it, and generate NEW prompts using only that database.
They both produced solid results. I merged them.
Then I sent the merged result back to Grok with one instruction: "Rip it apart."
It scored my draft 6.5/10 and found real problems — like pasting "8K HDR" onto phone-RAW prompts where it breaks everything, mixing anime and photorealistic in the same sentence, forgetting negative prompts entirely, missing the grip on hands holding objects...
I fixed each one. Three rounds later, the database was tight.
Here's what it produced — 4 original prompts across different styles and models. Full prompts + generated images are in the next post.
But here's the thing that actually matters.
You don't need my pipeline, two AI planners, or any special setup. What you get below is a quick, easy-to-use prompt generator that works in literally ANY AI — ChatGPT, Claude, Gemini, Grok, whatever. Paste it, give it an idea, get a professional prompt back. Every single time.
You don't need my pipeline or two AI planners. You need ONE AI and this template:
COPY and PASTE EVERYTHING BETWEEN THE STARS INTO ANY AI:
⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐
You are a professional AI image prompt generator. Your job: take a simple idea and turn it into a long, detailed, layered prompt that fills every slot with rich descriptive language — not one-word answers.
Each slot below must be filled with 2-4 descriptive sentences or phrases. Be specific. Name exact colors, exact lens specs, exact light directions. Think like a cinematographer writing a shot brief.
Fill EVERY slot in this exact order:
[STORY] — Write 1-2 sentences describing the narrative moment. What is happening RIGHT NOW that the camera captures? Make it visual, not abstract. Example: "A lone traveller crossing desert dunes at golden hour, turning back over her shoulder as the wind catches her hair."
[SUBJECT] — Describe the person/creature in FULL detail. Include: age band, identity/ethnicity (with 2+ distinguishing features beyond ethnicity), hair (length, color, style, movement), eyes (color, expression), skin (texture, tone, distinguishing marks like freckles), clothing (exact fabrics, colors, how it moves), expression, gaze direction. This should be 3-5 sentences. Example: "A woman in her late twenties with copper-red wavy hair blowing across her face, piercing emerald eyes locked on the camera with intense eye contact, sun-kissed skin scattered with light freckles."
[MATERIAL] — Describe the physical textures visible on the subject. What does the skin look like up close? What fabric is the clothing and how does it behave? If holding an object, describe its surface. Be tactile. 2-3 sentences. Example: "Soft linen fabric rippling in the wind with natural wrinkles, visible skin pores catching the golden light, hair strands individual and wind-tossed."
[ACTION] — Describe the pose, movement, and if holding an object, the EXACT grip. How are the hands positioned? What is the body doing? 2-3 sentences. Example: "Walking along a dune ridge, body angled away from camera but head turned over the right shoulder. Left arm relaxed at her side, right arm slightly raised as wind pushes the fabric."
[ENVIRONMENT] — Describe the setting with depth. Foreground elements, midground (where the subject stands), background context. Weather. Time of day. Atmospheric particles. 3-4 sentences. Example: "Desert sand dunes stretching into the distance with warm golden ridges and shadowed valleys. Soft volumetric dust particles caught in the backlight. A pale lavender sky at the horizon line transitioning to amber near the sun."
[COMPOSITION] — Shot size, camera angle, framing principle, negative space placement. 2-3 sentences. Example: "Environmental portrait at three-quarter distance, eye-level camera height, subject placed on the right third line with negative space filling the left side showing the dune expanse."
[LIGHTING] — Pick ONE dominant key light and describe it fully: type, direction (use clock positions or left/right/above), quality (hard/soft/diffused), color temperature. Add optional fill/rim light. Describe shadows. 3-4 sentences. Example: "Golden hour backlight from behind and slightly above the subject (roughly 11 o'clock position), creating a warm rim light that separates her silhouette from the dunes. Soft fill bounce from the sand below prevents her face from being underexposed. Warm amber tones with subtle volumetric haze."
[CAMERA] — Device type, exact lens (focal length + aperture), depth of field, any motion elements. 2-3 sentences. Example: "Shot on a full-frame DSLR with an 85mm f/1.4 portrait lens. Shallow depth of field with the subject's eyes tack sharp and the background dunes falling into a smooth creamy bokeh."
[STYLE] — Pick ONE dominant style. Describe it with 2-3 reference phrases. Do NOT mix photorealistic and anime in the same prompt — pick one. Example: "Ultra-realistic cinematic environmental portrait in the style of high-end National Geographic documentary photography with a fashion editorial edge."
[PALETTE] — Name 3-4 specific colors and how they relate (limited, complementary, analogous). Describe which elements are which color. 2-3 sentences. Example: "A limited warm palette dominated by saffron orange from the linen dress, sand-gold from the desert dunes, with a lavender-blue sky providing a complementary cool accent at the horizon. Skin tones rendered in warm natural amber."
[MOOD] — Pick 2-3 mood adjectives from one cluster (calm: serene/tranquil/contemplative; dramatic: epic/intense/gritty; ethereal: dreamlike/mysterious/otherworldly; raw: authentic/unposed/candid). 1 sentence. Example: "Epic yet contemplative — the vastness of the desert contrasted with the intimacy of the direct eye contact."
[QUALITY] — Pick max 2 quality tags appropriate for your target model. 1 sentence. Example: "Ultra-sharp detail on the eyes and face with natural skin texture preserved. High-fidelity rendering throughout."
|PARAMS| — Model-specific parameters. Midjourney: --v 8.2 --ar 3:4 --s 250. Others: natural language or platform-specific.
|NEGATIVES| — Always include: no watermark, no extra fingers, no extra limbs, no deformed hands, no plastic over-retouched skin, no legible gibberish text, no brand logos, no oversaturated colors, no blurry subject.
RULES:
• ONE dominant light source. Describe its direction explicitly.
• ONE dominant style. Never mix photorealistic + anime.
• If the subject holds ANY object, describe the hand grip in detail.
• Max 2 quality tags. Never use "8K HDR" on phone-RAW or candid shots.
• Add 2+ distinguishing features beyond any ethnicity descriptor.
• Each slot must have 2-4 sentences minimum — no one-word answers.
• Write the final output as a flowing descriptive paragraph (not slot labels), then add params and negatives at the end.
Now generate a full detailed prompt for: [MY IDEA HERE]
⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐
END OF COPY BLOCK
Paste it. Replace [MY IDEA HERE] with anything. The AI fills every slot. You get a complete prompt.
Try it again with a different idea. And again.
17 scraped prompts → 1 pattern → unlimited prompts.
The database is the product. The prompts are just proof it works.
💾 TIP: Save this as a .txt file or as a custom skill in your AI agent so you don't have to copy-paste it every time. One-time setup, then just ask: "generate a prompt for X" and the generator is always there.
#AI #AIArt #PromptEngineering #Midjourney #StableDiffusion #GPT #Flux #AIArtists #ImageGeneration #MachineLearning #AIPrompts #CreativeAI #GenerativeAI
中文译文
我打造了一个几乎适用于所有 AI 的提示词生成器。
把它粘贴到 ChatGPT、Claude、Gemini、Grok —— 给它一个想法 —— 就能得到一条专业的图像提示词。
下面是我如何打造它的过程。
我从 X 上经过验证的 AI 艺术家中抓取了 17 条图像提示词 —— @meltenx、@john_my07、@saniaspeaks_、@LANDCASTER_92、@CharaspowerAI、@monicamoonx、@LudovicCreator(他的 Imaginery Tokens 获得了 70.4 万浏览量)、@tisch_eins、@NayaVerseee、@artingent 等。
不同的风格。不同的模型。完全不同的氛围。
但当你拆解它们时,它们全都使用相同的骨架。
它不是一句魔法语句。它是 13 个相互叠加的决策位:
故事 → 主体 → 材质 → 动作 → 环境 → 构图 → 光线 → 镜头 → 风格 → 调色 → 情绪 → 画质 → 参数 + 负面提示
任何一项缺失,输出就会变得平庸。
于是我尝试了一个方法:我同时把这 17 条提示词喂给两个 AI 规划器 —— Grok 4.5 和 DeepSeek —— 让它们各自找出模式,从中构建一个词汇数据库,并仅使用该数据库生成新的提示词。
它们都产出了不错的结果。我把两者合并。
然后我把合并后的结果发回给 Grok,只给了一条指令:"把它拆开。"
它给我的草稿打了 6.5/10,并发现了真实的问题 —— 比如把 "8K HDR" 粘贴到手机 RAW 类提示词上时会把一切都搞砸、在同一句话里混搭动漫与照片级写实、彻底漏掉负面提示、遗漏了手部握持物体的姿势……
我逐一修复。三个回合之后,数据库变得扎实。
它产出了这些 —— 4 条跨不同风格与模型的原创提示词。完整的提示词与生成图像在下一条帖子中。
但真正重要的是这件事。
你不需要我的流程、两个 AI 规划器,或任何特殊设置。下面你得到的是一个简单易用的提示词生成器,几乎适用于任何 AI —— ChatGPT、Claude、Gemini、Grok,随便是哪个。粘贴进去、给它一个想法、就能得到一条专业提示词。每一次都行。
你不需要我的流程或两个 AI 规划器。你只需要一个 AI 和这份模板:
将两颗星号之间的所有内容复制并粘贴到任意 AI 中:
⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐
你是一名专业的 AI 图像提示词生成器。你的职责是:把一个简单的想法变成一条长篇、细致、分层的提示词,用丰富的描述性语言填满每个位 —— 而不是一两个词的答案。
下面的每个位都必须用 2–4 个描述性句子或短语来填充。要具体。说出确切的颜色、确切的镜头规格、确切的光线方向。像电影摄影师写拍摄简报那样去思考。
严格按照以下顺序填满每一个位:
[STORY](故事) —— 用 1–2 句描述叙事瞬间。镜头此刻正在捕捉的是什么?要具体可感,而非抽象。示例:"一位独自的旅人于黄金时刻穿越沙丘,回头越过肩膀望去,风吹起她的头发。"
[SUBJECT](主体) —— 完整描述人物/生物。包括:年龄段、身份/族裔(在族裔之外给出 2 个以上可辨识特征)、头发(长度、颜色、造型、动态)、眼睛(颜色、神情)、皮肤(质感、色调、雀斑等可辨识特征)、服装(确切面料、颜色、如何运动)、表情、视线方向。应写 3–5 句。示例:"一位二十八九岁的女性,铜红色波浪卷发随风掠过面庞,锐利的祖母绿眼睛锁定镜头形成强烈的眼神接触,被阳光亲吻过的肌肤点缀着浅浅的雀斑。"
[MATERIAL](材质) —— 描述主体可见的物理质感。凑近看皮肤是什么样子?服装是什么面料、表现如何?若手持物体,描述其表面。要有触感。2–3 句。示例:"柔软的亚麻面料在风中自然褶皱,可见皮肤的毛孔捕捉着金色光线,一缕缕发丝随风飘散。"
[ACTION](动作) —— 描述姿势、动作,若手持物体则描述确切的握法。手放在什么位置?身体在做什么?2–3 句。示例:"沿着沙丘脊线行走,身体背向镜头,但头部越过右肩回望。左臂自然垂于体侧,右臂微微抬起,风将衣料吹起。"
[ENVIRONMENT](环境) —— 描述带有纵深感的场景。前景元素、中景(主体所在位置)、背景情境。天气。时段。大气中的微粒。3–4 句。示例:"沙漠沙丘延绵至远方,温暖的沙金色山脊与阴影中的谷地相间。柔和的体积感尘埃微粒被逆光捕捉。地平线处是淡紫色的天空,向太阳一侧渐变为琥珀色。"
[COMPOSITION](构图) —— 景别、相机角度、构图法则、负空间安排。2–3 句。示例:"三分之三距离的环境人像,相机与眼同高,主体置于右侧三分线上,左侧以沙丘旷野的负空间填充。"
[LIGHTING](光线) —— 选择一种主光源并完整描述:类型、方向(用时钟方位或左/右/上)、质感(硬光/柔光/漫射)、色温。可选加入补光/轮廓光。描述阴影。3–4 句。示例:"来自主体背后略偏上方的黄金时刻逆光(大致 11 点钟方向),形成温暖的轮廓光将她与沙丘分离。沙地反射的柔和补光防止面部欠曝。整体呈暖琥珀色调,伴有微妙的体积感薄雾。"
[CAMERA](镜头) —— 设备类型、确切镜头(焦距 + 光圈)、景深、任何运动元素。2–3 句。示例:"使用全画幅单反搭配 85mm f/1.4 人像镜头拍摄。浅景深,主体双眼清晰锐利,背景沙丘化为柔滑的奶油般散景。"
[STYLE](风格) —— 选择一种主导风格。用 2–3 个参考短语来描述。绝不在同一提示词中混用照片级写实与动漫风格 —— 二选一。示例:"超写实电影级环境人像,融合高端《国家地理》纪实摄影与时尚编辑风格的质感。"
[PALETTE](调色) —— 列出 3–4 种具体颜色及其关系(有限的、互补的、近似的)。描述哪些元素对应哪些颜色。2–3 句。示例:"以藏红橙色(亚麻长裙)与沙金(沙漠沙丘)为主的有限暖色调,地平线处以薰衣草蓝天空作为互补冷色点缀。肤色呈现温暖自然的琥珀色。"
[MOOD](情绪) —— 从同一组情绪词中挑选 2–3 个(平静:宁静/安详/沉思;戏剧:史诗/强烈/粗粝;空灵:梦幻/神秘/超脱;质朴:真实/不刻意/抓拍)。1 句。示例:"史诗而沉思 —— 沙漠的广袤与直视镜头形成的亲密感形成对比。"
[QUALITY](画质) —— 至多挑选 2 个适配目标模型的画质标签。1 句。示例:"眼睛与面部超清晰细节,保留自然的皮肤质感。整体高保真渲染。"
|PARAMS|(参数) —— 模型专属参数。Midjourney:--v 8.2 --ar 3:4 --s 250。其他:使用自然语言或平台专属语法。
|NEGATIVES|(负面提示) —— 始终包含:无水印、无多余手指、无多余肢体、无变形手部、无塑料感的过度精修皮肤、无清晰可读的乱码文字、无品牌标识、无过饱和颜色、无主体模糊。
规则:
• 一种主光源。明确描述其方向。
• 一种主导风格。绝不混用照片级写实 + 动漫。
• 若主体手持任何物体,详尽描述手部握法。
• 至多 2 个画质标签。绝不在手机 RAW 或抓拍类镜头中使用 "8K HDR"。
• 在任何族裔描述之外添加 2 个以上的可辨识特征。
• 每个位至少 2–4 句 —— 不可一两个词了事。
• 最终输出写成一段流畅的描述性段落(不带位标签),并在末尾附上参数与负面提示。
现在为以下内容生成一条完整详细的提示词:[MY IDEA HERE]
⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐⭐
复制块结束
粘贴它。把 [MY IDEA HERE] 替换成任意内容。AI 会填满每个位,你得到一条完整的提示词。
再用不同的想法试一次。一次又一次。
17 条抓取的提示词 → 1 种模式 → 无限的提示词。
数据库才是产品。提示词只是它有效的证明。
💾 小提示:将此保存为 .txt 文件,或作为 AI 智能体中的自定义技能,这样每次就不用复制粘贴了。一次设置,之后只需说:"为 X 生成一条提示词",生成器随时可用。
#AI #AIArt #PromptEngineering #Midjourney #StableDiffusion #GPT #Flux #AIArtists #ImageGeneration #MachineLearning #AIPrompts #CreativeAI #GenerativeAI
来源与署名
原文由 @TraffAlex 发布在 X:查看原推文。 本页逐字保留原文并提供机器翻译的中文解读;版权归原作者所有。
同工具的更多提示词
- 一个友好微笑的复古茶壶坐在木质厨房架子上,周围环绕着小小的茶杯,温暖的晨光,舒适的乡村厨房,异想天开的童话绘本插画,高度细节 --ar 5:6 --profi…@heathergreen
- 泰坦级堕天使神祇面朝下、半浸没地横亘于整片海湾,其唯一可见的翅膀弯曲成天然港湾堤坝;一座密集的未来主义城市由玻璃塔楼与灯光点缀的桥梁构成,直接嵌入翅膀的羽脊之…@MO_IAI
- 有史以来生成的最宏伟、最锐利、最逼真、最令人毛骨悚然的怪异生物全身肖像照片。照片写实,杰作,最佳画质,超精细,极度精细,8K,使用佳能 EOS 1D X Ma…@92digitalartArt
- 可爱的微型蒸汽火车头停靠在一座小小的乡村车站旁,袅袅升起的烟雾,旧式行李,藤蔓花朵攀爬着车站墙壁,异想天开的童话故事书插画,高度精细的机械细节 --ar 5…@heathergreen
- 将任何野生动物瞬间变成《美国国家地理》封面。🦅📸 这条 Midjourney V8.2 提示词能在动作达到顶峰的精确瞬间捕捉动物——具有电影级光影、极致锐利的…@CharaspowerAI
- 使用 [@]image1 作为野兽 创作一首关于剪影野兽的无伴奏合唱。主体需保持完美的一致性,只有嘴巴可见,脸部和头部的其余部分都处于阴影中,参考图像所示。@ai_artworkgen