Image Prompt Director
Your idea + brand becomes a precise, model-aware image prompt – correct syntax per generator, consistent series, a rights check. It doesn't generate.
SKILL.md in the open agentskills.io standard — works directly in Claude, ChatGPT/Codex, Cursor, Copilot & more.
Image Prompt Director is an AI skill that turns your idea + brand into a precise, model-aware image prompt. It asks your target model (Midjourney, GPT Image/DALL·E, Gemini, FLUX, Stable Diffusion, Ideogram, Firefly) and emits the correct syntax instead of flowery adjectives, fills every pillar, builds a style/character bible for consistent series, and flags rights/labeling. It doesn't generate — verify-live.
What this skill does
- **Model + job first:** asks the generator/version + use-case, switches the syntax dialect and aspect ratio
- **Fills every pillar:** subject, action, environment, composition, camera, lighting, color, style, quality — and translates vibes into concrete descriptors
- **Consistency kit for series:** fixed style block + character bible + seed/reference discipline (`--oref`/`--sref`/Kontext/Nano Banana) — the #1 pain
- **Negative prompts done right:** real only in the SD family; elsewhere phrase exclusions positively, never over-stuff
- **Text & logos routed right:** legible text to Ideogram/GPT Image; exact logos/hex composited in a design tool
- **Aspect ratio + safe zones** per platform; **troubleshooting** (hands, text, bleeding, "AI look") with the one fix
- **Rights check:** license per model, copyright, trademark/likeness, labeling → hand off to KI-Transparenz
- **Two Node scripts:** prompt-forge (builds a model-correct prompt) and prompt-lint (QA against vagueness/dialect-mismatch/living artists)
Description
Image Prompt Director is a skill that turns an idea and your brand into a precise, **model-aware** image prompt — one you paste into your generator (Midjourney, GPT Image/DALL·E, Gemini/Nano Banana, FLUX, Stable Diffusion, Ideogram, Firefly). It's the art director, not the generator: it writes the best possible prompt (plus the negative prompt, parameters and a reusable style kit) and explains every choice. Founding rule: **model-aware, not flowery.** Prompt syntax doesn't port — `(word:1.3)` is Stable-Diffusion-only, Midjourney uses `--no`/`--sref`, GPT/Gemini/FLUX take plain natural language; generic bots that bolt adjectives onto an idea are why AI images all look the same. So it asks the model **and version** first, then emits the right dialect. It fills **every pillar** (subject → action → environment → composition → camera → lighting → color → style → quality — a missing pillar means generic slop), translates vibes into physical descriptors ("premium" → matte metal, soft studio light, negative space), directs the camera (shot/angle/focal length/depth of field) and **names the light** (golden hour, Rembrandt). For **series** it solves the #1 pain — consistency — with a fixed style block plus a **character bible**, seed discipline, and the right reference technique per model (`--oref`/`--sref`, FLUX Kontext, Gemini/Nano Banana, IP-Adapter). It handles **negative prompts** honestly (real only in the SD family; elsewhere phrase positively), **routes legible text** to the models that can render it (Ideogram/GPT Image/Recraft, not Midjourney), sets the **aspect ratio** to the destination (9:16 story, 4:5 feed, 16:9 thumbnail) and debugs **one variable at a time** (hands, faces, garbled text, prompt bleeding, the "AI look"). And it flags **rights and labeling** honestly: commercial license differs per model, a prompt alone doesn't confer copyright, don't clone living artists, composite exact logos/colors — and published AI images may need a label (hand off to the **KI-Transparenz** skill). All verify-live, because the models change monthly.
Examples
Tested with Claude Code
What it does not do
- **Doesn't generate** and doesn't upscale/inpaint or check the image — it writes the prompt; you generate in your own tool
- **Guarantees no** exact logos, exact text, exact brand hex or pixel-exact character identity — gives the reliable path (compositing/reference) instead
- **Not legal advice** — flags license/copyright/trademark/labeling; verify the model's live terms; labeling → KI-Transparenz
- **Verify-live** — models change monthly; it dates its claims rather than hardcoding "the latest is vX"
Compatibility & tech
- Tested (internal)
- 3 scenarios
- Recommended runtime
- Claude Opus or Sonnet; with web access to verify current model versions/parameters and "best model for X" live. The two scripts need Node (no extra packages); without Node the skill produces the prompt and lint by hand from the references.
- Modes
- Write an image prompt (model-aware) · Consistency kit for series (style/character bible) · Troubleshooting / iteration · Rights & labeling check
- Inputs
- idea/subject + target model + use-case · brand (colors, mood, style) + one-off or series · an existing prompt (to improve) or a failed result (to fix)
- Output format
- A Markdown deliverable per mode — the model-correct prompt (plus negative prompt/parameters) in clean copy-paste blocks, optionally the reusable style/character bible, plus a brief why, a rights flag, and one next iteration.
- Subcategory
- AI image generation, prompt direction & image consistency
- License
- Proprietär
Security profile
Runs entirely on your machine with your own AI — no external runtime, no running costs.
Contains executable code or runs actions/tools — take a quick look before using.
Accesses external sources / the network in normal use (e.g. live pages, search/data APIs).
What you get
- image-prompt-director-1.0.0/10 files
- .claude-plugin/marketplace.json
plugins/image-prompt-director/skills/image-prompt-director/10 files
- SKILL.md
- manifest.json
references/6 files
- image-prompt-director-anatomy.md
- image-prompt-director-consistency.md
- image-prompt-director-models.md
- image-prompt-director-rights.md
- image-prompt-director-troubleshooting.md
- image-prompt-director-usecases.md
scripts/2 files
- prompt-forge.mjs
- prompt-lint.mjs
- .agents/skills/image-prompt-director/→ universal — same content (Codex, Cursor, Copilot, Gemini, Windsurf, Cline)
- LICENSE.txt
Installation
Also works as a chat prompt
No AI tool? Paste it into Claude, ChatGPT or Gemini and use the method right away.
You get the full method. Only 2 helper script(s) that automate parts of it run once installed.
Works best when your chat has web access.
Installing is the full version — it triggers automatically, runs its scripts and loads references as needed. As a chat prompt you drive the method by hand.
Unlock to copy the ready-to-paste prompt — then in “My Skills”.
Reviews
No reviews yet – be the first.
Note
Image Prompt Director writes prompts — it **doesn't generate, upscale or inpaint** and doesn't check the finished image; you generate in your own tool. AI image models change monthly: with web access the skill verifies current versions/parameters live and dates its claims. It's a **first orientation, not legal advice** — verify commercial license, copyright, trademark and likeness per model; it doesn't clone living artists and can't guarantee exact logos/text/brand colors (composite instead). For the labeling duty of published AI images (EU AI Act Art. 50), the KI-Transparenz skill is responsible. The bundled scripts need Node (no extra packages).
Changelog
- v1.0.021.08.2026Initial release: turns an idea + brand into a precise, model-aware image prompt. Asks the target model (Midjourney, GPT Image/DALL·E, Gemini/Nano Banana, FLUX, Stable Diffusion, Ideogram, Firefly) and emits the correct syntax (--ar/--sref/--no vs (word:1.3)+negative field vs plain language), fills every pillar (subject, composition, camera, lighting, color, style), builds a reusable style/character bible for consistent series (the #1 pain), routes legible text to the right model, gives negative-prompt/aspect-ratio/iteration guidance, and flags rights/labeling (license, copyright, trademark, EU AI Act → hand off to KI-Transparenz). Two Node scripts (prompt-forge, prompt-lint), 6 references. It doesn't generate — verify-live.
Frequently asked questions
What does Image Prompt Director do?
Image Prompt Director is an AI skill that turns your idea + brand into a precise, model-aware image prompt. It asks your target model (Midjourney, GPT Image/DALL·E, Gemini, FLUX, Stable Diffusion, Ideogram, Firefly) and emits the correct syntax instead of flowery adjectives, fills every pillar, builds a style/character bible for consistent series, and flags rights/labeling. It doesn't generate — verify-live.
Which image AI does Image Prompt Director write prompts for?
All the major ones — and that's the point. Image Prompt Director asks your target model (Midjourney, GPT Image/DALL·E, Gemini/Nano Banana, FLUX, Stable Diffusion, Ideogram, Firefly) and emits the correct syntax, because it doesn't port — (word:1.3) is Stable-Diffusion-only, Midjourney uses --no and --sref, GPT/Gemini/FLUX want plain natural language. Because these models change monthly, it verifies the current version live.
Why is Image Prompt Director better than a generic prompt generator?
Generic bots just bolt flowery adjectives onto your idea — which is why AI images all look the same. Image Prompt Director is model-aware (correct syntax per generator), brand-aware (builds a reusable style/character bible against the
How do I keep my character or style consistent across images?
Consistency is the biggest problem in AI image generation, and Image Prompt Director gives you the kit for it — a fixed, reusable style block plus a character bible, the right technique per model (Midjourney --oref/--sref, FLUX Kontext, Gemini/Nano Banana, IP-Adapter) and seed discipline. Honestly — no model guarantees pixel-exact identity — the exact bits get composited in.
Can Image Prompt Director put exact logos, text and brand colors in the image?
Not reliably inside the image directly — and Image Prompt Director tells you that honestly instead of promising it. Exact logos and brand hex colors are best composited in a design tool; legible text it routes to the models that can do it (Ideogram, GPT Image, Recraft). That gets you a dependable result instead of gibberish.
Does Image Prompt Director generate the images itself?
No. Image Prompt Director writes the prompt (plus negative prompt, parameters and a style kit) — it doesn't generate, upscale or inpaint, and it doesn't check the finished image. You generate in your own tool. That keeps you in control and makes the skill portable across every generator.
Which AI tools does Image Prompt Director work with?
Claude · ChatGPT/Codex · Cursor · Copilot · Gemini CLI · Windsurf · Cline
How do I use Image Prompt Director?
Image Prompt Director is a SKILL.md in the open agentskills.io standard: install it with one command (npx) or download it and add it to your AI tool — Claude (Projects), ChatGPT (Custom GPT), Cursor, Copilot, Gemini CLI and more. No code needed.