Diagnosis
| Problem | Evidence in your prompt | Confidence |
|---|---|---|
| Three mixed styles | "cyberpunk watercolor oil painting" names three media. The model blends them. | Likely |
| Never night and day at once | "at night" and "bright sunny day" contradict each other. The model picks one or averages them. | Likely |
| Messy and cluttered | "minimalist but very intricate" contradicts itself. Three held objects plus a sign plus a forest is a lot for one image. | Likely |
| Too many fingers | The girl holds a sword, a bouquet and a lantern, which needs several complex hand grips. Many generators handle hands badly. | Likely |
| Gibberish sign text | Many generators render long text unreliably, and a 5-word all-caps phrase is a hard case. | Likely |
| "no bad hands" does nothing | Many tools read "hands" as something to draw and ignore the "no". It can even draw more attention to hands. | Possible |
| Filler words add nothing | "beautiful stunning masterpiece 8k ultra detailed" is abstract and gives the model nothing to draw. | Possible |
Minimal fix
Changes are in bold.
beautiful stunning masterpiece 8k ultra detailedgirl in a red dress standing in a forest at night,bright sunny day,holding a lantern only,a sword and a bouquet of flowers andwatercolor painting style,cyberpunk watercolor oil painting style, minimalist but very intricate,a wooden sign reading "WELCOME"that says 'WELCOME TO THE ENCHANTED WOODS', no bad hands
Clean rewrite
A girl in a red dress standing on a forest path at night, holding a glowing lantern in one hand, the other arm relaxed at her side. Full-body shot, centered, with tall dark trees behind her and a small wooden sign reading "WELCOME" to her left. Warm lantern light on her face and dress, deep blue moonlit shadows. Watercolor painting, soft washes, simple background.
If your tool has a negative prompt field, put this there: extra fingers, deformed hands, blurry text, mixed styles.
Variants to test
- Night and day together. Keep the rest, but change the light to "dusk, last orange sunset glow on the horizon, deep blue sky above, lantern lit". Dusk is the closest a single image gets to night and day at once. Look for whether the light now looks coherent.
- Keep the sword and flowers. Use "holding a lantern in her left hand, a sword sheathed at her belt", and drop the bouquet. Look for whether the hands stay correct when one hand is holding the lantern and the other is empty.
For a cyberpunk look, make that a separate test. Use "oil painting, neon pink and teal rim light" in place of watercolor. Don't combine it with the other styles.
Test plan
- Change one thing at a time and generate 3–4 images per version with the same settings.
- If hands are still wrong, crop or hide them: "hands behind her back", or a waist-up shot with the lantern at the edge of the frame.
- For the sign, either shorten it to one word (as above) or leave it blank ("a blank wooden sign") and add the text in an image editor afterward. This is the most reliable method.
What I couldn't tell
- Which tool you're using. This matters for several things:
- Midjourney and Stable Diffusion use negative prompts differently (
--noand a negative field). Tools like DALL·E or ChatGPT image generation usually have no negative field, so describe what you want instead. - Newer models handle text and hands much better than older ones, so your results may improve just by switching.
- Midjourney and Stable Diffusion use negative prompts differently (
- What your images look like. If you share one, I can tell you more precisely what went wrong.
- Check your tool's docs for how it handles negative prompts, text rendering and aspect ratio. A tall ratio like 2:3 suits a full-body standing figure.
Real run in an isolated folder with only this skill installed. Only the Skill and Read tools were enabled, so no command was run and no file was written. The scenario is fictional. No image was generated: the model only wrote text prompts, so how well they work depends on the image tool you use. The model could not see any generated image, so its diagnosis is a set of hypotheses to test.
Was es macht
Arbeitet wie ein Debugger: den genauen Prompt, das Tool und das Problem erfassen, den Prompt dann mit einer Tabelle häufiger Fehler abgleichen (ignorierte Teile, fades Ergebnis, Überladung, gemischte Stile, Widersprüche, falscher Bildausschnitt, flaches Licht, Farbübertragung, verformte Hände, unleserlicher Text, ungewollte Elemente, Unterschiede zwischen Durchläufen). Zu jedem Problem nennt es den Beleg im Prompt und die Sicherheit der Einschätzung und zeigt dann eine minimale Korrektur mit markierten Änderungen, eine saubere Neufassung und zwei Varianten, die je eine Sache testen.
Außerdem
Ein Testplan (eine Variable ändern, pro Version mehrere Bilder erzeugen, bei gleichen Einstellungen vergleichen), Hinweise, wann das Tool oder Modell die Grenze ist, und was nicht beurteilt werden konnte, weil Bild oder Tool-Verhalten nicht sichtbar sind.
Geeignet für
Einen Prompt, der „nicht funktioniert“, Bilder, denen immer ein Detail fehlt, und das Lernen, welche Wörter zählen.
Geringes Risiko: Reine Anweisungen ohne Skripte, ohne Netzwerkzugriff und ohne Dateischreiben. Er arbeitet mit dem Prompt und Ihrer Beschreibung des Ergebnisses; ohne geteiltes Bild sieht er es nicht und weiß nicht, wie Ihr Tool Wörter gewichtet, daher ist jede Diagnose eine zu prüfende Hypothese. Er hilft nicht, Sicherheitsfilter oder Nutzungsbedingungen eines Tools zu umgehen, reale Personen zur Täuschung nachzuahmen oder sexuelle Inhalte mit realen Personen oder Minderjährigen zu erzeugen. Zum Verhalten bei Negativ-Prompts und Textdarstellung prüfen Sie die Dokumentation Ihres Tools. Originalwerk von AIBars (MIT). Einmal mit einem erfundenen chaotischen Prompt ausprobiert; es wurde kein Bild erzeugt.