← Open Source
wuyoscar

GPT-Image2-Skill

GPT Image 2/2.5 prompt gallery, image prompt library, agentic skill, and CLI for OpenAI image generation/editing

ListsPrompt collectionsPython
Open on GitHub
Momentum
+-1stars in 24 hours-0.0%
5.63k
Stars
471
Forks
+46
This week
2
Contributors
Created 2026-04-22 · Updated 2026-10-05 · #16475 today
Top developers
README

GPT Image 2/2.5 Prompt Gallery + Agent Skills + CLI

Prompts, reference images, two agent skills and a CLI for GPT Image 2 and 2.5.

English · 中文

License: MIT PRs Welcome Models: GPT Image 2 / 2.5 Python ≥ 3.11

GPTImage2Skill banner

Overview · 2.5 samples · Example · Quick start · Install · CLI reference · Guides · Gallery · Contribute

🧭 What about this Repo

Item

Value

Gallery size

163 numbered entries across 31 categories, with selected images below

Surfaces

2 Agent Skills + CLI: Claude Code / Codex, OpenClaw, Hermes Agent and other skill-capable agent runtimes

Last update

2026-09-10

Docs

English + 中文

The repo keeps the GPT Image 2 prompt collection and gallery alongside 2.5 API support, reading material and task-specific references. The two skills handle image generation/editing and image-to-prompt extraction.

TBH, GPT Image 2.5 feels seriously capable. I prefer giving it a clear reference: a shape, a sketch or an image. Sometimes showing a layout from a PDF is more useful than describing it at length. It's how I like to work with GPT-6, too: minimize the prompt; make the reference clear.

I collect prompts, useful building blocks and references here to help you find a workflow that suits the job. Thanks for all the love this little gallery has received 🫶.

For the CLI, export the relevant PDF pages as PNG, WebP or JPG, then attach them with -i. See the supported image-reference formats. Keep required text and edit constraints explicit.

For PPT work, try vector-style diagrams, icons and slide layouts. The Image API outputs PNG, JPEG or WebP; editable SVG or PowerPoint shapes need a separate authoring step.

✨ Made with GPT Image 2.5

Two 2K samples: an exploded watch assembly with detailed callouts, and a multi-storey cafe cutaway built from a reference image. Both use gpt-image-2.5-sunburst, 2048x2048 and high.

  [![Sunburst exploded mechanical watch with numbered component callouts](docs/technical-illustration/meridian8-sunburst.png)](docs/technical-illustration/meridian8-sunburst.png)
  A · Meridian 8 exploded assembly  

2048x2048 · high · Curated

  [![Sunburst isometric cafe district with illuminated cutaway interiors](docs/isometric/cafe-cutaway-sunburst.png)](docs/isometric/cafe-cutaway-sunburst.png)
  B · Night cafe cutaway  

2048x2048 · high · Curated adaptation

A · No. 113 · Prompt

Create a premium technical exploded-view illustration of a fictional mechanical wristwatch called the Meridian 8, centered on a dark slate background with fine blueprint grid accents. Show the watch components separated vertically with precise spacing: sapphire crystal, dial, hands, chapter ring, movement plates, escapement, balance wheel, mainspring barrel, case, crown, and leather strap sections. Use realistic brushed steel, brass, ruby jewel accents, and deep navy dial details. Add crisp callouts and labels with the in-image text "Meridian 8", "Exploded Assembly", "42 mm Case", "25 Jewels", and "Power Reserve 72 h". Include numbered callouts "01" through "10" with short labels like "Balance Wheel", "Mainspring Barrel", and "Sapphire Crystal". The result should be highly detailed, technically believable, sharply rendered, and suitable for an industrial design plate with clean hierarchy, exact labeling, and refined material realism.

B uses the reference from No. 54, keeping the overall street layout while reworking the interiors, floors and lighting. Reference attribution: EvoLinkAI · Source.

Original isometric cafe district used as the edit reference

B · No. 54 · Edit prompt

Use the reference image as the layout anchor for a richly detailed isometric two-block cafe district at blue hour. Keep the street footprint, corner cafe, neighboring bookstore, bakery and fountain plaza recognizable. Transform it into a three-storey architectural cutaway diorama with coherent 30-degree isometric geometry.

Open the front-facing walls to reveal the cafe espresso bar and upstairs jazz lounge; bookshelves, reading nooks and a spiral staircase in the bookstore; pastry cases and a working oven in the bakery. Add a rooftop glass greenhouse, tiny terraces, copper plumbing, tiled stairs, balconies, hanging plants and warm lights visible through rain-speckled windows. At street level show wet cobbles, bicycles, the coffee cart, varied miniature pedestrians and reflections around the fountain. Every floor, doorway and staircase should connect plausibly.

Use warm amber interiors against deep teal evening shadows, tactile brick, glazed tiles, glass and brushed brass. Preserve crisp detail throughout the scene, with a clean dark navy background and room around the floating diorama. Give the scene depth through cutaway rooms and layered architecture. Use restrained, readable storefront lettering: "NIGHT OWL CAFE", "OPEN BOOKS", and "DAWN BAKERY". Keep the composition square and visually balanced.

See the sample record for settings, inputs and review notes.

🖼️ From reference image to result

Thanks to @LunarXuan for Get Prompt from Image. A vision-capable agent extracts a prompt from a reference image, then passes it to gpt-image or another generator. The contributor-provided reference and its generated result are shown below.

   ![Contributor-provided snowy urban alley reference image](docs/illustration/get-prompt-from-image-reference.jpg) 
  Reference image · Contributor-provided




   ![ImageGen result generated from the reverse-engineered prompt](docs/illustration/get-prompt-from-image-result.png) 
  Generated result · ImageGen output

Attach an image and invoke the skill with a slash command, $get-prompt-from-image, or plain language:

/get-prompt-from-image
Extract a reusable English positive prompt and a targeted negative prompt from this image, then recreate it with gpt-image.

📝 Extracted prompt used for the generated result

Positive Prompt

A highly polished semi-realistic Japanese narrative illustration rendered in a painterly digital style, using varied brush widths, a combination of hard edges and soft transitions, restrained contour lines, and controlled surface texture. The image should feel like a cold cinematic game-concept artwork. Use a wide 16:9 composition with strong depth in a snowy urban alley, where the snow-covered road narrows toward a distant vanishing point near the center. Place a large fluffy dark blue-gray wolfdog in the left foreground, shown in side profile facing right with its head raised, interacting with a hooded young woman kneeling near the center-right. She crouches in the snow facing left, gently touching the wolfdog’s muzzle or forehead with one gloved hand while the other rests near her knee for balance, creating a restrained and intimate gesture. She wears an oversized pale-gray winter hooded jacket with pointed ear-like details on top, dark gray panels, pockets, straps, and small muted red-orange accents, over black clothing, fitted black pants, and heavy dark boots. Short black or deep-brown hair falls from beneath the hood; her face is partly shadowed as she looks down at the wolfdog with a quiet, tired, yet gentle expression. Render the wolfdog’s fur with layered directional brushstrokes, making the back, neck, and tail thick and voluminous, with cool blue-gray shadows, pale highlights, and a subtle rim light along the silhouette. On the left, include metal fencing, utility boxes, and dense dark shrubs; in the distance, show tall urban buildings, street lamps, utility poles, and a blue-gray sky. On the right, include dark building facades, windows, snow-covered roof edges, evergreen branches, and foreground cardboard boxes and industrial clutter. Any environmental labels should remain blurred graphic marks with no readable text. Let the main light enter from the distant upper-left side of the alley, combining cold blue ambient shadows with warm golden reflections in the distance. Add subtle rim light to the snow, the woman, and the wolfdog, with medium-high contrast and warm orange clothing details acting as focal accents. Snow, slush, and shallow puddles in the foreground should show damp reflections. Use atmospheric perspective to soften distant buildings while keeping the woman and wolfdog clear. Establish depth through foreground, middle ground, background, occlusion, and perspective lines rather than strong blur. The mood is loneliness, trust, and a brief moment of tenderness in a frozen city. Preserve rough painterly strokes, cool-warm contrast, cinematic composition, and refined post-processing. Clearly remain a 2D semi-realistic painterly illustration, not photography, pure flat vector art, or 3D rendering.

Negative Prompt

photorealistic, 3D render, flat vector style, pure cel shading, watercolor bleed, oil painting impasto, chibi proportions, deformed anatomy, malformed hands, extra limbs, oversized wolf, sunny summer weather, cluttered composition, readable text, watermark

🚀 Quick start

Task Use
Generate or edit an image gpt-image
Extract a prompt from a reference get-prompt-from-image
Work in a terminal CLI examples below, with the selected --model

🎛️ Choose a model

Model Best starting point
gpt-image-2.5-flare Fast, high-quality everyday generation
gpt-image-2.5-sunburst Precise edits and reference-image workflows
gpt-image-2 Retain the model used by existing workflows

The Skill offers this menu when the model is missing or ambiguous (for example, “GPT 2.5”), then confirms the choice before generating. An exact supported model choice proceeds directly. It always passes an explicit --model; the standalone CLI still defaults to gpt-image-2. Changing models or adding outputs requires the user's approval.

Both 2.5 models add --quality xhigh / max and support transparent PNG/WebP output. Start with low drafts; higher quality can increase latency and cost. For example, after choosing Flare:

gpt-image --model gpt-image-2.5-flare \
  -p "An original flat leaf icon, centered with generous padding, transparent background" \
  --quality medium --background transparent --format png -f leaf.png

See model compatibility and verification notes and GPT Image 2.5 prompt templates. Local output validation for these templates is pending. Existing gallery images keep their original model and source credits.

After install, every gallery entry below can be copy-pasted as gpt-image --model -p "…" or requested from any skill-capable agent runtime in natural language, e.g. "generate the Boston Spring poster from the skill gallery".

Text → image

gpt-image --model gpt-image-2.5-flare -p "a photorealistic convenience store at 10pm" --size 1k --quality high -f store.png

Under the hood: POST /v1/images/generations with the explicitly selected model.

Text + reference image → image (edit)

# Single-reference edit / restyle
gpt-image --model gpt-image-2.5-sunburst -p "Make it a winter evening with heavy snowfall" \
  -i chess.png --quality high -f chess-winter.png

# Multi-reference edit: the edits endpoint accepts multiple input images
gpt-image --model gpt-image-2.5-sunburst -p "Place the dog from image 2 next to the woman in image 1. Match the same lighting, composition, and background. Do not change anything else." \
  -i woman.png -i dog.png --size portrait --quality medium -f woman-with-dog.png

# Mask-based inpaint: opaque = keep, transparent = regenerate
gpt-image --model gpt-image-2.5-sunburst -p "replace sky with aurora" \
  -i photo.jpg -m sky_mask.png -f aurora.png

Under the hood: POST /v1/images/edits (multipart form). GPT Image 2 and both 2.5 models use this endpoint, with multiple -i inputs and an optional -m mask. Read results from data[].b64_json and omit response_format. See model compatibility.

📥 Install

Choose either skill or install both: gpt-image generates and edits images, while get-prompt-from-image extracts prompts from reference images.

Check for an existing skill or CLI before installing. Preserve existing skill folders and API-key files. Use your runtime's skill list/status command when available, and ask before installing into a global or shared directory.

command -v gpt-image || true
command -v uv >/dev/null && uv tool list | grep -E '^gpt-image-cli([[:space:]]|$)' || true
test -n "${OPENAI_API_KEY:-}" && echo "OPENAI_API_KEY is already set (value hidden)"

Claude Code

/plugin marketplace add wuyoscar/gpt_image_2_skill
/plugin install gpt-image@wuyoscar-skills

Codex

Codex ships with built-in skill helpers such as $skill-installer and $skill-creator. Open Codex and invoke the built-in installer with the GitHub skill-folder URL for each skill you want:

# gpt-image
$skill-installer
Install this skill from GitHub:
https://github.com/wuyoscar/gpt_image_2_skill/tree/main/skills/gpt-image

# get-prompt-from-image
$skill-installer
Install this skill from GitHub:
https://github.com/wuyoscar/gpt_image_2_skill/tree/main/skills/get-prompt-from-image

The installer downloads each GitHub folder and places it under your Codex skills directory, usually:

~/.codex/skills/gpt-image
~/.codex/skills/get-prompt-from-image

Restart Codex after installation so the new skills are loaded.

If you prefer to install both manually, copy their skill folders into Codex's skills directory:

git clone https://github.com/wuyoscar/gpt_image_2_skill.git
cd gpt_image_2_skill

mkdir -p "${CODEX_HOME:-$HOME/.codex}/skills"
for skill in gpt-image get-prompt-from-image; do
  test -e "${CODEX_HOME:-$HOME/.codex}/skills/$skill" && echo "$skill already exists; stop before overwriting" && exit 1
  cp -R "skills/$skill" "${CODEX_HOME:-$HOME/.codex}/skills/"
done

AgentSkills / npx skills

For runtimes supported by the cross-agent skills installer, select either skill or install both together from GitHub:

# Change --agent to claude-code, codex, opencode, openclaw, or another supported runtime.
npx --yes skills@latest add wuyoscar/gpt_image_2_skill \
  --skill gpt-image \
  --skill get-prompt-from-image \
  --agent codex --copy

These examples intentionally avoid --global. Add --global only when you explicitly want this skill installed into that runtime's global/shared skills directory.

Other runtimes can use the manual agent-skill installation below.

Manual agent-skill install

Set AGENT_SKILLS_DIR to the skills directory used by your agent runtime, then symlink one or both skill folders into it.

git clone https://github.com/wuyoscar/gpt_image_2_skill.git
cd gpt_image_2_skill

# Choose the skill directory for your runtime.
# Examples:
#   Codex:      ~/.codex/skills
#   Claude Code / OpenClaw / Hermes Agent / other runtimes: use that runtime's documented skills directory.
export AGENT_SKILLS_DIR="/path/to/your/agent/skills"

mkdir -p "$AGENT_SKILLS_DIR"
for skill in gpt-image get-prompt-from-image; do
  test -e "$AGENT_SKILLS_DIR/$skill" && echo "$skill already exists; stop before overwriting" && exit 1
  ln -s "$PWD/skills/$skill" "$AGENT_SKILLS_DIR/$skill"
done

CLI

uvx --from git+https://github.com/wuyoscar/gpt_image_2_skill gpt-image --model gpt-image-2.5-flare -p "a cat astronaut"

# or install to PATH if not already installed
command -v gpt-image >/dev/null || uv tool install git+https://github.com/wuyoscar/gpt_image_2_skill
gpt-image --model gpt-image-2.5-flare -p "a cat astronaut"

Update

# plugin: use Claude Code's update flow
# codex skill: rerun the installer
# manual git clone
cd gpt_image_2_skill && git pull

# CLI
uv tool upgrade gpt-image-cli

Reads OPENAI_API_KEY from process env, then .env, then ~/.env without overriding an already-set env var.

API keys: This CLI reads the process environment, .env and ~/.env. To avoid using a local key, check all three locations; unset OPENAI_API_KEY clears the process variable only. Codex users can choose its built-in image tool when they prefer platform-managed generation.

🛠️ CLI reference

Parameters, quality settings and SDK examples

Parameters (complete)

Show full parameter reference

Flag Values Default Applies to Notes
-p, --prompt str required both Full prompt text.
--model exact model ID above gpt-image-2 both The Skill always passes the confirmed model explicitly; the CLI default stays backward-compatible.
-f, --file path ./fig/YYYY-MM-DD-HH-MM-SS-.png both Explicit output path.
-i, --image path (repeatable) omitted edits Presence routes through /v1/images/edits.
-m, --mask path (PNG, alpha) omitted edits Opaque = preserved, transparent = regenerated. Requires -i.
--input-fidelity low · high omitted edits Image 2 cannot set it, so the CLI omits it. For 2.5, explicit values pass through but behavior remains unverified; see model notes.
--size 1k · 2k · 4k · portrait · landscape · square · wide · tall · literal 1024x1024 etc. 1024x1024 both Literals must be 16-px multiples, max edge 3840, 3:1 cap, 655k-8.3M total pixels.
--quality auto · low · medium · high · xhigh · max high both xhigh / max require a 2.5 model; use low drafts to assess quality, latency, and cost first.
-n, --n 1-10 1 both Batch generation. n>1 suffixes filenames _0, _1, …
--background auto · opaque · transparent API default both Transparency requires PNG/WebP; Image 2 transparency remains in preview.
--moderation auto · low low generations low is the default here for broader prompt exploration; switch to auto if you want the stricter API-side default.
--format png · jpeg · webp png both Response encoding.
--compression 0-100 omitted both JPEG/WebP only.

Budget / quality guide

Use --quality to manage the generation budget.

  • low = cheap draft / collect / many variants
  • medium = normal exploration / style probing
  • high = final posters, Chinese text, diagrams, paper figures, banners

If you are generating dozens of candidates, start at low and only rerun finalists at high. On 2.5, compare xhigh / max for detail-critical assets. Agree on additional outputs and costs before comparing settings.

From gallery prompt → CLI / SDK

Each entry includes a prompt and metadata ("size" · "quality" · source). Use the shared CLI/SDK pattern below. This example uses "portrait" and "high":

# CLI
gpt-image --model gpt-image-2.5-flare -p "" --size portrait --quality high -f out.png
# OpenAI SDK — `size` is the literal pixels; the CLI shortcut maps to `1024x1536` for portrait
from openai import OpenAI
client = OpenAI()
result = client.images.generate(
    model="gpt-image-2.5-flare",
    prompt="",
    size="1024x1536",
    quality="high",
)

For reference-based edits, add -i ref.png (repeatable) and optionally -m mask.png on the CLI, or call client.images.edit(...) with image=[open(p, "rb") for p in refs]. Shared output options stay the same. --moderation applies to generation; edit calls omit it.

Exit codes: 0 success · 1 API/refusal error (full response body echoed to stderr) · 2 bad args or missing OPENAI_API_KEY.

📚 Guides and reading

[!CAUTION] Use generated images as references, workflow sketches or style targets for research figures. Prepare and verify the final figure separately before including it in a paper; directly publishing a raw generation can mislead readers.

📖 Prompting Fundamentals

Show prompting notes

Shared techniques from the historical Image 2 Cookbook, with current 2.5 guidance organized in a separate official-source reference router. Image 2 retains gallery-first guidance. A precise, model-confirmed 2.5 request needs no reference loading; otherwise choose one short task slice. Migration notes are separate. API parameters stay in model notes; community templates remain separately attributed.

  1. Structure, then goal. Use a consistent order: background/scene → subject → key details → constraints, and state the intended use (ad, UI mock, infographic) so the model picks the right mode and polish level.
  2. Any format works; consistency matters more. Minimal prompts, descriptive paragraphs, JSON-style structures, instruction-style prompts, and tag-based prompts all work. For production, choose a format you can easily read and maintain.
  3. Specificity + quality cues. Be concrete about materials, shapes, textures, and medium (photo, watercolor, 3D render). Add targeted levers only when they matter: film grain, textured brushstrokes, macro detail. For photorealism, say "photorealistic" directly; "real photograph", "taken on a real camera", and "iPhone photo" also help.
  4. Quote the exact text. Put required slogans, prices and labels in straight quotes. Copy the wording verbatim.
  5. Choose aspect ratio early. Decide 1:1 / 3:4 / 4:3 / 9:16 / 16:9 / 3:1 before writing the prompt. Keep the prompt's aspect ratio consistent with --size.
  6. One hero, supporting cast. Complex scenes work best when one subject is clearly primary and the rest is framed as supporting detail.
  7. Evaluate quality against the actual task. For 2.5 text, diagrams and multi-panel layouts, compare approved settings against legibility and layout requirements; high remains the CLI default. Judge each setting by the resulting text and layout.
What you need Reference
Gallery index gallery.md: choose a category, then read its prompts.
Concrete examples Load a relevant gallery-*.md file, such as products, UI or paper figures.
Prompt craft craft.md: text, composition, data relationships and edit constraints.
GPT Image 2.5 Reference index: separate generation, layout/text, editing and migration notes.
Historical Image 2 guidance OpenAI Cookbook, with original source and license. Use current model notes and the user's settings for API calls.

📚 GPT Image 2.5: understand the workflow

English-language official sources and established publications first, developer guides next; Chinese-language reading supplements them. Read these links when useful. OpenAI's documentation is the reference for API details.

🌐 English-first · 8 sources

Source Read What to take away
🏛️ 1 · OpenAI Official launch Reference-led editing, ChatGPT's new creative tools, and the Flare/Sunburst lineup.
📖 2 · OpenAI Docs Image prompting: GPT Image 2.5 Official model selection, reference roles, focused edits and output checks.
📰 3 · Axios Exclusive hands-on A reporter's experience with likeness preservation and visually targeted edits.
🧪 4 · TechRadar 24-hour hands-on Everyday editing controls, sketch-based creation and task templates.
✏️ 5 · The Verge Sketch walkthrough † Sketch coverage; full-text retrieval pending.
🛠️ 6 · Apidog API walkthrough · Model comparison Third-party implementation and migration perspectives; verify API details against official docs.
⚖️ 7 · OrcaRouter Flare vs Sunburst for builders ‡ Model-selection advice by workload; performance claims need independent checking.
🧩 8 · FindSkill Sketch step by step Use a drawing for layout and text for additional instructions.

🗂️ Chinese-language supplements · 2 sources

Source Read What to take away
🧪 9 · 人人都是产品经理 Reader-supplied article † Supplied as a 2 vs 2.5 hands-on comparison; content remains unverified.
🗞️ 10 · 軟體玩家 Feature roundup by 阿正老師 A feature roundup based on official announcements and media reports.

Source-access status, checked 2026-09-09: † Full text pending for The Verge and 人人都是产品经理. ‡ Indexed content only for OrcaRouter.

Using ChatGPT sketches with the CLI: Sketch, Templates and Comments operate inside ChatGPT. Export a sketch as an image and attach it with -i. For task-specific local notes, choose one slice from the optional 2.5 reference index.

🎨 Gallery

Open a category to view its images and prompts. The complete 163-prompt collection is indexed in gallery.md, with full entries in the linked gallery-*.md files.

Source labels. Curated means a repo-curated or substantially reworked prompt/image; outside-source items keep visible author/source links.

🎌
Anime & Manga
Full Gallery MD

🎮
Gaming
Full Gallery MD

🤖
Retro & Cyberpunk
Full Gallery MD

🎬
Cinematic & Animation
Full Gallery MD

👤
Character Design
Full Gallery MD

📝
Typography & Posters
Full Gallery MD

🎨
Illustration
Full Gallery MD

💧
Watercolor
Full Gallery MD

🖌️
Ink & Chinese
Full Gallery MD

🕹️
Pixel Art
Full Gallery MD

📐
Isometric
Full Gallery MD

📦
Product & Food
Full Gallery MD

🧩
Brand Systems & Identity
Full Gallery MD

📷
Photography
Full Gallery MD

🖥️
Screen Photography
Full Gallery MD

📊
Infographics & Field Guides
Full Gallery MD

📚
Research Paper Figures
Full Gallery MD

🏢
Official OpenAI Cookbook Examples
Full Gallery MD

✨
Edit Endpoint Showcase
Full Gallery MD

📱
UI/UX Mockups
Full Gallery MD

📊
Data Visualization
Full Gallery MD

⚙️
Technical Illustration
Full Gallery MD

🏛️
Architecture & Interior
Full Gallery MD

🔬
Scientific & Educational
Full Gallery MD

👗
Fashion Editorial
Full Gallery MD

🎨
Fine Art Painting
Full Gallery MD

✏️
More Illustration Styles
Full Gallery MD

🎥
Cinematic Film References
Full Gallery MD

💄
Beauty & Lifestyle
Full Gallery MD

🎟️
Events & Experience
Full Gallery MD

🖋️
Tattoo Design
Full Gallery MD


🎌 Anime & Manga

↑ Gallery index

Anime fashion portrait triptych

  [![Elegant cafe anime fashion portrait](docs/anime-manga/anime-cafe-stockings-fashion.png)](docs/anime-manga/anime-cafe-stockings-fashion.png)  

  A · Elegant cafe fashion  

"portrait" · "high" · "Curated"

  [![Neon arcade anime fashion portrait](docs/anime-manga/anime-arcade-stockings-fashion.png)](docs/anime-manga/anime-arcade-stockings-fashion.png)  

  B · Neon arcade fashion  

"portrait" · "high" · "Curated"

  [![Roadside mirror anime fashion selfie](docs/anime-manga/anime-roadside-mirror-fashion.png)](docs/anime-manga/anime-roadside-mirror-fashion.png)  

  C · Roadside mirror selfie  

"portrait" · "high" · "Curated"

Anime & Manga · 3-image portrait set · Curated

📝 Prompts for all three panels

Prompt A: Elegant cafe fashion

Create a tasteful portrait-oriented anime fashion illustration of an adult woman, age 24, with a cute playful expression, looking at the camera in a cozy European cafe at golden hour. She wears a cream blouse, charcoal pleated skirt, tailored cropped jacket, sheer black stockings, loafers, and a small ribbon hair clip; she is seated sideways at a small marble table with latte art, a sketchbook, and warm window light. Composition: three-quarter fashion portrait, elegant legs visible but relaxed and non-explicit, wholesome editorial mood, no nudity, no lingerie, no school uniform, no explicit pose, adult character only. Use polished modern anime rendering, crisp line art, luminous eyes, soft cel shading, subtle fabric texture, gentle blush, background bokeh, and a refined magazine-cover color palette.

Prompt B: Neon arcade fashion

Create a portrait-oriented anime fashion illustration of an adult woman, age 25, in a neon arcade district at night. She has a cute confident smile and looks directly at the viewer while standing beside glowing claw machines and retro game cabinets. Outfit: black turtleneck, red satin bomber jacket, high-waisted skirt, patterned dark stockings, platform shoes, small crossbody bag, star earrings. Composition: full-body fashion portrait with strong silhouette, neon reflections on wet pavement, vending machines, sticker-covered walls, colorful signage, and cinematic rim light. Keep the pose playful but non-explicit, no nudity, no lingerie, no fetish framing, adult character only. Use high-end anime key visual rendering, crisp line art, saturated magenta-cyan lighting, clean readable background details, and glossy cyber-pop atmosphere.

Prompt C: Roadside mirror selfie

Create a portrait-oriented anime fashion illustration of an adult woman, age 24, taking a playful roadside mirror selfie in the reflection of a parked scooter mirror on a quiet Tokyo side street. She looks into the mirror with a bright mischievous smile, one hand making a small peace sign near her cheek, the other holding a phone with a cute sticker case. Outfit: soft ivory knit cardigan, navy pleated skirt, sheer black stockings, loafers, small shoulder bag, ribbon hair clip, tasteful everyday street fashion. Composition: the mirror reflection is the main frame, with blurred street signs, vending machine glow, crosswalk stripes, and spring evening light around the mirror edge. Keep the pose cute, stylish, and non-explicit; no nudity, no lingerie, no fetish framing, adult character only. Use polished modern anime rendering, crisp line art, luminous eyes, soft cel shading, warm reflections, natural street-photo energy, and a charming slice-of-life mood.

MAPPA-style anime action still (Jujutsu-Kaisen aesthetic)

MAPPA-style anime action still (Jujutsu-Kaisen aesthetic)

"landscape" · "high" · "Curated"

📝 Prompt

An anime action still in the visual style of MAPPA's Jujutsu Kaisen (2020 TV anime). Landscape 16:9.

A silver-white-haired young man in a dark navy school-uniform jacket, a blue blindfold across his eyes, in a mid-fight stance — one palm extended outward releasing a swirling dense-blue energy sphere with lightning-like crackles around its edge. Opposite him, a demonic shadow creature made of liquid black mass with multiple eyes lunges from the right.

Backdrop: ruined urban street at dusk, shattered asphalt, cracked neon kanji sign "呪術" in split red LED, destroyed vehicles, rubble suspended mid-air by the shockwave, rain particles caught mid-flight.

Art direction: MAPPA-style digital 2D animation — heavy cel shading, crisp line-art, rim-light on both figures, motion-blur streaks around the energy sphere. Palette of deep navy, electric cyan, crimson splashes. Kinetic-impact composition in the tradition of JJK's Shibuya arc.

Shōnen battle key-visual (Naruto-Shippuden aesthetic)

Shōnen battle key-visual (Naruto-Shippuden aesthetic)

"landscape" · "high" · "Curated"

📝 Prompt

A shōnen anime battle key-visual in the visual style of Studio Pierrot's Naruto Shippuden. Landscape 16:9.

Two ninja figures clash mid-air at the exact instant their signature jutsu collide — a glowing blue spiral of swirling chakra on the left fighter's right palm, a crackling white lightning blade on the right fighter's right palm. The collision point sends a circular shockwave outward.

Both fighters wear hitai-ate forehead protectors, jounin-style tactical vests with scroll pouches, ninja sandals. Left: spiky blond hair, whisker cheek marks, focused snarl, blue eyes. Right: dark hair, one red sharingan-like eye with three tomoe, calm expression.

Backdrop: nighttime valley, cracked earth, giant uprooted trees mid-crash, moonlit clouds parting, sakura petals caught in the shockwave.

Art direction: Studio Pierrot Naruto-Shippuden aesthetic — dynamic perspective, strong speed lines radiating from the collision, anime-action key-frame quality, digital 2D cel shading, saturated but not neon, visible genga-quality line-art, dramatic backlight.

Manga / anime 1×2 panel

  [![Shōnen manga two-page spread (basketball slam dunk)](docs/anime-manga/manga-spread.png)](docs/anime-manga/manga-spread.png)  

  A · Shōnen manga two-page spread (basketball slam dunk)  

"landscape" · "high" · "Curated"

  [![Ten-panel anime character grid](docs/anime-manga/anime-ten-panel-character-grid.png)](docs/anime-manga/anime-ten-panel-character-grid.png)  

  B · Ten-panel anime character grid  

"landscape" · "high" · "Curated"

Anime & Manga · 1×2 panel · Curated

📝 Prompts for both manga/anime panels

Prompt A: Shōnen manga two-page spread (basketball slam dunk)

A black-and-white shōnen manga two-page spread (landscape 16:9 as a single composition, with a faint centre-gutter line). High-contrast ink plus screentone, Weekly Shōnen Jump basketball-manga tradition (Inoue's Slam Dunk / Fujimaki's Kuroko no Basuke).

Composition: 5 irregular panels plus one large diagonal panel spanning both pages at bottom-right for the climactic slam dunk.

- Top-left: close-up of the protagonist's intense eyes, sweat beading, headband tied tight
- Top-centre: wide shot of a packed high-school gymnasium, scoreboard reading "42 — 40 · 4Q 0:03"
- Top-right: rival team captain's shocked face, mouth agape
- Centre-left: protagonist leaping skyward with both hands gripping a basketball
- Centre-right-small: sound-effect katakana "バッ" in thick black letters
- Large diagonal bottom-right (half of both pages): protagonist slamming the ball through the hoop, rim bending, massive ink-brushed kanji "決" (decide) filling the negative space

Art direction: professional mangaka quality — confident inking, dramatic screentone gradients, speed lines radiating from the dunk, varied line-weights, off-white paper texture with faint page-edge shading.

Dialogue balloons intentionally blank; only the two sound effects are visible.

Prompt B: Ten-panel anime character grid

Create a single landscape image containing a clean 2×5 ten-panel anime character grid. Each panel shows a different adult young woman, age 22 to 26, designed as a cute gentle heroine archetype: bookish librarian, cheerful cafe barista, shy violinist, sporty tennis player, elegant student-council president, sleepy illustrator, flower-shop assistant, soft-spoken witch apprentice, city-pop singer, and cozy winter commuter. Keep all panels consistent in art direction: modern polished anime, crisp line art, soft cel shading, luminous eyes, pastel accent colors, tidy white gutters, small readable name tag at the bottom of each panel, and a balanced character-design-sheet feel. Every character should have a distinct hairstyle, outfit, prop, and expression. The overall board should feel like a collectible anime cast sheet / ten-grid poster, cute and wholesome, no nudity, no lingerie, no explicit pose, adult characters only.

16-panel anime expression grid

16-panel anime expression grid

"square" · "high" · "WeChat"

📝 Prompt

Create a 16-panel expression grid of a silver-haired, blue-eyed anime girl. Her face shape, hairstyle, and clothing must remain highly consistent across all panels. The 16 expressions should include: happy, sad, angry, surprised, shy, speechless, evil grin, contemplative, curious, proud, wronged, disdainful, confused, scared, crying, and a heart expression.

Tide Brothers 19-page manga proof sheet

Tide Brothers 19-page original manga proof sheet

"tall 2160×3840" · "high" · "Curated"

📝 Prompt

Create one tall manga chapter proof sheet containing 19 numbered miniature pages for an original shonen pirate manga, not based on any existing series. Title: "TIDE BROTHERS: THE STARFALL MAP". Main characters: Rune, a cheerful rubbery-armed young pirate captain with a straw-colored scarf but original costume; and Ash, his older flame-wielding brother with a red coat, freckles, and a calm smile. They are original characters, not existing IP. Show 19 small pages arranged as a readable contact sheet, each page with 1 to 3 manga panels, black-and-white ink, screentone, dynamic speed lines, expressive faces, and clear speech bubbles. Complete plot beats: 1 cover page with the brothers on a stormy deck; 2 reunion at a floating harbor; 3 discovery of a star-shaped map; 4 alien sea-beast emerges; 5 Rune jokes "Adventure found us first!"; 6 Ash replies "Then we answer together."; 7 rival sky pirates attack; 8 slapstick cooking scene; 9 quiet flashback promise; 10 double-page-style action pose compressed into one page; 11 map glows with alien constellations; 12 crew cheers; 13 villain captain steals the compass; 14 chase across rooftop sails; 15 Ash shields Rune with fire; 16 Rune launches a spring-like punch; 17 brothers laugh after victory; 18 cliffhanger: moon door opens; 19 final page text "NEXT: THE ISLAND ABOVE THE CLOUDS". Keep dialogue short, legible, and complete. Style: classic weekly shonen manga energy, original pirate adventure, wholesome brotherhood, no gore, no existing copyrighted characters.

🎮 Gaming

↑ Gallery index

Stealth and open-world action panel

  [![Hitman gameplay: OpenAI HQ](docs/gaming/hitman-openai.png)](docs/gaming/hitman-openai.png)  

  A · Hitman gameplay: OpenAI HQ  

"landscape" · "high" · "X"

  [![GTA 6 gameplay: Vice City beach](docs/gaming/gta6-beach.png)](docs/gaming/gta6-beach.png)  

  B · GTA 6 gameplay: Vice City beach  

"landscape" · "high" · "X"

Gaming · 2-image landscape gameplay panel

📝 Prompts for Stealth and open-world action panel

Prompt A: Hitman gameplay: OpenAI HQ

A Hitman level where you are in the OpenAI HQ and your mission is to steal GPT-6 without getting caught

Prompt B: GTA 6 gameplay: Vice City beach

GTA 6 in-game footage, very detailed, very realistic. Close-up shot taken from a stationary 4k monitor. (There's a slight blurriness in the image, as it feels like it was taken handheld). A wide, bright environment. Realistic details. The character is walking on the beach with /:dog.

Fantasy adventure panel

  [![Dark-fantasy swamp boss hunt](docs/gaming/dark-fantasy-hunt.png)](docs/gaming/dark-fantasy-hunt.png)  

  A · Dark-fantasy swamp boss hunt  

"landscape" · "high" · "Curated"

  [![Epic fellowship bridge approach](docs/gaming/epic-fellowship-bridge.png)](docs/gaming/epic-fellowship-bridge.png)  

  B · Epic fellowship bridge approach  

"landscape" · "high" · "Curated"

Gaming · 2-image landscape gameplay panel

📝 Prompts for Fantasy adventure panel

Prompt A: Dark-fantasy swamp boss hunt

Create an original AAA dark-fantasy action RPG screenshot. A silver-haired monster hunter in layered leather armor stands in a ruined marsh at blue hour, sword drawn toward a huge winged swamp beast rising from mist. Cinematic over-the-shoulder framing, believable HUD with health, stamina, potion icons, quest text, and minimap. Wet stones, dead trees, torchlight, moonlit fog, subtle alchemy glyphs, highly detailed materials, dramatic but readable composition, premium next-gen game look, 16:9 landscape.

Prompt B: Epic fellowship bridge approach

Create an original epic fantasy RPG key-art screenshot. A small fellowship of travelers crosses a colossal ancient stone bridge toward a luminous mountain city at sunrise. One ranger leads, a mage carries a lantern, a dwarf-like smith bears a hammer, and banners whip in the wind. Vast valley below, waterfalls, golden clouds, weathered masonry, cinematic scale, subtle HUD quest marker and compass, richly detailed armor and environment, AAA fantasy adventure tone, 16:9 landscape, highly detailed and uplifting.

Stylized game HUD panel

  [![Retro Japanese town pixel RPG](docs/gaming/retro-japan-rpg.png)](docs/gaming/retro-japan-rpg.png)  

  A · Retro Japanese town pixel RPG  

"landscape" · "high" · "Reddit"

  [![Cyberpunk Europe action HUD](docs/gaming/cyberpunk-europe-action.png)](docs/gaming/cyberpunk-europe-action.png)  

  B · Cyberpunk Europe action HUD  

"landscape" · "high" · "Reddit"

  [![Anime open-world adventure HUD](docs/gaming/anime-open-world.png)](docs/gaming/anime-open-world.png)  

  C · Anime open-world adventure HUD  

"landscape" · "high" · "Reddit"

  [![Mobile MOBA arena HUD](docs/gaming/mobile-moba-arena-hud.png)](docs/gaming/mobile-moba-arena-hud.png)  

  D · Mobile MOBA arena HUD  

"landscape" · "high" · "Curated"

Gaming · 2×2 landscape gameplay HUD panel

📝 Prompts for Stylized game HUD panel

Prompt A: Retro Japanese town pixel RPG

Create an isometric pixel-art RPG screenshot of a traditional Japanese village during cherry blossom season. Sakura petals drift through the air, a samurai player character practices sword moves in the square, villagers watch nearby, and the interface includes an inventory panel, stamina gauge, skill cooldown timers, and subtle quest UI. Cozy retro console feeling, soft ambient pastel lighting, crisp pixel details, 16:9 gameplay composition.

Prompt B: Cyberpunk Europe action HUD

Create a third-person cyberpunk action game screenshot set in a neon-soaked European capital at night. The protagonist has glowing cybernetic implants and stands on rain-slick streets near a famous landmark while holograms, drones, and flying traffic crowd the skyline. Add a polished game HUD with health bar, ammo count, radar, stealth/energy meters, and mission overlays. Vivid cyan-magenta palette, wet reflections, cinematic intensity, 16:9.

Prompt C: Anime open-world adventure HUD

Create a third-person over-the-shoulder screenshot from a nostalgic anime-style open-world adventure game. The protagonist stands in a lush forest with detailed foliage and vibrant shading, drawing a bow toward distant enemies. Add a clean on-screen HUD: quest log, compass at the top, character portrait and status effects at bottom left, subtle rain droplets on screen, and sun rays filtering through trees. Keep the composition dynamic, the forest immersive, and the UI believable like a premium action-RPG screenshot.

Prompt D: Mobile MOBA arena HUD

Create an original landscape mobile MOBA / action-RPG gameplay screenshot, inspired by competitive lane-battle games but not copying any existing franchise. 16:9 landscape, polished mobile game HUD. Scene: a bright fantasy arena at golden-hour dusk, three stylized heroes clash near a central river bridge and glowing crystal objective. Camera: slightly elevated isometric third-person gameplay view, readable battlefield lanes, minions, spell effects, terrain brush, turret silhouettes, and a boss-objective pit in the distance. HUD design: bottom-left translucent virtual joystick, bottom-right four circular ability buttons with cooldown numbers, ultimate button glowing but 87% charged, top-center score bar reading "12 - 11", match timer "08:42", team health bars, mini-map in the top-left, item quick slots, gold counter "3,420", clean mobile-safe margins, crisp icons, no real game logos. Art direction: premium anime-fantasy 3D mobile game, saturated teal / gold / violet palette, sharp readable UI, dynamic spell VFX, high-detail materials, readable text, screen-capture feel, not a poster, not a mockup board.

Nine-panel dark-fantasy worldbuilding set

Nine-panel dark-fantasy worldbuilding set

"square" · "high" · "X"

📝 Prompt

Create a square 3x3 worldbuilding set for an original dark-fantasy universe called "Saltwind Reach". Each panel is a distinct but consistent scene: a storm-battered coastal fortress at dawn, a foggy market street, a knight relic close-up, a handwritten map fragment, a monster silhouette study, a candlelit tavern interior, an alchemist kit flat lay, a moonlit harbor, and a faction banner concept. Keep one cohesive art direction across all nine panels: painterly realism, muted teal / rust / bone palette, cinematic weather, premium concept-art presentation, small caption labels, and strong consistency across costume motifs, architecture, symbols, and lighting. The full board should feel like a polished pre-production worldbuilding sheet rather than a collage of unrelated images.

🤖 Retro & Cyberpunk

↑ Gallery index

Cyberpunk mecha girl over sea fortress

Cyberpunk mecha girl over sea fortress

"landscape" · "high" · "GitHub archive"

📝 Prompt

A mecha girl mid-teens, pale skin smudged with soot and salt spray, sharp amber eyes with glowing HUD reticles, waist-length ash-white hair tied in a high ponytail whipping in the sea wind, matte gunmetal exoskeleton armor plating her shoulders, forearms and shins, exposed hydraulic pistons at the joints, chest rig with glowing cyan coolant lines, oversized oil-stained hangar jacket half slipping off one shoulder, a massive rail cannon resting on her right shoulder, dog tags and frayed red ribbon at her collar, standing off-center to the left on the rusted edge of a tilted steel platform jutting out over dark water, weight shifted onto one leg, left hand gripping the cannon strap, head turned slightly toward camera with a quiet defiant stare, steam venting from her back thrusters, her ponytail and jacket streaming sideways in the salt wind, a vast derelict sea-city at dusk, colossal megastructures of unknown purpose rising from the ocean in staggered silhouettes, bone-white monolithic towers fused with barnacled steel, cyclopean ring-shaped constructs canted at broken angles, rusted skeletal gantries threaded with dead cables, dark swells rolling between the pylons, shipwrecks half-swallowed at their feet, thick sea fog clinging to the bases while the upper structures pierce into a bruised sky, scattered faint lights blinking high in the towers like distant eyes, moody low-key lighting, cold teal ambient from the overcast sky, warm amber sodium glow leaking from a distant structure camera-right, hard backlight from a low sun behind the towers carving her silhouette, volumetric god rays cutting through sea mist, wet specular highlights on her armor, 35mm anamorphic lens, slight low angle looking up past her shoulder toward the structures, medium-wide shot, shallow depth of field with foreground rust in soft focus, horizontal lens flares, fine atmospheric haze compressing the distant megastructures into layered silhouettes, cinematic anime key visual, painterly digital illustration with crisp line art, desaturated oceanic palette of teal, bone-white and rust punched by small warm accent lights, film grain, high-contrast editorial poster aesthetic. Format 16:9.


Neon Orchid District design board

Neon Orchid District cyberpunk design board

"landscape" · "high" · "Curated"

📝 Prompt

Create a cyberpunk character-and-city design board in a premium magazine-layout format, landscape 16:9. Title text: "NEON ORCHID DISTRICT". The board is divided into five asymmetric panels: one large cinematic street scene of a rain-soaked elevated night market, two close-up portrait panels of original adult cyberpunk couriers with glowing orchid tattoos, one small isometric map panel showing alleys and drone routes, and one artifact panel showing encrypted transit passes, cybernetic gloves, and vending-machine stickers. Use layered neon magenta, cyan, acid green, wet asphalt reflections, holographic signage, dense but readable composition, editorial margins, small labels, and a cohesive retro-future anime/cyberpunk style. Original characters only, no existing IP, no explicit content.

Synth Moon Crew alien nightlife grid

Synth Moon Crew cyberpunk alien nightlife grid

"square" · "high" · "Curated"

📝 Prompt

Create a square cyberpunk alien nightclub catalog sheet called "SYNTH MOON CREW". Layout: a clean 3×3 grid of nine cards with thin chrome borders. Each card shows a different original alien or android nightlife character: glass-horn DJ, koi-scale bartender, moth-wing hacker, chrome geisha bassist, jellyfish courier, neon priestess, reptile fashion model, vending-machine oracle, and masked dancer. Each card has a tiny readable name tag and a unique color accent, but the whole grid shares a polished late-90s anime cyberpunk aesthetic, black background, fluorescent rim lights, glossy materials, sticker-like UI glyphs, playful stylish energy, no gore, no explicit content, original designs only.

🎬 Cinematic & Animation

↑ Gallery index

Pixar-style 3D animation still (kitten)

Pixar-style 3D animation still (kitten)

"landscape" · "high" · "Curated"

📝 Prompt

A Pixar-quality 3D animation still, landscape 16:9. Cinematic feature-film look, warm studio lighting.

Scene: a cozy apartment kitchen at dawn. A small orange tabby kitten sits on the countertop reaching a paw toward a rising soufflé in the oven; oven glow lighting the scene from below. Soft morning light through linen curtains. A wooden chopping board with a half-peeled lemon, a copper whisk with a small cloud of flour still airborne, a tiny succulent in a clay pot.

Character: kitten with expressive, slightly oversized eyes (classic Pixar proportions), individually sculpted whiskers, believable fur with micro-groom direction, curious-slightly-worried expression.

Art direction: full-CG Pixar aesthetic — subsurface scattering on ears and whiskers, physically based materials, soft shadow ambient occlusion, volumetric morning beam, shallow depth of field. Clean stylised shapes consistent with "Luca", "Soul", "Elemental" — not photoreal uncanny-valley.

1940s film-noir still

1940s film-noir still

"landscape" · "high" · "Curated"

📝 Prompt

A 1940s film-noir black-and-white movie still, landscape 16:9, high contrast. Shot on 35mm with visible grain.

Scene: a detective in trench coat and fedora stands alone at a rain-soaked street corner at 2 a.m., cigarette in hand, smoke curling upward. Wet cobblestones reflecting a single buzzing street lamp. A "HOTEL" neon sign on brick facade with letters "HOTE_" (the L flickered out). A vintage 1946 sedan parked at the curb, tail-lights glowing through drizzle.

Lighting: classic chiaroscuro — single hard key light above right, venetian-blind shadows on the wall behind him. Deep blacks, silvered highlights, full tonal range from pure white to pure black. No colour. Frame should feel lifted from "The Maltese Falcon", "Double Indemnity", or "The Third Man".

Professional 6-panel film storyboard

Professional 6-panel film storyboard

"landscape" · "high" · "Curated"

📝 Prompt

A 6-panel film storyboard laid out as a 3×2 grid, landscape 16:9 overall. Each panel is a rectangular pencil-and-marker sketch with a white margin border and a small information strip underneath.

Scene: a chase through a rainy Tokyo alleyway, ending in a rooftop jump.

Panel 1 — WIDE establishing: wet neon alleyway, runner entering from left; kanji signage on both walls. Info: "PANEL 1 · EXT. ALLEY · NIGHT · WIDE / static / 2s"
Panel 2 — OTS tracking: runner mid-stride from behind; pursuer silhouette 10 m back. Info: "PANEL 2 · OTS TRACKING / follow-cam / pan-L 45° / 3s"
Panel 3 — Close-up: runner's face, sweat, eyes darting up toward fire escape. Info: "PANEL 3 · CU RUNNER / static / 1.5s / SFX: breath"
Panel 4 — Low angle: runner leaping onto fire-escape ladder; rain streaks. Info: "PANEL 4 · LOW ANGLE / tilt-up 30° / 2s"
Panel 5 — Wide aerial: runner silhouetted against neon skyline, about to leap rooftops. Info: "PANEL 5 · WIDE AERIAL / crane-down / 4s"
Panel 6 — Match cut: runner's boots landing on wet rooftop; splash. Info: "PANEL 6 · MATCH CUT CU / static / 1s / SFX: splash"

Art direction: classic animation-school storyboard — pencil line-work, grey marker shading, red-pencil arrow annotations on panels 2 and 5 (camera move and action arc). Off-white paper texture background.

Studio-Ghibli-style animation still

Studio-Ghibli-style animation still

"landscape" · "high" · "Curated"

📝 Prompt

A Studio-Ghibli-style hand-painted animation still, landscape 16:9. A small wooden cottage sits on a grassy hillside overlooking a valley at golden hour. A child stands barefoot at the cottage doorway waving to a small furry forest spirit half-hidden in the meadow grass. A distant train cuts across the valley floor, swallows dip overhead.

Art direction: classic Miyazaki / Studio Ghibli watercolor-gouache style. Soft painterly edges, slightly desaturated greens and warm skin tones, visible brush texture in the clouds and grass. Thin ink line art on the characters. Gentle atmospheric perspective. The whole frame should feel like a cel from "My Neighbor Totoro" or "Kiki's Delivery Service", not a 3D render.

VHS grocery-store chaos still

VHS grocery-store chaos still

"landscape" · "high" · "Reddit"

📝 Prompt

Create a chaotic security-camera still from a 1990s grocery store. A man in full medieval armor is frozen mid-sprint stealing several rotisserie chickens past the dairy section. Overhead fluorescent lights reflect off the armor. The floor is baby-blue tile. Add a timestamp reading "08/13/96 04:44 AM" and a wall poster saying "NEW! TOASTER STRUDELS!". Make it low-fidelity, absurd, slightly intense, with motion blur, VHS color bleed, surveillance noise, and authentic analog-store lighting.

👤 Character Design

↑ Gallery index

Official character reference sheet

Official character reference sheet

"landscape" · "high" · "X"

📝 Prompt

Based on this character and background, please create a character reference sheet similar to official setting materials.
- Includes three-view drawings: front view, side view, and back view
- Add variations of the character's facial expressions
- Break down and display detailed parts of the clothing and equipment
- Add a color palette
- Include a brief explanation of the worldview setting
- Overall, use an organized layout (white background, illustration style)

Elven archer sketchbook concept sheet

Elven archer sketchbook concept sheet

"portrait" · "high" · "Reddit"

📝 Prompt

Create a fantasy concept art sketchbook page centered on a mystical elven archer with flowing robes. Render the main figure in loose graphite strokes with precise ink detailing. Surround the hero sketch with side views exploring cloak variations, a half-finished bow study with measurements, thumbnail action poses, handwritten annotations about enchanted embroidery patterns, and faint watercolor tests bleeding into the margins in forest-green and silver. The page should feel like a real art director's development sheet: exploratory, beautiful, readable, and richly tactile.

📝 Typography & Posters

↑ Gallery index

Poster 1×3 panel

  [![Chongqing rainy-night city promo poster](docs/typography-posters/city-tourism-promo-poster.png)](docs/typography-posters/city-tourism-promo-poster.png)  

  A · Chongqing rainy-night 山城雨夜  

"portrait" · "high" · "Xiaohongshu"

  [![Vogue-style fashion magazine cover](docs/typography-posters/vogue-cover.png)](docs/typography-posters/vogue-cover.png)  

  B · Vogue-style fashion magazine cover  

"portrait" · "high" · "Curated"

  [![1950s Astounding Stories pulp cover](docs/typography-posters/pulp-scifi-cover.png)](docs/typography-posters/pulp-scifi-cover.png)  

  C · 1950s Astounding Stories pulp  

"portrait" · "high" · "Curated"

Typography & Posters · 3-poster panel · Mixed original + community

📝 Prompts for all three posters

Prompt A: Chongqing rainy-night city promo poster

做一张 3:4 城市宣传海报,主题是"山城雨夜·重庆"。整体像高端城市文旅 campaign poster,不要廉价旅行社风格。画面中心是层叠山城建筑、轻轨穿楼、湿润街道、霓虹倒影、江边雾气和夜色中的坡道。用现代中文排版,加入少量准确标题与副标题:"山城雨夜" / "CHONGQING" / "8D 城市 / 江雾 / 火锅 / 轻轨 / 夜景"。信息密度适中,留白克制,色彩以深蓝、暖橙、湿润霓虹红为主,像一本设计年鉴里的城市品牌海报。

Prompt B: Vogue-style fashion magazine cover

A high-fashion magazine cover, 3:4 portrait, Vogue Paris / British Vogue editorial aesthetic.

Subject: a tall female model, medium-dark skin tone, mid-thirties, standing three-quarters to camera, direct piercing gaze. She wears a sculptural high-collared ivory wool coat over a silk slip dress in deep aubergine. Minimalist silver spiral earrings. Hair in a sleek low chignon with a single escaped strand. Makeup: matte bronze-warm, glossy plum lip.

Background: muted concrete-grey seamless paper backdrop, vertical shaft of cool daylight from upper left. Shallow depth of field.

Exact cover typography (all English, crisp, correctly spelled):
- Masthead, huge uppercase serif, white: "VOGUE"
- Date strip top-left, tiny caps: "NOVEMBER 2026 · PARIS EDITION · €9.00"
- Main cover line, bold sans-serif centered: "THE QUIET POWER ISSUE"
- Right-edge cover lines, stacked:
   "THE NEW MINIMALISTS — a 40-page portfolio"
   "HOW AI TOOLS ARE REWRITING THE ATELIER"
   "MARTIN MARGIELA'S UNREVEALED ARCHIVE"
   "SKIN · INVESTMENT · WHERE THE MONEY GOES NEXT"
- Bottom-left barcode with catalog code "VG1126"

Lighting: classic fashion editorial — soft single-source key, subtle fill, deep shadow on one cheek, fine film grain.

Prompt C: 1950s Astounding Stories pulp cover

A vintage sci-fi pulp magazine cover from the 1950s, 3:4 portrait. Classic "Astounding Science Fiction" / "Galaxy" aesthetic — painted gouache illustration with pulp-yellow paper texture, screen-printing registration slightly off, pale browned paper tone around edges.

Cover illustration: a chrome-silver rocket ship descending toward an alien red-desert planet with two Saturn-like ringed moons in a violet sky. A lone astronaut in a bulbous 1950s-style glass-dome space helmet stands foreground-left in a crimson pressurised suit, holding a ray-gun, facing a many-tentacled translucent green creature emerging from a fissure.

Exact typography:
- Masthead, huge yellow retro display serif arched across the top: "ASTOUNDING STORIES"
- Volume banner, red, under masthead: "VOL. XXXVII · NO. 5 · MARCH 1957 · 25¢"
- Featured story callout, bold red sans-serif bottom-left: "THE MEN FROM RIGEL — a novelette by E. A. KLEIN"

Art direction: painted gouache with visible brush strokes, saturated pulp palette (canary yellow, orange, red, electric violet, chrome silver), hand-lettered headlines, slightly rough paper texture, faint foxing on corners.

🎨 Illustration

↑ Gallery index

Vintage Amalfi Coast travel poster

Vintage Amalfi Coast travel poster

"portrait" · "high" · "X"

📝 Prompt

Modern pencil illustration of Vintage travel poster illustration of the Amalfi Coast, Italy, panoramic coastal cliff road scene, classic 1960s white car driving along a curved seaside road, deep blue Mediterranean sea with small sailboats, colorful pastel hillside village, bright blue sky with soft clouds, lemon tree branches with vibrant yellow lemons framing the foreground, warm summer sunlight, bold vibrant colors, retro 1950s travel poster style, cinematic composition, high detail, screen print texture, graphic illustration. Hand-drawn style, illustration with loose strokes and defined contours. High-contrast color palette, maintaining chromatic harmony between background and elements. Contemporary and decorative aesthetic.

Paper-cut forest night market

Paper-cut forest night market

"landscape" · "high" · "Curated"

📝 Prompt

Create a landscape editorial illustration in layered paper-cut style: a tiny forest night market hidden beneath giant mushrooms and fern leaves. Include warm lantern stalls selling acorn cakes, beetle taxis, a fox calligrapher, a badger tea vendor, children holding leaf umbrellas, and fireflies forming soft dotted paths. Style anchor: mid-century children’s book illustration meets contemporary layered paper diorama, visible cut-paper edges, soft shadows between layers, muted moss green, pumpkin orange, cream, and ink-blue palette. First glance: a cozy glowing market silhouette. Second glance: many small vendor stories. Third glance: handmade paper texture, tiny signage, and playful animal gestures. No photorealism, no 3D plastic look, no cluttered unreadable faces.

💧 Watercolor

↑ Gallery index

Rainy botanical greenhouse watercolor

Rainy botanical greenhouse watercolor

"landscape" · "high" · "Curated"

📝 Prompt

Create a delicate watercolor illustration of a rainy botanical greenhouse in early morning. Landscape composition, transparent washes, granulating pigments, soft wet-on-wet blooms, visible cold-pressed paper texture. Scene: arched glass greenhouse ribs, raindrops streaming down panes, hanging ferns, orchids, clay pots, a narrow stone path, a wooden bench with an open gardening notebook, and diffused silver daylight. Palette: sage green, eucalyptus gray, pale lavender, warm terracotta, and tiny yellow flower accents. Keep the image airy and poetic, with preserved white paper highlights, no hard digital gradients, no photorealistic lens effects, and no heavy outlines.

🖌️ Ink & Chinese

↑ Gallery index

Song dynasty night-market handscroll

Song dynasty night-market handscroll

"landscape" · "high" · "Curated"

📝 Prompt

Create a horizontal Chinese ink-and-wash handscroll scene of a Song dynasty riverside night market. Use gongbi-level architectural detail combined with loose ink atmosphere: arched stone bridge, lantern boats, teahouse balconies, book stalls, noodle steam, scholars reading under lamps, children chasing paper rabbits, and distant city walls fading into mist. Add small readable Chinese shop signs in brush style: "茶", "书", "面", "灯市". Palette: black ink, warm lantern ochre, muted cinnabar seals, and pale blue-gray moonlight. Composition should read as a continuous scroll with rhythmic clusters of people and negative-space water. Avoid modern objects, anime faces, fake calligraphy clutter, and overly saturated poster lighting.

🕹️ Pixel Art

↑ Gallery index

Pixel art 1×2 panel

  [![Pixel art car sprite sheet](docs/pixel-art/pixel-sprite-cars.png)](docs/pixel-art/pixel-sprite-cars.png)  

  A · Pixel art car sprite sheet  

"square" · "high" · "X"

  [![Pixel art breakfast still life](docs/pixel-art/pixel-breakfast.png)](docs/pixel-art/pixel-breakfast.png)  

  B · Pixel art breakfast still life  

"square" · "high" · "Reddit"

Pixel Art · 1×2 panel · Sources credited per panel

📝 Prompts for both pixel art panels

Prompt A: Pixel art car sprite sheet

A 10x10 pixel art sprite sheet of retro video game cars, 16-bit era aesthetic. Ten rows by ten columns of small vehicle sprites on a clean light-grey grid background, each cell 64x64 pixels. Variety across sprites: sedans, sports cars, muscle cars, SUVs, pickup trucks, vans, taxi cabs, police cruisers, convertibles, and hot rods, in a full rainbow of colors. All sprites rendered in a consistent 3/4 top-down perspective with matching shading, crisp pixel edges, no anti-aliasing, palette limited to ~16 tones per sprite, SNES / Super Nintendo cart-racing game tradition.

Prompt B: Pixel art breakfast still life

Create a nostalgic pixel-art breakfast still life. Show a tall stack of fluffy golden pancakes drizzled with glossy maple syrup, topped with strawberries and blueberries, with pixelated steam rising into the air. The plate sits on a pastel tablecloth and a hot cup of coffee rests in the background. Use rich breakfast colors, careful lighting, and delicious texture detail while staying true to clean, readable pixel art.

📐 Isometric

↑ Gallery index

Isometric fantasy village map

Isometric fantasy village map

"square" · "high" · "Reddit"

📝 Prompt

Create a vibrant isometric fantasy village map with a clean grid-based layout using 3x3 meter tiles. Include wooden houses with thatched roofs, cobblestone paths, and a central stone fountain. One corner of the map rises into a small grassy hill about 2 meters high with stairs connecting to the lower ground. Keep the isometric angle precise and game-ready. Warm sunlight sends clear rays and long shadows across the rooftops. Make the scene readable like a handcrafted strategy-game map, with crisp tile logic, charming environmental detail, and rich but controlled color.

📦 Product & Food

↑ Gallery index

Product & food 1×3 panel

  [![3D product box from dieline](docs/product-food/product-dieline-box.png)](docs/product-food/product-dieline-box.png)  

  A · 3D product box from dieline  

"portrait" · "high" · "X"

  [![Chocolate wafer product render (JSON-style)](docs/product-food/product-chocolate-wafer.png)](docs/product-food/product-chocolate-wafer.png)  

  B · Chocolate wafer (JSON-style)  

"portrait" · "high" · "X"

  [![Universal commercial poster template](docs/product-food/aurora-oolong-poster.png)](docs/product-food/aurora-oolong-poster.png)  

  C · Universal commercial poster (Aurora Oolong)  

"portrait" · "high" · "Xiaohongshu"

Product & Food · 3-image panel · Sources credited per panel

📝 Prompts for all three product & food panels

Prompt A: 3D product box from dieline

Assemble the dieline into a flawless 3D box with accurate panels, clean folds, undistorted type, and artwork preserved exactly. Shoot it upright at a refined three-quarter angle in a minimal premium studio setting with a soft neutral background, diffused light, subtle shadows, no props, true colours, matte paperboard texture, and realistic editorial detail. The box front reads "AURAE / COLD-BREW MATCHA / 12 fl oz" in clean sans-serif. Side panel shows small ingredient list in 8pt type, nutrition-facts-style block. Clean, editorial, award-winning packshot aesthetic.

Prompt B: Chocolate wafer product render (JSON-style)

/* PRODUCT_RENDER_CONFIG: Chocolate Wafer Hazelnut Edition
   VERSION: 2.0.1
   AESTHETIC: Premium Commercial Food Photography */

{
  "ENVIRONMENT": {
    "Background": "Gradient(Dark_Warm_Brown)",
    "Atmospheric_FX": ["Floating_Particles", "Depth_Blur", "Cinematic_Bokeh"],
    "Lighting": { "Type": "Directional_Studio_Warmer", "Highlights": "Specular_Glossy_Reflections", "Shadow_Softness": "High" }
  },
  "CORE_ASSETS": {
    "Primary_Subject": "Wafer_Rolls",
    "Physics": "Zero_Gravity_Diagonal_X_Composition",
    "Material_Properties": {
      "Outer": "Milk_Chocolate_Coating",
      "Surface_Texture": "Irregular_Nut_Clusters_Embedded",
      "Interior_Cross_Section": { "Structure": "Crispy_Hollow_Wafer", "Core": "Silky_Chocolate_Cream_Filling" }
    }
  },
  "PARTICLE_SYSTEMS": [
    { "Object": "Chocolate_Blocks", "Detail": "Rectangular_Embossed_Letter_B", "State": "Floating" },
    { "Object": "Hazelnuts", "State": "Halved_and_Fragmented", "Distribution": "Random_Orbit" }
  ],
  "FLUID_DYNAMICS": { "Element": "Chocolate_Splash", "Behavior": "Dynamic_Backdrop_Flow", "Viscosity": "Thick_Glossy" },
  "RENDER_OUTPUT": { "Resolution": "8K_UHD", "Aspect_Ratio": "3:4", "Quality_Flags": ["Hyper_Realistic", "Sharp_Foreground", "Indulgent_Mood"] }
}

Prompt C: Universal commercial poster template

Design a high-end commercial poster for a product called "Aurora Oolong Cold Brew". Minimalist style, clean frame, centered hero bottle and tea glass, soft studio lighting, realistic material textures, elegant condensation details, generous negative space, premium brand visual language, cinematic light and shadow, refined packaging typography, and ultra-detailed finish. Make it feel like a luxury beverage campaign that could run in a subway lightbox or fashion magazine.

🧩 Brand Systems & Identity

↑ Gallery index

Moss Radio brand identity showcase board

Moss Radio brand identity showcase board

"square" · "high" · "X"

📝 Prompt

Create a square high-end brand identity showcase board for a fictional brand called "Moss Radio". The brand should feel analog, cultured, warm, tactile, and design-forward. It operates in independent audio hardware and café-retail and should appeal to creative professionals and music obsessives. The overall mood should be nostalgic but modern. Design a polished modular grid of multiple tiles, each showing a different application of one cohesive visual identity system. Include logo explorations, wordmarks, app icon variations, editorial posters, product cards, landing page fragments, packaging concepts, typography specimens, interface snippets, color palette presentations, sticker systems, patterns, branded mockups, and small motion-inspired compositions. Use Swiss-inspired typography, rounded industrial shapes, and a moss green / parchment / charcoal / copper palette. Dense but elegant layout, sharp alignment, strong hierarchy, premium case-study presentation.

PS1 nostalgia reboot brand kit

PS1 nostalgia reboot brand kit

"square" · "high" · "X"

📝 Prompt

Create a clean brand kit presented as one square modular board for a fictional revival of the PlayStation One era called "PS1 1998 Reboot". The identity should merge Japanese editorial design, Y2K nostalgia, acid green accents, VHS texture, silver plastics, disc-menu UI motifs, retail stickers, controller packaging, startup-screen typography, and memory-card iconography. Show multiple coordinated tiles including posters, packaging, interface snippets, collectible cards, typography studies, icons, and branded mockups. Keep it polished, cohesive, art-directed, and emotionally nostalgic, like a real top-tier design studio case study rather than generic merch.

Playful brand kit: Mochi Metro

Playful brand kit: Mochi Metro

"square" · "high" · "X"

📝 Prompt

Playful brand kit for "Mochi Metro", bold colors, fun typography, modern layout, modular square board with logo studies, packaging snippets, posters, app icons, stickers, UI fragments, and a cheerful Tokyo-snack visual system. Crisp alignment, dense but clean, highly polished design presentation.

📷 Photography

↑ Gallery index

Photorealistic 2×2 panel

  [![RAW iPhone: 42nd Street subway](docs/photography/photoreal-subway.png)](docs/photography/photoreal-subway.png)  

  A · RAW iPhone: 42nd Street subway  

"landscape" · "high" · "X"

  [![Handwritten notebook flatlay](docs/photography/handwritten-notebook.png)](docs/photography/handwritten-notebook.png)  

  B · Handwritten notebook flatlay  

"landscape" · "high" · "X"

  [![Chess board mid-tournament game](docs/photography/chess-midgame.png)](docs/photography/chess-midgame.png)  

  C · Chess board mid-tournament game  

"landscape" · "high" · "X"

  [![360° equirectangular jungle panorama](docs/photography/panorama-jungle.png)](docs/photography/panorama-jungle.png)  

  D · 360° equirectangular jungle panorama  

"wide 2048×1152" · "high" · "X"

Photography · 2×2 panel · Sources credited per panel

📝 Prompts for all four photography panels

Prompt A: RAW iPhone: 42nd Street subway

Create a completely RAW quality, unprocessed, unedited image with full iPhone camera quality. A subway station in USA, a momentary blur. The subway is in motion. In front of the subway, there is an elderly woman and man.

Prompt B: Handwritten notebook flatlay

Amateur photo of an open notebook lying flat, filled with handwritten notes in black ballpoint pen. The handwriting is casual and slightly messy, like personal notes, natural imperfections, crossed out words, underlined headings. Shot from slightly above, natural daylight from a window, no flash. Casual desk setting, shot on iPhone

Prompt C: Chess board mid-tournament game

Generate a photorealistic photo of a chess board during the middle of a serious tournament game. Top-down three-quarter view, shallow depth of field. All pieces clearly distinguishable and correctly shaped: pawns, rooks, knights (with horse-head silhouette), bishops (mitre tops), queens, kings (with cross finials). The position is mid-game: several pieces already captured and set aside to the right of the board, some pawns advanced, pieces clustered around the central files d4-e5-f4.

Materials: polished wooden staunton-style pieces — dark side in rosewood, light side in maple. Board made of inlaid maple and walnut squares. A digital chess clock sits to the left showing "00:14:28 / 00:08:47". Soft overhead tournament lighting, blurred tournament-hall background. All pieces accurate, no mutants, no extra sets.

Prompt D: 360° equirectangular jungle panorama

360 equirectangular panorama of a dense prehistoric jungle scene. Cinematic detail. Strict 2:1 aspect ratio (e.g. 4096×2048). No distortion at the seams — the left and right edges must wrap seamlessly.

Scene: towering fern-covered trees, shafts of golden sunlight piercing the canopy, a slow river winding through the centre foreground, mist rising off the water. Scattered dinosaurs of varied species — a grazing Brachiosaurus neck visible among distant tree canopy, two small Gallimimus drinking at the river's edge, a Triceratops in the background underbrush. Tropical birds in flight, butterflies, dragonflies over the water.

Lighting: late-afternoon golden hour, warm directional backlight through the canopy. High dynamic range, slight atmospheric haze. Equirectangular projection suitable for spherical / 360 viewers.

Natural SNS portrait


Natural upper-body mirror selfie in ambient light

Photography · portrait · 1087×1447 · Author: @LunarXuan · Source: GitHub

English realism guidelines from the source skill, paired with a contributor-selected mirror-selfie example. Add your scene, outfit, pose and framing requirements.

📝 Prompt

# Realism Guidelines

Apply these as the default aesthetic target.

## Basic mindset

- Make the result feel like a natural photo of a clearly adult woman in her 20s that could plausibly appear on a real SNS feed.
- Suppress CG-like texture, aggressive beauty filtering, plastic skin, and unnatural shine.
- Preserve attractiveness without making the face or skin impossibly perfect.
- Avoid fashion-campaign or advertisement styling; favor ordinary daily-life photography.
- Prioritize real-world presence, lived-in detail, and natural imperfection over perfection.

## Photography and texture

- Prefer an everyday smartphone snapshot rather than a high-end camera look.
- Use natural window light or existing practical lighting in the scene.
- Allow mild hand-shake softness, smartphone sensor noise, modest sharpening artifacts, and slightly rough resolution.
- Keep framing a little imperfect: slight tilt, uneven negative space, minor cropping, or off-center composition.
- Avoid poses that look designed for a photoshoot; favor a caught-in-the-moment feeling.
- Do not overuse cinematic depth of field. Ordinary phones often keep more of the scene legible.

## Facial details

- Keep realistic facial proportions and bone structure.
- Do not enlarge the eyes. Preserve natural sclera ratio, eyelid thickness, and subtle left-right differences.
- Keep expressions restrained and spontaneous; a faint smile or relaxed neutral expression is preferable to a forced grin.
- Unless explicitly requested, avoid rigid straight-on passport framing.
- Preserve mild facial asymmetry around brows, eyes, cheeks, and mouth.

## Final touches

- Keep pores, fine facial hair, soft redness, tiny blemishes, and natural shadow variation where appropriate.
- Hands and fingers must have plausible count, joints, lengths, overlap, and grip.
- Keep the overall impression clean, approachable, and cute while remaining physically believable.
- Hair should contain flyaways, irregular clumps, slight frizz, and uneven strands rather than perfectly separated hair fibers.
- Use realistic body proportions consistent with the reference or request. Avoid exaggerated hourglass shaping or model-like body retouching.
- Favor the kind of cuteness that feels possible in real life rather than doll-like perfection.

## Common negative constraints

Avoid: CGI, 3D render, doll skin, porcelain skin, excessive skin smoothing, beauty-app face reshaping, huge eyes, perfectly symmetrical face, over-sharpened eyelashes, waxy highlights, glam studio lighting, commercial fashion campaign, artificial rim light, extreme bokeh, hyper-detailed pore texture, over-HDR, unreal anatomy, extra fingers, fused fingers, floating accessories, warped glasses, duplicated jewelry, fake text, malformed background objects.

🖥️ Screen Photography

↑ Gallery index

Real screen-photo prompt pair

  [![Music app + webcam preview](docs/screen-photography/laptop-music-webcam-screen.png)](docs/screen-photography/laptop-music-webcam-screen.png)  

  A · Music app + webcam preview  

"1152×1536" · "high" · "Reddit"

  [![Notes + FaceTime work screen](docs/screen-photography/laptop-notes-facetime-screen.png)](docs/screen-photography/laptop-notes-facetime-screen.png)  

  B · Notes + FaceTime work screen  

"1152×1536" · "high" · "Curated"

Screen Photography · 1×2 raw phone-photo-of-screen palette · A adapted from Reddit prompt structure, B Curated

📝 Prompts for both screen-photo panels

Prompt A: Music app + webcam preview

Create a raw smartphone photo of a laptop screen, not a screenshot. Aspect ratio 3:4, high-angle downward POV looking down at a laptop on a desk at night. The screen fills most of the frame with a thin strip of physical keyboard visible at the bottom. Emphasize visible RGB pixel grid, subtle moire bands, micro dust on glass, faint fingerprints, soft ambient reflections, handheld phone noise, slight perspective skew, imperfect glass. macOS dark mode. Background app: a generic music player in Liked Songs view with fictional visible tracks: "City Lights", "Late Night Walk", "Summer Static", "Blue Hour". Foreground app: a small webcam preview window floating center-right, showing only a cozy desk corner with a ceramic mug, notebook, small plush bear, warm desk lamp, and off-white wall. Make it look like an accidental real phone photo of a screen, candid and unpolished. No people, no faces, no celebrity names, no real-person likeness, no screenshot, no flat UI, no perfect clean glass, no studio lighting, no cartoon, no 3D render, no watermark.

Prompt B: Notes + FaceTime work screen

Create a raw smartphone photo of a laptop screen, not a screenshot. Aspect ratio 3:4, high-angle downward POV from someone standing over a desk at night. The laptop display fills most of the frame, with a narrow strip of black keyboard and trackpad visible at the bottom. Strong realism: visible RGB subpixel grid, subtle moire bands, small dust specks, faint fingerprints, uneven glass reflections, handheld phone noise, slight perspective skew, no studio polish. macOS dark mode. Background app: Apple Notes with a late-night study note titled "Design Critique" and short visible bullets: "layout", "lighting", "source links", "ship tomorrow". Foreground app: FaceTime live preview window floating lower-right, showing a fictional adult man in his 20s sitting at a cluttered desk, hoodie, tired but amused expression, warm desk lamp behind him, books and sticky notes in the room. A second small Finder window with image thumbnails is partly visible behind it. Make it feel like an accidental real phone photo of a working laptop screen. No real-person likeness, no beauty filter, no perfect UI, no screenshot, no watermark, no cartoon, no 3D render.

📊 Infographics & Field Guides

↑ Gallery index

Song Dynasty social-media feed

Song Dynasty social-media feed

"portrait" · "high" · "X"

📝 Prompt

"Song Dynasty People's Moments"/"SONG DYNASTY SOCIAL MEDIA FEED", Ancient and modern time-travel humor fusion interface design style, The image simulates a mobile phone social media interface, but the content is entirely Song Dynasty scenes, The avatar is a portrait of a Song Dynasty literati, Username "Su Dongpo SuShi_Official", Post content "Just arrived in Huangzhou, demoted but feeling okay. Made Dongpo pork myself today, tastes amazing, recipe attached:", The attached image is a close-up of Dongpo pork in Gongbi painting style, Likes list "Huang Tingjian, Qin Guan, Fo Yin etc. 126 people", Comments section "Wang Anshi: Hehe" "Sima Guang: Still the same taste", Interface elements such as the like icon are replaced with Song Dynasty patterns, The status bar shows "Great Song Mobile 5G" and "Third Year of Yuanfeng", The color scheme is mobile phone dark mode paired with elegant Song Dynasty tones, A masterpiece of fun collision between history and social media

Museum catalog disassembly infographic (唐代襦裙)

Museum catalog disassembly infographic (唐代襦裙)

"portrait" · "high" · "X"

📝 Prompt

Please automatically generate a "museum catalog-style Chinese disassembly infographic" based on the [Subject].

The entire image is required to combine a realistic main visual, structural disassembly, Chinese annotations, material descriptions, pattern meanings, color meanings, and core feature summaries. You need to automatically determine the most appropriate main subject, clothing system, artifact structure, era style, key components, material craftsmanship, color scheme, and layout structure based on the [Subject], and the user does not need to provide any other information.

The overall style should be: national museum exhibition boards, historical clothing catalogs, and cultural/museum thematic infographics, rather than ordinary posters, ancient-style portraits, e-commerce detail pages, or anime illustrations. The background uses paper textures such as off-white, silk white, and light tea color, making the overall look premium, restrained, professional, and collectible.

The layout is fixed as:
- Top: Chinese main title + subtitle + introduction
- Left: Structural disassembly area, with Chinese lead lines annotating key components, accompanied by close-up details
- Upper right: Material / craftsmanship / texture area, displaying real texture samples with descriptions
- Middle right: Pattern / color / meaning area, displaying the main color palette, pattern samples, and cultural explanations
- Bottom: Dressing order / composition flowchart + core feature summary

If the subject is suitable for character display, use a full-body standing posture of a real person as the central subject; if it is more suitable for artifacts or single structures, change it to a central subject disassembly diagram, but the overall form remains a complete Chinese infographic. All text must be in Simplified Chinese, clear, neat, and readable, without garbled characters, typos, English, or pinyin.

Avoid: poster feel, studio portrait feel, e-commerce feel, anime feel, cosplay feel, random annotations, incorrect structures, blurry text, fake materials, excessive decoration.

Encyclopedia field guide (Giant Panda)

Encyclopedia field guide (Giant Panda)

"portrait" · "high" · "X"

📝 Prompt

Generate a high-quality vertical encyclopedia-style infographic for [topic].

This should not be a normal poster or a simple illustration. It should feel like a modular educational infographic that combines the clarity of a field guide, the structure of an encyclopedia page, the polish of a lifestyle knowledge card, and the shareability of a strong social-media explainer.

The image should include:
- a clear and appealing main visual of the topic
- several enlarged detail callouts
- multiple rounded modular information sections
- strong title hierarchy and highlighted key labels
- concise but information-rich educational content
- visual scoring, quick takeaways, or a Top 5 module

Adapt the content sections automatically based on the topic. Useful categories include: basic profile, classification, appearance, habits or ecology, formation mechanism or structure, growth or usage conditions, care or maintenance advice, risks and cautions, suitable users or use cases, pros and cons, and a quick scorecard.

Visual requirements: use a clean light background, soft colors, subtle shadows, refined small icons, rounded information cards, and neat layout. The information density should be high but not crowded, and the final image should feel publishable, collectible, and repeatable as a knowledge-card format rather than an advertisement.

Do not make it look like a commercial promo poster. Emphasize knowledge organization, modular information, and a field-guide presentation.

Camera styles reference board for iPhone photographers

Camera styles reference board for iPhone photographers

"landscape" · "high" · "X"

📝 Prompt

Make me an image in 35 mm film style of a diagram showing the knowledge of camera styles, presets, and what to know about them as an aspiring iPhone photographer that wants to pursue their passion. Build it as a rich multi-panel reference board with labeled sections for film looks, digital presets, portrait approaches, street photography styles, color temperature, grain, contrast, flash, framing, and common mistakes. Each camera and preset style should appear in its actual style instead of being rendered uniformly in one style. Make it visually dense, highly educational, beautifully designed, and easy to scan.

📚 Research Paper Figures

↑ Gallery index

Research paper figure grid

  [![Patient cohort and multimodal biomarker workflow](docs/research-paper-figures/clinical-cohort-flow.png)](docs/research-paper-figures/clinical-cohort-flow.png)  

  A · Patient cohort and multimodal biomarker workflow  

"landscape" · "high" · "Curated"

  [![Single-cell immune atlas](docs/research-paper-figures/single-cell-immune-atlas.png)](docs/research-paper-figures/single-cell-immune-atlas.png)  

  B · Single-cell immune atlas  

"landscape" · "high" · "Curated"

  [![Multimodal medical-AI method](docs/research-paper-figures/multimodal-medical-ai-method.png)](docs/research-paper-figures/multimodal-medical-ai-method.png)  

  C · Multimodal medical-AI method  

"landscape" · "high" · "Curated"

  [![Therapeutic response statistics](docs/research-paper-figures/therapeutic-response-bar-forest.png)](docs/research-paper-figures/therapeutic-response-bar-forest.png)  

  D · Therapeutic response statistics  

"landscape" · "high" · "Curated"

  [![Transformer encoder-decoder architecture](docs/research-paper-figures/transformer-arch.png)](docs/research-paper-figures/transformer-arch.png)  

  E · Transformer encoder-decoder architecture  

"landscape" · "high" · "Curated"

  [![Multi-agent LLM system architecture](docs/research-paper-figures/agent-architecture.png)](docs/research-paper-figures/agent-architecture.png)  

  F · Multi-agent LLM system architecture  

"landscape" · "high" · "Curated"

  [![Denoising diffusion forward/reverse chain](docs/research-paper-figures/diffusion-chain.png)](docs/research-paper-figures/diffusion-chain.png)  

  G · Denoising diffusion forward/reverse chain  

"landscape" · "high" · "Curated"

  [![Empirical scaling laws plot](docs/research-paper-figures/scaling-curves.png)](docs/research-paper-figures/scaling-curves.png)  

  H · Empirical scaling laws plot  

"landscape" · "high" · "Curated"

  [![Benchmark comparison heatmap](docs/research-paper-figures/benchmark-heatmap.png)](docs/research-paper-figures/benchmark-heatmap.png)  

  I · Benchmark comparison heatmap  

"landscape" · "high" · "Curated"

  [![Ablation bar chart with error bars](docs/research-paper-figures/ablation-bars.png)](docs/research-paper-figures/ablation-bars.png)  

  J · Ablation bar chart with error bars  

"landscape" · "high" · "Curated"

  [![LLM pretraining data-mixture sankey](docs/research-paper-figures/data-sankey.png)](docs/research-paper-figures/data-sankey.png)  

  K · LLM pretraining data-mixture sankey  

"landscape" · "high" · "Curated"

  [![Multi-head attention heatmaps](docs/research-paper-figures/attention-heatmap.png)](docs/research-paper-figures/attention-heatmap.png)  

  L · Multi-head attention heatmaps  

"landscape" · "high" · "Curated"

  [![Frontier LLM family tree (2018-2026)](docs/research-paper-figures/model-timeline.png)](docs/research-paper-figures/model-timeline.png)  

  M · Frontier LLM family tree (2018-2026)  

"landscape" · "high" · "Curated"

  [![ReAct reasoning trace](docs/research-paper-figures/react-trace.png)](docs/research-paper-figures/react-trace.png)  

  N · ReAct reasoning trace  

"landscape" · "high" · "Curated"

  [![Frontier Safety Eval Loop](docs/research-paper-figures/frontier-safety-eval-loop.png)](docs/research-paper-figures/frontier-safety-eval-loop.png)  

  O · Frontier Safety Eval Loop  

"landscape" · "high" · "Curated"

  [![LLM Persona Atlas](docs/research-paper-figures/llm-persona-atlas.png)](docs/research-paper-figures/llm-persona-atlas.png)  

  P · LLM Persona Atlas  

"wide" · "high" · "Curated"

Research Paper Figures · 8×2 literature-science figure grid · Curated / source-attributed prompts retained below

📝 Prompts for all 16 research figures

Prompt A: Patient cohort and multimodal biomarker workflow

Create a Nature Medicine / Science Translational Medicine style research paper figure, landscape 3:2 (1536×1024), soft literature-science palette, minimal and elegant.

Figure title: "Patient cohort and multimodal biomarker workflow".

Layout: a clean 4-panel academic