Agent Skills

image-generation

Plan image creation or edits from prompts, products, or references. Resolve task type, model, and reference rules before handing ready requests to image-batch-runner.

Install

npx skills add https://github.com/postplusai/postplus-skills --skill image-generation
SKILL.md

Image Generation

Use When

  • The desired final asset is an image or image batch.
  • The request may include text prompts, uploaded images, previous outputs, product URLs, research handoffs, product images, persona images, banners, thumbnails, storyboard panels, or batch variants.
  • The next decision is task class, model/reference policy, and runner handoff.

Do Not Use When

  • The final asset is video, audio, subtitles, transcripts, or an edit plan.
  • The image request is already normalized and ready to execute. Use image-batch-runner.
  • The user needs hook or benchmark decoding first. Use reference-decode.

Core Boundary

This is the image generation controller. It does not submit jobs.

It must:

  1. identify the input type,
  2. classify the image task,
  3. select model and reference rules,
  4. create the handoff for image-batch-runner.

Task Classes

Task class Typical input Handoff
text_to_image prompt only write normalized image brief for image-batch-runner
image_edit uploaded image plus change request bind edit image and preserve/alter rules
reference_image benchmark frame or style board state its intended influence in the runner handoff
product_image product photo, URL, or product facts bind product identity and forbid invented claims
banner_thumbnail offer, hook, platform require aspect ratio and text/UI policy
storyboard_image panel plan or board spec hand storyboard panels to image-batch-runner
batch_variant many variants or personas preserve shared rules and vary only declared fields

Model And Reference Rules

  • Text-only drafts may use a text-to-image endpoint.
  • Edits require a bound source image and an edit-capable endpoint.
  • Persona, product, and brand identity are binding references unless the user explicitly marks them inspiration-only.
  • Benchmark clips, mood boards, and competitor references are inspiration-only unless the contract says otherwise.
  • Excluded references must not enter the runner request.

Routing Table

If not image-generation Send to
Needs media understanding first media-analysis
Needs reference meaning decoded reference-decode
Needs storyboard panels Prepare the panel plan here before image-batch-runner
Needs execution with normalized request image-batch-runner
Final output is video video-batch-runner
Final output is audio audio-generation

Output Shape

Return:

  • taskClass
  • inputType
  • modelRule
  • referencePolicy
  • requiredRunnerInputs
  • handoffSkill
  • mustNotDo

Stop Conditions

  • Stop when required user intent, source evidence, or owned input artifacts are missing and guessing would change the result.
  • Do not let image-batch-runner make creative classification decisions.

Public Command Boundary

  • Choose the smallest matching command or workflow from the user input and run it directly.

  • This public skill is instruction-driven. Produce the controller handoff artifact directly from the available evidence.

  • Do not call private provider/runtime paths or unpublished local tools.

  • If the CLI returns a quote-confirmation challenge, obtain user approval for its scope and cost before running postplus quote confirm --json --challenge-file <challenge.json> and retry with the returned token.

Related skills

ai-image-generationgenmedia-labs713KGenerate and edit images on RunComfy via the `runcomfy` CLI — a smart router across the full image-model catalog: FLUX 2 (Klein 9B/4B, Pro, Dev, Flash, Turbo, Max), Google Nano Banana 2 / Pro, OpenAI GPT Image 2, ByteDance Seedream 5 / 4-5 / 4-0 and Dreamina 4-0, Alibaba Qwen Image and Z-Image Turbo, Wan 2-7. Covers both text-to-image (t2i) and image-to-image / edit (i2i) endpoints — the skill picks the right model for the user's actual intent (typography precision, photoreal portraits, sub-secoai-image-generation101-skills547KGenerate AI images with GPT-Image-2, FLUX, Gemini, Grok, Seedream, Reve and 50+ models via inference.sh CLI. Models: GPT-Image-2, FLUX Dev LoRA, FLUX.2 Klein LoRA, Gemini 3 Pro Image, Grok Imagine, Seedream 4.5, Reve, ImagineArt. Capabilities: text-to-image, image-to-image, inpainting, LoRA, image editing, upscaling, text rendering. Use for: AI art, product mockups, concept art, social media graphics, marketing visuals, illustrations. Triggers: flux, image generation, ai image, text to image, stnano-banana-2prime-skills424KGenerate images with Google Nano Banana 2 (Gemini-family flash-tier text-to-image) on RunComfy — bundled with the model's documented prompting patterns so the skill gets sharper output than naive prompting against the same model. Documents Nano Banana 2's strengths (rapid iteration, in-image typography rendering, predictable framing, optional web-grounded context), the resolution-tier pricing, the safety-tolerance dial, and when to route to Nano Banana Pro / GPT Image 2 / Flux 2 / Seedream insteimage-editprime-skills424KEdit images on RunComfy — this skill is a smart router that matches the user's intent to the right edit model in the RunComfy catalog. Picks Nano Banana Edit (batch up to 20, identity-preserving default), OpenAI GPT Image 2 Edit (multilingual in-image text rewrite, multi-ref composition, layout precision), Flux Kontext Pro (single-ref high-fidelity local edit), or Z-Image Turbo Inpaint (mask-driven precise region edit). Bundles each model's documented prompting patterns so the skill gets sharper

Search skills and MCP servers

Fuzzy search across 23,137 skills and servers