Back to skills

Garden Skills: GPT Image 2

GPT Image 2 generation and editing with ~80 structured templates, adapting to whatever tools your agent has.

ConardLi

Designers / Developers / Content creators

Image promptingLicence: MITGitHub stars13k
GitHub card for ConardLi/garden-skills

Suggested prompts

Written by PicGens as starting points. Click one to copy it, then paste it to your agent once the skill is installed.

About this skill

This GPT Image 2 skill from ConardLi’s Garden Skills collection adapts to its environment. It detects one of three modes: a local mode that calls an OpenAI-compatible image API itself and saves prompts and images to a project folder; a host-native mode that hands the finished prompt to the agent’s own image tool; and an advisor mode that only writes prompts when no image tool is available.

The prompt library is broad — eighteen visual categories and around eighty structured templates covering posters, UI mockups, product shots, infographics, academic figures, technical diagrams, comics, avatars, storyboards and brand boards — with dedicated editing workflows and helper scripts. Other skills in the same collection cover web design, video presentations and article formatting.

What it makes

Generated or edited images, or structured GPT Image 2 prompts.

Before you install: For the local mode, an OpenAI-compatible image API key.

From the project README

Reproduced from ConardLi/garden-skills by ConardLi under the MIT licence; shortened here — read the rest on GitHub.

GPT Image 2 Skill

A focused image-generation / editing skill for GPT Image 2, with a single SKILL definition that adapts to three runtime modes — local generation, host-native delegation, and pure prompt advisor.

中文文档 · Back to collection root

GPT Image 2 Skill


What it does

This skill is a structured prompt-engineering and image-generation pack built around the GPT Image 2 model (and OpenAI-compatible image endpoints). It only does two image tasks — POST /images/generations and POST /images/edits — but it does them in three different runtime environments without changing user-facing behavior.

It bundles:

  • A mode-aware workflow so the same skill works whether the agent itself owns the image API key, the host has its own image tool, or there is no image tool at all.
  • A structured template library of 18 categories and 79 prompt templates covering posters, UI mockups, product visuals, infographics, academic figures, technical diagrams, comics, avatars, and editing workflows.
  • Reproducible prompt + image archival under garden-gpt-image-2/prompt/ and garden-gpt-image-2/image/ with task-slug + timestamp naming.

The three runtime modes

The very first thing this skill does on any task is run a tiny detection script:

node skills/gpt-image-2/scripts/check-mode.js
# or for structured output:
node skills/gpt-image-2/scripts/check-mode.js --json

The output picks one of three modes:

Mode Trigger Behavior
A — Garden local ENABLE_GARDEN_IMAGEGEN truthy AND OPENAI_API_KEY present End-to-end: pick template → render prompt → call generate.js / edit.js → image lands on disk
B — Host-native Garden disabled, but the host agent already has an image tool (image_generation, dalle, nano_banana, image MCP, etc.) Render the prompt, then delegate image generation to the host's own tool
C — Advisor Garden disabled, host has no image tool Skill degrades into a high-quality prompt writer — saves the rendered prompt to garden-gpt-image-2/prompt/ and instructs the user to paste it into ChatGPT / Midjourney / DALL·E / Sora / Nano Banana / their own gateway

In all three modes, prompt files are saved (mode A & C must save, mode B is recommended for reuse). Only mode A produces an image file; mode B leaves that to the host, mode C cannot.


Quick start

0. Detect the mode (always step 0)
node skills/gpt-image-2/scripts/check-mode.js

The commands below (1–4) only apply in Mode A.

1. Text-to-image
node skills/gpt-image-2/scripts/generate.js \
  --prompt "A cute baby sea otter" \
  --size 1024x1024 \
  --quality high
2. Generate from a saved prompt file
node skills/gpt-image-2/scripts/generate.js \
  --promptfile garden-gpt-image-2/prompt/poster-20260424-153045.md
3. Edit an existing image
node skills/gpt-image-2/scripts/edit.js \
  --image assets/source.png \
  --prompt "Replace the background with a clean studio scene"
4. Mask-based local edit
node skills/gpt-image-2/scripts/edit.js \
  --image assets/source.png \
  --mask  assets/mask.png \
  --prompt "Replace only the masked area with a glass vase"

For Mode B / C there is no CLI entry point — the skill just renders the final prompt and either hands it to the host's image tool (B) or shows it to the user (C).


Case Gallery

The public case library covers 18 categories, 79 templates, and 160+ generated / edited results. This gallery is a curated map of the most important capability families: each thumbnail opens the live case page, while the image itself is served from the dedicated ConardLi/gpt-image-2-101 case repository.

UI Mockups
Live commerce UI case
live-commerce-ui
Celebrity livestream commerce interface.
Social interface mockup case
social-interface-mockup
Official product announcement in a social feed.
Product card overlay case
product-card-overlay
Skincare landing-page hero with product, model, and badges.
Chat interface scene case
chat-interface-scene
Claude-style assistant screenshot with structured conversation.
Product And Branding
Exploded view poster case
exploded-view-poster
Vision Pro 2 optical and compute-module teardown.
Premium studio product case
premium-studio-product
Luxury skincare still life for editorial product pages.
Cosmetic packaging case
cosmetic-packaging
Premium skincare gift box with material polish.
Beverage label design case
beverage-label-design
Guochao sparkling-water bottle label and commercial scene.
Editing Workflows
Background replacement case
background-replacement
Portrait moved into Times Square night ambience.
Object removal case
object-removal
Remove unwanted people from a graduation group photo.
Product retouching case
product-retouching
Commerce-grade AirPods product cleanup.
Portrait local edit case
portrait-local-edit
Hair color and style edit while preserving identity.
Infographics And Visual Docs
Bento grid infographic case
bento-grid-infographic
iPhone 16 Pro feature breakdown in a compact grid.
Comparison infographic case
comparison-infographic
Phone comparison designed for decision support.
Dense explainer slide case
dense-explainer-slides
One-page AI Agent mechanism explainer.
Visual report page case
visual-report-page
Business summary page with KPI cards and chart rhythm.
Academic And Technical
Method pipeline overview case
method-pipeline-overview
RAG-based long-context QA pipeline for papers.
Neural network architecture case
neural-network-architecture
ViT-B/16 architecture figure with tensor flow.
System architecture case
system-architecture
Multi-tenant AI SaaS production architecture.
Sequence diagram case
sequence-diagram
OAuth 2.0 authorization code + PKCE sequence.
Story, Maps And Characters
Anime key visual case
anime-key-visual
Fantasy game launch key visual with crop-safe layout.
Food map case
food-map
Shanghai city-walk food map with illustrated landmarks.
Travel route map case
travel-route-map
Kyoto three-day route map with illustrated stops.
Professional portrait case
professional-portrait
Restrained executive portrait for company and media pages.

How to install Garden Skills: GPT Image 2

  1. 1.Skills CLI (whole collection)

    npx skills add ConardLi/garden-skills
  2. 2.Just this skill

    npx skills add ConardLi/garden-skills -s gpt-image-2

Licence and attribution

Garden Skills: GPT Image 2 is created and maintained by ConardLi and published at github.com/ConardLi/garden-skills under the MIT licence. The summary, suggested prompts and “About this skill” text are written by PicGens; the README section is reproduced under the project’s licence. The preview is GitHub’s own social card for the repository. PicGens is not affiliated with the author.

Related skills

All skills →
  • Example output of the GPT Image 2 Skill & Gallery agent skill

    GPT Image 2 Skill & Gallery

    Get GPT Image 2 and 2.5 images generated and edited from your agent, guided by a gallery of proven prompts.

    wuyoscarGitHub stars: 5.6k
  • Example output of the GPT Image 2 E-commerce Images agent skill

    GPT Image 2 E-commerce Images

    Get marketplace-ready product images — hero, lifestyle, A+ modules — from your product photos.

    buluslanGitHub stars: 394
  • GitHub card for anthropics/skills

    Canvas Design

    Get an original poster or art print drawn in code, plus the design philosophy behind it.

    anthropicsGitHub stars: 179k
  • GitHub card for anthropics/skills

    Theme Factory

    Get every deck, doc and page in a project wearing the same palette and font pairing.

    anthropicsGitHub stars: 179k

Looking for prompts instead of a skill?

Browse real gpt image 2 prompt templates with the images they made, or give your agent the free PicGens template skill.