canghe-image-gen
AI image generation with OpenAI, Google, DashScope and Canghe APIs. Supports text-to-image, reference images, aspect ratios. Sequential by default; parallel generation available on request. Use when user asks to generate, create, or draw images.
By freestylefly · 585 installs
npx skills add freestylefly/canghe-skills --skill canghe-image-gen
Source repository · Upstream listing
Image Generation (AI SDK)
Official API based image generation. Supports OpenAI, Google, DashScope (阿里通义万象), and Canghe providers.
Script Directory
Agent Execution :
1. SKILL DIR = this SKILL.md file's directory
2. Script path = ${SKILL DIR}/scripts/main.ts
Preferences (EXTEND.md)
Use Bash to check EXTEND.md existence (priority order):
┌──────────────────────────────────────────────────┬───────────────────┐
│ Path │ Location │
├──────────────────────────────────────────────────┼───────────────────┤
│ .canghe skills/canghe image gen/EXTEND.md │ Project directory │
├──────────────────────────────────────────────────┼───────────────────┤
│ $HOME/.canghe skills/canghe image gen/EXTEND.md │ User home │
└──────────────────────────────────────────────────┴───────────────────┘
┌───────────┬───────────────────────────────────────────────────────────────────────────┐
│ Result │ Action │
├───────────┼───────────────────────────────────────────────────────────────────────────┤
│ Found │ Read, parse, apply settings │
├───────────┼───────────────────────────────────────────────────────────────────────────┤
│ Not found │ Use defaults │
└───────────┴───────────────────────────────────────────────────────────────────────────┘
EXTEND.md Supports : Default provider Default quality Default aspect ratio Default image size Default models
Schema: references/config/preferences schema.md
Usage
Options
Option Description
prompt <text , p Prompt text
promptfiles <files... Read prompt from files (concatenated)
image <path Output image path (required)
provider google\ openai\ dashscope\ canghe Force provider (default: google)
model <id , m Model ID ( ref with OpenAI requires GPT Image model, e.g. gpt image 1.5 )
ar <ratio Aspect ratio (e.g., 16:9 , 1:1 , 4:3 )
size <WxH Size (e.g., 1024x1024 )
quality normal\ 2k Quality preset (default: 2k)
imageSize 1K\ 2K\ 4K Image size for Google (default: from quality)
ref <files... Reference images. Supported by Google multimodal, OpenAI edits (GPT Image models), and Canghe ( image url ). If provider omitted: Google first, then OpenAI, then Canghe
n <count Number of images
json JSON output
Environment Variables
Variable Description
OPENAI API KEY OpenAI API key
GOOGLE API KEY Google API key
DASHSCOPE API KEY DashScope API key (阿里云)
CANGHE API KEY Canghe API key
OPENAI IMAGE MODEL OpenAI model override
GOOGLE IMAGE MODEL Google model override
DASHSCOPE IMAGE MODEL DashScope model override (default: z image turbo)
CANGHE IMAGE MODEL Canghe model override (default: gemini 3 pro image preview)
OPENAI BASE URL Custom OpenAI endpoint
GOOGLE BASE URL Custom Google endpoint
DASHSCOPE BASE URL Custom DashScope endpoint
CANGHE BASE URL Custom Canghe endpoint (default: https://api.canghe.ai/v1 )
Load Priority : CLI args EXTEND.md env vars <cwd /.canghe skills/.env ~/.canghe skills/.env
Provider Selection
1. ref provided + no provider → auto select Google first, then OpenAI, then Canghe
2. provider specified → use it (if ref , must be google or openai or canghe )
3. Only one API key available → use that provider
4. Multiple available → default to Google
Quality Presets
Preset Google imageSize OpenAI Size Use Case
normal 1K 1024px Quick previews
2k (default) 2K 2048px Covers, illustrations, infographics
Google imageSize : Can be overridden with imageSize 1K 2K 4K
Aspect Ratios
Supported: 1:1 , 16:9 , 9:16 , 4:3 , 3:4 , 2.35:1
Google multimodal: uses imageConfig.aspectRatio
Google Imagen: uses aspectRatio parameter
OpenAI: maps to closest supported size
Generation Mode
Default : Sequential generation (one image at a time). This ensures stable output and easier debugging.
Parallel Generation : Only use when user explicitly requests parallel/concurrent generation.
Mode When to Use
Sequential (default) Normal usage, single images, small batches
Parallel User explicitly requests, large batches (10+)
Parallel Settings (when requested):
Setting Value
Recommended concurrency 4 subagents
Max concurrency 8 subagents
Use case Large batch generation when user requests parallel
Agent Implementation (parallel mode only):
Error Handling
Missing API key → error with setup instructions
Generation failure → auto retry once
Invalid aspect ratio → warning, proceed with default
Reference images with unsupported provider/model → error with fix hint (switch to Google multimodal or OpenAI GPT Image edits)
Extension Support
Custom configurations via EXTEND.md. See Preferences section for paths and supported options.