canghe-image-gen

AI image generation with OpenAI, Google, DashScope and Canghe APIs. Supports text-to-image, reference images, aspect ratios. Sequential by default; parallel generation available on request. Use when user asks to generate, create, or draw images.

By freestylefly · 585 installs

npx skills add freestylefly/canghe-skills --skill canghe-image-gen

Source repository · Upstream listing

Image Generation (AI SDK) Official API based image generation. Supports OpenAI, Google, DashScope (阿里通义万象), and Canghe providers. Script Directory Agent Execution : 1. SKILL DIR = this SKILL.md file's directory 2. Script path = ${SKILL DIR}/scripts/main.ts Preferences (EXTEND.md) Use Bash to check EXTEND.md existence (priority order): ┌──────────────────────────────────────────────────┬───────────────────┐ │ Path │ Location │ ├──────────────────────────────────────────────────┼───────────────────┤ │ .canghe skills/canghe image gen/EXTEND.md │ Project directory │ ├──────────────────────────────────────────────────┼───────────────────┤ │ $HOME/.canghe skills/canghe image gen/EXTEND.md │ User home │ └──────────────────────────────────────────────────┴───────────────────┘ ┌───────────┬───────────────────────────────────────────────────────────────────────────┐ │ Result │ Action │ ├───────────┼───────────────────────────────────────────────────────────────────────────┤ │ Found │ Read, parse, apply settings │ ├───────────┼───────────────────────────────────────────────────────────────────────────┤ │ Not found │ Use defaults │ └───────────┴───────────────────────────────────────────────────────────────────────────┘ EXTEND.md Supports : Default provider Default quality Default aspect ratio Default image size Default models Schema: references/config/preferences schema.md Usage Options Option Description prompt <text , p Prompt text promptfiles <files... Read prompt from files (concatenated) image <path Output image path (required) provider google\ openai\ dashscope\ canghe Force provider (default: google) model <id , m Model ID ( ref with OpenAI requires GPT Image model, e.g. gpt image 1.5 ) ar <ratio Aspect ratio (e.g., 16:9 , 1:1 , 4:3 ) size <WxH Size (e.g., 1024x1024 ) quality normal\ 2k Quality preset (default: 2k) imageSize 1K\ 2K\ 4K Image size for Google (default: from quality) ref <files... Reference images. Supported by Google multimodal, OpenAI edits (GPT Image models), and Canghe ( image url ). If provider omitted: Google first, then OpenAI, then Canghe n <count Number of images json JSON output Environment Variables Variable Description OPENAI API KEY OpenAI API key GOOGLE API KEY Google API key DASHSCOPE API KEY DashScope API key (阿里云) CANGHE API KEY Canghe API key OPENAI IMAGE MODEL OpenAI model override GOOGLE IMAGE MODEL Google model override DASHSCOPE IMAGE MODEL DashScope model override (default: z image turbo) CANGHE IMAGE MODEL Canghe model override (default: gemini 3 pro image preview) OPENAI BASE URL Custom OpenAI endpoint GOOGLE BASE URL Custom Google endpoint DASHSCOPE BASE URL Custom DashScope endpoint CANGHE BASE URL Custom Canghe endpoint (default: https://api.canghe.ai/v1 ) Load Priority : CLI args EXTEND.md env vars <cwd /.canghe skills/.env ~/.canghe skills/.env Provider Selection 1. ref provided + no provider → auto select Google first, then OpenAI, then Canghe 2. provider specified → use it (if ref , must be google or openai or canghe ) 3. Only one API key available → use that provider 4. Multiple available → default to Google Quality Presets Preset Google imageSize OpenAI Size Use Case normal 1K 1024px Quick previews 2k (default) 2K 2048px Covers, illustrations, infographics Google imageSize : Can be overridden with imageSize 1K 2K 4K Aspect Ratios Supported: 1:1 , 16:9 , 9:16 , 4:3 , 3:4 , 2.35:1 Google multimodal: uses imageConfig.aspectRatio Google Imagen: uses aspectRatio parameter OpenAI: maps to closest supported size Generation Mode Default : Sequential generation (one image at a time). This ensures stable output and easier debugging. Parallel Generation : Only use when user explicitly requests parallel/concurrent generation. Mode When to Use Sequential (default) Normal usage, single images, small batches Parallel User explicitly requests, large batches (10+) Parallel Settings (when requested): Setting Value Recommended concurrency 4 subagents Max concurrency 8 subagents Use case Large batch generation when user requests parallel Agent Implementation (parallel mode only): Error Handling Missing API key → error with setup instructions Generation failure → auto retry once Invalid aspect ratio → warning, proceed with default Reference images with unsupported provider/model → error with fix hint (switch to Google multimodal or OpenAI GPT Image edits) Extension Support Custom configurations via EXTEND.md. See Preferences section for paths and supported options.