Token导航 LogoToken导航TokenDH.com
图像处理敏感数据github未标认证来源可访问许可证需确认审计提醒

image-prompt-generator图像提示生成器

Agent Skill

用于辅助图像生成、图片编辑、视觉素材处理或图像模型工作流。它适合让 Agent 根据文本生成图片、处理背景、整理视觉提示词或调用相关图像工具。使用时需要确认输入图片、版权来源、输出格式和模型限制;涉及人物、品牌、商品或公开展示素材时,应额外核对授权、真实性和内容合规边界。

总安装

1,105

周安装

47

GitHub Stars

9

下载量

387
CodexClaudeCursorGemini CLI

安装说明

本站只整理中文说明和来源信息,不托管安装包,也不代用户安装。

GitHub

来源数

2

许可证

unknown

最后核验

2026-05-01

来源状态

来源可访问

安装方式

通过对话安装

复制提示词发给支持本地命令或 Skills 的 AI 助手,先确认命令和权限,再让它执行。

请帮我安装这个 Agent Skill:image-prompt-generator(图像提示生成器)
来源仓库:https://github.com/cdeistopened/skill-stack
仓库路径:skills/image-prompt-generator
安装命令:
npx skills add https://github.com/cdeistopened/skill-stack --skill image-prompt-generator
安装前请先检查当前环境是否支持对应 CLI,并向我确认将要执行的命令、安装目录、联网范围和文件读写权限;确认后再执行。

命令行安装

复制命令到本机终端执行。该命令会通过 npx skills 从第三方来源获取 Skill;本站只展示命令,不托管安装包,也不自动执行。

skills.shnpx skills
npx skills add https://github.com/cdeistopened/skill-stack --skill image-prompt-generator

简介

基于 Gemini API 生成专业级图像提示词,支持多种模型和参数配置。

  • 适用于图像生成、视觉素材处理和 AI 作画辅助场景。
  • 使用时需配置 API 密钥,确认输出格式和版权授权边界。
  • 通过 GitHub 安装,支持主流宿主环境,建议核对人物和品牌素材合规性。
  • image-prompt-generator 属于图像处理类 Skill,可作为该场景下的辅助能力补充。

SKILL.md

Image Prompt Generator

Generate professional, non-generic images using Google's Gemini API for image generation.

Prerequisites & Setup

Getting Your Gemini API Key

  1. Go to Google AI Studio
  2. Sign in with your Google account
  3. Click "Create API Key"
  4. Copy the generated key

Configuring the API Key

Option 1: Environment file (recommended)

Create a .env file in your project root:

GEMINI_API_KEY=your_api_key_here

Option 2: Direct environment variable

export GEMINI_API_KEY=your_api_key_here

Install Dependencies

pip install google-generativeai python-dotenv pillow

Available Models

ModelAPI NameBest For
Flashgemini-2.5-flash-imageSpeed, drafts, iteration
Progemini-3-pro-image-previewFinal assets, 16:9 aspect ratio, quality

Image Size (CRITICAL for quality)

SizeOutputUse For
1K~400KBDrafts, iteration
2K~2MBThumbnails, web assets (default)
4K~7MBPrint, high-resolution needs

CRITICAL: Always specify --size 2K or --size 4K for final assets. Without this, output defaults to 1K which looks low-quality.

Skill Stack Thumbnails - MANDATORY SETTINGS

For ANY Skill Stack thumbnail, you MUST use:

python scripts/images/generate_image.py "YOUR PROMPT" \
  --model pro \
  --size 2K \
  --aspect 16:9 \
  --output public/images/thumbnails \
  --name your-slug

And your prompt MUST include risograph style instructions:

STYLE: Risograph / screen print aesthetic. Visible halftone dots throughout.
Slight misregistration between color layers. Indie printmaking quality - warm, tactile, handmade.

COLOR: Warm cream base. Charcoal and gray tones. Terracotta (#D4654A) accent on ONE key element.

TEXTURE: Visible halftone pattern. Paper grain. NOT smooth digital gradients.

AVOID: Empty white borders, photorealism, glossy renders, smooth gradients, clip art aesthetic.

WHY: Without --size 2K, output is ~400KB garbage that looks like it was made with Flash. Without risograph style, output looks generic and AI-generated.


Workflow Overview

  1. Brainstorm Concepts - Generate 4-6 high-level visual ideas
  2. Select Direction - User picks the concept they like
  3. Optimize Prompt - Refine into a strong, detailed prompt
  4. Style Variations - Adapt to 2-3 different visual styles
  5. Generate Images - Run via Gemini API

Step 1: Brainstorm Concepts

When the user provides a topic or use case, generate 3 high-level visual concepts and indicate your pick. Each concept should be:

  • One sentence describing the visual idea
  • Concrete and immediate - you can picture it instantly
  • Conceptual but not abstract - a clear object/scene with meaning
  • Non-generic - avoid cliches (no lightbulbs for ideas, no handshakes for partnership)

Format:

**Concept A:** One sentence description of the visual concept and why it works.

**Concept B:** One sentence description...

**Concept C:** One sentence description...

**My pick: Concept X** — Brief reasoning why this one is strongest.

Example for "newsletter about personal productivity":

**Concept A:** A vintage compass where the needle points toward a coffee ring stain on a map—direction emerges from daily rituals.

**Concept B:** A clock where the 12 hours show seasonal changes—time management over long arcs, not just hours.

**Concept C:** One small key attached to dozens of decorative keychains—we overcomplicate simple solutions.

**My pick: Concept A** — The coffee stain is unexpected and personal, the compass is concrete without being cliché.

Wait for user to confirm or redirect before proceeding.

Step 2: Optimize the Prompt

Once the user selects a concept, develop it into a full prompt. Structure:

Create a [style type] illustration of [subject].

CONCEPT: [Expand the one-sentence idea into a clear visual description]

STYLE: [Artistic approach - load from references/styles/ if brand-specific]

COMPOSITION: [Framing, focal point, negative space, balance]

COLORS: [Palette - describe by name, not hex codes which may render as text]

TEXTURE: [Surface qualities, analog/digital feel]

AVOID: [What should NOT appear - be specific]

FORMAT: [Aspect ratio]

Key principles:

  • Natural language, full sentences - no tag soup
  • Describe colors by name (burnt orange, sky blue, near-black) not hex codes
  • Maximum 2-3 elements - if it feels busy, remove something
  • Favor metaphor over literal depiction

Step 3: Style Variations

MANDATORY for Skill Stack: Risograph style. Every thumbnail must include risograph style instructions in the prompt. See references/styles/risograph.md.

Key risograph elements to ALWAYS include in prompt:

  • "Visible halftone dots throughout"
  • "Slight misregistration between color layers"
  • "Indie printmaking quality - warm, tactile, handmade"
  • "NOT smooth digital gradients"

Available styles in references/styles/ (for non-Skill-Stack projects):

  • risograph.md - DEFAULT. Halftone dots, misregistration, indie printmaking aesthetic.
  • minimalist-ink.md - High-contrast black and white, crosshatching.
  • watercolor-line.md - Ink linework with watercolor washes.
  • editorial-conceptual.md - Conceptual, sophisticated, editorial wit.

Step 4: Generate via API

Running the Script

# Generate 2K thumbnail (recommended for web)
python scripts/images/generate_image.py "prompt here" --model pro --aspect 16:9 --size 2K

# Generate 4K for print/high-res
python scripts/images/generate_image.py "prompt here" --model pro --size 4K

# Save to specific folder
python scripts/images/generate_image.py "prompt" --output "./images" --name "my_image" --size 2K

Options:

  • --model pro (default, higher quality) or --model flash (faster)
  • --size 2K (default), 1K, or 4K - always use 2K or 4K for final assets
  • --aspect 16:9 (default), 1:1, 9:16, 3:4, 4:3
  • --variations N - generate N versions
  • --output./path - save location
  • --name prefix - filename prefix

Output location: Save images alongside the content they belong to - not a generic images dump.

Step 5: Iterate

After user reviews generated images:

  • 80% good? Request specific edits conversationally rather than regenerating
  • Composition off? Adjust framing or element placement in prompt
  • Wrong style? Try a different style reference
  • Too busy? Simplify to fewer elements
  • Colors wrong? Be more explicit about palette

Prompting Principles

Write Like a Creative Director

Brief the model like a human artist. Use proper grammar, full sentences, and descriptive adjectives.

Don'tDo
"Cool car, neon, city, night, 8k""A cinematic wide shot of a futuristic sports car speeding through a rainy Tokyo street at night. The neon signs reflect off the wet pavement and the car's metallic chassis."

Be specific about:

  • Subject: Instead of "a woman," say "a sophisticated elderly woman wearing a vintage chanel-style suit"
  • Materiality: Describe textures - "matte finish," "brushed steel," "soft velvet," "crumpled paper"
  • Setting: Define location, time of day, weather
  • Lighting: Specify mood and light source
  • Mood: Emotional tone of the image

Provide Context

Context helps the model make logical artistic decisions. Include the "why" or "for whom."

Example: "Create an image of a sandwich for a Brazilian high-end gourmet cookbook." *(Model infers: professional plating, shallow depth of field, perfect lighting)*

Keep It Simple

  • One clear focal point
  • Maximum 2-3 elements total
  • Generous negative space
  • If it feels busy, remove something

Avoid the Generic

Hard rules:

  • NEVER include text in images - Text renders poorly and looks amateur
  • No lightbulbs for "ideas"
  • No handshakes for "partnership"
  • No audio waveforms for "podcasts" or "audio"
  • No gears/cogs for "systems" or "process"
  • No happy stock photo poses
  • No glossy AI aesthetic
  • No puzzle pieces for "connection" or "integration"

Metaphors That Work

Strong concepts from past generations:

  • Orange cut open revealing neural network - organic exterior, systematic interior (knowledge systems, RAG)
  • Mail slot with origami bird emerging - delivery transforms into something alive (newsletters)
  • Exercise club crossed with quill pen - physical culture meets intellectual work (body + mind)
  • Typewriter key extreme close-up with radiating impact - decisive action, command (books, authority)
  • Medicine dropper with mandala bloom - clinical precision meets transcendence (therapy, transformation)
  • Straight razor on leather strop - precision tool, polishing, refinement (editing, production)
  • Prism dispersing light into distinct beams - one input, multiple outputs (frameworks)
  • Compass with coffee stain on map - direction from daily rituals (productivity)

Resources

references/styles/

Aesthetic style definitions:

  • risograph.md - DEFAULT - Halftone, misregistration, indie printmaking
  • minimalist-ink.md - Black and white ink illustration
  • watercolor-line.md - Ink with watercolor washes
  • editorial-conceptual.md - Conceptual editorial style

scripts/

  • generate_image.py - Gemini API image generation

Prompt Modifiers Reference

CategoryExamples
Lightinggolden hour, dramatic shadows, soft diffused light, neon glow, overcast
Stylecinematic, editorial, technical diagram, hand-drawn, photorealistic
Texturematte finish, brushed steel, soft velvet, crumpled paper, weathered wood
Compositionwide shot, close-up, bird's eye view, dutch angle, symmetrical
Moodenergetic, serene, dramatic, playful, sophisticated
Quality4K, high-fidelity, pixel-perfect, professional grade

Advanced Capabilities

Text Rendering & Infographics

Put exact text in quotes. Specify style: "polished editorial," "technical diagram," or "hand-drawn whiteboard."

Example prompts:

Earnings Report Infographic:
"Generate a clean, modern infographic summarizing the key financial highlights from this earnings report. Include charts for 'Revenue Growth' and 'Net Income', and highlight the CEO's key quote in a stylized pull-quote box."
Whiteboard Summary:
"Summarize the concept of 'Transformer Neural Network Architecture' as a hand-drawn whiteboard diagram suitable for a university lecture. Use different colored markers for the Encoder and Decoder blocks, and include legible labels for 'Self-Attention' and 'Feed Forward'."

Character Consistency & Thumbnails

Use reference images and state "Keep the person's facial features exactly the same as Image 1." Describe expression/action changes while maintaining identity.

Example prompt:

Viral Thumbnail:
"Design a viral video thumbnail using the person from Image 1.
Face Consistency: Keep the person's facial features exactly the same as Image 1, but change their expression to look excited and surprised.
Action: Pose the person on the left side, pointing their finger towards the right side of the frame.
Subject: On the right side, place a high-quality image of a delicious avocado toast.
Graphics: Add a bold yellow arrow connecting the person's finger to the toast.
Text: Overlay massive, pop-style text in the middle: 'Done in 3 mins!'. Use a thick white outline and drop shadow.
Background: A blurred, bright kitchen background. High saturation and contrast."

Image Reworking (Edit Existing Images)

The --input flag enables "rework mode" - pass an existing image to Gemini and describe the changes you want.

Key use cases:

  • Small tweaks - Adjust colors, add/remove elements, change lighting
  • Style transfer - Keep composition but change artistic style
  • Object manipulation - Remove, add, or modify specific objects
  • Seasonal/temporal changes - Same scene, different time/season

Running in rework mode:

# Basic edit - add something
python scripts/generate_image.py "Add snow to the roof and yard" \
  --input ./house.png \
  --model pro

# Color adjustment
python scripts/generate_image.py "Change the accent color from red to teal, keep everything else identical" \
  --input ./thumbnail.png \
  --model pro

# Style transfer - keep composition, change aesthetic
python scripts/generate_image.py "Convert this to risograph style with halftone dots and slight color misregistration" \
  --input ./photo.png \
  --model pro

# Generate variations of an edit
python scripts/generate_image.py "Make the lighting warmer, like golden hour" \
  --input ./portrait.png \
  --variations 3 \
  --model pro

Prompting tips for rework mode:

  1. Be specific about what to preserve:

- "Keep the person's facial features exactly the same" - "Maintain the composition and framing" - "Don't change the background"

  1. Be explicit about what to change:

- "Change ONLY the color of the shirt from blue to red" - "Add snow to the roof and nothing else" - "Remove the text overlay"

  1. Use comparative language:

- "Make the colors more vibrant" - "Increase the contrast slightly" - "Make the lighting softer and more diffused"

Output naming: Files from rework mode are named {prefix}_{timestamp}_edit_{model}.png to distinguish from generated images (_gen_).

Advanced Editing Examples

Object Removal:

python scripts/generate_image.py \
  "Remove the tourists from the background and fill with matching cobblestones and storefronts" \
  --input ./street-photo.png \
  --model pro

Seasonal Control:

python scripts/generate_image.py \
  "Turn this into winter. Add snow to the roof and yard. Change lighting to cold, overcast afternoon. Keep architecture identical." \
  --input ./house-summer.png \
  --model pro

Character Consistency (thumbnail series):

python scripts/generate_image.py \
  "Keep the person's face exactly the same. Change expression to surprised. Add a pointing gesture toward the right side of the frame." \
  --input ./person-reference.png \
  --model pro

Related Skills

  • youtube-title-creator - Pair generated images with optimized titles
  • social-content-creation - Use images in platform-optimized posts

*For custom brand styles, create new style files in references/styles/ following the existing format*

适合场景

01

用户想查找某类 Agent Skill 时

02

需要根据任务场景推荐可安装能力包时

03

需要对比不同来源的安装命令和来源信息时

能力概览

能力 1

按任务关键词查找相关 Skills

能力 2

展示可复制的安装命令

能力 3

保留来源站点、仓库和原始说明,方便继续核验

能力 4

展示第三方安全扫描或审计结果

安装后应在对应宿主中按原始 README 的触发条件使用;具体调用方式请以来源页面和 README 为准。

平台分布

Codex

34.92%
按下载量换算135

Claude

28.85%
按下载量换算112

Cursor

17.93%
按下载量换算69

Gemini CLI

8.44%
按下载量换算33

安全审计

Gen Agent Trust Hub

通过

Socket

通过

Snyk

可疑

权限和风险

敏感数据

该 Skill 可能接触密钥、Token、环境变量或敏感配置,应进入高风险复核队列,默认不自动发布。

安装前确认

本站仅展示第三方公开信息,不托管安装包,不提供自动安装或运行环境。安装前应自行审查源码、依赖和命令行为。来源安全扫描存在 warning/failed 结果,不能写成本站确认安全。当前只有一个来源,正式发布前建议补源仓库或其他目录站核验。

来源信息

继续浏览同类 Skills