重要(Desktop / 已注入 Host 工具时)

SkillMedia

Use this skill for drawing, image generation, or image editing only when the Host tools image_generation / image_edit are not available. If OpenDrSai Desktop has injected the above Host tools, do not load this skill; you must call image_generation or image_edit directly.

Instructions available. Your AI can read the instructions. Execution depends on the setup they require.

Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.

Then ask your AI: use the 重要(Desktop / 已注入 Host 工具时) skill

What this skill tells your AI

The instructions your AI receives, as published by hepai-lab/drsai in skills/skills_hepai/image-process/SKILL.md and read by ahel’s review.

若当前工具列表中已有 image_generation 或 image_edit:

  1. 不要加载本技能
  2. 不要运行下面的 Python 脚本或自写 HTTP 调用
  3. 直接调用 Host 工具 image_generation / image_edit(模型与凭证由 Host / Agent 设置处理)

只有在确认工具列表中没有这两个 Host 工具时,才使用下文脚本路径。


使用 scripts/image_generation_openai.py 进行图像生成(仅无 Host 工具时)。

快速使用示例

python image_generation_openai.py --prompt "一只可爱的海獭"
python image_generation_openai.py -p "日落海边" -n 2 -s 1024x1792
python image_generation_openai.py -p "科技风格logo" --quality high

查看可用模型

运行以下命令查看所有可用的图像生成模型:

python scripts/list_models.py

主要支持的图像生成模型

OpenAI 系列

  • openai/gpt-image-1 - 基础图像生成模型
  • openai/gpt-image-1-mini - 轻量版本
  • openai/gpt-image-1.5 - 增强版本
  • openai/gpt-image-2 - 最新版本(默认)
  • openai/chatgpt-image-latest - 最新ChatGPT图像
  • openai/dall-e-3 - DALL-E 3

字节跳动豆包系列

  • bytedance/doubao-seedream-5-0-260128 - 豆包5.0图像生成
  • bytedance/doubao-seedream-4-5-251128 - 豆包4.5图像生成
  • bytedance/doubao-seedream-3-0-t2i-250415 - 豆包3.0图像生成
  • bytedance/doubao-seedream-5-0-lite-260128 - 豆包5.0轻量版
  • bytedance/doubao-seedance-2-0-260128 - 豆包视频生成2.0

阿里云通义系列

  • aliyun/qwen-image-2.0 - 通义2.0图像生成
  • aliyun/qwen-image-2.0-pro - 通义专业版
  • aliyun/qwen-image-2.0-max - 通义最大值
  • aliyun/qwen-image-2.0-plus - 通义增强版
  • aliyun/qwen-image-edit - 通义图像编辑
  • aliyun/qwen-image-edit-plus - 通义图像编辑增强版

Google Gemini 系列

  • google/gemini-2.5-flash-image - Gemini 2.5 图像生成
  • google/gemini-3-pro-image-preview - Gemini 3 图像预览

xAI Grok 系列

  • xAI/grok-2-image - Grok 2 图像生成
  • xAI/grok-imagine-image - Grok 图像想象

使用不同模型生成图像

# 使用 OpenAI GPT-Image 1.5
python image_generation_openai.py -p "未来城市景观" -m "openai/gpt-image-1.5"

# 使用豆包5.0
python image_generation_openai.py -p "未来城市景观" -m "bytedance/doubao-seedream-5-0-260128"

# 使用通义专业版
python image_generation_openai.py -p "未来城市景观" -m "aliyun/qwen-image-2.0-pro"

API 信息

  • 图像生成端点: https://aiapi.ihep.ac.cn/apiv2/v1/images/generations
  • 文件上传端点: https://aiapi.ihep.ac.cn/apiv2
  • 环境变量: HEPAI_API_KEY

输出格式

生成完成后,图像会自动上传到 HepAI 并提供可访问的 URL。

必须使用以下格式输出图像,不要加```等符号:

其他参数

  • -n: 生成图像数量 (1-10)
  • -s: 图像尺寸 (1024x1024, 1536x1536, 1024x1792, 1792x1024)
  • -q: 图像质量 (low, medium, high, auto)
  • -b: 背景类型 (opaque, transparent)
  • -o: 自定义输出文件路径
  • --no-upload: 不上传到 HepAI(仅保存本地文件)

Signals

GitHub stars
24
Forks
5
Last commit
Sep 2026
Advanced
Item type
skill
Key
image-process
Source
github.com/hepai-lab/drsai