重要(Desktop / 已注入 Host 工具时)
SkillMediaUse this skill for drawing, image generation, or image editing only when the Host tools image_generation / image_edit are not available. If OpenDrSai Desktop has injected the above Host tools, do not load this skill; you must call image_generation or image_edit directly.
Instructions available. Your AI can read the instructions. Execution depends on the setup they require.
Account requirements not reviewed. Check the skill instructions before use; ahel provides instructions and does not run this skill.
Add ahel to your AI once: Claude, ChatGPT, Cursor, Claude Code or Codex. Then ask it to use this.
Then ask your AI: use the 重要(Desktop / 已注入 Host 工具时) skill
What this skill tells your AI
The instructions your AI receives, as published by hepai-lab/drsai in skills/skills_hepai/image-process/SKILL.md and read by ahel’s review.
若当前工具列表中已有 image_generation 或 image_edit:
- 不要加载本技能
- 不要运行下面的 Python 脚本或自写 HTTP 调用
- 直接调用 Host 工具
image_generation/image_edit(模型与凭证由 Host / Agent 设置处理)
只有在确认工具列表中没有这两个 Host 工具时,才使用下文脚本路径。
使用 scripts/image_generation_openai.py 进行图像生成(仅无 Host 工具时)。
快速使用示例
python image_generation_openai.py --prompt "一只可爱的海獭"
python image_generation_openai.py -p "日落海边" -n 2 -s 1024x1792
python image_generation_openai.py -p "科技风格logo" --quality high
查看可用模型
运行以下命令查看所有可用的图像生成模型:
python scripts/list_models.py
主要支持的图像生成模型
OpenAI 系列
openai/gpt-image-1- 基础图像生成模型openai/gpt-image-1-mini- 轻量版本openai/gpt-image-1.5- 增强版本openai/gpt-image-2- 最新版本(默认)openai/chatgpt-image-latest- 最新ChatGPT图像openai/dall-e-3- DALL-E 3
字节跳动豆包系列
bytedance/doubao-seedream-5-0-260128- 豆包5.0图像生成bytedance/doubao-seedream-4-5-251128- 豆包4.5图像生成bytedance/doubao-seedream-3-0-t2i-250415- 豆包3.0图像生成bytedance/doubao-seedream-5-0-lite-260128- 豆包5.0轻量版bytedance/doubao-seedance-2-0-260128- 豆包视频生成2.0
阿里云通义系列
aliyun/qwen-image-2.0- 通义2.0图像生成aliyun/qwen-image-2.0-pro- 通义专业版aliyun/qwen-image-2.0-max- 通义最大值aliyun/qwen-image-2.0-plus- 通义增强版aliyun/qwen-image-edit- 通义图像编辑aliyun/qwen-image-edit-plus- 通义图像编辑增强版
Google Gemini 系列
google/gemini-2.5-flash-image- Gemini 2.5 图像生成google/gemini-3-pro-image-preview- Gemini 3 图像预览
xAI Grok 系列
xAI/grok-2-image- Grok 2 图像生成xAI/grok-imagine-image- Grok 图像想象
使用不同模型生成图像
# 使用 OpenAI GPT-Image 1.5
python image_generation_openai.py -p "未来城市景观" -m "openai/gpt-image-1.5"
# 使用豆包5.0
python image_generation_openai.py -p "未来城市景观" -m "bytedance/doubao-seedream-5-0-260128"
# 使用通义专业版
python image_generation_openai.py -p "未来城市景观" -m "aliyun/qwen-image-2.0-pro"
API 信息
- 图像生成端点:
https://aiapi.ihep.ac.cn/apiv2/v1/images/generations - 文件上传端点:
https://aiapi.ihep.ac.cn/apiv2 - 环境变量:
HEPAI_API_KEY
输出格式
生成完成后,图像会自动上传到 HepAI 并提供可访问的 URL。
必须使用以下格式输出图像,不要加```等符号:
其他参数
-n: 生成图像数量 (1-10)-s: 图像尺寸 (1024x1024, 1536x1536, 1024x1792, 1792x1024)-q: 图像质量 (low, medium, high, auto)-b: 背景类型 (opaque, transparent)-o: 自定义输出文件路径--no-upload: 不上传到 HepAI(仅保存本地文件)
Signals
- GitHub stars
- 24
- Forks
- 5
- Last commit
- Sep 2026
Advanced
- Item type
- skill
- Key
image-process- Source
- github.com/hepai-lab/drsai