by shanliuling
AI image studio for DeepSeek Harness — generate, edit & compare images in chat, with 500+ prompts, gallery, multi-model workflows and ComfyUI.
# Add to your Claude Code skills
git clone https://github.com/shanliuling/dsh-image-genGuides for using ai agents skills like dsh-image-gen.
Last scanned: 8/19/2026
{
"issues": [],
"status": "PASSED",
"scannedAt": "2026-08-19T04:36:42.162Z",
"npmAuditRan": false,
"pipAuditRan": true,
"promptInjectionRan": true
}See how dsh-image-gen compares with popular alternatives.
dsh-image-gen is an open-source ai agents skill for AI coding assistants such as Claude Code, Codex CLI, and ChatGPT, built by shanliuling. AI image studio for DeepSeek Harness — generate, edit & compare images in chat, with 500+ prompts, gallery, multi-model workflows and ComfyUI. It has 558 GitHub stars.
Yes. dsh-image-gen passed SkillsLLM's automated security scan — a dependency vulnerability audit plus prompt-injection heuristics — with no high-severity issues. You can read the full report in the Security Report section on this page.
Clone the repository with "git clone https://github.com/shanliuling/dsh-image-gen" and add it to your Claude Code skills directory (see the Installation section above).
dsh-image-gen is primarily written in TypeScript. It is open-source under shanliuling on GitHub, so you can review or fork the full source.
Yes. SkillsLLM lists many other AI Agents skills you can browse and compare side by side. Open the AI Agents category from the badge at the top of this page, or use the Related Skills and comparison links further down to weigh dsh-image-gen against similar tools.
No comments yet. Be the first to share your thoughts!
⚠️ Third-Party Software Notice
This skill is third-party open-source software developed and hosted independently on GitHub. SkillsLLM is an informational directory and does not control or maintain the underlying repository.
Any security checks, ratings, or warnings displayed by SkillsLLM are automated and limited in scope. They do not constitute a security certification or guarantee that the software is safe, error-free, or free from malicious code, vulnerabilities, compromised dependencies, or prompt-injection risks.
Review the source code, permissions, dependencies, and configuration before installing or running any third-party skill. Use is at your own risk. To the maximum extent permitted by applicable law, SkillsLLM is not liable for losses arising from third-party software.
为 DeepSeek Harness 带来完整的 AI 图像创作工作流。
dsh-image-gen 不只是简单的对话生图,而是为 DSH 补齐从 对话生成与连续修图、AI 创作画布、Studio 批量创作、多模型横向对比,到 Prompt 灵感库 与 本地 ComfyUI 的完整图像创作能力。
支持主流云端图像模型与本地私有化工作流,既可使用 BYOK(自带 Key),也支持通过订阅账号直接使用,生成结果支持按工作区隔离存储。
已有 ChatGPT、Grok 或 Google 订阅?直接登录即可生图,无需额外购买 API Key。
支持:Gemini · OpenAI / Compatible · Seedream · DashScope · Grok Imagine · GLM-Image · 本地 ComfyUI
pnpm dsh plugin --profile web add dsh-image-gen@latest
版本更新提示: 本次版本变化较大,老用户请更新至最新版本。
| 入口 | 最适合 | 你可以做什么 |
|---|---|---|
| 💬 对话 | 快速表达想法 | 文生图、图生图、连续编辑、版本迭代 |
| ✏️ 画布 | 表达视觉创意 | 草稿生成、参考图组合、空间创作、持续修改 |
| 🎛️ 工作台 | 精细控制创作参数 | 批量生成、多图参考、高级参数调整 |
| ✨ 灵感 | 寻找创作方向 | Prompt 案例、风格探索、一键复用 |
| 🖼️ 图库 | 管理生成结果 | 搜索、收藏、下载、重新使用 |
环境要求:DeepSeek Harness 稳定版本,Node.js ^22.19.0 或 >= 24.0.0。
在你的 DeepSeek Harness 项目根目录下运行:
pnpm dsh plugin --profile web add dsh-image-gen@latest
💬 极客提示:你也可以直接把这句话发送给 DSH 对话中的 Agent:
帮我安装生图插件,在终端执行:pnpm dsh plugin --profile web add dsh-image-gen@latest
# 若已将 dsh 安装为系统全局命令:
dsh plugin --profile web add dsh-image-gen@latest
# 从 GitHub 仓库直接安装最新代码:
pnpm dsh plugin --profile web add git+https://github.com/shanliuling/dsh-image-gen.git
# 本地克隆源码开发安装:
git clone https://github.com/shanliuling/dsh-image-gen.git
pnpm dsh plugin --profile web add ./dsh-image-gen
重启 DSH 后进入:
设置 → 插件 → 图像生成
DSH 0.1.5 及更早版本的入口为「设置 → 插件 → 插件配置 → 图像生成」,插件已同时兼容两种入口。
选择 Provider,填写自己的 API Key,并按需调整模型、Endpoint / Base URL 与工作区保存选项。填好 Key 后可点击**「测试连接」验证可用性,或点击「拉取模型」**一键获取该厂商支持的全部生图模型,无需手动查文档。使用 ComfyUI 时,请填写 DSH Host 可访问的服务地址,并导入 API Format Workflow JSON。
已有 ChatGPT、Grok 或 Google 订阅?无需填写 API Key:展开对应的订阅 Provider 行,点击**「登录」**并在浏览器完成授权,即可直接开始文生图与图生图。
在聊天框中直接描述你想要的图片:
画一张雨夜霓虹街头的赛博朋克猫咪,电影感光线,16:9。
也可以直接上传参考图,让 Agent 进行风格重构或局部编辑:
保持角色与构图不变,给猫咪戴上一副黑色墨镜。
需要更细的参数控制时,点击会话顶部的 画廊 入口,进入 图库 / 工作台 / 灵感 / 收藏。
在无限画布中表达你的创意,通过对话将草稿、构图和想法转化为真实图片。
使用同一组 Prompt 和参考图并发调用多个模型,在一张画布中比较并保存结果。
灵活调用本地 GPU 算力,让私有化绘图无缝融入 Agent 对话。
ComfyUI 暂未接入 Studio 和多模型对比。
| Provider | 对话生图 | 对话编辑 | Studio | 多模型对比 |
|---|---|---|---|---|
| Google Gemini | ✅ | ✅ 多图 | ✅ | ✅ |
| OpenAI Images | ✅ | ✅ 多图 | ✅ | ✅ |
| OpenAI Compatible(中转站) | ✅ | ✅ 多图 | ✅ | ✅ |
| ByteDance Seedream / 火山方舟 | ✅ | ✅ 多图 | ✅ | ✅ |
| Aliyun DashScope / Qwen Image | ✅ | ✅ 多图 | ✅ | ✅ |
| xAI Grok Imagine | ✅ | ⚠️ 有限 | ✅ | ✅ |
| 智谱 GLM-Image | ✅ | — | ✅ | ✅ |
| Local ComfyUI | ✅ | ✅ 单图 | — | — |
| ChatGPT 订阅(免 Key) | ✅ | ✅ 多图 | ✅ | ✅ |
| Grok 订阅(免 Key) | ✅ | ✅ 多图 | ✅ | ✅ |
| Google 订阅(免 Key) | ✅ | ✅ 多图 | ✅ | ✅ |
Studio 与多模型对比目前只支持云端 Provider(含订阅通道);多模型对比调用的是各 Provider 在设置中已配置的模型。 智谱 GLM-Image 上游本身不支持图生图;xAI 图生图走 OpenAI 兼容协议(multipart),部分网关可能需等待后续适配。 订阅通道通过账号登录使用(免 API Key),支持对话文生图 / 图生图、Studio 批量生成与多模型对比;图生图走各订阅渠道的编辑接口(最多 5 张参考图),比例与清晰度可在 Studio 中直接选择。
| Provider | 默认模型 | 默认 Endpoint / Base URL |
|---|---|---|
| Google Gemini | gemini-3.1-flash-image |
https://generativelanguage.googleapis.com/v1beta/interactions |
| OpenAI Images | gpt-image-2 |
https://api.openai.com/v1 |
| OpenAI Compatible | 自定义 | 自定义 Base URL |
| ByteDance Seedream | doubao-seedream-5-0-260128 |
https://ark.cn-beijing.volces.com/api/v3 |
| Aliyun DashScope | qwen-image-3.0 |
https://dashscope.aliyuncs.com/api/v1 |
| xAI Grok Imagine | grok-imagine-image |
https://api.x.ai/v1 |
| 智谱 GLM-Image | glm-image |
https://open.bigmodel.cn/api/paas/v4 |
| Local ComfyUI | 用户导入的 API Workflow | http://127.0.0.1:8188 |
| ChatGPT 订阅 | gpt-image-2.5-flare(通道固定) |
账号登录,无需配置 |
| Grok 订阅 | grok-imagine-image-2.0(通道固定) |
账号登录,无需配置 |
| Google 订阅 | gemini-3.1-flash-image(通道固定) |
账号登录,无需配置 |