clothing-detail
服装工艺细节放大图。服装图 → 面料纹理、走线、织法的微距特写。当用户说「细节图」「特写」「面料放大」「工艺展示」「近景细节」时使用。
Install
npx skills add https://github.com/dlazy-ai/ecommerce-skills/tree/main/skills/clothing-detail
claude plugin marketplace add https://llmmart.ai/marketplace.json && claude plugin install dlazy-ai-ecommerce-skills@llmmart
git clone https://github.com/dlazy-ai/ecommerce-skills.git
The skills CLI installs just this skill, for any of its supported agents. Claude Code installs the whole dlazy-ai/ecommerce-skills collection as a plugin from our marketplace. Git is the plain clone.
Skill manifest
clothing-detail — 服装图生成细节放大图
一张服装图 → 局部微距特写。详情页里「证明这件衣服做得好」的那几张图。
为什么需要:转化率高的详情页通常有 2~3 张细节图(领口、袖口、面料纹理),但拍微距要专门的镜头和布光。本技能从常规商品图推出这些特写。
生成效果示例
| 输入:服装图 |
|---|
![]() |
garment-flatlay.jpg — 军绿麻花针织毛衣平铺图,800×800 |
实际执行的命令:
dlazy gpt-image-2 \
--prompt 'Macro detail shot for an e-commerce detail page. Zoom into the shoulder-and-collar area of this olive-green cable-knit sweater and render a photorealistic close-up that fills the frame. Show the ribbed crewneck collar meeting the raglan-style cable panel, individual yarn plies and the twist of the cable braid, the loft of the wool fibres, and soft directional light raking across the surface to reveal depth. Keep the colour and stitch pattern exactly as in the source. Shallow depth of field with the far edge softly out of focus, clean neutral background bokeh, no person, no text, no watermark.' \
--images docs/clothing-detail/garment-flatlay.jpg \
--size 1024x1024 --quality high --imageFormat jpeg \
--save docs/clothing-detail/example-output.jpg
输出
example-output.jpg — 1024×1024,60 credits。领口罗纹与麻花panel的交接、每根纱线的捻向、羊毛纤维的绒毛感都被解析出来,侧光让菱形提花的凹凸立体可见,远端落入柔和虚化。
1、能力边界
| 能力 | 说明 |
|---|---|
| 取景部位 | 领口罗纹 / 袖口 / 下摆 / 纽扣 / 拉链 / 口袋 / 刺绣 / 印花 / 织法结构 / 面料纤维 |
| 风格控制 | 参考图(照抄某张细节图的机位与光线)或自定义提示词 |
| 服装类型 | 帮助模型判断哪些部位值得放大 |
| 生成比例 | 1:1(方图细节位)/ 3:4(竖版详情页) |
不做:不改颜色、织法与结构;不添加原图没有的工艺(不存在的刺绣、不存在的拉链);不虚构面料成分。
2、输入素材规则
生成前先自检这几条硬性约束:
- 大小:20KB ~ 15MB
- 分辨率:大于 400×400
- 格式:jpg / jpeg / png / webp
输入建议
| 做法 | 说明 |
|---|---|
| ✅ 原图分辨率越高越好 | 微距是在放大原图信息,原图糊 = 细节图编 |
| ✅ 目标部位在原图里清晰可见 | 原图里看不清的部位,输出的是模型的想象 |
| ✅ 一次只放大一个部位 | 一张图里塞三个特写等于都不清楚 |
| ❌ 低分辨率 / 强压缩图 | 会放大出塑料感的假纹理 |
| ❌ 目标部位被遮挡 | 挡住的工艺只能靠编 |
3、取景部位 → prompt 写法
| 部位 | 取景描述 |
|---|---|
| 领口罗纹 | the ribbed crewneck collar meeting the body panel, showing rib wale spacing and the seam join |
| 袖口 | the ribbed cuff and the sleeve seam, showing rib elasticity and stitch density |
| 下摆 | the hem band and side seam, showing hem width and the finishing stitch |
| 纽扣 | a single button and its buttonhole, showing button material, thread cross and hole finishing |
| 拉链 | the zipper teeth and puller, showing tooth pitch, metal finish and the tape stitching |
| 刺绣 / 印花 | the [刺绣/印花] motif filling the frame, showing thread direction / print edge sharpness and substrate texture |
| 织法结构 | the [麻花/罗纹/提花] stitch structure, showing individual yarn plies and the twist of each loop |
| 面料纤维 | the fabric surface at extreme magnification, showing fibre halo and weave interlacing |
每条都要补三件事:
filling the frame ← 特写要占满画面
shallow depth of field with the far edge softly out of focus ← 微距的景深特征
soft directional light raking across the surface to reveal depth ← 侧光才能显出立体纹理
4、工具调用
本技能使用 dLazy 的 gpt-image-2(图像编辑模型 + --quality high;细节图的全部价值就是纹理保真度,这是本技能唯一不能省的地方)。
调用方式
两种等价写法,选一种。统一入口会自动选后端、失败重试、建目录落盘、估算成本:
# A. 统一入口(推荐):可切任意后端,加 --dry-run 不计费空跑
node scripts/gen.mjs --task clothing-detail \
--prompt '<见下方 Prompt 模板>' \
--images <按下表顺序> \
--save output/clothing-detail-<sku>.jpg
# B. 直接用 dLazy CLI(不想引入 Node 依赖时,效果等价)
dlazy gpt-image-2 --prompt '...' --images ... --save output/clothing-detail.jpg
参数约定(本技能固定用法)
| 参数 | 取值 | 理由 |
|---|---|---|
--images |
[原图];带风格参考时 [原图, 参考图] |
顺序即 prompt 中的 image 1 / 2 |
--size |
1024x1024(方图细节位)/ 1024x1536(竖版详情页) |
对应原站 1:1 / 3:4 |
--quality |
high(不要降) |
细节图靠纹理说话 |
--imageFormat |
jpeg |
通用格式 |
--batch |
2 |
取景位置有随机性 |
--save |
docs/clothing-detail/output-<sku>-<部位>.jpg |
按部位归档 |
Command Examples
# basic call: 领口细节
dlazy gpt-image-2 \
--prompt 'Macro detail shot for an e-commerce detail page. Zoom into the collar area of this garment and render a photorealistic close-up filling the frame. Show the ribbed crewneck collar meeting the body panel, individual yarn plies and the twist of each loop, soft directional light raking across the surface. Keep colour and stitch pattern exactly as in the source. Shallow depth of field, clean neutral background bokeh, no person, no text.' \
--images docs/clothing-detail/garment-flatlay.jpg \
--size 1024x1024 --quality high
# complex call: 一个 SKU 批量出 3 个部位的细节图
SRC=docs/clothing-detail/garment-flatlay.jpg
COMMON='Render a photorealistic macro close-up filling the frame. Keep the colour and stitch pattern exactly as in the source image. Shallow depth of field with the far edge softly out of focus, soft directional light raking across the surface to reveal depth, clean neutral background bokeh. No person, no text, no watermark.'
for P in collar cuff stitch; do
case $P in
collar) VIEW='Zoom into the collar area: the ribbed crewneck collar meeting the body panel, showing rib wale spacing and the seam join' ;;
cuff) VIEW='Zoom into the cuff area: the ribbed cuff and the sleeve seam, showing rib elasticity and stitch density' ;;
stitch) VIEW='Zoom into the cable-knit panel: the stitch structure, showing individual yarn plies and the twist of each loop' ;;
esac
dlazy gpt-image-2 \
--prompt "Macro detail shot for an e-commerce detail page. $VIEW. $COMMON" \
--images "$SRC" --size 1024x1024 --quality high --imageFormat jpeg \
--save "docs/clothing-detail/output-sku001-$P.jpg"
done
# 先估价不真跑
dlazy gpt-image-2 --dry-run --prompt '...' --images a.jpg --size 1024x1024 --quality high
延伸阅读
| 要查什么 | 去哪 |
|---|---|
| 认证、多后端配置、输出结构、错误码 | references/provider-cli.md |
gpt-image-2 的全部可用参数 |
references/model-flags.md |
| 统一入口的全部选项 | node scripts/gen.mjs --help |
5、Prompt 模板
Macro detail shot for an e-commerce detail page.
Zoom into [部位] of this [品类 + 颜色 + 面料] and render a photorealistic close-up
that fills the frame. Show [第三节表格里的取景描述].
Keep the colour and stitch pattern exactly as in the source.
Shallow depth of field with the far edge softly out of focus,
soft directional light raking across the surface to reveal depth,
clean neutral background bokeh.
No person, no text, no watermark.
按问题追加的修正句
| 问题 | 追加到 prompt 末尾 |
|---|---|
| 纹理像塑料 | Resolve individual [yarn plies / weave threads / fibre ends]; the surface must read as real textile, not plastic or CG. |
| 放大得不够 | Extreme magnification: the [部位] must occupy at least 70% of the frame. |
| 编出了不存在的工艺 | Do not invent any construction detail that is not visible in the source image. |
| 整张都很实、没有微距感 | Only the [部位] is in focus; everything beyond [X] must fall into smooth bokeh. |
| 颜色变了 | Sample the colour directly from the source image; no grading, no saturation boost. |
6、执行流程
- 挑原图:分辨率越高越好;确认目标部位清晰可见、无遮挡。
- 列部位清单:一个 SKU 通常出 2~3 张(领口 + 面料 + 一个特色工艺)。
- 每条 prompt 只放大一个部位,从第三节取景描述抄。
- 补齐三件事:占满画面 / 浅景深 / 侧光。
--quality high(不要降档)→--batch 2挑图,落盘到docs/clothing-detail/。- 质检:纹理是否真实(不是塑料感)、有没有编出不存在的工艺、颜色是否一致。
7、常见问题
| 现象 | 原因 | 处理 |
|---|---|---|
| 纹理塑料感 | 质量档位低或原图糊 | --quality high + 追加解析纹理句;换高分辨率原图 |
| 放大不够,还是半身 | 未写占比 | 追加 70% 画面占比句 |
| 编出了原图没有的拉链/刺绣 | 模型补全 | 追加禁止编造句 |
| 没有微距景深 | 未写景深 | 追加只有目标部位对焦的句子 |
| 颜色比原图艳 | 自动调色 | 追加取色约束句 |
| 三个部位挤在一张图 | 一条 prompt 写了多个部位 | 拆成多条,一条一个部位 |
Tips
Visit https://dlazy.com for more information.
Files (ecommerce-skills)
-
examples
-
brand.yaml 1.6 KB
# ⚠️ 由 scripts/build-skills.mjs 从 shared/examples/brand.yaml 同步生成,不要直接改这里。 # 店铺品牌视觉规范 —— 所有生图技能读这一份,保证几百个 SKU 看起来像同一家店。 # node scripts/brand.mjs --brand brand.yaml --for flat-lay # node scripts/gen.mjs --task flat-lay --brand brand.yaml --prompt '...' brand: name: 示例品牌 # 一句话概括调性,会原样进 prompt tone: quiet minimalist, warm and lived-in, never glossy or commercial model: # 锁模特:给一张脸的参考图,所有技能都会把它作为最后一张参考图传入 reference: assets/model/face-a.jpg description: East Asian woman, late twenties, natural makeup, shoulder-length black hair body: slim, height around 168cm photography: background: seamless off-white studio backdrop, RGB 248 248 246 lighting: soft large softbox from camera left, gentle fill, no hard shadows camera: 85mm equivalent, eye level, shallow depth of field grade: neutral white balance around 5200K, low contrast, slightly lifted blacks crop: full body with headroom, product centered layout: # 给带排版的技能(主图 / 详情页)用 margin: at least 8% empty margin on all sides typeface: clean sans-serif, no decorative fonts text_color: near-black on light background forbid: - no visible brand logos other than the product's own - no text or watermark - no exaggerated poses or dramatic wind effects - no oversaturated colors # 可选:把这些直接写进合规目标,生成时就按平台要求出图 compliance: platform: amazon
-
-
references
-
model-flags.md 2.1 KB
# `gpt-image-2` 参数清单 本技能默认用的模型的完整参数。日常只需要「参数约定」里那几个, 这份清单在需要用到非常规参数时再看。 **CRITICAL INSTRUCTION FOR AGENT**: Run the `dlazy gpt-image-2` command to get results. ```bash dlazy gpt-image-2 -h Options: --prompt <prompt> Prompt --images [images...] Images [image: url or local path] (max 5) --size <size> Size [default: auto] (choices: "1024x1024", "1536x1024", "1024x1536", "2048x2048", "2048x1152", "3840x2160", "2160x3840", "auto") --imageFormat <imageFormat> Image Format [default: jpeg] (choices: "jpeg", "png", "webp") --quality <quality> Quality [default: medium] (choices: "low", "medium", "high") --dry-run Print payload without executing the tool --no-wait Return generateId immediately for async tasks --timeout <seconds> Max seconds to wait for async completion (default: "1800") --input <jsonOrFile> Inline JSON or @path/to/file.json — merged under flag values (flags win) --save <path> Download the result asset to this local path (mkdir + retry handled for you). A destination path — NOT a response format; for stdout shape use --format --batch <n> Fan-out N parallel runs (cloud tools only) (default: "1") -h, --help display help for command ``` > Any flag also accepts pipe references — `-` (auto-pick from upstream stdin), `@N` (n-th output), `@N.path` (jsonpath into output), `@*` (all primary values), `@stdin` / `@stdin:path` (whole envelope). See `dlazy --help` for details. --- 换其他后端时参数由 `scripts/gen.mjs` 统一翻译,见 [`provider-cli.md`](provider-cli.md)。 -
provider-cli.md 4.7 KB
<!-- 由 scripts/build-skills.mjs 从 shared/references/provider-cli.md 同步生成,不要直接改这里。 --> # 后端调用参考 技能正文只写「要生成什么」。认证、计费、错误码、输出结构这些每个技能都一样的东西放在这里, **用到时再读**,不占技能的常驻上下文。 --- ## 一、认证 ### 默认后端 dLazy ```bash dlazy login # 设备码流程,远程 shell 也能用,自动写入本地配置 dlazy auth set <KEY> # 已有 key 时直接写入 ``` key 存在用户配置目录(macOS/Linux `~/.dlazy/config.json`,Windows `%USERPROFILE%\.dlazy\config.json`), 权限限本机用户。也可以每次调用用环境变量 `DLAZY_API_KEY` 传入。 手动获取:登录 [dlazy.com](https://dlazy.com) → [API Key 页面](https://dlazy.com/dashboard/organization/api-key)。 key 按组织隔离,可随时轮换或吊销。 ### 其他后端 本技能库不锁定单一厂商。配好任意一家的 key 即可跑: | 后端 | 环境变量 | 说明 | | --- | --- | --- | | `dlazy` | `dlazy login` 或 `DLAZY_API_KEY` | 默认,最省事 | | `openai` | `OPENAI_API_KEY` | 走 `/v1/images/edits` 与 `/v1/images/generations` | | `gemini` | `GEMINI_API_KEY` | Nano Banana 系列 | | `fal` | `FAL_KEY` | | | `replicate` | `REPLICATE_API_TOKEN` | | | `ark` | `ARK_API_KEY` + `ARK_MODEL` | 火山方舟,模型 ID 需按开通情况填 | 选路优先级:`--provider` 参数 > `PROVIDER` 环境变量 > 第一个配了 key 的 > `dlazy`。 ```bash node scripts/gen.mjs --doctor # 看当前哪个后端可用 ``` 各后端的模型 ID 可用 `GEN_MODEL_OPENAI` / `GEN_MODEL_GEMINI` / `GEN_MODEL_FAL` / `GEN_MODEL_REPLICATE` / `GEN_MODEL_ARK` 覆盖。**厂商目录会变,以各家最新文档为准。** --- ## 二、两种调用方式 ### 方式 A:统一入口(推荐) ```bash node scripts/gen.mjs --task <技能名> --prompt '...' --images a.jpg b.jpg --save out.jpg ``` 它负责:后端选路、默认尺寸档位、失败重试(429/5xx 指数退避)、落盘建目录、成本估算。 ```bash node scripts/gen.mjs --task flat-lay --prompt '...' --dry-run # 不调用不计费,只看要发什么 node scripts/gen.mjs --help ``` ### 方式 B:直接用 dLazy CLI 不想引入 Node 依赖时,技能正文里的 `dlazy ...` 命令可以原样执行,效果等价。 ```bash npx @dlazy/cli@1.2.3 <command> # 不装全局二进制 ``` - CLI 源码:[github.com/dlazy-ai/cli](https://github.com/dlazy-ai/cli) · npm 包 `@dlazy/cli` --- ## 三、数据流向 调用 dLazy 时:提示词与参数发往 `api.dlazy.com`;传入的本地图片会上传到 `files.dlazy.com` 供模型读取;产出 URL 同样托管在 `files.dlazy.com`。这是云端生成 API 的通用形态。 换成其他后端时,数据流向对应厂商,不经过 dLazy。 --- ## 四、输出结构 `gen.mjs`(加 `--json`): ```json { "ok": true, "task": "flat-lay", "provider": "dlazy", "model": "gpt-image-2", "files": ["docs/flat-lay/output-sku001.jpg"], "texts": [], "estimatedCredits": 60, "elapsedMs": 58213 } ``` dLazy CLI 原生: ```json { "ok": true, "result": { "tool": "gpt-image-2", "data": { "urls": ["https://files.dlazy.com/data/ai/....jpg"] }, "savedPath": "docs/flat-lay/example-output.jpg" } } ``` 加 `--no-wait` 的异步任务不返回 `data`,返回 `task: { generateId, status }`, 用 `dlazy status <generateId> --wait` 轮询。 文本类模型(如质检)产出在 `result.data.texts[0]`: ```bash dlazy claude-sonnet-5 --prompt '...' --images x.jpg \ | python3 -c 'import sys,json;print(json.load(sys.stdin)["result"]["data"]["texts"][0])' ``` --- ## 五、错误处理 | Code | 类型 | 示例 | | --- | --- | --- | | 401 | 未授权 / 无 key | `ok: false, code: "unauthorized"` | | 501 | 缺必填参数 | `error: required option '--prompt <prompt>' not specified` | | 502 | 本地文件读不到 | `Error: Image file not found: ...` | | 503 | 余额不足 | `ok: false, code: "insufficient_balance"` | | 503 | 服务端错误 | `HTTP status code error (500)` | | 504 | 异步任务失败 | `=== Generation Failed ===` / `Prompt violates safety policy` | **给 Agent 的硬性要求** 1. 命中 `insufficient_balance` → 明确告诉用户算力不足,并给出充值入口 <https://dlazy.com/dashboard/organization/settings?tab=credits> 2. 命中 `unauthorized` / 缺 key → 告诉用户去 <https://dlazy.com/dashboard/organization/api-key> 取 key,用 `dlazy auth set <key>` 存好再继续。 3. 用 `gen.mjs` 时,429 与 5xx 已自动重试;仍失败才向用户报错。 4. **不要**为了「跑通」而偷偷降级参数(尺寸、档位、批量),先问用户。
-
-
scripts
-
lib
-
miniyaml.mjs 2.8 KB · in bundle
-
providers.mjs 11.9 KB · in bundle
-
tasks.json 3.1 KB
{ "_note": "技能 → 默认模型与参数。dlazy 列为默认后端的模型名;其他后端走 providers.mjs 的通用映射,可用 GEN_MODEL_<PROVIDER> 覆盖。", "_credits": { "gpt-image-2": 60, "seedream-5.0": 30, "seedream-5.0-pro": 45, "banana-pro": 25, "claude-sonnet-5": 3 }, "tasks": { "flat-lay": { "model": "gpt-image-2", "size": "1024x1536", "quality": "high", "format": "jpeg" }, "wear-everything": { "model": "gpt-image-2", "size": "1024x1536", "quality": "medium", "format": "jpeg" }, "image-fusion": { "model": "seedream-5.0", "size": "3:4", "resolution": "2k" }, "one-shot": { "model": "gpt-image-2", "size": "1024x1536", "quality": "medium", "format": "jpeg" }, "fission-pattern": { "model": "gpt-image-2", "size": "1024x1536", "quality": "medium", "format": "jpeg" }, "item-detail": { "model": "seedream-5.0-pro", "size": "3:4", "resolution": "2k" }, "creative-scene": { "model": "banana-pro", "size": "1024x1536", "format": "jpeg" }, "batch-image": { "model": "seedream-5.0", "size": "3:4", "resolution": "2k" }, "to-3d": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "clothing-extraction": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "fabric-on-body": { "model": "gpt-image-2", "size": "1024x1536", "quality": "high", "format": "jpeg" }, "clothing-detail": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "clothing-grass-planting": { "model": "gpt-image-2", "size": "1024x1536", "quality": "medium", "format": "jpeg" }, "item-selling-point": { "model": "seedream-5.0-pro", "size": "1:1", "resolution": "2k" }, "item-change-background": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "remove-watermark": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "material-enhancement": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "item-repair": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "detect-task": { "model": "claude-sonnet-5", "text": true }, "listing-optimizer": { "model": "gpt-image-2", "size": "1024x1024", "quality": "high", "format": "jpeg" }, "cross-border-localize": { "model": "seedream-5.0-pro", "size": "1:1", "resolution": "2k" }, "brand-kit": { "model": "gpt-image-2", "size": "1024x1536", "quality": "high", "format": "jpeg" }, "platform-compliance": { "model": "claude-sonnet-5", "text": true }, "main-image-video": { "model": "$DLAZY_VIDEO_MODEL", "video": true }, "product-video-ad": { "model": "$DLAZY_VIDEO_MODEL", "video": true }, "ugc-testimonial": { "model": "$DLAZY_VIDEO_MODEL", "video": true } } }
-
-
brand.mjs 4.4 KB · in bundle
-
gen.mjs 9.3 KB · in bundle
-
-
skill.md 10.1 KB
--- name: clothing-detail description: 服装工艺细节放大图。服装图 → 面料纹理、走线、织法的微距特写。当用户说「细节图」「特写」「面料放大」「工艺展示」「近景细节」时使用。 --- # clothing-detail — 服装图生成细节放大图 一张服装图 → **局部微距特写**。详情页里「证明这件衣服做得好」的那几张图。 为什么需要:转化率高的详情页通常有 2~3 张细节图(领口、袖口、面料纹理),但拍微距要专门的镜头和布光。本技能从常规商品图推出这些特写。 --- ## 生成效果示例 | 输入:服装图 | | --- | | <img src="../../docs/clothing-detail/garment-flatlay.jpg" width="280"> | | `garment-flatlay.jpg` — 军绿麻花针织毛衣平铺图,800×800 | 实际执行的命令: ```bash dlazy gpt-image-2 \ --prompt 'Macro detail shot for an e-commerce detail page. Zoom into the shoulder-and-collar area of this olive-green cable-knit sweater and render a photorealistic close-up that fills the frame. Show the ribbed crewneck collar meeting the raglan-style cable panel, individual yarn plies and the twist of the cable braid, the loft of the wool fibres, and soft directional light raking across the surface to reveal depth. Keep the colour and stitch pattern exactly as in the source. Shallow depth of field with the far edge softly out of focus, clean neutral background bokeh, no person, no text, no watermark.' \ --images docs/clothing-detail/garment-flatlay.jpg \ --size 1024x1024 --quality high --imageFormat jpeg \ --save docs/clothing-detail/example-output.jpg ``` **输出** <img src="../../docs/clothing-detail/example-output.jpg" width="320"> `example-output.jpg` — 1024×1024,60 credits。领口罗纹与麻花panel的交接、每根纱线的捻向、羊毛纤维的绒毛感都被解析出来,侧光让菱形提花的凹凸立体可见,远端落入柔和虚化。 --- ## 1、能力边界 | 能力 | 说明 | | --- | --- | | 取景部位 | 领口罗纹 / 袖口 / 下摆 / 纽扣 / 拉链 / 口袋 / 刺绣 / 印花 / 织法结构 / 面料纤维 | | 风格控制 | 参考图(照抄某张细节图的机位与光线)或自定义提示词 | | 服装类型 | 帮助模型判断哪些部位值得放大 | | 生成比例 | `1:1`(方图细节位)/ `3:4`(竖版详情页) | **不做**:不改颜色、织法与结构;不添加原图没有的工艺(不存在的刺绣、不存在的拉链);不虚构面料成分。 --- ## 2、输入素材规则 生成前先自检这几条硬性约束: - 大小:**20KB ~ 15MB** - 分辨率:**大于 400×400** - 格式:**jpg / jpeg / png / webp** **输入建议** | 做法 | 说明 | | --- | --- | | ✅ 原图分辨率越高越好 | 微距是在放大原图信息,原图糊 = 细节图编 | | ✅ 目标部位在原图里清晰可见 | 原图里看不清的部位,输出的是模型的想象 | | ✅ 一次只放大一个部位 | 一张图里塞三个特写等于都不清楚 | | ❌ 低分辨率 / 强压缩图 | 会放大出塑料感的假纹理 | | ❌ 目标部位被遮挡 | 挡住的工艺只能靠编 | --- ## 3、取景部位 → prompt 写法 | 部位 | 取景描述 | | --- | --- | | 领口罗纹 | `the ribbed crewneck collar meeting the body panel, showing rib wale spacing and the seam join` | | 袖口 | `the ribbed cuff and the sleeve seam, showing rib elasticity and stitch density` | | 下摆 | `the hem band and side seam, showing hem width and the finishing stitch` | | 纽扣 | `a single button and its buttonhole, showing button material, thread cross and hole finishing` | | 拉链 | `the zipper teeth and puller, showing tooth pitch, metal finish and the tape stitching` | | 刺绣 / 印花 | `the [刺绣/印花] motif filling the frame, showing thread direction / print edge sharpness and substrate texture` | | 织法结构 | `the [麻花/罗纹/提花] stitch structure, showing individual yarn plies and the twist of each loop` | | 面料纤维 | `the fabric surface at extreme magnification, showing fibre halo and weave interlacing` | **每条都要补三件事**: ```text filling the frame ← 特写要占满画面 shallow depth of field with the far edge softly out of focus ← 微距的景深特征 soft directional light raking across the surface to reveal depth ← 侧光才能显出立体纹理 ``` --- ## 4、工具调用 本技能使用 dLazy 的 **`gpt-image-2`**(图像编辑模型 + `--quality high`;细节图的全部价值就是纹理保真度,这是本技能唯一不能省的地方)。 ### 调用方式 两种等价写法,选一种。统一入口会自动选后端、失败重试、建目录落盘、估算成本: ```bash # A. 统一入口(推荐):可切任意后端,加 --dry-run 不计费空跑 node scripts/gen.mjs --task clothing-detail \ --prompt '<见下方 Prompt 模板>' \ --images <按下表顺序> \ --save output/clothing-detail-<sku>.jpg # B. 直接用 dLazy CLI(不想引入 Node 依赖时,效果等价) dlazy gpt-image-2 --prompt '...' --images ... --save output/clothing-detail.jpg ``` **参数约定(本技能固定用法)** | 参数 | 取值 | 理由 | | --- | --- | --- | | `--images` | `[原图]`;带风格参考时 `[原图, 参考图]` | 顺序即 prompt 中的 image 1 / 2 | | `--size` | `1024x1024`(方图细节位)/ `1024x1536`(竖版详情页) | 对应原站 1:1 / 3:4 | | `--quality` | `high`(**不要降**) | 细节图靠纹理说话 | | `--imageFormat` | `jpeg` | 通用格式 | | `--batch` | `2` | 取景位置有随机性 | | `--save` | `docs/clothing-detail/output-<sku>-<部位>.jpg` | 按部位归档 | ### Command Examples ```bash # basic call: 领口细节 dlazy gpt-image-2 \ --prompt 'Macro detail shot for an e-commerce detail page. Zoom into the collar area of this garment and render a photorealistic close-up filling the frame. Show the ribbed crewneck collar meeting the body panel, individual yarn plies and the twist of each loop, soft directional light raking across the surface. Keep colour and stitch pattern exactly as in the source. Shallow depth of field, clean neutral background bokeh, no person, no text.' \ --images docs/clothing-detail/garment-flatlay.jpg \ --size 1024x1024 --quality high # complex call: 一个 SKU 批量出 3 个部位的细节图 SRC=docs/clothing-detail/garment-flatlay.jpg COMMON='Render a photorealistic macro close-up filling the frame. Keep the colour and stitch pattern exactly as in the source image. Shallow depth of field with the far edge softly out of focus, soft directional light raking across the surface to reveal depth, clean neutral background bokeh. No person, no text, no watermark.' for P in collar cuff stitch; do case $P in collar) VIEW='Zoom into the collar area: the ribbed crewneck collar meeting the body panel, showing rib wale spacing and the seam join' ;; cuff) VIEW='Zoom into the cuff area: the ribbed cuff and the sleeve seam, showing rib elasticity and stitch density' ;; stitch) VIEW='Zoom into the cable-knit panel: the stitch structure, showing individual yarn plies and the twist of each loop' ;; esac dlazy gpt-image-2 \ --prompt "Macro detail shot for an e-commerce detail page. $VIEW. $COMMON" \ --images "$SRC" --size 1024x1024 --quality high --imageFormat jpeg \ --save "docs/clothing-detail/output-sku001-$P.jpg" done # 先估价不真跑 dlazy gpt-image-2 --dry-run --prompt '...' --images a.jpg --size 1024x1024 --quality high ``` ### 延伸阅读 | 要查什么 | 去哪 | | --- | --- | | 认证、多后端配置、输出结构、错误码 | [`references/provider-cli.md`](references/provider-cli.md) | | `gpt-image-2` 的全部可用参数 | [`references/model-flags.md`](references/model-flags.md) | | 统一入口的全部选项 | `node scripts/gen.mjs --help` | ## 5、Prompt 模板 ```text Macro detail shot for an e-commerce detail page. Zoom into [部位] of this [品类 + 颜色 + 面料] and render a photorealistic close-up that fills the frame. Show [第三节表格里的取景描述]. Keep the colour and stitch pattern exactly as in the source. Shallow depth of field with the far edge softly out of focus, soft directional light raking across the surface to reveal depth, clean neutral background bokeh. No person, no text, no watermark. ``` **按问题追加的修正句** | 问题 | 追加到 prompt 末尾 | | --- | --- | | 纹理像塑料 | `Resolve individual [yarn plies / weave threads / fibre ends]; the surface must read as real textile, not plastic or CG.` | | 放大得不够 | `Extreme magnification: the [部位] must occupy at least 70% of the frame.` | | 编出了不存在的工艺 | `Do not invent any construction detail that is not visible in the source image.` | | 整张都很实、没有微距感 | `Only the [部位] is in focus; everything beyond [X] must fall into smooth bokeh.` | | 颜色变了 | `Sample the colour directly from the source image; no grading, no saturation boost.` | --- ## 6、执行流程 1. **挑原图**:分辨率越高越好;确认目标部位清晰可见、无遮挡。 2. **列部位清单**:一个 SKU 通常出 2~3 张(领口 + 面料 + 一个特色工艺)。 3. **每条 prompt 只放大一个部位**,从第三节取景描述抄。 4. **补齐三件事**:占满画面 / 浅景深 / 侧光。 5. **`--quality high`**(不要降档)→ `--batch 2` 挑图,落盘到 `docs/clothing-detail/`。 6. **质检**:纹理是否真实(不是塑料感)、有没有编出不存在的工艺、颜色是否一致。 --- ## 7、常见问题 | 现象 | 原因 | 处理 | | --- | --- | --- | | 纹理塑料感 | 质量档位低或原图糊 | `--quality high` + 追加解析纹理句;换高分辨率原图 | | 放大不够,还是半身 | 未写占比 | 追加 70% 画面占比句 | | 编出了原图没有的拉链/刺绣 | 模型补全 | 追加禁止编造句 | | 没有微距景深 | 未写景深 | 追加只有目标部位对焦的句子 | | 颜色比原图艳 | 自动调色 | 追加取色约束句 | | 三个部位挤在一张图 | 一条 prompt 写了多个部位 | 拆成多条,一条一个部位 | --- ## Tips Visit https://dlazy.com for more information.
Comments (0)
Sign in to join the conversation.
Reviews (0)
No reviews yet.

No comments yet.