Agent skill

Linkfox Multimodal Generate Image

by linkfox-ai in linkfox-ai/linkfox-skills

AI驱动的图片生成与编辑工具,用于制作高质量产品图。当用户要求生成图片、制作图片、编辑照片、文生图、图生图、换背景、变换风格、替换图片中的物体、将产品合成到场景中、换模特、制作任何类型的AI生成视觉内容、AI drawing, image generation, text-to-image, image-to-image, background replacement, style…

MITAuto-check passedMedia & Creative

Install Linkfox Multimodal Generate Image

skills CLI
$ npx skills add linkfox-ai/linkfox-skills --skill linkfox-multimodal-generate-image -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install linkfox-ai/linkfox-skills linkfox-multimodal-generate-image --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/linkfox-ai/linkfox-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/linkfox-multimodal-generate-image .claude/skills/linkfox-multimodal-generate-image && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
linkfox-multimodal-generate-image
GitHub stars
107
Token cost
~1.9k tokens
SKILL.md length
852 words
Files
6 (incl. scripts, references)
Skills in repo
177
Repo updated
First seen
Licence
MIT

At a glance

AI驱动的图片生成与编辑工具,用于制作高质量产品图。当用户要求生成图片、制作图片、编辑照片、文生图、图生图、换背景、变换风格、替换图片中的物体、将产品合成到场景中、换模特、制作任何类型的AI生成视觉内容、AI drawing, image generation, text-to-image, image-to-image, background replacement, style…

  • Works in 4 steps: Be specific and descriptive: Clearly… → Reference images by number: When using… → State the operation explicitly: Use… → …
  • Tasks that involve Image generation
  • SKILL.md covers Core Concepts, Parameter Guide, Local Image Upload and 调用方式, plus 5 more sections
  • Runs Python scripts from its folder; calls python; needs LINKFOX_AGENT_API_KEY and LINKFOXAGENT_API_KEY

What it does

Linkfox Multimodal Generate Image is an agent skill from linkfox-ai/linkfox-skills. AI驱动的图片生成与编辑工具,用于制作高质量产品图。当用户要求生成图片、制作图片、编辑照片、文生图、图生图、换背景、变换风格、替换图片中的物体、将产品合成到场景中、换模特、制作任何类型的AI生成视觉内容、AI drawing, image generation, text-to-image, image-to-image, background replacement, style transfer, product image creation, AI image editing时触发此技能。即使用户未明确说"AI图片",只要其请求涉及生成、修改或变换图片,也应触发此技能。

Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. The skill folder holds 7 other files, including scripts and reference files (for example `references/api.md`, `references/onboarding.md` and `scripts/multimodal_generate_image.py`).

It sits in Media & Creative, covering Image generation. The licence is MIT.

When your agent uses it

  • Tasks that involve Image generation

Example prompts

  • “/linkfox-multimodal-generate-image”

Requirements

  • Python 3
  • A credential in LINKFOX_AGENT_API_KEY
  • A credential in LINKFOXAGENT_API_KEY

Workflow steps

4 steps, taken from the first numbered list in SKILL.md.

  1. Be specific and descriptive: Clearly describe the subject, scene, lighting, style, and mood you want.
  2. Reference images by number: When using reference images, refer to them as "image 1", "image 2", etc., in the order they appear in…
  3. State the operation explicitly: Use clear action verbs like "replace", "change", "put", "combine", "generate".
  4. Keep within 1000 characters: Prompts have a maximum length of 1000 characters.

What it can do on your machine

Read from SKILL.md and the folder at commit 38fef04. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 3 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • skill.linkfox.com

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • LINKFOX_AGENT_API_KEY
    • LINKFOXAGENT_API_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Linkfox Multimodal Generate Image loads about 1.9k tokens when it runs, and up to ~3.3k if it reads all its reference files. Until then it costs about 81 tokens; SKILL.md has 852 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~81
When it runs · the whole SKILL.md, loaded when a task matches
~1.9k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.3k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from linkfox-ai/linkfox-skills at commit 38fef04, republished under its MIT licence (© linkfox-ai). 852 words, ~1,948 tokens.

Download SKILL.mdSave it as .claude/skills/linkfox-multimodal-generate-image/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
linkfox-multimodal-generate-image
description
AI驱动的图片生成与编辑工具,用于制作高质量产品图。当用户要求生成图片、制作图片、编辑照片、文生图、图生图、换背景、变换风格、替换图片中的物体、将产品合成到场景中、换模特、制作任何类型的AI生成视觉内容、AI drawing, image generation, text-to-image, image-to-image, background replacement, style transfer, product image creation, AI image editing时触发此技能。即使用户未明确说"AI图片",只要其请求涉及生成、修改或变换图片,也应触发此技能。

AI Image Generation

This skill guides you on how to generate and edit images using the AI image generation service, helping users create high-quality product images, modify existing images, and perform creative visual transformations.

Core Concepts

The AI Image Generation tool produces new images based on a text prompt and optional reference images. It supports a wide range of use cases:

  • Text-to-image: Generate a brand-new image purely from a text description.
  • Image-to-image: Provide one or more reference images and a prompt to generate a new image that preserves elements from the references.
  • Image editing: Modify specific elements, colors, backgrounds, or styles in an existing image.
  • Product compositing: Place a product from one image into a scene from another image.
  • Model swapping: Replace the model or mannequin in a product photo.

Reference images are strongly recommended when the user wants the output to closely resemble an existing product or scene. Up to 3 reference image URLs can be provided, separated by commas.

Parameter Guide

ParameterRequiredDescriptionDefault
promptYesText description of the desired image. Supports text-to-image, image-to-image, editing, model swapping, and more. Max 1000 characters.--
referenceImageUrlNoURL(s) of reference image(s). Separate multiple URLs with commas. Up to 3 images supported. Max 1000 characters.--
aspectRatioNoAspect ratio of the output image.1:1
Supported Aspect Ratios
ValueDescription
1:1Square (default)
3:4Portrait
4:3Landscape
9:16Vertical fullscreen
16:9Horizontal fullscreen
Prompt Writing Tips
  1. Be specific and descriptive: Clearly describe the subject, scene, lighting, style, and mood you want.
  2. Reference images by number: When using reference images, refer to them as "image 1", "image 2", etc., in the order they appear in referenceImageUrl.
  3. State the operation explicitly: Use clear action verbs like "replace", "change", "put", "combine", "generate".
  4. Keep within 1000 characters: Prompts have a maximum length of 1000 characters.
Prompt Examples by Scenario

Object replacement:

Replace the vase on the table in image 1 with a potted plant

Background color change:

Change the background color of image 1 to pure white

Product compositing:

Place the product from image 2 onto the marble countertop in image 1

Style transfer:

Transform image 1 into the artistic style shown in image 2

Text-to-image (no reference):

A professional product photo of a sleek black wireless headphone on a gradient blue background, studio lighting, 8K quality

Model swapping:

Replace the model in image 1 with a different model while keeping the same clothing and pose

Local Image Upload

This tool requires publicly accessible image URLs for reference images. If the user provides a local image file path (e.g., C:\Users\...\photo.png, /home/.../image.jpg), you must upload it first to obtain a public URL.

Run the upload script:

bash
python scripts/upload_image.py /path/to/local/image.png

The script will return a public URL (valid for 24 hours) that can be used as the reference image URL parameter.

调用方式

  • API 端点:POST /multimodal/generateImage(完整参数/响应/错误码见 references/api.md)
  • Python 脚本:python scripts/multimodal_generate_image.py '<JSON 参数>' [--inline]
  • 成本约束:本工具会消耗算力;同一会话同一参数组合默认只调用一次,脚本带 24h 本地缓存。失败/空结果不得自动换关键词、翻页或改邮编连续试探;需要继续检索时先向用户说明会产生额外消耗。

输出策略(脚本默认行为):

  • 始终将完整响应写入 <cwd>/linkfox/<YYYY-MM-DD>/<session>/data/linkfox-multimodal-generate-image-<timestamp>.json(<cwd> 为脚本执行时的工作目录,在 Claude Code 里即当前项目目录;<session> 取自环境变量 SESSION_ID,按用户任务自动聚合;禁止写入 /tmp,当前目录不可写则报错)
  • 响应体 ≤ 8 KB:落盘后把完整 JSON 打印到 stdout
  • 响应体 > 8 KB:落盘后 stdout 只输出摘要(顶层字段、常见计数如 total/costToken、最大列表字段的长度 + 前 3 条样本)
  • 加 --inline 强制全量打印到 stdout(同样落盘)

读数据建议:先看摘要判断是否足够;需要具体字段时优先用 jq或ConvertFrom-Json 从保存的 json 文件按需抽取,避免整份 JSON 进入上下文。

解决认证和算力问题

发生以下异常情况时,采用 references/onboarding.md 引导解决问题:

异常情况
  • 未配置API Key:环境变量未配置 LINKFOX_AGENT_API_KEY,也未配置 LINKFOXAGENT_API_KEY。
  • 响应401或402状态码
  • 响应提示算力或余额不足:消息含"算力余额不足/计费不足/余额不足/quota exceeded/insufficient balance/套餐到期/需充值/请充值",或类似含义的内容。
Show full SKILL.md (407 more words)Show less

Display Rules

  1. Show the generated image: When the response contains image content in the text field, display it directly to the user using markdown image syntax.
  2. Status reporting: Check the status and finished fields. If image generation is still in progress, inform the user and advise waiting.
  3. Prompt transparency: Briefly describe what prompt and parameters were sent so the user understands what was requested.
  4. Aspect ratio confirmation: If the user does not specify dimensions, use the default 1:1 ratio but mention it so they can request a different ratio if needed.
  5. Reference image guidance: If the user wants a result close to an existing image but did not provide a reference URL, proactively suggest they provide one for better fidelity.
  6. Error handling: When generation fails, explain the issue based on the response status field and suggest adjustments (e.g., simplify the prompt, check reference image URLs, try a different aspect ratio).

Important Limitations

  • Reference image limit: A maximum of 3 reference image URLs can be provided per request.
  • Prompt length: The prompt must not exceed 1000 characters.
  • URL validity: Reference image URLs must be publicly accessible. Private or expired URLs will cause failures.
  • Aspect ratio options: Only 1:1, 3:4, 4:3, 9:16, and 16:9 are supported.

User Expression & Scenario Quick Reference

Applicable -- Requests involving image generation or editing:

User SaysScenario
"Generate an image", "Create a picture"Text-to-image generation
"Edit this photo", "Modify the image"Image editing
"Change the background", "Make it white background"Background replacement
"Put the product on this scene"Product compositing
"Make it look like this style"Style transfer
"Swap the model", "Change the person"Model swapping
"Create a product photo"Product image generation
"Make a vertical/landscape version"Aspect ratio adjustment

Not applicable -- Needs beyond image generation:

  • Image analysis or recognition (reading text from images, identifying objects)
  • Video generation or editing
  • Image file format conversion
  • Batch processing of hundreds of images
  • Image hosting or storage

算力消耗规则

按动态规则计费:消耗算力 = rawTokens × 0.045。

重要:本技能的服务按倍数动态计算,可能一次性消耗大量算力,必须提醒用户,由用户决定是否继续。

Feedback:

Auto-detect and report feedback via the Feedback API when any of the following apply:

  1. The functionality or purpose described in this skill does not match actual behavior
  2. The skill's results do not match the user's intent
  3. The user expresses dissatisfaction or praise about this skill
  4. Anything you believe could be improved

Call the feedback API as specified in references/api.md. Do not interrupt the user's flow.


For more high-quality, professional cross-border e-commerce skills, set LinkFox Skills.

© linkfox-ai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (scripts, references) in skills/linkfox-multimodal-generate-image of linkfox-ai/linkfox-skills.

  • SKILL.md
  • references/api.md
  • references/onboarding.md
  • scripts/multimodal_generate_image.py
  • scripts/onboarding.py
  • scripts/upload_image.py

Open the folder on GitHubat commit 38fef04

Compare with similar skills

Linkfox Multimodal Generate Image next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Linkfox Multimodal Generate Image compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Linkfox Multimodal Generate Image this skilllinkfox-ai/linkfox-skills107—~1.9kAutomated safety check: PassMIT
AI Image Generation and Editingzhayujie/CowAgent47k—~1.3kAutomated safety check: PassMIT
Structured Image Generationbytedance/deer-flow84k4 repos~2.9kAutomated safety check: PassMIT
Canghe Comicfreestylefly/canghe-skills4618 repos~3.2kAutomated safety check: PassNone
Generate Imageynulihao/AgentSkillOS61810 repos~1.7kAutomated safety check: NotesNone
GPT Image Generation CLIwuyoscar/GPT-Image2-Skill5.7k—~2.5kAutomated safety check: NotesMIT

Similar skills

  • Generates or edits images from text prompts through a Python script that picks an image backend based on which API keys are configured.

    47k GitHub stars~1.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Structured Image Generation

    bytedance/deer-flow

    Turns an image request into a structured JSON prompt and runs a bundled Python script to generate the picture, optionally guided by reference images.

    84k GitHub starsUsed in 4 repos~2.9k tokens
    Media & CreativeAuto-check passed
  • Canghe Comic

    freestylefly/canghe-skills

    Knowledge comic creator supporting multiple art styles and tones.

    461 GitHub starsUsed in 8 repos~3.2k tokens
    Media & CreativeAuto-check passed
  • Generate Image

    ynulihao/AgentSkillOS

    Generate or edit images using AI models (FLUX, Gemini). An agent skill from ynulihao/AgentSkillOS.

    618 GitHub starsUsed in 10 repos~1.7k tokens
    Media & CreativeAuto-check: notes
  • GPT Image Generation CLI

    wuyoscar/GPT-Image2-Skill

    Generates and edits images with GPT Image 2 or 2.5 through a packaged CLI and a prompt gallery, after settling which model fits the request.

    5.7k GitHub stars~2.5k tokensUpdated 10 days ago
    Media & CreativeAuto-check: notes
  • Minimal Zine Poster Generator

    LiamGvchi/gc-minimal-zine-poster

    Creates or analyzes quiet, paper-texture zine posters with big negative space, one color accent and experimental type, returning an image prompt and the generated poster.

    7.3k GitHub stars~2.9k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed

More from linkfox-ai/linkfox-skills

All 177 skills in this repo
  • Linkfox 1688 Search By Image

    linkfox-ai/linkfox-skills

    1688平台以图搜图,通过商品图片精准检索外观相似或同款的1688货源,返回标题、价格、起批量、月销量、复购率、交易评分等核心数据。当用户提到1688以图搜图、1688找货源、以图找同款、跨境找工厂、1688识图、图片找货源、找相似货源、image search 1688、find supplier by…

    107 GitHub starsUsed in 1 repo~2.4k tokens
    Auto-check passed
  • Linkfox Aba Intelligent Query

    linkfox-ai/linkfox-skills

    亚马逊ABA(品牌分析)搜索词数据的查询与分析,涵盖15个站点近3年的周维度数据。当用户提到ABA数据、亚马逊搜索词分析、关键词挖掘、搜索排名趋势、市场机会分析、季节性关键词、高点击低转化分析、蓝海词发现、竞品关键词分析、ABA data, search term report, keyword mining, search ranking trends, blue ocean…

    107 GitHub starsUsed in 1 repo~2.2k tokens
    Auto-check passed
  • Linkfox Amazon Alexa Search

    linkfox-ai/linkfox-skills

    通过亚马逊前台的 Alexa 购物助手发起自然语言问答,获取与问题相关的导购回答、推荐商品分组、ASIN 列表,以及可继续追问的问题。每次调用仅支持 1 条 prompt,如需追问须由 agent 总结上下文后拼接新问题发起新请求。可用 url 补充亚马逊页面上下文。当用户提到亚马逊 Alexa、Alexa 购物助手、亚马逊智能助手、AI…

    107 GitHub starsUsed in 1 repo~3k tokens
    Auto-check passed
  • 亚马逊反向选品:基于历史商业洞察报告沉淀的指标数据池,按 30+ 项商业维度(市场规模与增长、价格区间与档位份额、竞争密度与头部集中度、人群画像如年龄/性别/收入、评论卖点与痛点等)反向筛选亚马逊赛道与关键词。当用户提到反向选品、指标筛选、细分市场反查、蓝海赛道挖掘、低竞争赛道、新人友好赛道、品牌分散市场、痛点切入、卖点反查、定价档位机会、人群画像选品、Amazon niche reverse…

    107 GitHub starsUsed in 1 repo~3k tokens
    Auto-check passed
  • Linkfox Amazon Product Detail

    linkfox-ai/linkfox-skills

    通过ASIN获取亚马逊商品详细信息,包括标题、图片、五点描述、规格参数、A+页面、价格、评分评论、变体等;可在取得原始HTML时尝试提取Item Highlights(商品亮点)。当用户提到亚马逊商品详情、ASIN查询、商品页面数据、Listing分析、五点描述提取、Item…

    107 GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed
  • Linkfox Amazon Reviews List

    linkfox-ai/linkfox-skills

    按ASIN获取并分析亚马逊商品评论,支持15个站点(含美国站),按星级筛选评论。当用户提到亚马逊评论、美国站评论、商品评价、买家投诉、差评、好评、星级评分、评论分析、评论情感、产品改良建议、Vine评论、已验证购买评论、竞品评论研究、Amazon reviews, US reviews, Amazon.com reviews, product feedback, negative review…

    107 GitHub starsUsed in 1 repo~2.5k tokens
    Auto-check passed

Questions about Linkfox Multimodal Generate Image

What does Linkfox Multimodal Generate Image do?

AI驱动的图片生成与编辑工具,用于制作高质量产品图。当用户要求生成图片、制作图片、编辑照片、文生图、图生图、换背景、变换风格、替换图片中的物体、将产品合成到场景中、换模特、制作任何类型的AI生成视觉内容、AI drawing, image generation, text-to-image, image-to-image, background replacement, style…. Linkfox Multimodal Generate Image is an agent skill from linkfox-ai/linkfox-skills.

When should I use Linkfox Multimodal Generate Image?

Linkfox Multimodal Generate Image fits situations like: tasks that involve Image generation.

How do I install Linkfox Multimodal Generate Image in Claude Code?

Run `npx skills add linkfox-ai/linkfox-skills --skill linkfox-multimodal-generate-image -a claude-code`. Or copy the skill folder (skills/linkfox-multimodal-generate-image in linkfox-ai/linkfox-skills) into .claude/skills/linkfox-multimodal-generate-image in your project. Claude Code loads it when a task matches its description.

How do I install Linkfox Multimodal Generate Image in Codex?

Run `npx skills add linkfox-ai/linkfox-skills --skill linkfox-multimodal-generate-image -a codex`. Or copy the skill folder (skills/linkfox-multimodal-generate-image in linkfox-ai/linkfox-skills) into .agents/skills/linkfox-multimodal-generate-image in your project. Codex loads it when a task matches its description.

Can I use Linkfox Multimodal Generate Image in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add linkfox-ai/linkfox-skills --skill linkfox-multimodal-generate-image -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/linkfox-multimodal-generate-image, .gemini/skills/linkfox-multimodal-generate-image, .github/skills/linkfox-multimodal-generate-image and .opencode/skills/linkfox-multimodal-generate-image in your project.

What does Linkfox Multimodal Generate Image need to run?

Going by SKILL.md and its folder, Linkfox Multimodal Generate Image needs Python for the scripts in its folder, the command-line tools its instructions call (python) and credentials named LINKFOX_AGENT_API_KEY and LINKFOXAGENT_API_KEY. Our summary lists: Python 3; A credential in LINKFOX_AGENT_API_KEY; A credential in LINKFOXAGENT_API_KEY.

Does Linkfox Multimodal Generate Image access the network?

SKILL.md names 1 domain. As links in the text: skill.linkfox.com. This is read from the text; nothing was executed.

Is Linkfox Multimodal Generate Image safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Linkfox Multimodal Generate Image use?

Linkfox Multimodal Generate Image is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Linkfox Multimodal Generate Image use?

About 1.9k tokens (SKILL.md is roughly 7.8k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 1.4k tokens, read only when the agent opens those files.

What are the alternatives to Linkfox Multimodal Generate Image?

Skills that share tags, products or a category with Linkfox Multimodal Generate Image: AI Image Generation and Editing (zhayujie/CowAgent, 47k stars), Structured Image Generation (bytedance/deer-flow, 84k stars), Canghe Comic (freestylefly/canghe-skills, 461 stars) and Generate Image (ynulihao/AgentSkillOS, 618 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Linkfox Multimodal Generate Image?

linkfox-ai (a GitHub user) maintains it in linkfox-ai/linkfox-skills, which has 107 GitHub stars. The repository holds 177 skills in this directory. The repository was last updated on September 14, 2026.

Source: linkfox-ai/linkfox-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.