Agent skill

Provider Model Refresh

by calesthio in calesthio/OpenMontage

Use the September 2026 image, video, speech and Avatar V adapters with explicit model/host contracts.

AGPL-3.0Auto-check passedMedia & Creative

Install Provider Model Refresh

skills CLI
$ npx skills add calesthio/OpenMontage --skill provider-model-refresh -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install calesthio/OpenMontage provider-model-refresh --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/calesthio/OpenMontage.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/provider-model-refresh .claude/skills/provider-model-refresh && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
provider-model-refresh
GitHub stars
65k
Token cost
~1.4k tokens
SKILL.md length
695 words
Files
1
Skills in repo
41
Repo updated
First seen
Licence
AGPL-3.0

At a glance

Use the September 2026 image, video, speech and Avatar V adapters with explicit model/host contracts.

  • Tasks that involve Text to speech and voice
  • SKILL.md covers Routing, Image requests, Video and jobs and Speech and music, plus 2 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Tasks that involve Image generation

What it does

Provider Model Refresh is an agent skill from calesthio/OpenMontage. Use the September 2026 image, video, speech and Avatar V adapters with explicit model/host contracts.

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Text to speech and voice and Image generation. The repository describes itself as: World's first open-source, agentic video production system. 12 production pipelines, 100+ tools, 700+ agent skill and production-knowledge files. Turn your AI coding assistant… The licence is AGPL-3.0.

When your agent uses it

  • Tasks that involve Text to speech and voice
  • Tasks that involve Image generation

Example prompts

  • “/provider-model-refresh”

What it can do on your machine

Read from SKILL.md and the folder at commit 9327439. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Provider Model Refresh loads about 1.4k tokens when it runs. Until then it costs about 31 tokens; SKILL.md has 695 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~31
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from calesthio/OpenMontage at commit 9327439, republished under its AGPL-3.0 licence (© calesthio). 695 words, ~1,378 tokens.

Download SKILL.mdSave it as .claude/skills/provider-model-refresh/SKILL.md (or your agent's skills folder).
name
provider-model-refresh
description
Use the September 2026 image, video, speech and Avatar V adapters with explicit model/host contracts.

Provider model refresh

For media production, first follow AGENT_GUIDE.md and the selected pipeline. These adapters are discoverable through the normal BaseTool registry. Inspect their input schema and status before making a request. An API credential indicates configuration, not verified account entitlement. No paid live test is implied.

Routing

Use preferred_tool to require an exact adapter and hosting_provider to require an API host. preferred_provider remains a legacy ranking preference. Select an exact model for images/video or model_id for speech. An unavailable exact model must fail rather than silently substitute a different model.

ModelDirect toolAggregator tools
GPT Image 2.5 Flare / Sunburstopenai_imageopenai_fal_image, atlas_refresh_image
Gemini 3 Pro Image / Nano Banana Progoogle_imagengemini_fal_image, gemini_replicate_image, atlas_refresh_image
Ideogram 4.5 / Precise Editideogram_imageideogram_fal_image
Wan 3.0—wan_fal_video, wan_atlas_video, wan_replicate_video
H3 Max—h3_max_video (fal)
LTX-2.5 Fast / Proltx_api_video—
Eleven v4 / v4 Turboelevenlabs_ttsfal_elevenlabs_tts (v4 only)
Gemini 3.8 Flash / Flash-Lite TTSgemini_tts—
Lyria 3.5google_music—
Sonic 3.6cartesia_tts—
Inworld TTS-2 / Flashinworld_tts—
Qwen Image 2.1qwen_image_21 (local)—
HeyGen Avatar Vheygen_avatar—

Image requests

GPT Image 2.5 uses separate gpt-image-2.5-flare and gpt-image-2.5-sunburst IDs. Use Flare for speed and Sunburst for demanding edits. Describe what changes and what must remain. Direct edits use local image_path/image_paths and optional mask_path. Keep alpha output in PNG or WebP. The old GPT Image 2 default remains available for compatibility.

Google Pro's exact ID is gemini-3-pro-image; it is distinct from gemini-3.1-flash-image (Nano Banana 2). Use up to 14 reference images, and 1K/2K/4K on supported routes. Replicate fallback is disabled.

Ideogram direct generation_mode=precise_edit preserves unaffected pixels; use image_path followed by up to four references in image_paths. With a mask, only three references are allowed. Black edits, white preserves. Precise Edit cannot specify output size. dry_run=true validates and prices the exact request without generating. On fal, precise_edit sets edit_precision=high.

Qwen Image 2.1 requires the actual local model directory and compatible Diffusers. It never downloads weights implicitly. Check the Qwen Research License for the intended use. Transparent generation validates the returned alpha channel rather than adding an opaque alpha channel.

Video and jobs

operation selects text_to_video, image_to_video or reference_to_video. Replicate Wan supports text and first-frame image only. Wan on Atlas uses refers objects and duration=-1 for automatic length; fal uses reference URL arrays and nullable duration. Do not transfer native parameters between hosts. Pass route-specific controls in provider_params; checked-in schemas under tools/provider_contracts validate the complete request before submission.

LTX-2.5 supports text/image/audio generation, not retake or extend. Frame rates are 24, 25, 48, 50; long durations require Fast at 720p/1080p and 24/25 FPS. duration=null lets the model choose length, but cannot combine with a last frame. H3 Max is separate from the older H3 endpoint.

Set job_path to persist a checkpoint. On timeout, pass the returned resume_job to the same tool; this polls the existing job without a second paid submission. A submission timeout without an ID is ambiguous: check the provider dashboard before submitting again. Preserve provenance and outputs. Unknown pricing is reported explicitly, never treated as a free service.

Show full SKILL.md (199 more words)Show less

Speech and music

Direct Eleven IDs are eleven_v4 / eleven_v4_turbo; fal uses eleven-v4. fal v4 accepts at most 5,000 characters and has no speed/style control. Direct v4 accepts up to 10,000. Voice catalogs and rights differ by host. Gemini TTS uses Interactions with speech annotations; configure each dialogue speaker's voice and turn text separately. Its current adapter outputs WAV. Cartesia uses API version 2026-08-14 and WAV. Inworld uses a portal-issued Base64 credential with Basic auth, MP3, and at most 2,000 UTF-16 code units. Lyria duration is a prompt target, not guaranteed exact timing.

Avatar V

Use heygen_avatar in the avatar-spokesperson pipeline. First list_looks or inspect_look; generation checks avatar_type=digital_twin and supported_api_engines before POST. Avatar V is explicitly requested with engine=avatar_v; it never falls back to IV. Provide exactly one of script (+ voice_id), audio_url, or audio_asset_id. A reference look must be a digital twin in the same verified group. Use an authorized existing presenter and preserve HeyGen's consent/account requirements.

Evidence and limitations

Contracts were inspected 2026-10-03; each schema records its official source. See docs/provider-update-plan-2026-10-03.md for release research. Public Kling 4 generation endpoints remain unverified, so do not invent a model ID. Live voice agents require a separate streaming/session integration beyond batch narration.

© calesthio, AGPL-3.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in .agents/skills/provider-model-refresh of calesthio/OpenMontage.

Open the folder on GitHubat commit 9327439

Compare with similar skills

Provider Model Refresh next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Provider Model Refresh compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Provider Model Refresh this skillcalesthio/OpenMontage65k—~1.4kAutomated safety check: PassAGPL-3.0
Wedding Video Guided Wizardaaronyi97/wedding-video-guided-wizard291—~1kAutomated safety check: PassMIT
Hbg Life SimulationMr-funny/hbg-life-simulation140—~5.8kAutomated safety check: PassMIT
RunninghubHM-RunningHub/OpenClaw_RH_Skills141—~1.6kAutomated safety check: PassApache-2.0
Media Genclacky-ai/openclacky1.2k—~7.3kAutomated safety check: PassMIT
Qwen Editdigitalsamba/claude-code-video-toolkit2.2k1 repos~711Automated safety check: PassMIT

Similar skills

  • Wedding Video Guided Wizard

    aaronyi97/wedding-video-guided-wizard

    Guide a creator through a real couple's custom wedding video, from a shareable story intake card and Kimi writing pack through narration, external GPT image prompts, image-to-video packs, music and…

    291 GitHub stars~1k tokensUpdated 28 days ago
    Media & CreativeAuto-check passed
  • Hbg Life Simulation

    Mr-funny/hbg-life-simulation

    Create Chinese HBG “模拟人生 / 人生副本” narrative videos with a consistent comic IP, a rapid multi-life opening, continuous natural-speed narration, synchronized short captions, dense static manga…

    140 GitHub stars~5.8k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Runninghub

    HM-RunningHub/OpenClaw_RH_Skills

    Generate images, videos, audio, and 3D models via RunningHub API (420 endpoints) and run any RunningHub AI Application (custom ComfyUI workflow) by webappId.

    141 GitHub stars~1.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Media Gen

    clacky-ai/openclacky

    Generate or edit images, videos, or audio in the current task.

    1.2k GitHub stars~7.3k tokensUpdated today
    Media & CreativeAuto-check passed
  • Qwen Edit

    digitalsamba/claude-code-video-toolkit

    AI image editing prompting patterns for Qwen-Image-Edit. An agent skill from digitalsamba/claude-code-video-toolkit.

    2.2k GitHub starsUsed in 1 repo~711 tokens
    Media & CreativeAuto-check passed
  • Moonlit Book Workflow

    lincwang123-bot/moonlit-stories

    Build or resume a reusable 12-page personalized English picture book in the Moonlit watercolor house style, using Codex built-in image generation for a per-book character sheet plus illustrations…

    174 GitHub stars~1.4k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed

More from calesthio/OpenMontage

All 41 skills in this repo
  • Video Understand

    calesthio/OpenMontage

    Understand video content locally using ffmpeg frame extraction and Whisper transcription.

    65k GitHub stars~841 tokensUpdated 5 days ago
    Auto-check passed
  • Avatar Video

    calesthio/OpenMontage

    Create AI avatar videos with precise control over avatars, voices, scripts, scenes, and backgrounds using HeyGen's v2 API.

    65k GitHub stars~1.6k tokensUpdated 5 days ago
    Auto-check passed
  • D3 Viz

    calesthio/OpenMontage

    Creating interactive data visualisations using d3.js. An agent skill from calesthio/OpenMontage.

    65k GitHub starsUsed in 3 repos~5.4k tokens
    Auto-check passed
  • Create Video

    calesthio/OpenMontage

    Create videos from a text prompt using HeyGen's Video Agent.

    65k GitHub stars~1.3k tokensUpdated 5 days ago
    Auto-check passed
  • Threejs World Generation

    calesthio/OpenMontage

    Build deterministic, editable, free-viewpoint Three.js worlds from text or structured briefs.

    65k GitHub stars~2k tokensUpdated 5 days ago
    Auto-check passed
  • Video Edit

    calesthio/OpenMontage

    Edit videos locally using ffmpeg. An agent skill from calesthio/OpenMontage.

    65k GitHub stars~855 tokensUpdated 5 days ago
    Auto-check: notes

Questions about Provider Model Refresh

What does Provider Model Refresh do?

Use the September 2026 image, video, speech and Avatar V adapters with explicit model/host contracts. Provider Model Refresh is an agent skill from calesthio/OpenMontage. Use the September 2026 image, video, speech and Avatar V adapters with explicit model/host contracts.

When should I use Provider Model Refresh?

Provider Model Refresh fits situations like: tasks that involve Text to speech and voice; tasks that involve Image generation.

How do I install Provider Model Refresh in Claude Code?

Run `npx skills add calesthio/OpenMontage --skill provider-model-refresh -a claude-code`. Or copy the skill folder (.agents/skills/provider-model-refresh in calesthio/OpenMontage) into .claude/skills/provider-model-refresh in your project. Claude Code loads it when a task matches its description.

How do I install Provider Model Refresh in Codex?

Run `npx skills add calesthio/OpenMontage --skill provider-model-refresh -a codex`. Or copy the skill folder (.agents/skills/provider-model-refresh in calesthio/OpenMontage) into .agents/skills/provider-model-refresh in your project. Codex loads it when a task matches its description.

Can I use Provider Model Refresh in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add calesthio/OpenMontage --skill provider-model-refresh -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/provider-model-refresh, .gemini/skills/provider-model-refresh, .github/skills/provider-model-refresh and .opencode/skills/provider-model-refresh in your project.

What does Provider Model Refresh need to run?

SKILL.md names no scripts, command-line tools or credentials: Provider Model Refresh is instructions for the agent only.

Does Provider Model Refresh access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Provider Model Refresh safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Provider Model Refresh use?

Provider Model Refresh is published under the AGPL-3.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Provider Model Refresh use?

About 1.4k tokens (SKILL.md is roughly 5.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Provider Model Refresh?

Skills that share tags, products or a category with Provider Model Refresh: Wedding Video Guided Wizard (aaronyi97/wedding-video-guided-wizard, 291 stars), Hbg Life Simulation (Mr-funny/hbg-life-simulation, 140 stars), Runninghub (HM-RunningHub/OpenClaw_RH_Skills, 141 stars) and Media Gen (clacky-ai/openclacky, 1.2k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Provider Model Refresh?

calesthio (a GitHub user) maintains it in calesthio/OpenMontage, which has 65,101 GitHub stars. The repository holds 41 skills in this directory. The repository was last updated on October 3, 2026.

Source: calesthio/OpenMontage on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.