Agent skill

Script Evaluator

by AnastasiyaW in AnastasiyaW/codex-claude-code-config

Evaluate video scripts and presentations for flatness, tension, and emotional impact.

MITAuto-check passedMedia & Creative

Install Script Evaluator

skills CLI
$ npx skills add AnastasiyaW/codex-claude-code-config --skill script-evaluator -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install AnastasiyaW/codex-claude-code-config script-evaluator --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/AnastasiyaW/codex-claude-code-config.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/video-production/script-evaluator .claude/skills/script-evaluator && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
script-evaluator
GitHub stars
154
Token cost
~2.1k tokens
SKILL.md length
959 words
Files
1
Skills in repo
50
Repo updated
First seen
Licence
MIT

At a glance

Evaluate video scripts and presentations for flatness, tension, and emotional impact.

  • Works in 11 steps: TENSION - Does the viewer feel something? → SPECIFICITY - Is it concrete or vague? → EMOTIONAL ARC - Does it go somewhere? → …
  • : is this script good
  • SKILL.md covers Preserve facts before scoring…, Evaluation: 6 Dimensions…, Scoring and 5 Common Flatness Patterns, plus 4 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Script Evaluator is an agent skill from AnastasiyaW/codex-claude-code-config. Evaluate video scripts and presentations for flatness, tension, and emotional impact. Use when: 'is this script good', 'review script', 'evaluate video', 'why is this boring', 'flatness check', 'script review', 'improve script', 'rate this video'. Scores 6 dimensions (tension, specificity, emotional arc, hook, customer voice, visual variety), identifies specific problems, and suggests concrete fixes with examples. Do NOT use to generate a new script or scene structure from scratch (use video-narrative-arc), to…

Its SKILL.md is about 2.1k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Media & Creative, covering Video production and Video scripts and shorts. It works with Remotion. The repository describes itself as: Claude Code, Codex, and multi-agent configuration system: principles, hooks, skills, and workflow patterns for AI-assisted development. The licence is MIT.

When your agent uses it

  • : is this script good
  • Why is this boring
  • Rate this video
  • Generate a new script

Example prompts

  • “is this script good”
  • “review script”
  • “evaluate video”
  • “/script-evaluator”

Workflow steps

11 steps, taken from the step headings in SKILL.md.

  1. TENSION - Does the viewer feel something?
  2. SPECIFICITY - Is it concrete or vague?
  3. EMOTIONAL ARC - Does it go somewhere?
  4. HOOK STRENGTH - Will they keep watching?
  5. CUSTOMER VOICE - Does it sound human?
  6. VISUAL VARIETY - Is it visually dynamic?
  7. "Feature Parade"
  8. "Logo-First"
  9. "Generic Superlatives"
  10. "Missing Middle"
  11. "Uniform Energy"

What it can do on your machine

Read from SKILL.md and the folder at commit e71c6a8. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Links to these hosts (documentation or services it may open):

    • ftc.gov

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Script Evaluator loads about 2.1k tokens when it runs. Until then it costs about 181 tokens; SKILL.md has 959 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~181
When it runs · the whole SKILL.md, loaded when a task matches
~2.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from AnastasiyaW/codex-claude-code-config at commit e71c6a8, republished under its MIT licence (© AnastasiyaW). 959 words, ~2,101 tokens.

Download SKILL.mdSave it as .claude/skills/script-evaluator/SKILL.md (or your agent's skills folder).
name
script-evaluator
description
Evaluate video scripts and presentations for flatness, tension, and emotional impact. Use when: 'is this script good', 'review script', 'evaluate video', 'why is this boring', 'flatness check', 'script review', 'improve script', 'rate this video'. Scores 6 dimensions (tension, specificity, emotional arc, hook, customer voice, visual variety), identifies specific problems, and suggests concrete fixes with examples. Do NOT use to generate a new script or scene structure from scratch (use video-narrative-arc), to build the product brief (use product-meaning-extractor), or to render/finish the video (use remotion-production-guide or video-post-production); this only critiques an existing script/scene.

Script Evaluator / Flatness Detector

Review a video script, presentation, or rendered scene code and identify exactly WHY it's boring and HOW to fix it. Works on scripts, storyboards, and Remotion/code-based video scenes.

Preserve facts before scoring style

Check claims against the supplied brief and its sources. Keep factual numbers, comparisons, product capabilities, and customer quotations tied to their evidence and limitations. The examples below are illustrative, not reusable product facts. Never invent a statistic, customer, testimonial, or guarantee to raise a score. Without evidence, use a supported qualitative example or flag the exact claim as unverified while continuing the rest of the review. A creative score is a heuristic, not a measured engagement prediction, factual approval, or authority to publish.

Evaluation: 6 Dimensions (score each 1-10)

1. TENSION - Does the viewer feel something?
ScoreDescription
1-3No enemy/problem. Features listed without context. "We do X, Y, Z" tone
4-6Problem mentioned but vague ("saves time"). Some contrast but generic
7-10Specific visceral enemy. Clear before→after. Viewer thinks "that's me!"

Fix: Name the specific, observed pain. Use attributed customer language or clearly identified paraphrases. Show the "before" state without exaggerating it; add a number only when its source supports the claim.

2. SPECIFICITY - Is it concrete or vague?
Red FlagFix
"Saves time"HOW MUCH? "Saves 45 minutes per photo"
"Better quality"WHAT metric? "Original DPI, pixel-to-pixel"
"Innovative solution"WHAT innovation? "Neural retouching with layer output"
"AI-powered"What does the AI DO? "Detects and removes 12 types of artifacts"
"Trusted by thousands"WHO? "Used by 500+ jewelry studios including [name]"

Rule: Prefer a supported concrete example over a vague adjective. Numbers are optional; retain scope and uncertainty rather than forcing numerical specificity.

3. EMOTIONAL ARC - Does it go somewhere?
ScoreDescription
1-3Same energy start to finish. Feels like a slideshow of facts
4-6Some variation but no clear peak/valley
7-10Identifiable beats. At least one DOWN (tension) and one UP (relief). Last scene feels different from first

Fix: Map scenes to an emotional curve. Must have at least one DOWN (problem) and one UP (amazement/relief). The CTA should feel like resolution, not another slide.

4. HOOK STRENGTH - Will they keep watching?
ScoreDescription
1-3Starts with logo, "We are [company]...", or decorative intro with no information
4-6Has information but doesn't create urgency or curiosity
7-10First frame has surprising information. Creates curiosity gap. Uses customer language. MUST see next scene

Fix: The hook is your ad for your ad. Read the first 3 seconds aloud. If it sounds like a corporate intro, try a relevant question, supported fact, attributed customer quote, or clear visual contrast. A stronger hook must not strengthen a claim beyond its evidence.

5. CUSTOMER VOICE - Does it sound human?
ScoreDescription
1-3"Leveraging cutting-edge technology." "Seamless integration." "World-class results."
4-6Reasonable language but still "written by marketing" feeling
7-10Actual customer phrases. Sounds like describing to a friend. Simple, direct, concrete

Fix: Use relevant available customer reviews with source, attribution, date/context, and faithful wording. No fixed quote quota. If no reviews are available, use plain explanatory copy; never fabricate or present inferred language as a testimonial.

6. VISUAL VARIETY - Is it visually dynamic?
ScoreDescription
1-3Same layout every scene. All text, no imagery. Same animation everywhere
4-6Some variation but key moment doesn't stand out visually
7-102+ visual styles. Pacing matches emotion. Most important scene looks DIFFERENT

Fix: Alternate scene types. After text-heavy scenes, do a visual scene. The most important scene should have a unique visual treatment.

Show full SKILL.md (374 more words)Show less

Scoring

TENSION:        _/10
SPECIFICITY:    _/10
EMOTIONAL ARC:  _/10
HOOK:           _/10
CUSTOMER VOICE: _/10
VISUAL VARIETY: _/10
────────────────────
TOTAL:          _/60

VERDICT:
  50-60: Strong on this creative rubric; check factual and task acceptance separately
  40-49: Good, minor tweaks needed  
  30-39: Needs work on weakest dimensions
  20-29: Major rewrite - go back to Product Brief
  <20:   Start over with Product Meaning Extractor

5 Common Flatness Patterns

1. "Feature Parade"

Symptom: Scene 1: Feature A. Scene 2: Feature B. Scene 3: Feature C. CTA. Why flat: No narrative tension. It's a list, not a story. Fix: Add a problem scene BEFORE features. Features answer a question - ask the question first.

2. "Logo-First"

Symptom: Opens with 3-5 seconds of logo animation. Why flat: Nobody cares about your logo yet. They care about their problem. Fix: Move logo to the END. Open with the hook.

3. "Generic Superlatives"

Symptom: "The best solution." "Revolutionary technology." "World-class results." Why flat: Every product says this. These words carry zero information. Fix: Replace with supported specifics. A measured speed comparison needs its source and conditions; otherwise explain the actual workflow improvement without a made-up multiplier.

4. "Missing Middle"

Symptom: Good hook, good CTA, flat middle that walks through features. Why flat: No emotional peak. The demo section needs a WOW moment. Fix: Find the single most impressive thing the product does. Build to it. Make it visually different from everything else.

5. "Uniform Energy"

Symptom: Every scene has same pacing, animation speed, text size. Why flat: No rhythm. Like music at one volume. Fix: Vary tempo: fast cuts → slow hero shot → medium features → fast proof → slow CTA.

Usage

After writing a script or coding scenes:

"Evaluate this script/video using the script-evaluator skill.
Score each dimension 1-10 and give specific fixes for anything below 7."

Also useful for comparing versions:

"Evaluate V1 and V2 side by side. Which is stronger and why?"

Gotchas

  • A high total with one dimension at 2 = video still feels broken. Fix the weakest link first.
  • Choose the weakest relevant dimension from the brief and observed evidence; tension is not a universally established top predictor of engagement.
  • Don't optimize "interesting" at the expense of truth or clarity. A confusing or exaggerated script is not an improvement.
  • This evaluator works on: scripts, storyboards, Remotion TSX code, and even finished videos (describe what you see).

Troubleshooting

  • High creative score but unsupported claim: keep the score separate, identify the exact unsupported claim, and revise or remove that claim before calling the content ready for its requested use.
  • No testimonials or numeric proof: do not penalize truthful copy into fabrication; evaluate supported examples and the intended audience instead.

Source basis

  • FTC: advertisement endorsements — a testimonial cannot supply experience or substantiation that does not exist. Applied here as an evidence-preservation rule, not a legal-compliance certification.

© AnastasiyaW, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/video-production/script-evaluator of AnastasiyaW/codex-claude-code-config.

Open the folder on GitHubat commit e71c6a8

Compare with similar skills

Script Evaluator next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Script Evaluator compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Script Evaluator this skillAnastasiyaW/codex-claude-code-config154—~2.1kAutomated safety check: PassMIT
ShortsAgriciDaniel/claude-shorts218—~3.2kAutomated safety check: NotesMIT
AI Video Short Makerhassancs91/claude-faceless-shorts-creator271—~1.6kAutomated safety check: NotesMIT
Layered-Collage Documentary Shortshassancs91/claude-faceless-shorts-creator271—~1.5kAutomated safety check: PassMIT
Video ScriptbozhouDev/video-skills-toolkit150—~1.8kAutomated safety check: PassMIT
Stitch to Remotion Walkthrough Videosgoogle-labs-code/stitch-skills8.4k6 repos~3.2kAutomated safety check: NotesApache-2.0

Similar skills

  • Shorts

    AgriciDaniel/claude-shorts

    Interactive longform-to-shortform video creator. An agent skill from AgriciDaniel/claude-shorts.

    218 GitHub stars~3.2k tokensUpdated 2 mo ago
    Media & CreativeAuto-check: notes
  • AI Video Short Maker

    hassancs91/claude-faceless-shorts-creator

    Produces a vertical generative-video short end to end: a locked recurring character animated by a fal video model, ElevenLabs voice, Remotion captions and a clean loop.

    271 GitHub stars~1.6k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes
  • Layered-Collage Documentary Shorts

    hassancs91/claude-faceless-shorts-creator

    Builds a vertical collage-style documentary short from script to render: scene dissection into image layers, layer production, TSX assembly on a collage kit, frame checks and voice.

    271 GitHub stars~1.5k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Video Script

    bozhouDev/video-skills-toolkit

    使用已经锁定的完整口播音频与带时间码字幕,把确认后的逐字稿转成管线无关的视频导演方案和 Beat Graph。适用于“根据最终配音做分镜”“按字幕设计整片画面”“规划 B-roll/证据/素材”“做 HyperFrames 或 Remotion…

    150 GitHub stars~1.8k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed
  • Stitch to Remotion Walkthrough Videos

    google-labs-code/stitch-skills

    Official

    Builds walkthrough videos from Stitch design projects using Remotion, with transitions, zoom effects and text overlays on each screen.

    8.4k GitHub starsUsed in 6 repos~3.2k tokens
    Media & CreativeAuto-check: notes
  • Faceless Explainer Video

    heygen-com/hyperframes

    Turns an article, notes or a topic brief into an explainer video whose visuals are invented per scene, built frame by frame in HyperFrames with no footage.

    59k GitHub starsUsed in 3 repos~7.7k tokens
    Media & CreativeAuto-check: notes

More from AnastasiyaW/codex-claude-code-config

All 50 skills in this repo
  • Bug Reproducer

    AnastasiyaW/codex-claude-code-config

    Find likely software bugs in a codebase, rank concrete bug candidates, and prove or reject them with focused regression tests before proposing a fix.

    154 GitHub stars~4.1k tokensUpdated today
    Auto-check passed
  • Motion Framer

    AnastasiyaW/codex-claude-code-config

    A skill your agent uses when implementing Motion or Framer Motion in React/JavaScript: interactive UI components, micro-interactions, gestures, layout or page transitions, and scroll-based animation.

    154 GitHub starsUsed in 1 repo~5.2k tokens
    Auto-check passed
  • Proof Verify

    AnastasiyaW/codex-claude-code-config

    Plan-based verification - freeze acceptance criteria before building, then verify after with an independent fresh-context agent (the builder must not verify their own work).

    154 GitHub stars~2.6k tokensUpdated today
    Auto-check passed
  • Workflow Orchestration

    AnastasiyaW/codex-claude-code-config

    Написание и запуск Claude Code dynamic workflows (JS-оркестратор субагентов).

    154 GitHub stars~3.8k tokensUpdated today
    Auto-check passed
  • Notebooklm Grounded Research

    AnastasiyaW/codex-claude-code-config

    A skill your agent uses when: NotebookLM, notebooklm MCP, large documentation sets, courses, books, papers, or citation-backed research are mentioned.

    154 GitHub stars~2.4k tokensUpdated today
    Auto-check: warnings
  • Deepseek Provider Contract

    AnastasiyaW/codex-claude-code-config

    Validate a proposed DeepSeek API integration before any key or project context is sent: check thinking-mode tool-call history, strict-schema assumptions, bounded output, and provider data boundaries.

    154 GitHub stars~1.2k tokensUpdated today
    Auto-check passed

Works with

Questions about Script Evaluator

What does Script Evaluator do?

Evaluate video scripts and presentations for flatness, tension, and emotional impact. Script Evaluator is an agent skill from AnastasiyaW/codex-claude-code-config. Evaluate video scripts and presentations for flatness, tension, and emotional impact.

When should I use Script Evaluator?

Script Evaluator fits situations like: : is this script good; why is this boring; rate this video; generate a new script.

How do I install Script Evaluator in Claude Code?

Run `npx skills add AnastasiyaW/codex-claude-code-config --skill script-evaluator -a claude-code`. Or copy the skill folder (skills/video-production/script-evaluator in AnastasiyaW/codex-claude-code-config) into .claude/skills/script-evaluator in your project. Claude Code loads it when a task matches its description.

How do I install Script Evaluator in Codex?

Run `npx skills add AnastasiyaW/codex-claude-code-config --skill script-evaluator -a codex`. Or copy the skill folder (skills/video-production/script-evaluator in AnastasiyaW/codex-claude-code-config) into .agents/skills/script-evaluator in your project. Codex loads it when a task matches its description.

Can I use Script Evaluator in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add AnastasiyaW/codex-claude-code-config --skill script-evaluator -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/script-evaluator, .gemini/skills/script-evaluator, .github/skills/script-evaluator and .opencode/skills/script-evaluator in your project.

What does Script Evaluator need to run?

SKILL.md names no scripts, command-line tools or credentials: Script Evaluator is instructions for the agent only.

Does Script Evaluator access the network?

SKILL.md names 1 domain. As links in the text: ftc.gov. This is read from the text; nothing was executed.

Is Script Evaluator safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Script Evaluator use?

Script Evaluator is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Script Evaluator use?

About 2.1k tokens (SKILL.md is roughly 8.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Script Evaluator?

Skills that share tags, products or a category with Script Evaluator: Shorts (AgriciDaniel/claude-shorts, 218 stars), AI Video Short Maker (hassancs91/claude-faceless-shorts-creator, 271 stars), Layered-Collage Documentary Shorts (hassancs91/claude-faceless-shorts-creator, 271 stars) and Video Script (bozhouDev/video-skills-toolkit, 150 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Script Evaluator?

AnastasiyaW (a GitHub user) maintains it in AnastasiyaW/codex-claude-code-config, which has 154 GitHub stars. The repository holds 50 skills in this directory. The repository was last updated on October 7, 2026.

Source: AnastasiyaW/codex-claude-code-config on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.