Agent skill

Multi Modal Input

by Owl-Listener in Owl-Listener/inclusive-design-skills

Design interfaces that offer multiple input methods so users can choose what works for their abilities and context.

MITAuto-check passedFrontend & Design

Install Multi Modal Input

skills CLI
$ npx skills add Owl-Listener/inclusive-design-skills --skill multi-modal-input -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install Owl-Listener/inclusive-design-skills multi-modal-input --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/Owl-Listener/inclusive-design-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/inclusive-interaction/skills/multi-modal-input .claude/skills/multi-modal-input && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
multi-modal-input
GitHub stars
103
Token cost
~721 tokens
SKILL.md length
314 words
Files
1
Skills in repo
55
Repo updated
First seen
Licence
MIT

At a glance

Design interfaces that offer multiple input methods so users can choose what works for their abilities and context.

  • Works in 5 steps: Can every task be completed through… → Can every task be completed through… → Are there complex interactions (drag,… → …
  • Designing any interactive system where users provide input — forms
  • SKILL.md covers Core Principle, The Input Spectrum, Design Patterns and Assessment Questions
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Multi Modal Input is an agent skill from Owl-Listener/inclusive-design-skills. Design interfaces that offer multiple input methods so users can choose what works for their abilities and context. Use when designing any interactive system where users provide input — forms, search, editors, creative tools, communication interfaces. Triggers on: multi-modal, input methods, alternative input, how people interact, mouse alternative, touch alternative, input flexibility, switch access, eye tracking, head pointer.

Its SKILL.md is about 720 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

It sits in Frontend & Design. The repository describes itself as: Inclusive design skills for AI coding agents — from cognitive accessibility to adaptive interfaces, inclusive research, and accessibility decision-making. The licence is MIT.

When your agent uses it

  • Designing any interactive system where users provide input — forms
  • Communication interfaces
  • Alternative input
  • How people interact

Example prompts

  • “/multi-modal-input”

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Can every task be completed through keyboard alone?
  2. Can every task be completed through pointer alone?
  3. Are there complex interactions (drag, gesture, drawing) that
  4. Is paste enabled in all text fields?
  5. Do all form controls have clickable labels?

What it can do on your machine

Read from SKILL.md and the folder at commit 6e0740f. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Multi Modal Input loads about 721 tokens when it runs. Until then it costs about 113 tokens; SKILL.md has 314 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~113
When it runs · the whole SKILL.md, loaded when a task matches
~721

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from Owl-Listener/inclusive-design-skills at commit 6e0740f, republished under its MIT licence (© Owl-Listener). 314 words, ~721 tokens.

Download SKILL.mdSave it as .claude/skills/multi-modal-input/SKILL.md (or your agent's skills folder).
name
multi-modal-input
description
Design interfaces that offer multiple input methods so users can choose what works for their abilities and context. Use when designing any interactive system where users provide input — forms, search, editors, creative tools, communication interfaces. Triggers on: multi-modal, input methods, alternative input, how people interact, mouse alternative, touch alternative, input flexibility, switch access, eye tracking, head pointer.

Multi-Modal Input Design

Design systems where users can accomplish any task through whichever input method works for them — keyboard, mouse, touch, voice, switch, eye tracking, or any combination.

Core Principle

Never assume how someone will interact with your interface. Offer choices. Let the user decide.

The Input Spectrum

People interact with technology through many methods, often combining several at once:

  • Keyboard — physical, on-screen, or switch-activated
  • Mouse / trackpad — standard pointer devices
  • Touch — fingers, stylus, or assistive touch
  • Voice — speech commands, dictation
  • Switch devices — single or dual switches scanning through options
  • Eye tracking — gaze-based selection
  • Head pointers — head movement controlling a cursor
  • Sip-and-puff — breath-controlled switches

Design Patterns

Input Equivalence
  • Every action must be possible through at least keyboard AND pointer (mouse/touch)
  • Voice input should be available as a third option where practical
  • Never lock a feature to a single input method
  • Test: can someone complete this task using ONLY keyboard? ONLY touch? ONLY voice?
Flexible Text Entry
  • Support physical keyboard, on-screen keyboard, voice dictation, and paste from clipboard
  • Auto-complete and suggestions reduce typing burden
  • Don't disable paste in form fields (password managers, assistive tools depend on this)
  • Allow scanning and OCR for filling in reference numbers or codes
Selection Without Precision
  • Radio buttons and checkboxes: make the label clickable, not just the control
  • Dropdowns: allow type-ahead search for long lists
  • Date pickers: always offer a text field alternative alongside the calendar widget
  • Colour pickers: provide text input for hex/RGB values
Complex Interactions
  • Drag-and-drop: always provide button-based reordering
  • Drawing/annotation: offer text description as alternative
  • Map interactions: provide address search alongside map selection
  • Gestures (swipe, pinch): always provide button equivalents

Assessment Questions

  1. Can every task be completed through keyboard alone?
  2. Can every task be completed through pointer alone?
  3. Are there complex interactions (drag, gesture, drawing) that lack simpler alternatives?
  4. Is paste enabled in all text fields?
  5. Do all form controls have clickable labels?

© Owl-Listener, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in inclusive-interaction/skills/multi-modal-input of Owl-Listener/inclusive-design-skills.

Open the folder on GitHubat commit 6e0740f

Compare with similar skills

Multi Modal Input next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Multi Modal Input compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Multi Modal Input this skillOwl-Listener/inclusive-design-skills103—~721Automated safety check: PassMIT
Web Artifacts Builderanthropics/skills180k40 repos~769Automated safety check: PassApache-2.0
React Doctormakeplane/plane60k12 repos~657Automated safety check: PassAGPL-3.0
React Composition Patternsvercel-labs/openreview1.7k58 repos~721Automated safety check: PassMIT
Impeccablebestofjs/bestofjs3.1k27 repos~2.6kAutomated safety check: PassMIT
Figma Design System Builderwarpdotdev/warp65k2 repos~4.4kAutomated safety check: PassAGPL-3.0

Similar skills

  • Web Artifacts Builder

    anthropics/skills

    Official

    Builds multi-component claude.ai HTML artifacts as a small React, TypeScript and Tailwind project, then bundles it into one shareable HTML file.

    180k GitHub starsUsed in 40 repos~769 tokens
    Frontend & DesignAuto-check passed
  • React Doctor

    makeplane/plane

    Scans React code for lint, accessibility, bundle size and architecture issues, reports a health score and checks that changes do not lower it.

    60k GitHub starsUsed in 12 repos~657 tokens
    Frontend & DesignAuto-check passed
  • React Composition Patterns

    vercel-labs/openreview

    Official

    Rules for structuring React components with composition instead of boolean props, covering compound components, lifted state, variants and React 19 changes.

    1.7k GitHub starsUsed in 58 repos~721 tokens
    Frontend & DesignAuto-check passed
  • Impeccable

    bestofjs/bestofjs

    A skill your agent uses when the user wants to design, redesign, shape, critique, audit, polish, clarify, distill, harden, optimize, adapt, animate, colorize, extract, or otherwise improve a…

    3.1k GitHub starsUsed in 27 repos~2.6k tokens
    Frontend & DesignAuto-check passed
  • Builds or updates a design system in Figma from a codebase in ordered phases: discovery, variables and tokens, components, theming and documentation, with checkpoints.

    65k GitHub starsUsed in 2 repos~4.4k tokens
    Frontend & DesignAuto-check passed
  • Tailwindcss Development

    anonaddy/anonaddy

    Always invoke when the user's message includes 'tailwind' in any form.

    4.9k GitHub starsUsed in 10 repos~865 tokens
    Frontend & DesignAuto-check passed

More from Owl-Listener/inclusive-design-skills

All 55 skills in this repo
  • Assistive Technology Scenarios

    Owl-Listener/inclusive-design-skills

    Writes usage scenarios, use cases and storyboards that show real people using screen readers, switches, voice control and other assistive technology to finish tasks.

    103 GitHub stars~1.1k tokensUpdated 3 mo ago
    Auto-check passed
  • Error Prevention and Recovery

    Owl-Listener/inclusive-design-skills

    Designs forgiving forms and flows: prevent input errors, write messages that say what happened and what to do, and add undo, confirmation and recovery paths.

    103 GitHub stars~739 tokensUpdated 3 mo ago
    Auto-check passed
  • Accessible Heading Structure

    Owl-Listener/inclusive-design-skills

    Designs heading hierarchies for screen reader navigation and cognitive accessibility on pages, articles, dashboards and forms.

    103 GitHub stars~763 tokensUpdated 3 mo ago
    Auto-check passed
  • Ability Spectrum Mapping

    Owl-Listener/inclusive-design-skills

    Maps a feature across a range of vision, hearing, motor and cognitive ability to find where the design starts to fail, and to explain accessibility scope to stakeholders.

    103 GitHub stars~917 tokensUpdated 3 mo ago
    Auto-check passed
  • Accessibility Debt Tracking

    Owl-Listener/inclusive-design-skills

    Track and manage accessibility debt — known accessibility issues that have been deferred.

    103 GitHub stars~869 tokensUpdated 3 mo ago
    Auto-check passed
  • Accessibility Testing Strategy

    Owl-Listener/inclusive-design-skills

    Plan what to test, how to test, and who should test for accessibility.

    103 GitHub stars~1k tokensUpdated 3 mo ago
    Auto-check passed

Questions about Multi Modal Input

What does Multi Modal Input do?

Design interfaces that offer multiple input methods so users can choose what works for their abilities and context. Multi Modal Input is an agent skill from Owl-Listener/inclusive-design-skills. Design interfaces that offer multiple input methods so users can choose what works for their abilities and context.

When should I use Multi Modal Input?

Multi Modal Input fits situations like: designing any interactive system where users provide input — forms; communication interfaces; alternative input; how people interact.

How do I install Multi Modal Input in Claude Code?

Run `npx skills add Owl-Listener/inclusive-design-skills --skill multi-modal-input -a claude-code`. Or copy the skill folder (inclusive-interaction/skills/multi-modal-input in Owl-Listener/inclusive-design-skills) into .claude/skills/multi-modal-input in your project. Claude Code loads it when a task matches its description.

How do I install Multi Modal Input in Codex?

Run `npx skills add Owl-Listener/inclusive-design-skills --skill multi-modal-input -a codex`. Or copy the skill folder (inclusive-interaction/skills/multi-modal-input in Owl-Listener/inclusive-design-skills) into .agents/skills/multi-modal-input in your project. Codex loads it when a task matches its description.

Can I use Multi Modal Input in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add Owl-Listener/inclusive-design-skills --skill multi-modal-input -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/multi-modal-input, .gemini/skills/multi-modal-input, .github/skills/multi-modal-input and .opencode/skills/multi-modal-input in your project.

What does Multi Modal Input need to run?

SKILL.md names no scripts, command-line tools or credentials: Multi Modal Input is instructions for the agent only.

Does Multi Modal Input access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Multi Modal Input safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Multi Modal Input use?

Multi Modal Input is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Multi Modal Input use?

About 721 tokens (SKILL.md is roughly 2.9k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Multi Modal Input?

Skills that share tags, products or a category with Multi Modal Input: Web Artifacts Builder (anthropics/skills, 180k stars), React Doctor (makeplane/plane, 60k stars), React Composition Patterns (vercel-labs/openreview, 1.7k stars) and Impeccable (bestofjs/bestofjs, 3.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Multi Modal Input?

Owl-Listener (a GitHub user) maintains it in Owl-Listener/inclusive-design-skills, which has 103 GitHub stars. The repository holds 55 skills in this directory. The repository was last updated on June 9, 2026.

Source: Owl-Listener/inclusive-design-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.