Agent skill

End Of Speech Integration

by rapidaai in rapidaai/voice-ai

Add or modify end-of-speech integrations in assistant-api with strict separation from VAD internals.

Custom licenceAuto-check passedAI & LLM Engineering

Install End Of Speech Integration

skills CLI
$ npx skills add rapidaai/voice-ai --skill end-of-speech-integration -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install rapidaai/voice-ai end-of-speech-integration --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/rapidaai/voice-ai.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.codex/skills/end-of-speech-integration .claude/skills/end-of-speech-integration && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
end-of-speech-integration
GitHub stars
745
Token cost
~877 tokens
SKILL.md length
318 words
Files
6 (incl. scripts, references)
Skills in repo
26
Repo updated
First seen
Licence
Custom licence

At a glance

Add or modify end-of-speech integrations in assistant-api with strict separation from VAD internals.

  • Works in 3 steps: EOS signal mode: transcript-only,… → Priority: lower latency or lower… → Deployment/model constraints.
  • Transcript/audio/history-aware turn-finalization logic
  • SKILL.md covers Mission, Hard boundaries, Inputs expected from user and Packet contract, plus 5 more sections
  • Runs Shell scripts from its folder; calls go and yarn

What it does

End Of Speech Integration is an agent skill from rapidaai/voice-ai. Add or modify end-of-speech integrations in assistant-api with strict separation from VAD internals. Use for transcript/audio/history-aware turn-finalization logic, provider wiring, and EOS UI config.

Its SKILL.md is about 880 tokens, which your agent loads only when the skill is triggered. The skill folder holds 9 other files, including scripts and reference files (for example `agents/openai.yaml`, `examples/sample.md` and `references/checklist.md`).

It sits in AI & LLM Engineering, covering Speech recognition and synthesis. The repository describes itself as: Rapida is an open-source, end-to-end voice AI orchestration platform for building real-time conversational voice agents with audio streaming, STT, TTS, VAD, multi-channel…

When your agent uses it

  • Transcript/audio/history-aware turn-finalization logic
  • Provider wiring

Example prompts

  • “/end-of-speech-integration”

Requirements

  • A Bash shell

Workflow steps

3 steps, taken from the first numbered list in SKILL.md.

  1. EOS signal mode: transcript-only, audio-model, or history-aware.
  2. Priority: lower latency or lower false-finalization.
  3. Deployment/model constraints.

What it can do on your machine

Read from SKILL.md and the folder at commit 3c9caac. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 1 file in scripts/ (Shell), which the agent can run.

    Shell commands in SKILL.md call:

    • go
    • yarn

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use yarn, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

End Of Speech Integration loads about 877 tokens when it runs, and up to ~1.1k if it reads all its reference files. Until then it costs about 57 tokens; SKILL.md has 318 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~57
When it runs · the whole SKILL.md, loaded when a task matches
~877
With references · SKILL.md plus every file in references/, read only if the agent opens them
~1.1k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 318 words (~877 tokens).

“Implement EOS that finalizes each user turn exactly once, at low latency, without mutating VAD behavior.”

— opening of SKILL.md by rapidaai, Custom licence
name
end-of-speech-integration

Read the full SKILL.md on GitHub

Files

SKILL.md and 5 other files (scripts, references) in .codex/skills/end-of-speech-integration of rapidaai/voice-ai.

  • SKILL.md
  • agents/openai.yaml
  • examples/sample.md
  • references/checklist.md
  • references/eos-checklist.md
  • scripts/validate.sh

Open the folder on GitHubat commit 3c9caac

Compare with similar skills

End Of Speech Integration next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

End Of Speech Integration compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
End Of Speech Integration this skillrapidaai/voice-ai745—~877Automated safety check: PassCustom licence
Yichen Asrmcncarl/yichen-skills4.4k—~780Automated safety check: PassCustom licence
Youtube FetcherJimmySadek/youtube-fetcher-to-markdown485—~3.1kAutomated safety check: PassMIT
Video Understandingzenstory-ai/video-recap-skills561—~1.1kAutomated safety check: PassMIT
Hriterrense/ros2-multimodal-robot-collab111—~268Automated safety check: PassMIT
Ax Audiodosco/aithy107—~2.5kAutomated safety check: PassApache-2.0

Similar skills

  • Yichen Asr

    mcncarl/yichen-skills

    逸尘自用的统一音视频转写入口,在 StepFun Step ASR 与火山引擎豆包 ASR 之间按输出需求、安全边界和可用状态路由。用于本地音频或视频的纯文本转写、时间戳、SRT 字幕、口播粗剪,以及转写前体检;用户明确指定服务商时不得静默切换。Use when a local audio or video file needs transcription and the correct…

    4.4k GitHub stars~780 tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check passed
  • Youtube Fetcher

    JimmySadek/youtube-fetcher-to-markdown

    Retrieve transcripts from YouTube, Instagram, TikTok, X, Vimeo and other video sites, summarize or analyze what was said (and shown on screen), or save an Obsidian-ready Markdown knowledge-base note…

    485 GitHub stars~3.1k tokensUpdated 4 days ago
    AI & LLM EngineeringAuto-check passed
  • Video Understanding

    zenstory-ai/video-recap-skills

    把视频分析为结构化理解索引:场景检测、ASR 转写、逐场景 VLM 观察、静音窗口、融合时间线和写作 brief. An agent skill from zenstory-ai/video-recap-skills.

    561 GitHub stars~1.1k tokensUpdated 7 days ago
    AI & LLM EngineeringAuto-check passed
  • Hri

    terrense/ros2-multimodal-robot-collab

    A skill your agent uses when an Agent needs to speak to the operator through TTS, interpret ASR text, request clarification, or confirm a robot delivery action.

    111 GitHub stars~268 tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • Ax Audio

    dosco/aithy

    This skill helps an LLM generate correct audio code with @ax-llm/ax.

    107 GitHub stars~2.5k tokensUpdated 1 mo ago
    AI & LLM EngineeringAuto-check passed
  • Daily

    sickn33/agentic-awesome-skills

    Documentation and capabilities reference for Daily. An agent skill from sickn33/agentic-awesome-skills.

    47k GitHub starsUsed in 3 repos~3.6k tokens
    AI & LLM EngineeringAuto-check passed

More from rapidaai/voice-ai

All 26 skills in this repo
  • Local Setup And Run

    rapidaai/voice-ai

    Explain and validate local setup paths for this repository with Docker and without Docker.

    745 GitHub stars~595 tokensUpdated 4 days ago
    Auto-check passed
  • System Understanding

    rapidaai/voice-ai

    Build a code-grounded implementation plan before coding. An agent skill from rapidaai/voice-ai.

    745 GitHub stars~721 tokensUpdated 4 days ago
    Auto-check passed
  • Vad Integration

    rapidaai/voice-ai

    Add or modify VAD providers and tuning in assistant-api with strict separation from EOS internals.

    745 GitHub stars~758 tokensUpdated 4 days ago
    Auto-check passed
  • Local Setup And Run

    rapidaai/voice-ai

    Explain and validate local setup paths for this repo with Docker and without Docker.

    745 GitHub stars~683 tokensUpdated 4 days ago
    Auto-check passed
  • System Understanding

    rapidaai/voice-ai

    Build a code-grounded implementation plan before coding. An agent skill from rapidaai/voice-ai.

    745 GitHub stars~727 tokensUpdated 4 days ago
    Auto-check passed
  • LLM Integration

    rapidaai/voice-ai

    Add or modify integration-api LLM providers with caller factory wiring, unified provider routing, streaming behavior, and metric/audit compatibility.

    745 GitHub stars~749 tokensUpdated 4 days ago
    Auto-check passed

Questions about End Of Speech Integration

What does End Of Speech Integration do?

Add or modify end-of-speech integrations in assistant-api with strict separation from VAD internals. End Of Speech Integration is an agent skill from rapidaai/voice-ai. Add or modify end-of-speech integrations in assistant-api with strict separation from VAD internals.

When should I use End Of Speech Integration?

End Of Speech Integration fits situations like: transcript/audio/history-aware turn-finalization logic; provider wiring.

How do I install End Of Speech Integration in Claude Code?

Run `npx skills add rapidaai/voice-ai --skill end-of-speech-integration -a claude-code`. Or copy the skill folder (.codex/skills/end-of-speech-integration in rapidaai/voice-ai) into .claude/skills/end-of-speech-integration in your project. Claude Code loads it when a task matches its description.

How do I install End Of Speech Integration in Codex?

Run `npx skills add rapidaai/voice-ai --skill end-of-speech-integration -a codex`. Or copy the skill folder (.codex/skills/end-of-speech-integration in rapidaai/voice-ai) into .agents/skills/end-of-speech-integration in your project. Codex loads it when a task matches its description.

Can I use End Of Speech Integration in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add rapidaai/voice-ai --skill end-of-speech-integration -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/end-of-speech-integration, .gemini/skills/end-of-speech-integration, .github/skills/end-of-speech-integration and .opencode/skills/end-of-speech-integration in your project.

What does End Of Speech Integration need to run?

Going by SKILL.md and its folder, End Of Speech Integration needs a shell for the scripts in its folder and the command-line tools its instructions call (go and yarn). Our summary lists: A Bash shell.

Does End Of Speech Integration access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is End Of Speech Integration safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does End Of Speech Integration use?

End Of Speech Integration has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does End Of Speech Integration use?

About 877 tokens (SKILL.md is roughly 3.5k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 247 tokens, read only when the agent opens those files.

What are the alternatives to End Of Speech Integration?

Skills that share tags, products or a category with End Of Speech Integration: Yichen Asr (mcncarl/yichen-skills, 4.4k stars), Youtube Fetcher (JimmySadek/youtube-fetcher-to-markdown, 485 stars), Video Understanding (zenstory-ai/video-recap-skills, 561 stars) and Hri (terrense/ros2-multimodal-robot-collab, 111 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains End Of Speech Integration?

rapidaai (a GitHub organization) maintains it in rapidaai/voice-ai, which has 745 GitHub stars. The repository holds 26 skills in this directory. The repository was last updated on October 6, 2026.

Source: rapidaai/voice-ai on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.