Agent skill

Call Transcript Reliability Gate

by CALLE-AI in CALLE-AI/awesome-phone-call-agents

Offline experimental CALL-E transcript auditor that grades a returned transcript RELIABLE, SUSPECT, or UNUSABLE from text-visible ASR-hallucination symptoms before anything acts on it, and crafts…

MITAuto-check passedMedia & Creative

Install Call Transcript Reliability Gate

skills CLI
$ npx skills add CALLE-AI/awesome-phone-call-agents --skill call-transcript-reliability-gate -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install CALLE-AI/awesome-phone-call-agents call-transcript-reliability-gate --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/CALLE-AI/awesome-phone-call-agents.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/call-transcript-reliability-gate .claude/skills/call-transcript-reliability-gate && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
call-transcript-reliability-gate
GitHub stars
106
Token cost
~1.3k tokens
SKILL.md length
585 words
Files
8 (incl. scripts, references)
Skills in repo
98
Repo updated
First seen
Licence
MIT

At a glance

Offline experimental CALL-E transcript auditor that grades a returned transcript RELIABLE, SUSPECT, or UNUSABLE from text-visible ASR-hallucination symptoms before anything acts on it, and crafts…

  • Tasks that involve Meeting notes and agendas
  • SKILL.md covers When To Use, When Not To Use, Workflow and Scientific Foundation, plus 1 more section
  • Runs Python scripts from its folder; calls python3
  • Tasks that involve Speech recognition and synthesis

What it does

Call Transcript Reliability Gate is an agent skill from CALLE-AI/awesome-phone-call-agents. Offline experimental CALL-E transcript auditor that grades a returned transcript RELIABLE, SUSPECT, or UNUSABLE from text-visible ASR-hallucination symptoms before anything acts on it, and crafts ASR-risk-aware goals. It is not proof of hallucination, not a transcription accuracy certificate, and not authorization to act.

Its SKILL.md is about 1.3k tokens, which your agent loads only when the skill is triggered. The skill folder holds 9 other files, including scripts and reference files (for example `references/example-transcript-clean.json`, `references/example-transcript-unusable.json` and `references/example-transcript.json`).

It sits in Media & Creative, covering Meeting notes and agendas, Speech recognition and synthesis and Transcription. The repository describes itself as: Portable phone-call Agent Skills, apps, examples, adapters, and scheduler recipes for AI agents. The licence is MIT.

When your agent uses it

  • Tasks that involve Meeting notes and agendas
  • Tasks that involve Speech recognition and synthesis
  • Tasks that involve Transcription

Example prompts

  • “/call-transcript-reliability-gate”

Requirements

  • Python 3

What it can do on your machine

Read from SKILL.md and the folder at commit ae5f78f. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Ships 2 files in scripts/ (Python), which the agent can run.

    Shell commands in SKILL.md call:

    • python3

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Call Transcript Reliability Gate loads about 1.3k tokens when it runs, and up to ~3.8k if it reads all its reference files. Until then it costs about 89 tokens; SKILL.md has 585 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~89
When it runs · the whole SKILL.md, loaded when a task matches
~1.3k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~3.8k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); the scripts in this folder are not scanned.

SKILL.md

The full file from CALLE-AI/awesome-phone-call-agents at commit ae5f78f, republished under its MIT licence (© CALLE-AI). 585 words, ~1,317 tokens.

Download SKILL.mdSave it as .claude/skills/call-transcript-reliability-gate/SKILL.md (or your agent's skills folder). This skill also uses 7 other files; get the full folder from GitHub.
name
call-transcript-reliability-gate
description
Offline experimental CALL-E transcript auditor that grades a returned transcript RELIABLE, SUSPECT, or UNUSABLE from text-visible ASR-hallucination symptoms before anything acts on it, and crafts ASR-risk-aware goals. It is not proof of hallucination, not a transcription accuracy certificate, and not authorization to act.
license
MIT

call-transcript-reliability-gate

Before you trust what the phone heard, check how it was written down.

Every verification skill in this repository - call-review, verity-verification-core, provenance-grade, exact-ref - starts from the same assumption: the transcript is ground truth. This skill audits that assumption. Speech-to-text systems hallucinate fluent text with no basis in the audio, and those hallucinations disproportionately appear as loops, caption-credit boilerplate, and even harmful phantom content. A downstream verdict built on a hallucinated confirmation is wrong with perfect confidence.

When To Use

  • after any CALL-E call whose result will be written somewhere (booking, record update, payment) and before other transcript skills consume it
  • when a summary contains values the transcript turns seem to repeat oddly, or boilerplate no phone caller would say
  • before placing a number-critical call, to craft a goal that reduces transcription risk in the first place

When Not To Use

  • to prove the provider hallucinated; text signals are advisory reasons to re-confirm, not verdicts about the audio
  • during a call; this is strictly post-call transcript analysis plus pre-call goal crafting, because CALL-E exposes transcripts, not live audio
  • as a replacement for word-level confidence scores; CALL-E does not expose them, and this skill says so
  • to authorize any action; verdicts route work to humans, they never permit anything

Workflow

Gate a finished call
bash
python3 scripts/transcript_reliability_gate.py analyze --transcript path/to/call-result.json

Reads the real get_call_run result shape ({status, result: {transcript}}) or the flat shape used by sibling skill fixtures. Emits a card:

  • verdict: RELIABLE / SUSPECT / UNUSABLE
  • evidence: turn index, masked span, matched rules
  • signals_summary: counts per rule - loop_repetition (same 3+-word phrase repeated 3+ times in one turn), boilerplate_phantom (caption credits and video boilerplate that non-speech audio triggers), harm_violence / harm_extremism / harm_slur_prefix (documented hallucination harm categories - always routed to human review), non_english_insertion (script switch mid-call), empty_turn, no_callee_turns, single_turn_call, extreme_turn_length, empty_word_content
  • fields_to_reconfirm: numbers and date words inside suspect turns, masked
  • recommended_action: proceed_with_caution, reverify_key_fields, or do_not_act_on_transcript

Confidence is not claimed. Labels are fixed heuristic outcomes, not empirically calibrated probabilities, and every card says so.

Craft an ASR-risk-aware goal
bash
python3 scripts/transcript_reliability_gate.py craft --scenario number-critical-call

Emits the plan_call inputs JSON whose goal instructs digit-by-digit values, read-back requests, and keep-talking-during-holds behavior, so analysis and the next call stay consistent.

Show full SKILL.md (236 more words)Show less

Scientific Foundation

ResearchRelevance
Careless Whisper: Speech-to-Text Hallucination Harms (Koenecke et al., ACM FAccT 2024, arXiv 2402.08021)Documents hallucination rates and the harm taxonomy (38% of studied hallucinations contain explicit harms) our harm rules approximate
Lost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models (Atwany et al., ACL 2025, arXiv 2502.12414)Grounds the distribution-shift framing behind the language-switch and extreme-shape signals
Investigation of Whisper ASR Hallucinations Induced by Non-Speech Audio (Baranski et al., 2025, arXiv 2501.11378)Grounds the boilerplate-phantom rules: music and silence trigger caption-credit text
From Text Metrics to Model Internals: A Study of Whisper ASR Hallucination Detection (Jasinski et al., Interspeech 2026, arXiv 2606.23060)Establishes text-based hallucination detection as a paradigm on human-annotated data; our detectors are a text-only approximation of it

CALL-E exposes transcripts without word-level confidence or audio, so this skill implements the text-side approximation and labels every output analysis_mode: "heuristic". Citation notes: the FAcct paper's exact title says "Speech-to-Text", and the Interspeech study's HALAS dataset is human-annotated - our rules were not trained on it.

Differences from sibling skills

  • call-review audits whether a trusted transcript supports the structured result; this skill gates whether the transcript itself is trustworthy enough to audit.
  • provenance-grade grades how the callee knew what they said; this skill grades whether what they said was even transcribed faithfully.
  • conversation-clarify resolves ambiguity by placing a call; this skill reduces ambiguity creation in the next call's goal.

© CALLE-AI, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 7 other files (scripts, references) in skills/call-transcript-reliability-gate of CALLE-AI/awesome-phone-call-agents.

  • SKILL.md
  • references/example-transcript-clean.json
  • references/example-transcript-unusable.json
  • references/example-transcript.json
  • references/examples.md
  • references/safety.md
  • scripts/test_transcript_reliability_gate.py
  • scripts/transcript_reliability_gate.py

Open the folder on GitHubat commit ae5f78f

Compare with similar skills

Call Transcript Reliability Gate next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Call Transcript Reliability Gate compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Call Transcript Reliability Gate this skillCALLE-AI/awesome-phone-call-agents106—~1.3kAutomated safety check: PassMIT
Transcript Fixerdaymade/claude-code-skills1.4k—~10kAutomated safety check: PassMIT
Summarize Callreysu/ai-life-skills270—~3.8kAutomated safety check: NotesMIT
Watch Videocoreyhaines31/makerskills848—~3.7kAutomated safety check: PassMIT
Transcribegnekt/My-Brain-Is-Full-Crew3.9k—~5kAutomated safety check: PassCustom licence
Audio Transcriptionmitsuhiko/agent-stuff3.2k—~1kAutomated safety check: PassApache-2.0

Similar skills

  • Transcript Fixer

    daymade/claude-code-skills

    Corrects ASR/STT transcription errors — homophones, garbled terms, person-name errors, mixed Chinese/English — with dictionary rules plus Claude's built-in AI, no external API key required.

    1.4k GitHub stars~10k tokensUpdated today
    Productivity & AutomationAuto-check passed
  • Summarize Call

    reysu/ai-life-skills

    Transcribe a call recording with speaker diarization, summarize it, and create Obsidian vault notes (call note, transcript, person notes for participants).

    270 GitHub stars~3.8k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes
  • Watch Video

    coreyhaines31/makerskills

    When you want to extract content from a video — YouTube, Loom, Vimeo, Riverside, Zoom recording, local MP4, X/IG video, anything yt-dlp supports.

    848 GitHub stars~3.7k tokensUpdated 1 mo ago
    Media & CreativeAuto-check passed
  • Transcribe

    gnekt/My-Brain-Is-Full-Crew

    Process audio recordings, meeting transcripts, podcasts, or lectures.

    3.9k GitHub stars~5k tokensUpdated 3 mo ago
    Media & CreativeAuto-check passed
  • Audio Transcription

    mitsuhiko/agent-stuff

    Transcribe local audio/video and Apple Voice Memos quickly with cached MLX Whisper models, including bad/low-quality audio.

    3.2k GitHub stars~1k tokensUpdated 10 days ago
    Media & CreativeAuto-check passed
  • Speech Recognition

    dpearson2699/swift-ios-skills

    Transcribe speech to text using Apple's Speech framework. An agent skill from dpearson2699/swift-ios-skills.

    1.2k GitHub stars~3.7k tokensUpdated 2 mo ago
    Media & CreativeAuto-check passed

More from CALLE-AI/awesome-phone-call-agents

All 98 skills in this repo
  • Accessible Outing Verifier

    CALLE-AI/awesome-phone-call-agents

    Demonstrates advisory accessibility-planning checks with offline fixtures and a proposed bounded CALL-E workflow; use for exploring unknown or qualified venue claims without making calls.

    106 GitHub stars~1.4k tokensUpdated today
    Auto-check passed
  • Ground Truth Gate

    CALLE-AI/awesome-phone-call-agents

    A skill your agent uses when an agent holds some evidence for a physical-world claim but the evidence is broader, narrower, or older than the exact question asked, and it must first decide whether a…

    106 GitHub stars~3.3k tokensUpdated today
    Auto-check passed
  • Is It Accessible

    CALLE-AI/awesome-phone-call-agents

    Call a venue and ask the accessibility questions that matter to one specific person — step-free entry, hearing loop, guide dogs, quiet hours, changing places — then return a per-need verdict backed…

    106 GitHub stars~4.3k tokensUpdated today
    Auto-check passed
  • Landmark Navigation Assist

    CALLE-AI/awesome-phone-call-agents

    Turns a pre-written, building-level location config into a CALL-E outbound phone-call task that guides a delivery driver through the last few hundred metres to a specific building using landmarks…

    106 GitHub stars~1.6k tokensUpdated today
    Auto-check passed
  • Research Gap Call Verifier

    CALLE-AI/awesome-phone-call-agents

    Turn cited business research into a bounded, approval-gated phone-call plan that asks only unresolved factual questions, then reconcile CALL-E-compatible results without treating voicemail, refusal…

    106 GitHub stars~1.7k tokensUpdated today
    Auto-check passed
  • Structured Outcome Followup Call

    CALLE-AI/awesome-phone-call-agents

    Place a goal-driven CALL-E call that collects specific structured answers, score those answers against a deterministic rubric you supply, and conditionally trigger a follow-up action — all runnable…

    106 GitHub stars~1.4k tokensUpdated today
    Auto-check passed

Questions about Call Transcript Reliability Gate

What does Call Transcript Reliability Gate do?

Offline experimental CALL-E transcript auditor that grades a returned transcript RELIABLE, SUSPECT, or UNUSABLE from text-visible ASR-hallucination symptoms before anything acts on it, and crafts…. Call Transcript Reliability Gate is an agent skill from CALLE-AI/awesome-phone-call-agents. Offline experimental CALL-E transcript auditor that grades a returned transcript RELIABLE, SUSPECT, or UNUSABLE from text-visible ASR-hallucination symptoms before anything acts on it, and crafts ASR-risk-aware goals.

When should I use Call Transcript Reliability Gate?

Call Transcript Reliability Gate fits situations like: tasks that involve Meeting notes and agendas; tasks that involve Speech recognition and synthesis; tasks that involve Transcription.

How do I install Call Transcript Reliability Gate in Claude Code?

Run `npx skills add CALLE-AI/awesome-phone-call-agents --skill call-transcript-reliability-gate -a claude-code`. Or copy the skill folder (skills/call-transcript-reliability-gate in CALLE-AI/awesome-phone-call-agents) into .claude/skills/call-transcript-reliability-gate in your project. Claude Code loads it when a task matches its description.

How do I install Call Transcript Reliability Gate in Codex?

Run `npx skills add CALLE-AI/awesome-phone-call-agents --skill call-transcript-reliability-gate -a codex`. Or copy the skill folder (skills/call-transcript-reliability-gate in CALLE-AI/awesome-phone-call-agents) into .agents/skills/call-transcript-reliability-gate in your project. Codex loads it when a task matches its description.

Can I use Call Transcript Reliability Gate in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add CALLE-AI/awesome-phone-call-agents --skill call-transcript-reliability-gate -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/call-transcript-reliability-gate, .gemini/skills/call-transcript-reliability-gate, .github/skills/call-transcript-reliability-gate and .opencode/skills/call-transcript-reliability-gate in your project.

What does Call Transcript Reliability Gate need to run?

Going by SKILL.md and its folder, Call Transcript Reliability Gate needs Python for the scripts in its folder and the command-line tools its instructions call (python3). Our summary lists: Python 3.

Does Call Transcript Reliability Gate access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Call Transcript Reliability Gate safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. The check reads SKILL.md only: the scripts in the folder are not scanned, so read them before running anything.

What licence does Call Transcript Reliability Gate use?

Call Transcript Reliability Gate is published under the MIT licence (declared in SKILL.md). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Call Transcript Reliability Gate use?

About 1.3k tokens (SKILL.md is roughly 5.3k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.5k tokens, read only when the agent opens those files.

What are the alternatives to Call Transcript Reliability Gate?

Skills that share tags, products or a category with Call Transcript Reliability Gate: Transcript Fixer (daymade/claude-code-skills, 1.4k stars), Summarize Call (reysu/ai-life-skills, 270 stars), Watch Video (coreyhaines31/makerskills, 848 stars) and Transcribe (gnekt/My-Brain-Is-Full-Crew, 3.9k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Call Transcript Reliability Gate?

CALLE-AI (a GitHub organization) maintains it in CALLE-AI/awesome-phone-call-agents, which has 106 GitHub stars. The repository holds 98 skills in this directory. The repository was last updated on October 8, 2026.

Source: CALLE-AI/awesome-phone-call-agents on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.