Agent skill

Lilly Community Research

by ssaaffaakk in ssaaffaakk/Lilly

Lilly community-research skill. An agent skill from ssaaffaakk/Lilly.

MITAuto-check passedAI & LLM Engineering

Install Lilly Community Research

skills CLI
$ npx skills add ssaaffaakk/Lilly --skill lilly-community-research -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install ssaaffaakk/Lilly lilly-community-research --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/ssaaffaakk/Lilly.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.claude/skills/lilly-community-research .claude/skills/lilly-community-research && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
lilly-community-research
GitHub stars
171
Token cost
~1.4k tokens
SKILL.md length
570 words
Files
6 (incl. references)
Skills in repo
1
Repo updated
First seen
Licence
MIT

At a glance

Lilly community-research skill. An agent skill from ssaaffaakk/Lilly.

  • Working on any of those four components
  • SKILL.md covers When to use, When NOT to use, Known mistake → candidate fix… and Files
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • The user mentions Croatian drift

What it does

Lilly Community Research is an agent skill from ssaaffaakk/Lilly. Lilly community-research skill. Community-sourced (Reddit) fixes, tools and failure-mode warnings for Lilly's four components: Whisper ASR (Croatian drift, KenLM rescoring, low-resource fine-tuning), MarianMT translation (chrF2 plateau, BCS lexicon drift, glossary/constrained decoding), OCR (invented words, Cyrillic blindness, diacritics, synthetic-to-real gap) and Piper TTS (voice ceiling, phonemizer/espeak problems). Use when working on any of those four components, or when the user mentions Croatian drift…

Its SKILL.md is about 1.4k tokens, which your agent loads only when the skill is triggered. The skill folder holds 6 other files, including reference files (for example `references/ocr.md`, `references/sources.md` and `references/speech-asr.md`).

It sits in AI & LLM Engineering, covering Speech recognition and synthesis, Text to speech and voice and Translation. It works with Reddit, Kaggle and FastAPI. The repository describes itself as: Bosnian-first offline assistant: type, speak, and snap street text → English. FastAPI app + Kaggle-trained listen/read models. Fail-stop training. The licence is MIT.

When your agent uses it

  • Working on any of those four components
  • The user mentions Croatian drift
  • Invented OCR words
  • KenLM rescoring

Example prompts

  • “community tips”
  • “/lilly-community-research”

What it can do on your machine

Read from SKILL.md and the folder at commit 4ebfde6. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Lilly Community Research loads about 1.4k tokens when it runs, and up to ~9.7k if it reads all its reference files. Until then it costs about 226 tokens; SKILL.md has 570 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~226
When it runs · the whole SKILL.md, loaded when a task matches
~1.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~9.7k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from ssaaffaakk/Lilly at commit 4ebfde6, republished under its MIT licence (© ssaaffaakk). 570 words, ~1,351 tokens.

Download SKILL.mdSave it as .claude/skills/lilly-community-research/SKILL.md (or your agent's skills folder). This skill also uses 5 other files; get the full folder from GitHub.
name
lilly-community-research
description
Lilly community-research skill. Community-sourced (Reddit) fixes, tools and failure-mode warnings for Lilly's four components: Whisper ASR (Croatian drift, KenLM rescoring, low-resource fine-tuning), MarianMT translation (chrF2 plateau, BCS lexicon drift, glossary/constrained decoding), OCR (invented words, Cyrillic blindness, diacritics, synthetic-to-real gap) and Piper TTS (voice ceiling, phonemizer/espeak problems). Use when working on any of those four components, or when the user mentions Croatian drift, invented OCR words, Cyrillic OCR, diacritics, rs_cyrillic, KenLM rescoring, whisper speech improvement, language-token mislabelling, low-resource NMT, backtranslation, glossary injection, constrained decoding, chrF not moving, Piper voice quality, TTS espeak phonemes, or "community tips". Provides concrete, measured-mistake-mapped techniques with Reddit sources.

Lilly Community Research — community-sourced fixes, mapped to our measured mistakes

A permanent index of Reddit-sourced techniques for Lilly's four components. Built 13 Sep 2026 from a deep Reddit research pass. Each reference maps every technique to the measured failure it fixes, and each has its provenance in references/sources.md.

GATE RULE — read first. This skill is advice, not a verdict. No technique in here ships, changes a pre-registered number, or becomes a new baseline unless it runs the same pre-registered gate as any other arm: PREREGISTRATION.md, score through app.translate.Engine / app.ocr.scan, on the held-out rulers, never past a metric you carved for it. The proper shape of every idea here is: new arm in training/PREREGISTRATION.md, run it, write RESULTS, then judge. If a technique below contradicts a result file already in the repo, the result file wins until a new arm beats it at a gate.

When to use

  • Working on any of: speech listener (Whisper), translation (Marian/opus-mt, en↔bs), camera reader (OCR), or TTS voice (Piper).
  • User mentions: Croatian drift, invented OCR words, Cyrillic, diacritics, KenLM, language-token mislabelling, chrF plateau, glossary/constrained decoding, espeak, WER not improving, "community tips", "what does the community do".

When NOT to use

  • Nothing here replaces docs/* (what-moves-the-model, diacritic-gate-literature, OCR-ROADMAP, HANDOFF.md). The repo docs are the settled record; this skill is the un-assessed idea shelf.
  • If a technique contradicts a repo result file, the result file wins (see gate rule).
  • Do not cite a Reddit thread as proof for a claim about Lilly's own numbers. Threads are evidence that someone tried the technique; our numbers decide if it works here.
Show full SKILL.md (308 more words)Show less

Known mistake → candidate fix map (the whole shelf at a glance)

ComponentMeasured mistake / numberCandidate fix (details in reference)
SpeechCroatian substitution 1.1%→6.1%, large-v3KenLM/Ijekavian n-gram rescoring — no retrain
SpeechCroatian spelling via decode (vreme/dete)initial_prompt + suppress_tokens on Croatian forms
Speech<bs> token in front of Croatian text (root cause)Force language=bs + task=transcribe per segment
Speechdrift cascades window→windowcondition_on_previous_text=False + VAD gate
Speechsmall-data fine-tune regressed WER1000–3000 steps, checkpoint-per-bucket eval, bs holdout
Speechdecoder picked up hr prior (11h bs vs 91h hr)freeze encoder / tune only decoder (Distil-Whisper lesson)
Speechappended hr data hurt old perfone consolidated pass, bs upweighted, bs holdout stops
Translationfine-tune does not move chrF2 (−0.16 tie)data must be genuinely out-of-corpus; re-check overlap
TranslationCroatian lexicon losses (vlak/voz, travnja/aprila…)glossary injection + constrained decoding at inference
TranslationCroatian-lexicon drift is "dictionary not GPU"zero-training decode bias against hr pairs (lever #2)
OCR450 invented words at floor 0.9word-level gating + two-engine agreement + charset whitelist
OCRCyrillic unreadable (272/1,702 crops)route crops to rs_cyrillic / cyrillic_g2, then transliterate
OCRđ = 8 real examples (weak diacritic column)synth đ/d pair crops, sign-font family, diacritic-clear fonts
OCR100% on postcards, 53% on real photossynthetic backbone → real-data fine-tune; measure camera noise
OCRdetector false positives (39 empty crops)box textness/score filter + glare pre-processing
TTSall voices ~52% vs sr_RS 22.3%treat TTS source as eval confound; eval buckets per phoneme set
TTSespeak sr phonemes on 72% of bs utterancesverify espeak.voice + base model; phoneme-cache diagnostic
TTS~52% ceiling / mel plateaugrapheme-VITS run skips espeak (bs is phonemic)

Every one of those has a Reddit thread in references/sources.md.

Files

  • references/speech-asr.md — Whisper, Croatian drift, KenLM, low-resource recipe
  • references/translation.md — NMT, chrF2, back-translation, lexicon/constrained decoding
  • references/ocr.md — invented words, Cyrillic, diacritics, synthetic-to-real
  • references/tts.md — Piper voice, espeak phonemes, eval confounds
  • references/sources.md — every Reddit URL with the finding it supports

© ssaaffaakk, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 5 other files (references) in .claude/skills/lilly-community-research of ssaaffaakk/Lilly.

  • SKILL.md
  • references/ocr.md
  • references/sources.md
  • references/speech-asr.md
  • references/translation.md
  • references/tts.md

Open the folder on GitHubat commit 4ebfde6

Compare with similar skills

Lilly Community Research next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Lilly Community Research compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Lilly Community Research this skillssaaffaakk/Lilly171—~1.4kAutomated safety check: PassMIT
Piper Tts Trainingsammcj/agentic-coding162—~1.4kAutomated safety check: PassApache-2.0
Video Translatorshang-zhu/violin1.1k—~1kAutomated safety check: NotesMIT
App Builderqualcomm/qai-appbuilder247—~2.6kAutomated safety check: PassCustom licence
Groq Core Workflow Bjeremylongshore/tons-of-skills-marketplace2.8k—~1.4kAutomated safety check: PassMIT
Agentstadaspetra/loop2961 repos~2.5kAutomated safety check: PassMIT

Similar skills

  • Piper Tts Training

    sammcj/agentic-coding

    Train custom TTS voices for Piper (ONNX format) using fine-tuning or from-scratch approaches.

    162 GitHub stars~1.4k tokensUpdated 2 days ago
    AI & LLM EngineeringAuto-check passed
  • Video Translator

    shang-zhu/violin

    Dub a video into another language and generate subtitles using the default Together + Cartesia stack.

    1.1k GitHub stars~1k tokensUpdated 1 mo ago
    Media & CreativeAuto-check: notes
  • App Builder

    qualcomm/qai-appbuilder

    App Builder — generate complete, runnable fullstack WebUI applications (FastAPI backend + pure HTML/CSS/JS frontend) around on-device AI Model Packs (OCR, TTS, ASR, Super-Resolution, etc.).

    247 GitHub stars~2.6k tokensUpdated yesterday
    Backend & APIsAuto-check passed
  • Groq Core Workflow B

    jeremylongshore/tons-of-skills-marketplace

    A skill your agent uses when you need Groq's non-chat endpoints — transcribing or translating audio with Whisper, understanding images with Llama 4 vision, generating speech (TTS), or benchmarking…

    2.8k GitHub stars~1.4k tokensUpdated yesterday
    Media & CreativeAuto-check passed
  • Agents

    tadaspetra/loop

    Build voice AI agents with ElevenLabs. An agent skill from tadaspetra/loop.

    296 GitHub starsUsed in 1 repo~2.5k tokens
    AI & LLM EngineeringAuto-check passed
  • Elevenlabs Agents

    jezweb/claude-skills

    Build conversational AI voice agents on the ElevenLabs platform.

    1.1k GitHub stars~3.3k tokensUpdated 2 days ago
    AI & LLM EngineeringAuto-check passed

Questions about Lilly Community Research

What does Lilly Community Research do?

Lilly community-research skill. An agent skill from ssaaffaakk/Lilly. Lilly Community Research is an agent skill from ssaaffaakk/Lilly. Lilly community-research skill.

When should I use Lilly Community Research?

Lilly Community Research fits situations like: working on any of those four components; the user mentions Croatian drift; invented OCR words; kenLM rescoring.

How do I install Lilly Community Research in Claude Code?

Run `npx skills add ssaaffaakk/Lilly --skill lilly-community-research -a claude-code`. Or copy the skill folder (.claude/skills/lilly-community-research in ssaaffaakk/Lilly) into .claude/skills/lilly-community-research in your project. Claude Code loads it when a task matches its description.

How do I install Lilly Community Research in Codex?

Run `npx skills add ssaaffaakk/Lilly --skill lilly-community-research -a codex`. Or copy the skill folder (.claude/skills/lilly-community-research in ssaaffaakk/Lilly) into .agents/skills/lilly-community-research in your project. Codex loads it when a task matches its description.

Can I use Lilly Community Research in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add ssaaffaakk/Lilly --skill lilly-community-research -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/lilly-community-research, .gemini/skills/lilly-community-research, .github/skills/lilly-community-research and .opencode/skills/lilly-community-research in your project.

What does Lilly Community Research need to run?

SKILL.md names no scripts, command-line tools or credentials: Lilly Community Research is instructions for the agent only.

Does Lilly Community Research access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Lilly Community Research safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Lilly Community Research use?

Lilly Community Research is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Lilly Community Research use?

About 1.4k tokens (SKILL.md is roughly 5.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 8.3k tokens, read only when the agent opens those files.

What are the alternatives to Lilly Community Research?

Skills that share tags, products or a category with Lilly Community Research: Piper Tts Training (sammcj/agentic-coding, 162 stars), Video Translator (shang-zhu/violin, 1.1k stars), App Builder (qualcomm/qai-appbuilder, 247 stars) and Groq Core Workflow B (jeremylongshore/tons-of-skills-marketplace, 2.8k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Lilly Community Research?

ssaaffaakk (a GitHub user) maintains it in ssaaffaakk/Lilly, which has 171 GitHub stars. The repository was last updated on October 7, 2026.

Source: ssaaffaakk/Lilly on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.