Agent skill

Training Honest Models

by flyrank-bih in flyrank-bih/flyrank-ml-internship-starter

Trains a first model the honest way — method chosen to fit the question, compared against the baseline on the same split and metric, errors read before scores are believed.

Custom licenceAuto-check passed

Install Training Honest Models

skills CLI
$ npx skills add flyrank-bih/flyrank-ml-internship-starter --skill training-honest-models -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install flyrank-bih/flyrank-ml-internship-starter training-honest-models --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/flyrank-bih/flyrank-ml-internship-starter.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/training-honest-models .claude/skills/training-honest-models && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
training-honest-models
GitHub stars
140
Token cost
~588 tokens
SKILL.md length
300 words
Files
1
Skills in repo
13
Repo updated
First seen
Licence
Custom licence

At a glance

Trains a first model the honest way — method chosen to fit the question, compared against the baseline on the same split and metric, errors read before scores are believed.

  • Moving from a rule baseline to a learned model
  • SKILL.md covers Choose the method to fit the…, The comparison table…, Read the errors and Reproducibility basics, plus 1 more section
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md
  • Reviewing a model that reports only a single score

What it does

Training Honest Models is an agent skill from flyrank-bih/flyrank-ml-internship-starter. Trains a first model the honest way — method chosen to fit the question, compared against the baseline on the same split and metric, errors read before scores are believed. Use when moving from a rule baseline to a learned model, or when reviewing a model that reports only a single score.

Its SKILL.md is about 590 tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.

The repository describes itself as: Starter repo for the FlyRank ML Internship - a runnable ML pipeline on real anonymized Google Search data, with Colab notebooks. Fork it, build your capstone in it.

When your agent uses it

  • Moving from a rule baseline to a learned model
  • Reviewing a model that reports only a single score

Example prompts

  • “Use the training-honest-models skill to train a first model the honest way — method chosen to fit the question, compared against the baseline on the…”
  • “/training-honest-models”

What it can do on your machine

Read from SKILL.md and the folder at commit 882b73e. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md.

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Training Honest Models loads about 588 tokens when it runs. Until then it costs about 78 tokens; SKILL.md has 300 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~78
When it runs · the whole SKILL.md, loaded when a task matches
~588

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

Its licence (Custom licence) doesn't allow us to republish the file, so here is its outline and opening line. It has 300 words (~588 tokens).

“The model is the easy part. The honesty is the work: same data, same split, same metric as the baseline — then read the errors before believing the score.”

— opening of SKILL.md by flyrank-bih, Custom licence
name
training-honest-models

Read the full SKILL.md on GitHub

Files

Just SKILL.md in skills/training-honest-models of flyrank-bih/flyrank-ml-internship-starter.

Open the folder on GitHubat commit 882b73e

Compare with similar skills

Training Honest Models next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Training Honest Models compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Training Honest Models this skillflyrank-bih/flyrank-ml-internship-starter140—~588Automated safety check: PassCustom licence
Ito Trainingaffaan-m/ECC277k1 repos~1.5kAutomated safety check: PassMIT
Train Poseruvnet/RuView97k—~504Automated safety check: PassMIT
Ray Train Distributed TrainingOrchestra-Research/AI-Research-SKILLs13k2 repos~2.7kAutomated safety check: PassMIT
Fal Trainnexu-io/open-design100k—~293Automated safety check: PassApache-2.0
Neural Trainingruvnet/ruflo74k1 repos~432Automated safety check: PassMIT

Similar skills

  • Ito Training

    affaan-m/ECC

    Inspect the availability of ML training on a completed Itô compute booking and, when the canonical backend becomes available, hand off an explicitly confirmed training manifest.

    277k GitHub starsUsed in 1 repo~1.5k tokens
    AI & LLM EngineeringAuto-check passed
  • Train Pose

    ruvnet/RuView

    Train/evaluate WiFi pose models honestly — camera-supervised (MediaPipe + CSI) and camera-free (WiFlow), always checked against the mean-pose baseline before any PCK is quoted.

    97k GitHub stars~504 tokensUpdated today
    DevelopmentAuto-check passed
  • Ray Train Distributed Training

    Orchestra-Research/AI-Research-SKILLs

    Scales PyTorch, TensorFlow and Hugging Face training from a single GPU to multi-node clusters with Ray Train, including Ray Tune sweeps and checkpoint recovery.

    13k GitHub starsUsed in 2 repos~2.7k tokens
    AI & LLM EngineeringAuto-check passed
  • Fal Train

    nexu-io/open-design

    Train custom AI models (LoRA) on fal.ai for personalized image generation tailored to a brand, character, or style.

    100k GitHub stars~293 tokensUpdated yesterday
    AI & LLM EngineeringAuto-check passed
  • Neural Training

    ruvnet/ruflo

    Neural pattern training with SONA (Self-Optimizing Neural Architecture), MoE (Mixture of Experts), and EWC++ for knowledge consolidation.

    74k GitHub starsUsed in 1 repo~432 tokens
    Auto-check passed
  • Neural Train

    ruvnet/ruflo

    Train SONA + MicroLoRA neural patterns from successful task completions; runs the DISTILL + CONSOLIDATE phases of the 4-step pipeline

    74k GitHub stars~1.2k tokensUpdated yesterday
    Auto-check: notes

More from flyrank-bih/flyrank-ml-internship-starter

All 13 skills in this repo
  • Querying Big Datasets

    flyrank-bih/flyrank-ml-internship-starter

    Works with datasets far too big to download or load in pandas — SQL over remote Parquet with DuckDB, aggregate-then-model, iterate on samples.

    140 GitHub stars~750 tokensUpdated 1 mo ago
    Auto-check passed
  • Building Baselines

    flyrank-bih/flyrank-ml-internship-starter

    Builds the transparent rule-based baseline every model must beat — a hand-written score with reason codes, ranked output, and precision@K evaluation.

    140 GitHub stars~587 tokensUpdated 1 mo ago
    Auto-check passed
  • Deploying Static Pages

    flyrank-bih/flyrank-ml-internship-starter

    Deploys a static page (research paper, portfolio piece) for free from a GitHub repo using GitHub Pages — setup, file layout, verification, and recording the final URL.

    140 GitHub stars~564 tokensUpdated 1 mo ago
    Auto-check passed
  • Framing ML Problems

    flyrank-bih/flyrank-ml-internship-starter

    Frames a data/ML problem before any modeling — the decision, the action, the cost of a wrong call, task type, target, and success metric.

    140 GitHub stars~700 tokensUpdated 1 mo ago
    Auto-check passed
  • Writing Honest Claims

    flyrank-bih/flyrank-ml-internship-starter

    Writes findings in language the evidence can carry — the claim ladder (observed → directional → decision-support, never causal without a design), effect sizes over drama, banned phrasings.

    140 GitHub stars~665 tokensUpdated 1 mo ago
    Auto-check passed
  • Writing Research Papers

    flyrank-bih/flyrank-ml-internship-starter

    Structures and writes a public research page — canonical sections (abstract through limitations), storytelling that carries findings, chart hygiene, referencing for credibility, repurposing for…

    140 GitHub stars~787 tokensUpdated 1 mo ago
    Auto-check passed

Questions about Training Honest Models

What does Training Honest Models do?

Trains a first model the honest way — method chosen to fit the question, compared against the baseline on the same split and metric, errors read before scores are believed. Training Honest Models is an agent skill from flyrank-bih/flyrank-ml-internship-starter. Trains a first model the honest way — method chosen to fit the question, compared against the baseline on the same split and metric, errors read before scores are believed.

When should I use Training Honest Models?

Training Honest Models fits situations like: moving from a rule baseline to a learned model; reviewing a model that reports only a single score.

How do I install Training Honest Models in Claude Code?

Run `npx skills add flyrank-bih/flyrank-ml-internship-starter --skill training-honest-models -a claude-code`. Or copy the skill folder (skills/training-honest-models in flyrank-bih/flyrank-ml-internship-starter) into .claude/skills/training-honest-models in your project. Claude Code loads it when a task matches its description.

How do I install Training Honest Models in Codex?

Run `npx skills add flyrank-bih/flyrank-ml-internship-starter --skill training-honest-models -a codex`. Or copy the skill folder (skills/training-honest-models in flyrank-bih/flyrank-ml-internship-starter) into .agents/skills/training-honest-models in your project. Codex loads it when a task matches its description.

Can I use Training Honest Models in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add flyrank-bih/flyrank-ml-internship-starter --skill training-honest-models -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/training-honest-models, .gemini/skills/training-honest-models, .github/skills/training-honest-models and .opencode/skills/training-honest-models in your project.

What does Training Honest Models need to run?

SKILL.md names no scripts, command-line tools or credentials: Training Honest Models is instructions for the agent only.

Does Training Honest Models access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Training Honest Models safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Training Honest Models use?

Training Honest Models has a licence file (the repository's licence) that doesn't match a standard licence. Read it on GitHub before reusing the skill.

How many tokens does Training Honest Models use?

About 588 tokens (SKILL.md is roughly 2.4k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Training Honest Models?

Skills that share tags, products or a category with Training Honest Models: Ito Training (affaan-m/ECC, 277k stars), Train Pose (ruvnet/RuView, 97k stars), Ray Train Distributed Training (Orchestra-Research/AI-Research-SKILLs, 13k stars) and Fal Train (nexu-io/open-design, 100k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Training Honest Models?

flyrank-bih (a GitHub organization) maintains it in flyrank-bih/flyrank-ml-internship-starter, which has 140 GitHub stars. The repository holds 13 skills in this directory. The repository was last updated on August 20, 2026.

Source: flyrank-bih/flyrank-ml-internship-starter on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.