Search

AI & LLM Engineering · By Archive228

4 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Get JSON out of the model reliably. An agent skill from Archive228/loopkit.

Archive228/loopkit755—~830Automated safety check: PassMIT2 mo ago
2

Build a repeatable eval loop that grades agent output with an LLM judge, so prompt/skill changes get scored against a baseline instead of eyeballed.

Archive228/loopkit755—~876Automated safety check: PassMIT2 mo ago
3

Cache the parts of the prompt that don't change so a long-running loop stops paying full price on every turn.

Archive228/loopkit755—~735Automated safety check: PassMIT2 mo ago
4

Split the Plan/Act/Verify loop across three model tiers — frontier planner, cheap executor, frontier judge — via env vars read by run.sh.

Archive228/loopkit755—~447Automated safety check: PassMIT2 mo ago