Fal Vision
nexu-io/open-design
Analyze images — segment objects, detect, run OCR, describe, and answer visual questions via fal.ai vision models.
Builds project visions through interactive guided conversation.
$ npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install YougLin-dev/Aha-Loop vision-builder --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/YougLin-dev/Aha-Loop.git skills-src && mkdir -p .claude/skills && cp -r skills-src/.agents/skills/vision-builder .claude/skills/vision-builder && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "vision-builder" agent skill from https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builder into .claude/skills/vision-builder/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-builder", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builderType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install YougLin-dev/Aha-Loop vision-builder --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/YougLin-dev/Aha-Loop.git skills-src && mkdir -p .agents/skills && cp -r skills-src/.agents/skills/vision-builder .agents/skills/vision-builder && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "vision-builder" agent skill from https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builder into .agents/skills/vision-builder/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-builder", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install YougLin-dev/Aha-Loop vision-builder --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/YougLin-dev/Aha-Loop.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/.agents/skills/vision-builder .cursor/skills/vision-builder && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "vision-builder" agent skill from https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builder into .cursor/skills/vision-builder/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-builder", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/YougLin-dev/Aha-Loop.git --path .agents/skills/vision-builder--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install YougLin-dev/Aha-Loop vision-builder --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/YougLin-dev/Aha-Loop.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/.agents/skills/vision-builder .gemini/skills/vision-builder && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "vision-builder" agent skill from https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builder into .gemini/skills/vision-builder/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-builder", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install YougLin-dev/Aha-Loop vision-builderInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/YougLin-dev/Aha-Loop.git skills-src && mkdir -p .github/skills && cp -r skills-src/.agents/skills/vision-builder .github/skills/vision-builder && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "vision-builder" agent skill from https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builder into .github/skills/vision-builder/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-builder", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install YougLin-dev/Aha-Loop vision-builder --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/YougLin-dev/Aha-Loop.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/.agents/skills/vision-builder .opencode/skills/vision-builder && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "vision-builder" agent skill from https://github.com/YougLin-dev/Aha-Loop/tree/main/.agents/skills/vision-builder into .opencode/skills/vision-builder/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "vision-builder", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
vision-builderBuilds project visions through interactive guided conversation.
Vision Builder is an agent skill from YougLin-dev/Aha-Loop. Builds project visions through interactive guided conversation. Use when users have vague ideas needing structure. Triggers on: build vision, I have an idea, start new project, new idea.
Its SKILL.md is about 1.9k tokens, which your agent loads only when the skill is triggered. It is a single SKILL.md file with no bundled scripts.
The repository describes itself as: [MVP] Aha Loop is a fully autonomous AI development system, extended from the core ideas of Ralph. It's not just an execution engine, but a complete AI development framework with… The licence is MIT.
8 steps, taken from the step headings in SKILL.md.
Read from SKILL.md and the folder at commit 8d799b2. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
No scripts in the folder and no shell commands in SKILL.md (its code samples are markdown).
From the folder's file list and the shell code blocks in SKILL.md.
No URLs in SKILL.md.
From URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Vision Builder loads about 1.9k tokens when it runs. Until then it costs about 50 tokens; SKILL.md has 614 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from YougLin-dev/Aha-Loop at commit 8d799b2, republished under its MIT licence (© YougLin-dev). 614 words, ~1,914 tokens.
.claude/skills/vision-builder/SKILL.md (or your agent's skills folder).Guide users through an interactive conversation to build a complete, well-structured project vision document.
project.vision.md documentActivate this skill when user says things like:
Start by understanding what kind of project this is.
Question: What type of project is this?
Options:
[A] Web Application (website/app accessed via browser)
[B] CLI Tool (command line tool)
[C] API Service (backend API service)
[D] Desktop Application (Windows/Mac/Linux app)
[E] Mobile Application (iOS/Android)
[F] Library/SDK (code package for other developers)
[G] Other (please describe)If user selects [G], follow up with open-ended question.
Understand the problem being solved.
Question: What problem does this project solve?
Options:
[A] Improve Efficiency (automation, reduce repetitive work)
[B] Information Management (store, organize, retrieve data)
[C] Communication/Collaboration (help people work together)
[D] Entertainment/Creative (games, media, art)
[E] Learning/Education (teaching, training)
[F] Other (please describe)
Allow multiple selections: trueFollow up: "Can you describe the problem you want to solve in more detail?"
Identify who will use the product.
Question: Who will use this product?
Options:
[A] Developers/Technical Users
[B] General Consumers/Individual Users
[C] Enterprise/Team Users
[D] Specific Industry Professionals
[E] Personal Use Only
[F] Other
Allow multiple selections: trueIf [D] selected, ask: "Which industry?"
Understand the scope.
Question: What is the scale and ambition of this project?
Options:
[A] Small Project - Quick idea validation, completed in days
[B] Medium Project - Complete features, completed in weeks
[C] Large Project - Full product, requires months
[D] Uncertain - Help me evaluateDefine what success looks like.
Question: What defines success for this project? (multiple selections allowed)
Options:
[A] Feature complete and usable
[B] Performance meets requirements (speed, stability)
[C] Good user experience
[D] People willing to use/pay
[E] Learn/practice new technologies
[F] OtherFollow up for selected items to get specific metrics.
Gather technical constraints.
Question: Do you have technology stack preferences?
Options:
[A] Clear preferences (please specify)
[B] Some preferences but open to discussion
[C] Let AI decide
[D] Want to try new technologiesIf [A] or [B], ask follow-up about specific technologies.
Question: Are there any technologies you want to avoid?
Options:
[A] No, open to all
[B] Avoid overly complex frameworks
[C] Avoid paid/commercial components
[D] Specific technologies to avoid (please specify)Identify limitations.
Question: Are there any constraints? (multiple selections allowed)
Options:
[A] Limited time
[B] Limited budget/free only
[C] Must work offline
[D] Specific platform requirements
[E] High security/privacy requirements
[F] No special constraintsGet reference points.
Question: Are there similar products for reference?
Options:
[A] Yes, I can name a few
[B] Have a general direction but unsure of specific products
[C] No, this is a completely new idea
[D] Help me find some referencesAfter completing the conversation, generate project.vision.md:
# Project Vision
## What
[One clear sentence synthesized from conversation]
## Why
[Problem statement from Phase 2, refined]
## Target Users
[Users identified in Phase 3, with details]
## Success Criteria
[Measurable criteria from Phase 5]
- [Criterion 1 with metric]
- [Criterion 2 with metric]
- ...
## Constraints
[Constraints from Phase 7]
- [Constraint 1]
- ...
## Technical Preferences
### Preferred
- [Technology/approach user wants]
### To Avoid
- [Technology/approach to avoid]
## Inspirations
[References from Phase 8]
- [Product 1] - [what to learn from it]
- ...
## Non-Goals
[Explicitly out of scope based on conversation]
- [Non-goal 1]
- ...
---
*Vision built interactively on [date]*
*Ready for: Orchestrator → Architecture → Roadmap*Add follow-up prompts:
Offer to help:
Allow skipping optional sections:
Never skip: What, Why, Target Users, Success Criteria
User: I want to build some kind of expense tracker
AI: Let me help you refine this idea into a complete project vision.
First, what type of project is this? [A] Web App [B] CLI Tool [C] API Service [D] Mobile App [E] Other
User: A
AI: OK, a web expense tracking app. What problem does it solve? [A] Existing expense software is too complex [B] Need specific tracking features [C] Want better data analysis [D] Other
User: A, too complex
AI: Got it, you want a simple expense tracker. Who will use it? [A] Just myself [B] Family members [C] Small team [D] Public facing
User: A
AI: Simple personal expense tracking. What defines success? [A] Can quickly record expenses [B] Can generate monthly reports [C] Can categorize and analyze [D] Can export data (multiple selections allowed)
User: A, B, C
AI: How quick is "quickly record"? [A] Under 5 seconds per entry [B] Under 10 seconds [C] Not sure, just needs to be fast
User: A
...continues until vision is complete...
Before saving vision:
After vision is built:
project.vision.mdYou are a professional product consultant.
Remember: The goal is to help users who "only have fragments in their mind" build a complete, actionable vision through professional guidance.
© YougLin-dev, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
Just SKILL.md in .agents/skills/vision-builder of YougLin-dev/Aha-Loop.
Open the folder on GitHubat commit 8d799b2
Vision Builder next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Vision Builder this skillYougLin-dev/Aha-Loop | 181 | — | ~1.9k | Automated safety check: Pass | MIT | |
| Fal Visionnexu-io/open-design | 100k | — | ~295 | Automated safety check: Pass | Apache-2.0 | |
| Vision Sftwshobson/agents | 40k | — | ~2k | Automated safety check: Pass | MIT | |
| Firecrawl Interact Integrationfirecrawl/firecrawl | 190k | 1 repos | ~731 | Automated safety check: Pass | ISC | |
| Visiongridaco/grida | 2.7k | — | ~1.5k | Automated safety check: Pass | Apache-2.0 | |
| Team Builderaffaan-m/ECC | 275k | 1 repos | ~1.8k | Automated safety check: Pass | MIT |
nexu-io/open-design
Analyze images — segment objects, detect, run OCR, describe, and answer visual questions via fal.ai vision models.
wshobson/agents
Fine-tune vision-language models (VLMs) with supervised learning on image+text data.
firecrawl/firecrawl
Guides adding Firecrawl's /interact endpoint to product code for pages that need clicks, forms, pagination or logged-in flows beyond plain scraping.
gridaco/grida
Query images with a local Ollama vision model without loading the image into the main agent context.
affaan-m/ECC
Interactive picker that discovers available agent personas via the claude agents command and agents/ markdown globs, groups them into domains, has the user select up to five, dispatches them in…
PostHog/posthog
Build reusable conversion models — funnel/step conversion rates, drop-off, and time-to-convert — on either PostHog data-warehouse views (HogQL) or an external dbt project.
YougLin-dev/Aha-Loop
Defines God Committee member behavior and responsibilities with oversight authority.
YougLin-dev/Aha-Loop
Designs system architecture and selects technology stack based on vision analysis.
YougLin-dev/Aha-Loop
Reviews and cleans up outdated documentation. An agent skill from YougLin-dev/Aha-Loop.
YougLin-dev/Aha-Loop
Guides God Committee members through consensus-building for collective decisions.
YougLin-dev/Aha-Loop
Guides God Committee members through executing interventions.
YougLin-dev/Aha-Loop
Guides parallel exploration of multiple implementation approaches using git worktrees.
Builds project visions through interactive guided conversation. Vision Builder is an agent skill from YougLin-dev/Aha-Loop. Builds project visions through interactive guided conversation.
Vision Builder fits situations like: users have vague ideas needing structure; start new project.
Run `npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a claude-code`. Or copy the skill folder (.agents/skills/vision-builder in YougLin-dev/Aha-Loop) into .claude/skills/vision-builder in your project. Claude Code loads it when a task matches its description.
Run `npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a codex`. Or copy the skill folder (.agents/skills/vision-builder in YougLin-dev/Aha-Loop) into .agents/skills/vision-builder in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add YougLin-dev/Aha-Loop --skill vision-builder -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/vision-builder, .gemini/skills/vision-builder, .github/skills/vision-builder and .opencode/skills/vision-builder in your project.
SKILL.md names no scripts, command-line tools or credentials: Vision Builder is instructions for the agent only.
SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Vision Builder is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 1.9k tokens (SKILL.md is roughly 7.7k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.
Skills that share tags, products or a category with Vision Builder: Fal Vision (nexu-io/open-design, 100k stars), Vision Sft (wshobson/agents, 40k stars), Firecrawl Interact Integration (firecrawl/firecrawl, 190k stars) and Vision (gridaco/grida, 2.7k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
YougLin-dev (a GitHub user) maintains it in YougLin-dev/Aha-Loop, which has 181 GitHub stars. The repository holds 15 skills in this directory. The repository was last updated on February 3, 2026.
Source: YougLin-dev/Aha-Loop on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.