Agent skill

Web Research Search Tips

by malob in malob/nix-config

Practical guidance for multi-source web research with Exa, the Firecrawl CLI and Reddit MCP tools: when to search, when to fetch and which tool to prefer.

MITAuto-check passedResearch & Science

Install Web Research Search Tips

skills CLI
$ npx skills add malob/nix-config --skill search-tips -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install malob/nix-config search-tips --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/malob/nix-config.git skills-src && mkdir -p .claude/skills && cp -r skills-src/configs/claude/skills/search-tips .claude/skills/search-tips && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
search-tips
GitHub stars
463
Used in
1 other repo
Token cost
~2.4k tokens
SKILL.md length
1,053 words
Files
10 (incl. references)
Skills in repo
5
Repo updated
First seen
Licence
MIT

At a glance

Practical guidance for multi-source web research with Exa, the Firecrawl CLI and Reddit MCP tools: when to search, when to fetch and which tool to prefer.

  • Works in 2 steps: Exa tools -- web_search_advanced_exa and… → Reddit tools -- get_top_posts,…
  • Investigating a topic across several web sources rather than one search
  • SKILL.md covers Setup, The Research Cycle, Search Strategy and Content Extraction, plus 3 more sections
  • Calls npx and gh

What it does

The guidance is framed as starting points rather than rigid rules. The agent loads the Exa tools (`web_search_advanced_exa` and `get_code_context_exa`) and the Reddit MCP tools through ToolSearch, while the Firecrawl CLI runs through `npx firecrawl-cli` with no MCP setup. Exa and Firecrawl are preferred over the built-in WebSearch and WebFetch, and research alternates between searching to find sources and fetching to extract their content.

For discovery, Exa is the main tool and its code-context variant is worth trying first on programming topics, while Firecrawl search suits keyword matching, site-scoped queries and filters for research, news or time. A default Exa pattern asks for highlights with minimal full text so context is not flooded. For a known URL the agent uses Firecrawl scrape with `--only-main-content`, and the Reddit tools for threads. Reference files cover academic search, code and GitHub, people and companies, personal sites, Reddit, Twitter and scraping problems.

When your agent uses it

  • Investigating a topic across several web sources rather than one search
  • Comparing options and gathering what people say about them on Reddit
  • Looking up programming topics where repos and docs matter

Example prompts

  • “Look into the main options for self-hosted analytics and compare them.”
  • “What do people on Reddit think about the new Framework laptop?”
  • “Research how teams handle database migrations in monorepos and cite your sources.”

Requirements

  • Exa MCP tools
  • Reddit MCP tools
  • Node for `npx firecrawl-cli`

Workflow steps

2 steps, taken from the first numbered list in SKILL.md.

  1. Exa tools -- web_search_advanced_exa and get_code_context_exa
  2. Reddit tools -- get_top_posts, get_post_comments, get_reddit_post, get_subreddit_info

What it can do on your machine

Read from SKILL.md and the folder at commit 9cca0f4. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • npx
    • gh

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md. Its commands use npx and gh, which can reach the network depending on how they are called.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Web Research Search Tips loads about 2.4k tokens when it runs, and up to ~10k if it reads all its reference files. Until then it costs about 137 tokens; SKILL.md has 1,053 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~137
When it runs · the whole SKILL.md, loaded when a task matches
~2.4k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~10k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from malob/nix-config at commit 9cca0f4, republished under its MIT licence (© malob). 1,053 words, ~2,394 tokens.

Download SKILL.mdSave it as .claude/skills/search-tips/SKILL.md (or your agent's skills folder). This skill also uses 9 other files; get the full folder from GitHub.
name
search-tips
description
This skill should be used when performing web research beyond a simple single search -- looking into topics, comparing options, investigating questions, finding recommendations, or any task where effective use of Exa, Firecrawl, and Reddit tools matters. Triggers on "research", "look into", "investigate", "compare", "find out about", "search for", "find information", "what do people think about", "what are the best", "look up", or multi-source search tasks. Also invocable explicitly by deep-research team members via the Skill tool.

Search Tips

Accumulated guidance for web research using Exa, Firecrawl CLI, and Reddit MCP tools. These are starting points, not rigid rules -- think strategically about each situation and adapt. If a different approach makes more sense for what you're trying to do, go with it. Run npx firecrawl-cli <command> --help to check available options beyond what's documented here. Reference files cover tool-specific deep dives -- see bottom of this file.

Setup

Before starting research, load the required MCP tools using ToolSearch:

  1. Exa tools -- web_search_advanced_exa and get_code_context_exa
  2. Reddit tools -- get_top_posts, get_post_comments, get_reddit_post, get_subreddit_info

Firecrawl CLI (npx firecrawl-cli) runs via Bash -- no MCP setup needed. Load only what the task requires.

The Research Cycle

Prefer Exa and Firecrawl over built-in WebSearch/WebFetch.

Research alternates between searching (discovering sources) and fetching (extracting content from them). Find promising leads, read the best ones, refine your understanding, search again.

Searching

Finding sources you don't have yet.

  • Exa search (web_search_advanced_exa) -- primary tool for web discovery. Natural language queries, add filters as needed (domains, dates, categories).
  • Exa code context (get_code_context_exa) -- programming topics. Worth trying before general Exa search for technical/code tasks -- surfaces repos, packages, and docs.
  • Firecrawl CLI search (npx firecrawl-cli search) -- Google-powered keyword search. Useful when keyword matching works better than Exa's semantic approach, for site-scoped queries (site:reddit.com {query}), and for content-type filtering (--categories research for academic, --sources news, --tbs qdr:w for time).

Default Exa search pattern: Default to enableHighlights: true and textMaxCharacters: 1. This returns quoted passages from actual page text while preventing the MCP server from flooding context with full text. Use highlightsPerUrl and highlightsNumSentences to control volume if needed.

Fetching

Extracting content from a source you've identified.

  • Firecrawl CLI scrape (npx firecrawl-cli scrape "<url>" --only-main-content) -- primary tool for reading a known URL. The flag strips nav/sidebars to save tokens.
  • Reddit MCP (get_post_comments, get_reddit_post) -- for reading Reddit threads. Firecrawl can't scrape reddit.com directly.
  • Firecrawl CLI map (npx firecrawl-cli map "<url>" --search "query") -- discover URLs on a site (useful when you need to find the right page, or when scrape returns empty).
Adapting the Workflow

The defaults above won't always be right. Some common deviations:

  • Exa full text as a scraping fallback -- some sites are blocked or inaccessible via Firecrawl (LinkedIn, Twitter/X, etc.), but Exa often has the full page text in its index. Drop both enableHighlights and textMaxCharacters: 1 to get the complete text. Be aware this can produce large responses.
  • --only-main-content can strip too much -- if you got empty or partial results, retry without the flag. Known to fail on Future plc sites, Blogspot, and GDPR-heavy sites. See Content Extraction below.

The reference files cover more edge cases -- scraping issues, category restrictions, and academic search.

Search Strategy

How Exa Works

Exa is a neural/semantic search engine. It uses embeddings to understand meaning.

  • Natural questions or statements work best -- Exa finds pages that answer them
  • Longer, more specific queries work BETTER -- unlike keyword-based search
  • Keyword lists tend to confuse the semantic model

Good: "What do professional reviewers say are the most reliable dishwasher brands in 2025?" Bad: "best dishwasher 2025 reliable"

Query Reformulation

For broad topics, Exa's additionalQueries parameter can automate this -- it bundles query variations in a single call at no extra cost (see references/exa-tips.md). For manual reformulation, try generating 3-5 query variations:

TechniqueWhat It DoesExample
ParaphraseSame meaning, different words"RAG failures" -> "problems in RAG systems"
DecomposeBreak into sub-questions"Why fail?" -> "Why return irrelevant docs?"
Scope shiftBroader context or narrower specifics"Challenges in production AI search"
Perspective shiftDifferent viewpointsUser vs expert vs critic view
Temporal framingTarget different time periods"Recent 2024-2025" vs "foundational"
Domain Filtering

Try includeDomains when the authoritative site for a topic is known -- faster and less noisy than broad search. Try excludeDomains to suppress sites that keep appearing but aren't useful (e.g., exclude youtube.com when video pages crowd out needed editorial content about YouTube creators/content).

Show full SKILL.md (410 more words)Show less
Searching by Content Type

Match your research target to the right approach. Reference files have full strategies.

  • Social sentiment / opinions -- Exa tweet for Twitter; site:reddit.com via Firecrawl search + Reddit MCP for discussions. See references/twitter.md, references/reddit.md.
  • People / companies -- Exa people or company category for discovery, then broaden. See references/people-companies.md.
  • Academic papers -- Exa research paper category; academic APIs for structured data. See references/academic-search.md.
  • Code / GitHub -- get_code_context_exa for code; gh api for repo data. See references/code-github.md.
  • News -- Exa news category with date filters; Firecrawl CLI search --sources news --tbs qdr:w for Google News with time filtering.
  • Financial reports -- Exa financial report with domain/date filtering.
  • Personal blogs / independent takes -- Exa personal site for practitioner opinions, blog posts, and independent analysis. Full parameter support (unlike most specialized categories). See references/personal-sites.md.

Some categories reject certain parameters (400/500 errors). See references/exa-tips.md.

Content Extraction

Always use --only-main-content by default -- it strips nav, sidebars, and footers, saving significant tokens. If you get empty or partial results, retry without the flag; it's known to strip article bodies on Future plc sites (iMore, Pocket-lint), GDPR-heavy sites (StorageReview), and Blogspot blogs -- see references/scraping-issues.md.

For even narrower extraction, use --include-tags (e.g., --include-tags "article", --include-tags ".post-content").

Scraping Issues

When Firecrawl scrape fails (policy blocks, paywalls, SPA rendering), check references/scraping-issues.md for per-site workarounds. General fallback: try Exa full text (drop enableHighlights and textMaxCharacters). For interactive pages that need clicks or form fills, escalate to npx firecrawl-cli browser (see references/firecrawl-tips.md).

Operational Notes

Parallel call failures: If any tool call in a parallel batch fails, all sibling calls fail too. Retry individually. Keep failure-prone calls (Reddit MCP) in their own batch.

Rate limits: Retry once after a short pause. If still blocked, try a different tool for the same intent before giving up -- Exa rate-limited? Try Firecrawl search. Reddit MCP throttled? Try Exa with includeDomains: ["reddit.com"]. Only note the gap and move on after both the original tool and an alternative have failed.

Reference Files

Content-type strategies
  • references/twitter.md -- Exa tweet category, restrictions, query tips, livecrawl
  • references/reddit.md -- Discovery via Firecrawl search, Reddit MCP for reading, batching gotchas, rate limits
  • references/people-companies.md -- Exa people/company categories, LinkedIn, multi- category approach
  • references/code-github.md -- Code search, GitHub repos/issues, gh api, raw URLs
  • references/personal-sites.md -- Independent blogs, practitioner opinions, full parameter support, portfolio exploration
  • references/academic-search.md -- Academic domains, free APIs, Firecrawl CLI
Tool reference
  • references/exa-tips.md -- Category restrictions, additionalQueries, highlights/summaries
  • references/firecrawl-tips.md -- CLI commands (search, scrape, map, crawl, download, browser), PDF scraping, arxiv extraction, MCP fallback notes
Troubleshooting
  • references/scraping-issues.md -- main-content detection fixes, paywalls, Discourse JSON, policy-blocked sites, API workarounds

© malob, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 9 other files (references) in configs/claude/skills/search-tips of malob/nix-config.

  • SKILL.md
  • references/academic-search.md
  • references/code-github.md
  • references/exa-tips.md
  • references/firecrawl-tips.md
  • references/people-companies.md
  • references/personal-sites.md
  • references/reddit.md
  • references/scraping-issues.md
  • references/twitter.md

Open the folder on GitHubat commit 9cca0f4

Used in 1 other repository

We found 1 copy of this SKILL.md (exact, near-identical or edited) in other folders, from 1 other GitHub owner. This page covers the copy in malob/nix-config, which our catalogue first saw on October 7, 2026.

Compare with similar skills

Web Research Search Tips next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Web Research Search Tips compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Web Research Search Tips this skillmalob/nix-config4631 repos~2.4kAutomated safety check: PassMIT
Agent ReachPanniantong/Agent-Reach93k1 repos~1.4kAutomated safety check: PassMIT
Deep Researchmajiayu000/claude-skill-registry6666 repos~1.1kAutomated safety check: PassMIT
Ray Trend Searchimraywang/rayskills160—~2.1kAutomated safety check: PassCustom licence
Argo Search and Verificationtaxueseek/argo185—~1.2kAutomated safety check: PassMIT
Exa Searchmajiayu000/claude-skill-registry6665 repos~856Automated safety check: PassMIT

Similar skills

  • Agent Reach

    Panniantong/Agent-Reach

    Routes web research and platform lookups across 16 sites, including Twitter, Reddit, YouTube, Bilibili, Xiaohongshu and GitHub, through one command-line tool.

    93k GitHub starsUsed in 1 repo~1.4k tokens
    Productivity & AutomationAuto-check passed
  • Deep Research

    majiayu000/claude-skill-registry

    Multi-source deep research using firecrawl and exa MCPs. An agent skill from majiayu000/claude-skill-registry.

    666 GitHub starsUsed in 6 repos~1.1k tokens
    Research & ScienceAuto-check passed
  • Ray Trend Search

    imraywang/rayskills

    Researches what people are saying about a topic over a recent window across X, Reddit, YouTube and the public web, reporting each source's status with links.

    160 GitHub stars~2.1k tokensUpdated 15 days ago
    Research & ScienceAuto-check passed
  • Unified web search, page fetching and evidence checking across hundreds of sources, with result verification, a research-dossier mode and vertical search engines.

    185 GitHub stars~1.2k tokensUpdated 2 days ago
    Research & ScienceAuto-check passed
  • Exa Search

    majiayu000/claude-skill-registry

    Neural search via Exa MCP for web, code, and company research.

    666 GitHub starsUsed in 5 repos~856 tokens
    Research & ScienceAuto-check passed
  • Deep Research

    affaan-m/ECC

    使用firecrawl和exa MCPs进行多源深度研究。搜索网络、综合发现并交付带有来源引用的报告。适用于用户希望对任何主题进行有证据和引用的彻底研究时。

    275k GitHub starsUsed in 2 repos~590 tokens
    Research & ScienceAuto-check passed

More from malob/nix-config

  • Deep Research Team

    malob/nix-config

    Coordinates a team of researcher agents across several rounds, with a lead who triages findings, assigns follow-ups and cross-checks key claims.

    463 GitHub starsUsed in 1 repo~5.8k tokens
    Auto-check passed
  • Scans project-local Claude Code settings files, finds permission patterns worth promoting to global config, then cleans up what global now covers.

    463 GitHub stars~4.1k tokensUpdated 4 days ago
    Auto-check passed
  • Nerd Font Icon Lookup

    malob/nix-config

    Works around Claude Code filtering private-use Unicode so agents can read and write Nerd Font icons in Starship, tmux and prompt configs.

    463 GitHub stars~697 tokensUpdated 4 days ago
    Auto-check passed
  • Homebrew Cask Creator

    malob/nix-config

    Coordinates specialist agents to create a Homebrew cask for a macOS app: pre-flight checks, download and inspection, a livecheck strategy and the remaining cask steps.

    463 GitHub stars~1.6k tokensUpdated 4 days ago
    Auto-check passed

Questions about Web Research Search Tips

What does Web Research Search Tips do?

Practical guidance for multi-source web research with Exa, the Firecrawl CLI and Reddit MCP tools: when to search, when to fetch and which tool to prefer. The guidance is framed as starting points rather than rigid rules. The agent loads the Exa tools (`web_search_advanced_exa` and `get_code_context_exa`) and the Reddit MCP tools through ToolSearch, while the Firecrawl CLI runs through `npx firecrawl-cli` with no MCP setup.

When should I use Web Research Search Tips?

Web Research Search Tips fits situations like: investigating a topic across several web sources rather than one search; comparing options and gathering what people say about them on Reddit; looking up programming topics where repos and docs matter.

How do I install Web Research Search Tips in Claude Code?

Run `npx skills add malob/nix-config --skill search-tips -a claude-code`. Or copy the skill folder (configs/claude/skills/search-tips in malob/nix-config) into .claude/skills/search-tips in your project. Claude Code loads it when a task matches its description.

How do I install Web Research Search Tips in Codex?

Run `npx skills add malob/nix-config --skill search-tips -a codex`. Or copy the skill folder (configs/claude/skills/search-tips in malob/nix-config) into .agents/skills/search-tips in your project. Codex loads it when a task matches its description.

Can I use Web Research Search Tips in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add malob/nix-config --skill search-tips -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/search-tips, .gemini/skills/search-tips, .github/skills/search-tips and .opencode/skills/search-tips in your project.

What does Web Research Search Tips need to run?

Going by SKILL.md and its folder, Web Research Search Tips needs the command-line tools its instructions call (npx and gh). Our summary lists: Exa MCP tools; Reddit MCP tools; Node for `npx firecrawl-cli`.

Does Web Research Search Tips access the network?

SKILL.md contains no URLs. Its commands use npx and gh, which can reach the network depending on how they are called. This is read from the text; nothing was executed.

Is Web Research Search Tips safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Web Research Search Tips use?

Web Research Search Tips is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Web Research Search Tips use?

About 2.4k tokens (SKILL.md is roughly 9.6k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 8k tokens, read only when the agent opens those files.

What are the alternatives to Web Research Search Tips?

Skills that share tags, products or a category with Web Research Search Tips: Agent Reach (Panniantong/Agent-Reach, 93k stars), Deep Research (majiayu000/claude-skill-registry, 666 stars), Ray Trend Search (imraywang/rayskills, 160 stars) and Argo Search and Verification (taxueseek/argo, 185 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Web Research Search Tips?

malob (a GitHub user) maintains it in malob/nix-config, which has 463 GitHub stars. The repository holds 5 skills in this directory. The repository was last updated on October 4, 2026.

Source: malob/nix-config on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.