Agent skill

Perplexity Web MCP Bridge

by jacob-bd in jacob-bd/perplexity-web-mcp

Searches the web and queries premium AI models through Perplexity's own web interface, tracking quota carefully across quick, Pro and Deep Research tiers.

MITAuto-check passedProductivity & Automation

Install Perplexity Web MCP Bridge

skills CLI
$ npx skills add jacob-bd/perplexity-web-mcp --skill perplexity-web-mcp -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install jacob-bd/perplexity-web-mcp perplexity-web-mcp --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/jacob-bd/perplexity-web-mcp.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/perplexity-web-mcp .claude/skills/perplexity-web-mcp && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
perplexity-web-mcp
GitHub stars
197
Token cost
~7.6k tokens
SKILL.md length
1,771 words
Files
4 (incl. references)
Skills in repo
1
Repo updated
First seen
Licence
MIT

At a glance

Searches the web and queries premium AI models through Perplexity's own web interface, tracking quota carefully across quick, Pro and Deep Research tiers.

  • Works in 5 steps: Authenticate first: Run pwm login before… → Tokens last ~30 days: Re-run pwm login… → Check quota before your first query… → …
  • Searching the web through Perplexity when ordinary search isn't enough
  • SKILL.md covers Quick Reference, Critical Rules, Quota-Aware Usage Protocol… and Tool Detection, plus 7 more sections
  • Instructions only: no scripts, shell commands, URLs or credentials in SKILL.md

What it does

Wraps Perplexity's web interface as both a CLI (`pwm ask`, `pwm chat`, `pwm research`) and a set of MCP tools, with a dedicated AI-oriented reference command recommended as the first step for commands, models, auth flows and error recovery. Login happens once up front and tokens last roughly a month, so a 403 error means re-running login rather than debugging further.

A fixed cost model governs every query: quick or Sonar 2 queries and standard Pro searches each cost one Pro search from a shared weekly pool, a council query costs one search per model plus one synthesis search from that same pool, and Deep Research draws from a much smaller monthly pool instead. Quota is checked before the first query of every session, with Max-only models excluded automatically on a Pro subscription and usage restricted to quick or Sonar 2 once the weekly Pro pool drops below a fifth remaining.

Two rules are treated as absolute: default to the cheapest quick tier and escalate to Pro only when a query genuinely needs it, and never invoke Deep Research on its own initiative, only when the user explicitly asks for it, since a single Deep Research call can consume a meaningful share of the small monthly allowance.

When your agent uses it

  • Searching the web through Perplexity when ordinary search isn't enough
  • Querying a specific premium model like Claude or Grok through Perplexity
  • Checking remaining Pro and Deep Research quota before a research session
  • Running a council query across several models with one synthesized answer

Example prompts

  • “pwm ask what's the latest on this ongoing news story, using quick search.”
  • “Check my Perplexity quota before we start today's research session.”
  • “Run a Deep Research query on this topic — I know it'll use a monthly credit.”

Requirements

  • The pwm CLI, logged in via pwm login
  • A Perplexity account with an active subscription

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Authenticate first: Run pwm login before any queries
  2. Tokens last ~30 days: Re-run pwm login on 403 errors
  3. Check quota before your first query every session (see protocol below)
  4. Default to quick/Sonar 2 — only escalate when the query genuinely needs Pro
  5. Never use Deep Research autonomously — only when the user explicitly asks

What it can do on your machine

Read from SKILL.md and the folder at commit e34b508. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    No scripts in the folder and no shell commands in SKILL.md (its code samples are bash).

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    No URLs in SKILL.md.

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names no API keys, tokens, secrets or passwords.

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Perplexity Web MCP Bridge loads about 7.6k tokens when it runs, and up to ~14k if it reads all its reference files. Until then it costs about 131 tokens; SKILL.md has 1,771 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~131
When it runs · the whole SKILL.md, loaded when a task matches
~7.6k
With references · SKILL.md plus every file in references/, read only if the agent opens them
~14k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from jacob-bd/perplexity-web-mcp at commit e34b508, republished under its MIT licence (© jacob-bd). 1,771 words, ~7,625 tokens.

Download SKILL.mdSave it as .claude/skills/perplexity-web-mcp/SKILL.md (or your agent's skills folder). This skill also uses 3 other files; get the full folder from GitHub.
name
perplexity-web-mcp
description
Search the web and query AI models via Perplexity AI using perplexity-web-mcp-cli. Supports CLI commands (pwm ask, pwm chat, pwm research, pwm config), MCP tools (pplx_*), and Anthropic/OpenAI-compatible API server. Use when the user mentions "perplexity", "pplx", "pwm", "web search with AI", "chat with Perplexity", "deep research", "search the internet", or wants to query premium models like GPT-6 Sol, GPT-5.6 Terra, Grok 4.7, Claude, Gemini, GLM, Kimi, or Nemotron through Perplexity's web interface.
metadata.version
0.16.1
metadata.author
Jacob BD

Perplexity Web MCP

Search the web and query premium AI models through Perplexity AI.

Quick Reference

Run pwm --ai for comprehensive AI-optimized documentation covering all commands, models, MCP tools, auth flows, and error recovery.

bash
pwm --ai                # Full AI reference (RECOMMENDED first step)
pwm --help              # CLI help
pwm login --check       # Check auth status

Critical Rules

  1. Authenticate first: Run pwm login before any queries
  2. Tokens last ~30 days: Re-run pwm login on 403 errors
  3. Check quota before your first query every session (see protocol below)
  4. Default to quick/Sonar 2 — only escalate when the query genuinely needs Pro
  5. Never use Deep Research autonomously — only when the user explicitly asks

Quota-Aware Usage Protocol (MANDATORY)

Perplexity has hard quota limits. Wasting Pro queries on simple lookups exhausts the weekly pool fast, leaving nothing for questions that actually need it.

Cost Model
TierWhat It CostsResetsTypical Pool
Sonar 2 / quick1 Pro SearchWeekly~300/week
Pro Search (standard/detailed, pplx_ask, pplx_query, all model-specific tools)1 Pro Search queryWeekly~300/week
Council (pplx_council, pwm council)N+1 Pro Searches (1 per model + 1 Sonar 2 synthesis)Weekly~300/week (shared)
Deep Research (pplx_deep_research, research intent)1 Deep Research queryMonthly~5-10/month
Before Every Session
  1. Check quota first: Call pplx_usage() (MCP) or pwm usage (CLI) before your first query.
  2. Review the remaining Pro and Research counts and the Subscription line.
  3. If Subscription is Pro, exclude Max-only models (gpt56_sol, claude_opus, claude_opus55) from model selection and councils.
  4. If Pro < 20% remaining, restrict yourself to quick/Sonar 2 for everything except user-requested Pro queries.
Before Every Query: Choose the Lowest Sufficient Tier

Ask yourself: "Can Sonar 2 answer this?" If yes, use quick. Only escalate if the answer is no.

Use quick (Sonar 2 — 1 Pro Search, cheapest option) when the query is:

  • A factual lookup: "What is the capital of France?"
  • A definition: "What does CORS stand for?"
  • A simple current-event check: "Who won the Super Bowl?"
  • A quick status/version check: "What is the latest version of React?"
  • A straightforward how-to that's well-documented: "How do I create a venv in Python?"
  • A single-fact retrieval: "What is the population of Tokyo?"
  • A simple translation or conversion: "How many meters in a mile?"

Use standard (1 Pro Search) when the query:

  • Needs synthesis across multiple web sources: "Compare Next.js and Remix for SSR"
  • Requires very current data from multiple sources: "What happened in AI this week?"
  • Asks for a how-to with nuance: "Best practices for PostgreSQL indexing in 2026"
  • Needs cited sources for credibility: "What are the side effects of metformin?"
  • Involves a real comparison or tradeoff analysis

Use detailed (1 Pro Search, premium model) when the query:

  • Requires complex multi-step reasoning: "Analyze the pros/cons of microservices vs monolith for a 10-person startup"
  • Demands deep technical analysis: "Explain the differences between Raft and Paxos consensus algorithms"
  • Needs authoritative synthesis with reasoning: "What are the economic implications of the new EU AI Act?"

Use research (1 Deep Research — scarce) ONLY when:

  • The user explicitly asks for "deep research", "comprehensive report", or similar
  • Never use autonomously — always ask the user first
  • Falls back to premium Pro Search if research quota is exhausted

Use council (N+1 Pro Searches — expensive) when:

  • The user needs high-confidence answers validated across multiple AI providers
  • Important decisions, fact-checking, or complex analysis
  • BEFORE calling: ASK the user which models and how many (each = 1 Pro Search)
  • Available models: sonar, gpt56_terra, gpt56_sol, gpt6_sol, grok45, grok47, claude_sonnet, claude_opus, claude_opus55, gemini_pro, gemini38, nemotron, glm52, glm53, kimi_k26, kimi_k3
  • Max-only models: gpt56_sol, claude_opus, claude_opus55. Do not use these for Pro subscriptions.
  • The catalog is server-driven: pwm models prints the live list (free, no query quota) and every identifier it shows works with -m.
  • Default: 3 Pro-compatible models (GPT-5.6 Terra, Claude Sonnet, Gemini Pro) + synthesis = 4 Pro Searches
Decision Flowchart
You want to query Perplexity...
│
├─ Is this a simple fact, definition, or well-known how-to?
│  └─ YES → intent='quick' (Sonar 2, 1 Pro Search)
│
├─ Does it need multiple current web sources or cited synthesis?
│  └─ YES → intent='standard' (1 Pro Search)
│
├─ Does it need deep reasoning, complex analysis, or premium model quality?
│  └─ YES → intent='detailed' (1 Pro Search, premium model)
│
├─ Does the user need high-confidence answers from multiple AI providers?
│  └─ YES → pplx_council / pwm council (N+1 Pro Searches — ASK USER which models first!)
│
├─ Did the user explicitly request deep research / comprehensive report?
│  └─ YES → intent='research' (1 Deep Research)
│
└─ When in doubt → intent='quick' (Sonar 2, upgrade later if insufficient)
Smart Routing

The tool includes quota-aware routing. Instead of choosing a model manually, use the smart query interface and let it pick the best option:

MCP:  pplx_smart_query(query, intent="quick")       # default for most lookups
MCP:  pplx_smart_query(query, intent="standard")    # when quick isn't enough
CLI:  pwm ask "query"                                # auto routes via smart logic
CLI:  pwm ask "query" --intent quick                 # explicit intent hint
Automatic Quota Protection

The smart router automatically protects you:

  • Healthy quota: Uses the ideal model for your intent
  • Low quota (<20% pro remaining): Response footer warns you to conserve
  • Critical quota (<10% pro remaining): Downgrades detailed→auto to conserve
  • Exhausted quota: Falls back to Sonar 2 for everything except research (Sonar 2 is forced to concise mode to ensure grounded responses using search results)
  • Research exhausted: Falls back to premium Pro Search
  • Response metadata shows what model was used, why, and remaining quota
When to Use Explicit Models Instead

Only use model-specific tools (pplx_gpt56_terra, pplx_claude_sonnet, etc.) when:

  • The user explicitly requests a specific model
  • You're comparing outputs across models
  • The smart router's choice isn't working for the specific use case

Each explicit model call costs 1 Pro Search query — there is no free tier for these.

Tool Detection

Check which interface is available before proceeding:

has_mcp = check for tools starting with "pplx_"
has_cli = can run "pwm" commands via shell

if has_mcp and has_cli:
    Ask user which they prefer, or use MCP for programmatic access
elif has_mcp:
    Use pplx_* MCP tools directly
else:
    Use pwm CLI via shell

Workflow Decision Tree

User wants to...
|
+-- Search the web / ask a question (RECOMMENDED: smart routing)
|   +-- CLI:  pwm ask "query"                    # smart routing (default)
|   +-- MCP:  pplx_smart_query(query)            # smart routing (default)
|   +-- Explicit model: pwm ask "query" -m gpt56_terra  or  pplx_query(query, model="gpt56_terra")
|
+-- Chat interactively in the terminal (one thread per session)
|   +-- CLI:  pwm chat                              # sonar default; /model switches, /new resets, /exit quits
|
+-- Browse past conversations (FREE, no quota)
|   +-- CLI:  pwm threads                        # list recent threads
|   +-- CLI:  pwm threads --search "topic"       # search threads
|   +-- MCP:  pplx_list_threads()               # list threads
|   +-- MCP:  pplx_list_threads(search_term="X") # search threads
|
+-- Read or resume a past conversation (FREE, no quota)
|   +-- CLI:  pwm threads --search "topic"       # find slug
|   +-- MCP:  pplx_get_thread(slug)             # read full history
|   +-- MCP:  pplx_smart_query(query, conversation_id=slug) # resume
|
+-- Export full library to JSON (FREE, no quota)
|   +-- CLI:  pwm export                        # all threads → pplx-export-<date>.json
|   +-- CLI:  pwm export --search "ai"          # filtered export
|
+-- Query multiple models at once (Model Council)
|   +-- CLI:  pwm council "query"                         # default 3 models
|   +-- CLI:  pwm council "query" -m gpt56_terra,claude_sonnet  # custom models
|   +-- MCP:  pplx_council(query)                         # ASK USER which models first!
|
+-- Deep research on a topic
|   +-- CLI:  pwm research "query"
|   +-- MCP:  pplx_deep_research(query)
|
+-- Use a specific model
|   +-- CLI:  pwm ask "query" -m gpt56_terra --thinking
|   +-- MCP:  pplx_gpt56_terra_thinking(query)  or  pplx_query(query, model="gpt56_terra", thinking=True)
|
+-- Check remaining quotas
|   +-- CLI:  pwm usage
|   +-- MCP:  pplx_usage()
|
+-- Authenticate / re-authenticate
|   +-- Interactive:      pwm login
|   +-- Non-interactive:  pwm login --email EMAIL, then pwm login --email EMAIL --code CODE [--totp-code CODE]
|   +-- MCP (no shell):   pplx_auth_request_code(email), then pplx_auth_complete(email, code[, totp_code])
|
+-- Start MCP server
|   +-- pwm-mcp
|
+-- Start API server (for Claude Code / OpenAI SDK)
|   +-- pwm api [--port PORT]

CLI Commands

Querying
bash
pwm ask "What is quantum computing?"

Choose a specific model with -m:

bash
pwm ask "Compare React and Vue" -m gpt56_terra
pwm ask "Explain attention mechanism" -m claude_sonnet

Enable extended thinking with -t:

bash
pwm ask "Prove sqrt(2) is irrational" -m claude_sonnet --thinking

Focus on specific sources with -s:

bash
pwm ask "review this code for bugs" -s none            # Model only, no web search
pwm ask "transformer improvements 2025" -s academic   # Scholarly papers
pwm ask "best mechanical keyboard" -s social           # Reddit/Twitter
pwm ask "Apple revenue Q4 2025" -s finance             # SEC EDGAR filings
pwm ask "latest AI news" -s all                        # All sources
pwm connectors list                                    # List connector source IDs
pwm ask "private company funding" -s pitchbook_mcp_cashmere

Connector source IDs:

  • CLI: run pwm connectors list, then pass the source ID with -s.
  • MCP: call pplx_connectors(), then pass the source ID as source_focus.
  • Do not guess connector IDs. If no connector is listed, use normal source focus values.
  • Connector access inherits the authenticated Perplexity account's permissions and may expose private data to the calling agent.
  • Set PWM_CONNECTORS_ENABLED=0 to disable connector queries, or use PWM_CONNECTOR_ALLOWLIST with exact comma-separated connector IDs.
  • When local connector policy is unset, reported connectors remain available for backward compatibility.
  • Unknown source values fail instead of falling back to web search.

Output options:

bash
pwm ask "What is Rust?" --json            # JSON (for piping)
pwm ask "What is Rust?" --no-citations    # Answer only, no URLs

Combine flags:

bash
pwm ask "protein folding advances" -m gemini_pro -s academic --json

Read long prompts from stdin or a file, and attach files:

bash
pwm ask - < question.txt              # prompt from stdin
pwm ask --prompt-file question.txt    # prompt from a UTF-8 file
pwm ask "Summarize the key risks" --file report.pdf   # attachment (repeatable)

Attachments are validated before the query is sent; on Free accounts an attachment may count as a Pro Search.

Interactive Chat

Keep one Perplexity thread across turns in the terminal:

bash
pwm chat                              # defaults to Sonar 2 (free-tier friendly)
pwm chat -m auto                      # quota-aware routing per message
pwm chat -m claude_sonnet --thinking  # pin a model for the session

In-session commands: /new starts a new thread, /model [NAME] shows or switches the model (auto = quota-aware routing), /exit (or /quit, or Ctrl-D) quits. The thread continues across model switches.

Saved Defaults

Store a preferred model, thinking mode, and source once:

bash
pwm config set --model grok47 --thinking --source web
pwm config show
pwm config clear                # or clear one key: pwm config clear model

Defaults apply to pwm ask, pwm chat, and MCP pplx_query whenever the matching option is omitted; explicit flags always win (--no-thinking overrides a saved thinking default).

Shared MCP daemon

Use the shared daemon when multiple MCP clients or Codex sessions should connect to one local MCP process:

bash
pwm serve-mcp --transport streamable-http  # HTTP MCP on 127.0.0.1:8000/mcp
pwm serve-mcp --status                # Check the daemon
pwm serve-mcp --stop                  # Stop the daemon
pwm setup add codex --http            # Configure Codex for Streamable HTTP
pwm serve-mcp --transport sse         # Start the legacy SSE transport

The HTTP transports bind to loopback by default. Keep the daemon on a loopback address unless you put an authenticated reverse proxy in front of it.

Show full SKILL.md (714 more words)Show less
Model Council

Query multiple models in parallel and get a synthesized consensus. Each model in the council costs 1 Pro Search, plus 1 for Sonar 2 synthesis. Default: 3 Pro-compatible models + synthesis = 4 Pro Searches. Before selecting models, check pplx_usage() or pwm usage. If the subscription is Pro, exclude Max-only models (gpt56_sol, claude_opus).

bash
pwm council "What are the best practices for microservices?"           # default 3 models
pwm council "Compare Rust and Go for backend" -m gpt56_terra,claude_sonnet  # custom 2 models
pwm council "Explain quantum computing" -s academic                   # with source focus
pwm council "Prove the Pythagorean theorem" --thinking                # extended thinking
pwm council "AI trends 2026" --chairman claude_sonnet                 # premium synthesis (+1 Pro)
pwm council "Is React or Vue better?" --no-synthesis                  # skip synthesis
pwm council "AI trends 2026" --json                                   # JSON output
Thread Library (FREE — no quota)

Browse and export past Perplexity conversations:

bash
pwm threads                          # list most recent 20 threads
pwm threads --limit 50              # get 50 threads
pwm threads --search "quantum"      # filter threads by keyword
pwm threads --offset 20             # page 2 (skip first 20)
pwm threads --json                  # JSON output (for piping)

Export full library to a JSON file (no browser required):

bash
pwm export                              # all threads → pplx-export-<date>.json
pwm export --output ./my-backup.json   # custom path
pwm export --search "ai"               # filtered export
pwm export --limit 50                  # cap at 50 threads
Deep Research

Uses a separate monthly quota. Produces in-depth reports with extensive sources.

bash
pwm research "agentic AI trends 2026"
pwm research "climate policy impact" -s academic
pwm research "NVIDIA competitive landscape" -s finance --json
Authentication
bash
pwm login                                                # Interactive
pwm login --check                                        # Check status
pwm login --email user@example.com                       # Send code
pwm login --email user@example.com --code 123456         # Complete
pwm login --email user@example.com --code 123456 --totp-code 654321  # Complete with TOTP

Set PWM_SAVE_TO_LIBRARY=1 to save shared CLI and MCP queries to the Perplexity thread library. Queries are incognito by default.

Usage
bash
pwm usage                   # Cached limits
pwm usage --refresh         # Force-refresh from server

MCP Tools Summary

ToolCostPurpose
pplx_smart_queryVaries by intentUSE THIS BY DEFAULT — quota-aware auto routing
pplx_list_threadsFREEBrowse past conversations — paginated, searchable. Use before spending quota.
pplx_get_threadFREEFull history for any past thread. Also enables conversation resumption via conversation_id.
pplx_sonar1 Pro SearchPerplexity Sonar 2
pplx_query1 ProExplicit model selection with thinking toggle; omitted arguments follow saved pwm config defaults
pplx_ask1 ProQuick Q&A (auto model)
pplx_councilN+1 Pro (1 per model + 1 synthesis)Model Council — ASK USER which models first! Check subscription first; exclude Max-only gpt56_sol/claude_opus on Pro. Supports thinking=True and chairman for synthesis model.
pplx_gpt56_terra / _thinking1 ProOpenAI GPT-5.6 Terra (versatile)
pplx_gpt56_sol / _thinking1 ProOpenAI GPT-5.6 Sol (latest, Max tier)
pplx_grok45 / _thinking1 ProxAI Grok 4.5
pplx_claude_sonnet / _think1 ProAnthropic Claude Sonnet 5
pplx_claude_opus / _think1 ProAnthropic Claude 4.8 Opus
pplx_gemini_pro_think1 ProGoogle Gemini 3.1 Pro (thinking always on)
pplx_nemotron_thinking1 ProNVIDIA Nemotron 3 Ultra (thinking always on)
pplx_glm521 ProZ.ai GLM 5.2 (thinking always on)
pplx_kimi_k26 / _thinking1 ProMoonshot Kimi K2.6
pplx_deep_research1 ResearchIn-depth reports (scarce monthly quota)
pplx_usageFREECheck remaining quotas
pplx_connectorsFREEList account connector source IDs for source_focus
pplx_auth_statusFREECheck auth status
pplx_auth_request_codeFREESend verification code
pplx_auth_completeFREEComplete email and optional TOTP authentication

All query tools accept source_focus: "none", "web", "academic", "social", "finance", "all", or a connector source ID from pplx_connectors(). Use source_focus="none" for model-only queries without web search.

Multi-Turn Conversations: All query tools accept an optional conversation_id parameter. The server returns [Conversation ID: <uuid>] at the end of each response. Extract this UUID and pass it to the next query (any query tool, including pplx_smart_query) to maintain context across multiple turns.

For full MCP tool parameters: See references/mcp-tools.md

Models

CLI NameProviderThinkingNotes
autoPerplexityNoAuto-selects best
sonarPerplexityNoSonar 2 (API id experimental); concise mode keeps answers grounded
deep_researchPerplexityNoMonthly quota
gpt56_terraOpenAIToggleGPT-5.6 Terra
gpt56_solOpenAIToggleGPT-5.6 Sol (Max tier)
gpt6_solOpenAIToggleGPT-6 Sol
grok45xAIToggleGrok 4.5
grok47xAIToggleGrok 4.7
claude_sonnetAnthropicToggleClaude Sonnet 5
claude_opusAnthropicToggleClaude Opus 4.8 (Max tier)
claude_opus55AnthropicToggleClaude Opus 5.5 (Max tier)
gemini_proGoogleAlwaysGemini 3.1 Pro (thinking only)
gemini38GoogleToggleGemini 3.8 Flash
nemotronNVIDIAAlwaysNemotron 3 Ultra 550B (thinking only)
glm52Z.aiAlwaysGLM 5.2 (thinking only)
glm53Z.aiAlwaysGLM 5.3 (thinking only)
kimi_k26MoonshotToggleKimi K2.6
kimi_k3MoonshotAlwaysKimi K3 (thinking only)

For full model details: See references/models.md

Source Focus Options

OptionDescriptionExample Use Case
noneNo search — model training data only. Note: still costs 1 Pro Search for premium modelsCode review, writing, analysis without web
webGeneral web search (default)News, general questions
academicAcademic papers, journalsResearch, citations, scientific topics
socialReddit, Twitter, forumsOpinions, recommendations, community
financeSEC EDGAR filingsCompany financials, regulatory filings
allWeb + Academic + SocialBroad coverage across all sources

Error Recovery

ErrorCauseSolution
403 ForbiddenToken expiredpwm login
429 Rate limitQuota exhaustedWait, check pwm usage
"No token found"Not authenticatedpwm login
"LIMIT REACHED"Quota at zeroWait for reset or upgrade

Common Patterns

Thread Library & Conversation Resumption
bash
# List recent threads
pwm threads

# Search before spending quota
pwm threads --search "python packaging"

# Export full library to JSON (no browser needed)
pwm export

MCP — quota-free thread browsing:

pplx_list_threads()                        # recent 20 threads
pplx_list_threads(search_term="topic")     # search first
pplx_get_thread("<slug>")                  # read full history

Resume pattern — continue any past conversation:

# 1. Find the thread
pplx_list_threads(search_term="quantum")
# → returns slug: "f1f6562c-91be-47e9-..."

# 2. Read it for context (optional)
pplx_get_thread("f1f6562c-91be-47e9-...")

# 3. Continue right where it left off
pplx_smart_query("follow-up question", conversation_id="f1f6562c-91be-47e9-...")

MCP Resources (if your MCP client supports resources):

perplexity://library                        # your thread library
perplexity://thread/<slug>                  # a specific thread
bash
pwm ask "What happened in AI today?"
bash
pwm ask "Explain the visitor pattern in OOP" -s none
pwm ask "Write a Python decorator for retry logic" -m claude_sonnet -s none
Specific model
bash
pwm ask "Compare React and Vue" -m gpt56_terra
Model with thinking
bash
pwm ask "Prove sqrt(2) is irrational" -m claude_sonnet -t
Academic research
bash
pwm ask "transformer improvements 2025" -m gemini_pro -s academic
Financial analysis
bash
pwm ask "Apple revenue Q4 2025" -s finance
Launch Claude Code seamlessly (Integration)
bash
pwm hack claude
Deep research pipeline
bash
pwm research "quantum computing breakthroughs 2026" --json > research.json
Check everything before heavy use
bash
pwm login --check && pwm usage
Re-authenticate (non-interactive, for AI agents)
bash
pwm login --email user@example.com
# wait for email, then:
pwm login --email user@example.com --code 123456
# For a TOTP-enabled account, include: --totp-code 654321

API Server

For API server setup and model name mapping, see references/api-endpoints.md.

© jacob-bd, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

SKILL.md and 3 other files (references) in skills/perplexity-web-mcp of jacob-bd/perplexity-web-mcp.

  • SKILL.md
  • references/api-endpoints.md
  • references/mcp-tools.md
  • references/models.md

Open the folder on GitHubat commit e34b508

Compare with similar skills

Perplexity Web MCP Bridge next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Perplexity Web MCP Bridge compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Perplexity Web MCP Bridge this skilljacob-bd/perplexity-web-mcp197—~7.6kAutomated safety check: PassMIT
Zapier Workflowsdavila7/claude-code-templates32k2 repos~4.9kAutomated safety check: PassMIT
Huge AI Searchwangwingzero/huge-ai-search144—~563Automated safety check: PassMIT
Mysearchskernelx/MySearch-Proxy159—~3kAutomated safety check: NotesNone
Agent Searchlennney/agent-search-mcp110—~1.9kAutomated safety check: PassApache-2.0
Bright Data MCPbrightdata/skills2641 repos~3.7kAutomated safety check: PassMIT

Similar skills

  • Zapier Workflows

    davila7/claude-code-templates

    Manage and trigger pre-built Zapier workflows and MCP tool orchestration.

    32k GitHub starsUsed in 2 repos~4.9k tokens
    Productivity & AutomationAuto-check passed
  • Huge AI Search

    wangwingzero/huge-ai-search

    Searches the live web through Google AI Mode via Huge AI Search (MCP or CLI).

    144 GitHub stars~563 tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check passed
  • Mysearch

    skernelx/MySearch-Proxy

    Install, verify, debug, and use MySearch MCP/Skill. An agent skill from skernelx/MySearch-Proxy.

    159 GitHub stars~3k tokensUpdated 6 mo ago
    Productivity & AutomationAuto-check: notes
  • Agent Search

    lennney/agent-search-mcp

    Use Agent Search MCP for free-first English and Chinese web search with compact evidence and minimal tool calls.

    110 GitHub stars~1.9k tokensUpdated 17 days ago
    Productivity & AutomationAuto-check passed
  • Bright Data MCP

    brightdata/skills

    Bright Data MCP handles ALL web data operations. An agent skill from brightdata/skills.

    264 GitHub starsUsed in 1 repo~3.7k tokens
    Productivity & AutomationAuto-check passed
  • Skill

    redf0x1/camofox-mcp

    Anti-detection browser automation MCP skill for OpenClaw agents with 47 tools for navigation, interaction, observation, extraction, downloads, profiles, sessions, and stealth web search.

    117 GitHub stars~3.1k tokensUpdated 1 mo ago
    Productivity & AutomationAuto-check passed

Questions about Perplexity Web MCP Bridge

What does Perplexity Web MCP Bridge do?

Searches the web and queries premium AI models through Perplexity's own web interface, tracking quota carefully across quick, Pro and Deep Research tiers. Wraps Perplexity's web interface as both a CLI (`pwm ask`, `pwm chat`, `pwm research`) and a set of MCP tools, with a dedicated AI-oriented reference command recommended as the first step for commands, models, auth flows and error recovery. Login happens once up front and tokens last roughly a month, so a 403 error means re-running login rather than debugging further.

When should I use Perplexity Web MCP Bridge?

Perplexity Web MCP Bridge fits situations like: searching the web through Perplexity when ordinary search isn't enough; querying a specific premium model like Claude or Grok through Perplexity; checking remaining Pro and Deep Research quota before a research session; running a council query across several models with one synthesized answer.

How do I install Perplexity Web MCP Bridge in Claude Code?

Run `npx skills add jacob-bd/perplexity-web-mcp --skill perplexity-web-mcp -a claude-code`. Or copy the skill folder (skills/perplexity-web-mcp in jacob-bd/perplexity-web-mcp) into .claude/skills/perplexity-web-mcp in your project. Claude Code loads it when a task matches its description.

How do I install Perplexity Web MCP Bridge in Codex?

Run `npx skills add jacob-bd/perplexity-web-mcp --skill perplexity-web-mcp -a codex`. Or copy the skill folder (skills/perplexity-web-mcp in jacob-bd/perplexity-web-mcp) into .agents/skills/perplexity-web-mcp in your project. Codex loads it when a task matches its description.

Can I use Perplexity Web MCP Bridge in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add jacob-bd/perplexity-web-mcp --skill perplexity-web-mcp -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/perplexity-web-mcp, .gemini/skills/perplexity-web-mcp, .github/skills/perplexity-web-mcp and .opencode/skills/perplexity-web-mcp in your project.

What does Perplexity Web MCP Bridge need to run?

SKILL.md names no scripts, command-line tools or credentials: Perplexity Web MCP Bridge is instructions for the agent only. Our summary lists: The pwm CLI, logged in via pwm login; A Perplexity account with an active subscription.

Does Perplexity Web MCP Bridge access the network?

SKILL.md contains no URLs. Any network use would come from the scripts or tools the agent runs. This is read from the text; nothing was executed.

Is Perplexity Web MCP Bridge safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Perplexity Web MCP Bridge use?

Perplexity Web MCP Bridge is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Perplexity Web MCP Bridge use?

About 7.6k tokens (SKILL.md is roughly 31k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 6.1k tokens, read only when the agent opens those files.

What are the alternatives to Perplexity Web MCP Bridge?

Skills that share tags, products or a category with Perplexity Web MCP Bridge: Zapier Workflows (davila7/claude-code-templates, 32k stars), Huge AI Search (wangwingzero/huge-ai-search, 144 stars), Mysearch (skernelx/MySearch-Proxy, 159 stars) and Agent Search (lennney/agent-search-mcp, 110 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Perplexity Web MCP Bridge?

jacob-bd (a GitHub user) maintains it in jacob-bd/perplexity-web-mcp, which has 197 GitHub stars. The repository was last updated on October 4, 2026.

Source: jacob-bd/perplexity-web-mcp on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.