Agent skill

Venice API Overview

by veniceai in veniceai/skills

High-level map of the Venice.ai API: base URL, auth modes per endpoint, endpoint categories, response headers, pricing model, error shape and versioning.

MITAuto-check passedBackend & APIs

Install Venice API Overview

skills CLI
$ npx skills add veniceai/skills --skill venice-api-overview -a claude-code

Project install by default; add -g for ~/.claude/skills/.

GitHub CLI
$ gh skill install veniceai/skills venice-api-overview --agent claude-code

Project scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).

Manual copy
$ git clone --depth 1 https://github.com/veniceai/skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/venice-api-overview .claude/skills/venice-api-overview && rm -rf skills-src

Use ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.

Claude Code skills documentation · loads skills from .claude/skills/

Facts

Skill name
venice-api-overview
GitHub stars
143
Token cost
~3.5k tokens
SKILL.md length
1,271 words
Files
1
Skills in repo
22
Repo updated
First seen
Licence
MIT

At a glance

High-level map of the Venice.ai API: base URL, auth modes per endpoint, endpoint categories, response headers, pricing model, error shape and versioning.

  • Works in 5 steps: Read venice-auth and choose Bearer vs… → GET /models?type=… - pick a model and… → Wire up one happy-path call from the… → …
  • Writing code against api.venice.ai for the first time
  • SKILL.md covers Use when, Base URL, Authentication and Endpoint map, plus 6 more sections
  • Calls curl; reaches api.venice.ai; needs VENICE_API_KEY and WALLET_KEY

What it does

This skill is a starting reference for anyone writing code against api.venice.ai. Venice is described as an OpenAI-compatible inference platform for text, image, audio, video, embeddings and typed decisions, with one API and two ways to pay: an API key from a Venice account, or a wallet through x402 using USDC on Base or Solana, with no account needed.

It gives the base URL and the location of the OpenAPI spec, then which authentication each endpoint family accepts: a bearer key or Sign-In-With-X for inference, bearer only for API key, billing and character routes, wallet only for x402 balance and transactions, and none for model listings and quotes. A call with no credentials gets a 402 with payment discovery instead of a 401. The skill also covers response headers such as rate limit, payment-required and deprecation headers, the pricing model, the error shape and versioning, and points to a venice-auth skill for authentication detail.

When your agent uses it

  • Writing code against api.venice.ai for the first time
  • Choosing between API-key and wallet authentication
  • Finding which Venice endpoint handles a given task
  • Interpreting rate limit and payment response headers

Example prompts

  • “Which Venice endpoints accept wallet authentication and which need a bearer API key?”
  • “Give me an overview of the Venice API before I start integrating it.”
  • “What do the rate-limit and payment-required headers from api.venice.ai mean?”

Requirements

  • A Venice API key, or an x402-capable wallet holding USDC on Base or Solana

Workflow steps

5 steps, taken from the first numbered list in SKILL.md.

  1. Read venice-auth and choose Bearer vs x402.
  2. GET /models?type=… - pick a model and note its model_spec.constraints and model_spec.pricing.
  3. Wire up one happy-path call from the matching skill.
  4. Add error handling using venice-errors (402, 422, 429).
  5. Hook up observability via x-ratelimit-* / x-venice-balance-* headers, /billing/usage-history (ADMIN key only), or…

What it can do on your machine

Read from SKILL.md and the folder at commit 5eaeac5. It shows what the files ask for, not the result of running them.

  • Tool permissions

    Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.

    From allowed-tools in the SKILL.md frontmatter.

  • Runs code

    Shell commands in SKILL.md call:

    • curl

    From the folder's file list and the shell code blocks in SKILL.md.

  • Network

    Hosts in commands or code, which the agent is likely to contact:

    • api.venice.ai

    From URLs in SKILL.md, links to its own repository left out.

  • Credentials

    Names these keys or tokens, usually read from environment variables:

    • VENICE_API_KEY
    • WALLET_KEY

    From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.

Context cost

Venice API Overview loads about 3.5k tokens when it runs. Until then it costs about 93 tokens; SKILL.md has 1,271 words of instructions outside code blocks.

Always · name and description, kept in context so the agent knows when to use it
~93
When it runs · the whole SKILL.md, loaded when a task matches
~3.5k

Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.

Safety

Auto-check passed

The automated check found no risky patterns in SKILL.md.

Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.

SKILL.md

The full file from veniceai/skills at commit 5eaeac5, republished under its MIT licence (© veniceai). 1,271 words, ~3,482 tokens.

Download SKILL.mdSave it as .claude/skills/venice-api-overview/SKILL.md (or your agent's skills folder).
name
venice-api-overview
description
High-level map of the Venice.ai API - base URL, which auth mode each endpoint accepts (API key, x402 wallet, or none), endpoint categories (including decisions, voice changer, and retired routes), response headers (rate limit, balance, deprecation, x402), pricing model, error shape, and versioning. Load this first when starting any Venice integration.

Venice API Overview

Venice.ai is an OpenAI-compatible inference platform for text, image, audio, video, embeddings, and typed decisions. One API - two ways to pay: a traditional API key (Venice account), or a wallet (x402, USDC on Base or Solana, no account required).

Use when

  • You're writing code against api.venice.ai for the first time.
  • You need to decide between API-key and x402/wallet authentication.
  • You want a quick map of which endpoint to call for which task.
  • You need to understand the common response headers (x-ratelimit-*, PAYMENT-REQUIRED, deprecation headers, etc.).

Base URL

All endpoints live under:

https://api.venice.ai/api/v1

The OpenAPI spec is served at https://api.venice.ai/api/v1/swagger.yaml (info.version is a YYYYMMDD.HHMMSS timestamp; read it from the live spec).

Authentication

SchemeHeaderBest for
BearerAuthAuthorization: Bearer $VENICE_API_KEYServer-side apps, account management, usage analytics, DIEM / bundled credits
siwx (x402)SIGN-IN-WITH-X: <base64 SIWX JSON> (legacy X-Sign-In-With-X also accepted)No account, pay-as-you-go with USDC on Base or Solana, serverless / agents

Not every endpoint accepts both:

EndpointsAccepts
All inference: chat, responses, embeddings, decisions, image, audio (speech, transcriptions, voices, queue/retrieve/complete, voice-changer queue/retrieve/complete), video queue/retrieve/complete, augment, POST /crypto/rpc/{network}Bearer or SIWX
/api_keys/* (except generate_web3_key), /billing/* (except the retired /billing/usage), /characters/*Bearer only (a SIGN-IN-WITH-X header alone gets 401 Authentication failed)
/x402/balance/{wallet}, /x402/transactions/{wallet}SIWX only (signer must own the wallet)
/models*, /image/styles, /video/quote (except upscale models such as topaz-video-upscale: Bearer key required, so wallets can't quote them), /audio/quote, /audio/voice-changer/quote, /crypto/rpc/networks, /tee/attestation, /tee/signature, /api_keys/generate_web3_key, POST /x402/top-upNo auth needed

On any route that needs credentials - Bearer-only ones included - a request with no Authorization and no SIGN-IN-WITH-X header gets 402 with an x402 discovery body (payment options + SIWX challenge + authOptions), not 401. See venice-auth.

bash
# Bearer
curl https://api.venice.ai/api/v1/chat/completions \
  -H "Authorization: Bearer $VENICE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"zai-org-glm-5-2","messages":[{"role":"user","content":"hi"}]}'
ts
// x402 wallet via the SDK (EVM / Base wallets)
import { VeniceClient } from 'venice-x402-client'
const v = new VeniceClient(process.env.WALLET_KEY!)
await v.models()

Endpoint map

Inference
CategoryEndpointsSkill
ChatPOST /chat/completionsvenice-chat
Responses (Alpha)POST /responsesvenice-responses
EmbeddingsPOST /embeddingsvenice-embeddings
Decisions (Beta)POST /decisions (alias POST /systemone)venice-decisions
Image genPOST /image/generate, POST /images/generations, GET /image/stylesvenice-image-generate
Image editPOST /image/edit, POST /image/multi-edit, POST /image/upscale, POST /image/background-removevenice-image-edit
TTSPOST /audio/speech, POST /audio/voices (voice cloning)venice-audio-speech
STTPOST /audio/transcriptionsvenice-audio-transcription
Music / audio (async)POST /audio/quote, /audio/queue, /audio/retrieve, /audio/completevenice-audio-music
Voice changer (async)POST /audio/voice-changer/quote, /queue, /retrieve, /completevenice-audio-voice-changer
Video (async)POST /video/quote, /video/queue, /video/retrieve, /video/completevenice-video

Voice changer: the endpoints are in the spec, but no voice-changer model is publicly listed today - check GET /models?type=music for a model with voice_changer: true first.

Catalog
CategoryEndpointsSkill
ModelsGET /models, /models/traits, /models/compatibility_mappingvenice-models
CharactersGET /characters, /characters/{slug}, /characters/{slug}/reviewsvenice-characters

GET /models?type= accepts text (default), image, video, music, tts, asr, embedding, upscale, inpaint, decision, plus the filters all and code.

Account, billing, wallet
CategoryEndpointsSkill
API keysGET/POST/PATCH/DELETE /api_keys, GET /api_keys/{id}, /api_keys/rate_limits, /api_keys/rate_limits/log, /api_keys/generate_web3_keyvenice-api-keys
Billing (Beta)GET /billing/balance, /billing/usage-history, /billing/usage-analyticsvenice-billing
x402 walletPOST /x402/top-up, GET /x402/balance/{wallet}, GET /x402/transactions/{wallet}venice-x402
Utility
CategoryEndpointsSkill
Crypto RPC proxyGET /crypto/rpc/networks, POST /crypto/rpc/{network}venice-crypto-rpc
AugmentPOST /augment/text-parser, /augment/scrape, /augment/searchvenice-augment
TEE verificationGET /tee/attestation, GET /tee/signature (public, 10 req/min per IP)venice-text-routing
Retired (return 410 Gone)
EndpointReplacement
GET /billing/usage - sunset 2026-09-16GET /billing/usage-history (cursor pagination: pageSize + nextCursor, startTimestamp / endTimestamp)
POST /video/transcriptionsPOST /chat/completions with a video_url part on a model whose capabilities.supportsVideoInput is true

Both answer with 410 plus Deprecation and Link: <…>; rel="successor-version" headers (/billing/usage also sends Sunset). They need no auth; a per-IP limit of 60 requests/minute returns 429 to clients that keep polling.

Response headers to watch

HeaderWhenMeaning
x-ratelimit-limit-requests / -remaining-requests / -reset-requestsMost inference responsesRequest window for your account - per model (the tighter of per-minute and per-day), or per endpoint on video / audio-generation routes and /augment/scrape / /augment/search. Reset is a Unix timestamp in milliseconds.
x-ratelimit-limit-tokens / -remaining-tokens / -reset-tokensToken-limited text modelsTokens-per-minute window (reset in ms).
x-ratelimit-remaining / x-ratelimit-resetsMost responses on routes that need credentialsThe error budget (failed requests allowed in the 30 s window), not your request quota. Read before the current response is counted, so a failed response showing 1 means none are left. Reset in ms.
x-venice-balance-usd / x-venice-balance-diemInference responsesSpendable balance when the request started (x402 callers see their wallet credit here). Omitted when that balance is zero; the USD figure excludes bundled and earned credits.
x-venice-versionInference responsesServer revision - handy in bug reports.
x-venice-deprecated, x-venice-model-deprecation-date, x-venice-model-deprecation-warning, x-venice-deprecated-replacementRequests to a model scheduled for retirement (chat, image, video queue)Retirement date and suggested replacement.
PAYMENT-REQUIREDEvery x402 402: no credentials, wallet below the minimum balance, /x402/top-up discovery, /x402/balance / /x402/transactions without SIWXBase64 JSON of the x402 v2 payment-required object (accepts[], plus the sign-in-with-x challenge except on /x402/top-up).
Retry-After429 "model overloaded"Seconds to wait (default 30).
Deprecation / Sunset / LinkRetired endpointsSee table above.
Content-EncodingWhen you send Accept-Encoding (gzip, br, deflate)Compressed responses on any route, once the body is large enough to be worth compressing. The spec documents it on chat, embeddings, decisions and image generation.

The spec also documents an X-Balance-Remaining header on x402 responses, but current server code does not set it - read x-venice-balance-usd or call GET /x402/balance/{wallet} instead.

Show full SKILL.md (461 more words)Show less

Pricing model at a glance

  • Pricing is dynamic per request, metered in USD. Paid endpoints in the spec carry an x-payment-info block (price.mode: dynamic, min: "0.001", max: "10.00" USD); POST /x402/top-up is 5-10000. Read-only routes (/models, quotes) have none.
  • API-key accounts charge each request to a single currency, picked in the order DIEM → earned credits → bundled credits → USD (see venice-billing). Per-key consumptionLimits can cap USD / DIEM spend.
  • x402 wallets spend a prepaid USDC credit balance topped up on Base or Solana (minimum top-up $5; a request needs at least $0.10 of balance to start). An EVM wallet linked to a Venice account with staked DIEM spends DIEM first.
  • The per-model price is on GET /models → model_spec.pricing, already including any promotion active for your account. Video has no price there - use POST /video/quote; music / voice changer have exact quotes via /audio/quote and /audio/voice-changer/quote. See venice-models.

Standard error shape

Most errors are:

json
{ "error": "Human-readable message" }

Schema validation failures (400) add a details tree and an issues array:

json
{ "error": "Invalid request parameters", "details": { "_errors": [], "type": { "_errors": ["Invalid enum value…"] } }, "issues": [ … ] }

Exceptions worth knowing: context-length overflows on chat return an OpenAI-style object ({ "error": { "message", "type": "invalid_request_error", "param": "messages", "code": "context_length_exceeded" } }), upstream provider rejections return the provider's message (plus request_id on routes that track one, such as chat), and 402 on x402 returns structured top-up data. A POST whose Content-Type is neither application/json nor multipart/form-data gets 400 "'Content-Type' must be 'application/json'" before auth runs. See venice-errors for the full table and retry strategy.

OpenAI compatibility - what works and what doesn't

  • Drop-in: /chat/completions, /responses, /embeddings, /images/generations, /audio/speech, /audio/transcriptions, /models.
  • Accepted but ignored for compat: user, store (on chat). user is not an alias of Venice's anon_user_id (it does partition the error budget - see venice-errors).
  • Venice-only extensions live under venice_parameters - the full set on /chat/completions, a seven-field subset (character_slug, enable_e2ee, enable_web_search, enable_web_scraping, enable_web_citations, include_venice_system_prompt, include_search_results_in_stream) on /responses.
  • Model feature suffixes (e.g. zai-org-glm-5-1:enable_web_search=on, kimi-k2-6:strip_thinking_response=true&disable_thinking=true) flip venice_parameters via the model ID - see venice-chat.
  • A trait name (default_reasoning) or a legacy alias (gpt-4o) can be sent as model; Venice resolves it. See venice-models.

Versioning

  • info.version in swagger.yaml is a timestamp (YYYYMMDD.HHMMSS). There is no /v2; features roll forward on the single /api/v1 surface and are guarded by:
    • Alpha/Beta labels in endpoint descriptions (Responses is Alpha; Decisions and Billing are Beta).
    • betaModel / deprecation metadata and capability flags on /models.
  • Retired endpoints answer 410 with Deprecation / Link headers rather than disappearing silently.
  • Always check the model's model_spec.capabilities (supportsWebSearch, supportsReasoning, supportsE2EE, supportsXSearch, supportsMultipleImages, supportsFunctionCalling, supportsAudioInput, supportsVideoInput, …) before relying on a feature.

Fast start checklist

  1. Read venice-auth and choose Bearer vs x402.
  2. GET /models?type=… - pick a model and note its model_spec.constraints and model_spec.pricing.
  3. Wire up one happy-path call from the matching skill.
  4. Add error handling using venice-errors (402, 422, 429).
  5. Hook up observability via x-ratelimit-* / x-venice-balance-* headers, /billing/usage-history (ADMIN key only), or /x402/transactions/{wallet} (wallet callers).

© veniceai, MIT. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file

Files

Just SKILL.md in skills/venice-api-overview of veniceai/skills.

Open the folder on GitHubat commit 5eaeac5

Compare with similar skills

Venice API Overview next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.

Venice API Overview compared with similar skills
SkillStarsUsed inTokensAuto-checkLicenceRepo updated
Venice API Overview this skillveniceai/skills143—~3.5kAutomated safety check: PassMIT
Frappe Core APIImpertio-Studio/Frappe_Claude_Skill_Package188—~3.2kAutomated safety check: PassMIT
Memstack Automation API Integrationcwinvestments/memstack423—~2.8kAutomated safety check: PassProprietary
Shopify APIMicrock/ordinary-claude-skills404—~4.3kAutomated safety check: PassCustom licence
Agentic Walletcoinbase/agentic-wallet-skills1272 repos~1kAutomated safety check: PassMIT
Passport Developmenttrypostit/trypost685—~1.9kAutomated safety check: PassMIT

Similar skills

  • Frappe Core API

    Impertio-Studio/Frappe_Claude_Skill_Package

    A skill your agent uses when building ERPNext/Frappe API integrations (v14/v15/v16) including REST API, RPC API, authentication, webhooks, and rate limiting.

    188 GitHub stars~3.2k tokensUpdated 22 days ago
    Backend & APIsAuto-check passed
  • A skill your agent uses when the user says 'API integration', 'connect APIs', 'sync data', 'data mapping', 'rate limiting', or needs system-to-system connectors with authentication, rate limit…

    423 GitHub stars~2.8k tokensUpdated 12 days ago
    Backend & APIsAuto-check passed
  • Shopify API

    Microck/ordinary-claude-skills

    Complete API integration guide for Shopify including GraphQL Admin API, REST Admin API, Storefront API, Ajax API, OAuth authentication, rate limiting, and webhooks.

    404 GitHub stars~4.3k tokensUpdated 1 mo ago
    Backend & APIsAuto-check passed
  • Agentic Wallet

    coinbase/agentic-wallet-skills

    Crypto wallet operations via the awal CLI — sign in, check balances, send USDC/ETH/POL/SOL, trade tokens, fund the wallet, and use the x402 payment protocol to discover paid services, pay for API…

    127 GitHub starsUsed in 2 repos~1k tokens
    Backend & APIsAuto-check passed
  • Passport Development

    trypostit/trypost

    Develops OAuth2 API authentication with Laravel Passport. An agent skill from trypostit/trypost.

    685 GitHub stars~1.9k tokensUpdated today
    Backend & APIsAuto-check passed
  • Integrates applications with the SoundCloud HTTP API using OAuth 2.1, OpenAPI, and developer docs.

    259 GitHub stars~787 tokensUpdated 9 days ago
    Backend & APIsAuto-check passed

More from veniceai/skills

All 22 skills in this repo
  • Picks which Venice text model to call for a prompt based on privacy tier, input modality, capabilities and cost, and decides when to escalate from a local agent.

    143 GitHub stars~5.2k tokensUpdated 3 days ago
    Auto-check passed
  • Venice Models API

    veniceai/skills

    Documents Venice's model discovery endpoints, GET /models, /models/traits and /models/compatibility_mapping, so an agent can pick a model by capability, constraint or price.

    143 GitHub stars~4.3k tokensUpdated 3 days ago
    Auto-check passed
  • Venice API Keys

    veniceai/skills

    Manages Venice API keys through the /api_keys endpoints: create, list, update and revoke keys, set spending limits, and read rate limits.

    143 GitHub stars~3.8k tokensUpdated 3 days ago
    Auto-check passed
  • Venice Audio Music

    veniceai/skills

    Async music, sound-effect and long-form voice generation via Venice.

    143 GitHub stars~3.1k tokensUpdated 3 days ago
    Auto-check passed
  • Venice Audio Speech

    veniceai/skills

    Generate speech from text via POST /audio/speech, and clone a voice via POST /audio/voices.

    143 GitHub stars~3.6k tokensUpdated 3 days ago
    Auto-check passed
  • Transcribe audio files to text via POST /audio/transcriptions.

    143 GitHub stars~1.7k tokensUpdated 3 days ago
    Auto-check passed

Works with

Questions about Venice API Overview

What does Venice API Overview do?

High-level map of the Venice.ai API: base URL, auth modes per endpoint, endpoint categories, response headers, pricing model, error shape and versioning. ai. Venice is described as an OpenAI-compatible inference platform for text, image, audio, video, embeddings and typed decisions, with one API and two ways to pay: an API key from a Venice account, or a wallet through x402 using USDC on Base or Solana, with no account needed.

When should I use Venice API Overview?

Venice API Overview fits situations like: writing code against api.venice.ai for the first time; choosing between API-key and wallet authentication; finding which Venice endpoint handles a given task; interpreting rate limit and payment response headers.

How do I install Venice API Overview in Claude Code?

Run `npx skills add veniceai/skills --skill venice-api-overview -a claude-code`. Or copy the skill folder (skills/venice-api-overview in veniceai/skills) into .claude/skills/venice-api-overview in your project. Claude Code loads it when a task matches its description.

How do I install Venice API Overview in Codex?

Run `npx skills add veniceai/skills --skill venice-api-overview -a codex`. Or copy the skill folder (skills/venice-api-overview in veniceai/skills) into .agents/skills/venice-api-overview in your project. Codex loads it when a task matches its description.

Can I use Venice API Overview in Cursor, Gemini CLI or GitHub Copilot?

Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add veniceai/skills --skill venice-api-overview -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/venice-api-overview, .gemini/skills/venice-api-overview, .github/skills/venice-api-overview and .opencode/skills/venice-api-overview in your project.

What does Venice API Overview need to run?

Going by SKILL.md and its folder, Venice API Overview needs the command-line tools its instructions call (curl) and credentials named VENICE_API_KEY and WALLET_KEY. Our summary lists: A Venice API key, or an x402-capable wallet holding USDC on Base or Solana.

Does Venice API Overview access the network?

SKILL.md names 1 domain. In commands or code: api.venice.ai; the agent is likely to contact it when it follows the instructions. This is read from the text; nothing was executed.

Is Venice API Overview safe to install?

Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.

What licence does Venice API Overview use?

Venice API Overview is published under the MIT licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.

How many tokens does Venice API Overview use?

About 3.5k tokens (SKILL.md is roughly 14k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full.

What are the alternatives to Venice API Overview?

Skills that share tags, products or a category with Venice API Overview: Frappe Core API (Impertio-Studio/Frappe_Claude_Skill_Package, 188 stars), Memstack Automation API Integration (cwinvestments/memstack, 423 stars), Shopify API (Microck/ordinary-claude-skills, 404 stars) and Agentic Wallet (coinbase/agentic-wallet-skills, 127 stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.

Who maintains Venice API Overview?

veniceai (a GitHub organization) maintains it in veniceai/skills, which has 143 GitHub stars. The repository holds 22 skills in this directory. The repository was last updated on October 5, 2026.

Source: veniceai/skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.