Gemini Live API Dev
JetBrains/skills
A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API.
A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live.
$ npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a claude-codeProject install by default; add -g for ~/.claude/skills/.
$ gh skill install google-gemini/gemini-skills gemini-live-api-dev --agent claude-codeProject scope by default; add --scope user for a personal install. Needs GitHub CLI 2.90.0 or later (public preview).
$ git clone --depth 1 https://github.com/google-gemini/gemini-skills.git skills-src && mkdir -p .claude/skills && cp -r skills-src/skills/gemini-live-api-dev .claude/skills/gemini-live-api-dev && rm -rf skills-srcUse ~/.claude/skills/ instead of .claude/skills for a personal install. The folder must contain SKILL.md.
Claude Code skills documentation · loads skills from .claude/skills/
Install the "gemini-live-api-dev" agent skill from https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-dev into .claude/skills/gemini-live-api-dev/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gemini-live-api-dev", then confirm the skill loads.Claude Code copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$skill-installer install https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-devType this inside Codex. $skill-installer <name> installs a curated skill from openai/skills. The installer writes to $CODEX_HOME/skills (default ~/.codex/skills). Restart Codex if the skill does not show up.
$ npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a codexProject install goes to .agents/skills/; add -g for ~/.codex/skills/.
$ gh skill install google-gemini/gemini-skills gemini-live-api-dev --agent codexProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/google-gemini/gemini-skills.git skills-src && mkdir -p .agents/skills && cp -r skills-src/skills/gemini-live-api-dev .agents/skills/gemini-live-api-dev && rm -rf skills-srcUse ~/.agents/skills/ instead of .agents/skills for a personal install.
Codex skills documentation · loads skills from .agents/skills/
Install the "gemini-live-api-dev" agent skill from https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-dev into .agents/skills/gemini-live-api-dev/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gemini-live-api-dev", then confirm the skill loads.Codex copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a cursorProject install goes to .agents/skills/; add -g for ~/.cursor/skills/.
$ gh skill install google-gemini/gemini-skills gemini-live-api-dev --agent cursorProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/google-gemini/gemini-skills.git skills-src && mkdir -p .cursor/skills && cp -r skills-src/skills/gemini-live-api-dev .cursor/skills/gemini-live-api-dev && rm -rf skills-srcUse ~/.cursor/skills/ instead of .cursor/skills for a personal install.
Cursor skills documentation · loads skills from .cursor/skills/, .agents/skills/, .claude/skills/, .codex/skills/
Install the "gemini-live-api-dev" agent skill from https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-dev into .cursor/skills/gemini-live-api-dev/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gemini-live-api-dev", then confirm the skill loads.Cursor copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gemini skills install https://github.com/google-gemini/gemini-skills.git --path skills/gemini-live-api-dev--scope user (default) or --scope workspace; --path is the subfolder of the repo that holds the skill; --consent skips the security confirmation prompt.
$ npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a gemini-cliProject install goes to .agents/skills/; add -g for ~/.gemini/skills/.
$ gh skill install google-gemini/gemini-skills gemini-live-api-dev --agent gemini-cliProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/google-gemini/gemini-skills.git skills-src && mkdir -p .gemini/skills && cp -r skills-src/skills/gemini-live-api-dev .gemini/skills/gemini-live-api-dev && rm -rf skills-srcUse ~/.gemini/skills/ instead of .gemini/skills for a personal install, then run /skills reload.
Gemini CLI skills documentation · loads skills from .gemini/skills/, .agents/skills/
Install the "gemini-live-api-dev" agent skill from https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-dev into .gemini/skills/gemini-live-api-dev/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gemini-live-api-dev", then confirm the skill loads.Gemini CLI copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ gh skill install google-gemini/gemini-skills gemini-live-api-devInstalls for Copilot at project scope by default; add --scope user for a personal install. Preview a skill first with gh skill preview. Needs GitHub CLI 2.90.0 or later (public preview).
$ npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a github-copilotProject install goes to .agents/skills/; add -g for ~/.copilot/skills/.
$ git clone --depth 1 https://github.com/google-gemini/gemini-skills.git skills-src && mkdir -p .github/skills && cp -r skills-src/skills/gemini-live-api-dev .github/skills/gemini-live-api-dev && rm -rf skills-srcUse ~/.copilot/skills/ instead of .github/skills for a personal install. Commit .github/skills so cloud agent and code review can use it.
GitHub Copilot skills documentation · loads skills from .github/skills/, .claude/skills/, .agents/skills/
Install the "gemini-live-api-dev" agent skill from https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-dev into .github/skills/gemini-live-api-dev/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gemini-live-api-dev", then confirm the skill loads.GitHub Copilot copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
$ npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a opencodeOpenCode documents no install command of its own. Project install goes to .agents/skills/; add -g for ~/.config/opencode/skills/.
$ gh skill install google-gemini/gemini-skills gemini-live-api-dev --agent opencodeProject scope by default (.agents/skills/); add --scope user for a personal install.
$ git clone --depth 1 https://github.com/google-gemini/gemini-skills.git skills-src && mkdir -p .opencode/skills && cp -r skills-src/skills/gemini-live-api-dev .opencode/skills/gemini-live-api-dev && rm -rf skills-srcUse ~/.config/opencode/skills/ instead of .opencode/skills for a personal install.
OpenCode skills documentation · loads skills from .opencode/skills/, .claude/skills/, .agents/skills/
Install the "gemini-live-api-dev" agent skill from https://github.com/google-gemini/gemini-skills/tree/main/skills/gemini-live-api-dev into .opencode/skills/gemini-live-api-dev/ in this project. Copy the whole folder (SKILL.md and every file beside it), keep the folder name "gemini-live-api-dev", then confirm the skill loads.OpenCode copies the folder itself, the same result as the manual copy. Check what it changed before you commit it.
gemini-live-api-devA skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live.
Gemini Live API Dev is an agent skill from google-gemini/gemini-skills, published by the product's own GitHub organization. Use this skill when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live. Covers WebSocket-based audio/video/text streaming, voice activity detection (VAD), background reasoning (extended thinking), asynchronous function calling, session management, ephemeral tokens, live transcription, and live translation. SDKs covered - google-genai (Python), @google/genai (JavaScript/TypeScript).
Its SKILL.md is about 4.6k tokens, which your agent loads only when the skill is triggered. The skill folder holds 2 other files, including reference files (for example `references/migration.md`).
It sits in Backend & APIs, covering Realtime and WebSockets, Transcription and Authentication. It works with Google Gemini, Python, JavaScript and TypeScript. The repository describes itself as: Skills for the Gemini API, SDK and model/agent interactions. The licence is Apache-2.0.
9 steps, taken from the first numbered list in SKILL.md.
Read from SKILL.md and the folder at commit 832c8f9. It shows what the files ask for, not the result of running them.
Pre-approves nothing: there is no allowed-tools line, so your agent's usual permission prompts apply.
From allowed-tools in the SKILL.md frontmatter.
Shell commands in SKILL.md call:
pipnpmFrom the folder's file list and the shell code blocks in SKILL.md.
Hosts in commands or code, which the agent is likely to contact:
ai.google.devAlso links to:
docs.livekit.iodocs.pipecat.aidocs.fishjam.iovisionagents.aivoximplant.comfirebase.google.comcolab.research.google.comFrom URLs in SKILL.md, links to its own repository left out.
Names no API keys, tokens, secrets or passwords.
From names ending in _API_KEY, _TOKEN, _SECRET, _KEY or _PASSWORD in SKILL.md.
Gemini Live API Dev loads about 4.6k tokens when it runs, and up to ~6.9k if it reads all its reference files. Until then it costs about 125 tokens; SKILL.md has 1,309 words of instructions outside code blocks.
Estimates: characters ÷ 4, the usual rule of thumb; real counts depend on the model's tokenizer. Scripts and assets cost tokens only if the agent reads them.
The automated check found no risky patterns in SKILL.md.
Automated static check — not a guarantee. Review scripts before installing. It scans the text of SKILL.md for risky patterns (piping downloads into a shell, reading credential files, hidden Unicode, destructive commands); files beside SKILL.md are not scanned.
The full file from google-gemini/gemini-skills at commit 832c8f9, republished under its Apache-2.0 licence (© google-gemini). 1,309 words, ~4,645 tokens.
.claude/skills/gemini-live-api-dev/SKILL.md (or your agent's skills folder). This skill also uses 1 other file; get the full folder from GitHub.The Live API enables low-latency, real-time voice and video interactions with Gemini over WebSockets. It processes continuous streams of audio, video, or text to deliver immediate, human-like spoken responses and background reasoning.
Key capabilities:
[!NOTE] The Live API connects directly via WebSockets. For WebRTC support or simplified integration, use a partner integration.
gemini-3.8-live — Default option for most low-latency voice agent experiences and real-time dialogue without reasoning delays. Supports interleaved reasoning, asynchronous function calling by default (behavior: NON_BLOCKING), and full-session client content updates.gemini-3.8-live-extended-thinking — High-reasoning audio-to-audio model recommended when higher background reasoning is required during live interactions. Processes background reasoning and async tool calls (behavior: NON_BLOCKING required) while streaming continuous spoken conversational fillers; lifecycle managed via interaction_status (IN_PROGRESS vs IDLE).gemini-3.5-transcribe-live — Real-time streaming speech-to-text with interim hypotheses, finalized transcripts, smart formatting, and Hybrid VAD.gemini-3.5-live-translate-preview — Real-time speech-to-speech streaming translation across 70+ languages.[!WARNING] Legacy Models (
gemini-3.1-flash-live-preview,gemini-2.5-flash-native-audio-*,gemini-live-2.5-flash-preview,gemini-2.0-flash-live-001): Readreferences/migration.mdfor breaking protocol changes (behavior: "NON_BLOCKING",thinking_level,interaction_status,send_client_content).
google-genai >= 2.3.0 — pip install -U google-genai@google/genai >= 2.3.0 — npm install @google/genai[!WARNING] Legacy SDKs
google-generativeai(Python) and@google/generative-ai(JS) are deprecated. Never use them.
To streamline real-time audio/video app development, use a third-party integration supporting the Gemini Live API over WebRTC or WebSockets:
audio/pcm;rate=16000[!IMPORTANT] Use
send_realtime_input/sendRealtimeInputfor all real-time streaming user input (audio, video, and text). On Gemini 3.8 models,send_client_content/sendClientContentis supported across the full session lifecycle with explicit roles (userormodel) to inject conversation context (turn_complete=trueunconditionally interrupts active generation).
[!WARNING] Do not use
mediainsendRealtimeInput. Use the specific keys:audiofor audio data,videofor images/video frames, andtextfor text input.
from google import genai
client = genai.Client(api_key="YOUR_API_KEY")import { GoogleGenAI } from '@google/genai';
const ai = new GoogleGenAI({ apiKey: 'YOUR_API_KEY' });from google.genai import types
config = types.LiveConnectConfig(
response_modalities=[types.Modality.AUDIO],
system_instruction=types.Content(
parts=[types.Part(text="You are a helpful assistant.")]
)
)
async with client.aio.live.connect(model="gemini-3.8-live", config=config) as session:
pass # Session is activeconst session = await ai.live.connect({
model: 'gemini-3.8-live',
config: {
responseModalities: ['audio'],
systemInstruction: { parts: [{ text: 'You are a helpful assistant.' }] }
},
callbacks: {
onopen: () => console.log('Connected'),
onmessage: (response) => console.log('Message:', response),
onerror: (error) => console.error('Error:', error),
onclose: () => console.log('Closed')
}
});await session.send_realtime_input(text="Hello, how are you?")session.sendRealtimeInput({ text: 'Hello, how are you?' });await session.send_realtime_input(
audio=types.Blob(data=chunk, mime_type="audio/pcm;rate=16000")
)session.sendRealtimeInput({
audio: { data: chunk.toString('base64'), mimeType: 'audio/pcm;rate=16000' }
});# frame: raw JPEG-encoded bytes
await session.send_realtime_input(
video=types.Blob(data=frame, mime_type="image/jpeg")
)session.sendRealtimeInput({
video: { data: frame.toString('base64'), mimeType: 'image/jpeg' }
});[!IMPORTANT] A single server event can contain multiple content parts simultaneously (e.g., audio chunks and transcript). Always process all parts in each event to avoid missing content.
async for response in session.receive():
content = response.server_content
if content:
# Audio — process ALL parts in each event
if content.model_turn:
for part in content.model_turn.parts:
if part.inline_data:
audio_data = part.inline_data.data
# Transcription
if content.input_transcription:
print(f"User: {content.input_transcription.text}")
if content.output_transcription:
print(f"Gemini: {content.output_transcription.text}")
# Interruption
if content.interrupted is True:
pass # Stop playback, clear audio queue// Inside the onmessage callback
const content = response.serverContent;
if (content?.modelTurn?.parts) {
for (const part of content.modelTurn.parts) {
if (part.inlineData) {
const audioData = part.inlineData.data; // Base64 encoded
}
}
}
if (content?.inputTranscription) console.log('User:', content.inputTranscription.text);
if (content?.outputTranscription) console.log('Gemini:', content.outputTranscription.text);
if (content?.interrupted) { /* Stop playback, clear audio queue */ }Use gemini-3.8-live-extended-thinking when your voice agent must evaluate complex data, plan multiple steps, or handle long-running tools. The model speaks natural conversational fillers (e.g. "Checking flight options now...") while executing asynchronous tools in the background.
Key requirements:
thinking_config=types.ThinkingConfig(thinking_level="low") ("minimal" | "low" | "medium" | "high").behavior="NON_BLOCKING". Synchronous blocking mode is not supported and returns an error.interaction_status): Do not rely on turn_complete=True alone to detect turn completion. Monitor message.interaction_status (Python) / message.interactionStatus (JS):"IN_PROGRESS": Server is reasoning, speaking conversational fillers, or waiting for async tool responses."IDLE": Server has completed all background reasoning and tool calls; session is ready for user input.See references/migration.md and the Thinking in Live API Guide for complete Python and JavaScript implementation examples.
The Live API supports real-time, low-latency streaming translation of speech (audio) across 70+ languages. For full details on options and capabilities, see the Live Translate Guide.
gemini-3.5-live-translate-preview — The recommended translation model for all Live Translate use cases.TranslationConfig)To enable translation, specify a TranslationConfig object inside your live session setup:
translation_config on LiveConnectConfig:config = types.LiveConnectConfig(
response_modalities=[types.Modality.AUDIO],
translation_config=types.TranslationConfig(
target_language_code="es", # Target language code (e.g. es, fr, pl)
echo_target_language=True,
),
input_audio_transcription=types.AudioTranscriptionConfig(),
output_audio_transcription=types.AudioTranscriptionConfig(),
)translationConfig inside generationConfig:{
"setup": {
"model": "models/gemini-3.5-live-translate-preview",
"generationConfig": {
"responseModalities": ["AUDIO"],
"translationConfig": {
"targetLanguageCode": "es",
"echoTargetLanguage": true
}
}
}
}The Live API supports real-time streaming speech-to-text over WebSockets with low-latency interim hypotheses, finalized transcripts, and Hybrid VAD. For full details, see the Live Transcription Guide and Colab Cookbook.
gemini-3.5-transcribe-livesmart: cleans up filler words, resolves inline self-corrections, and structures formatting.verbatim (default): exact word-for-word transcript.config = types.LiveConnectConfig(
response_modalities=["TEXT"],
input_audio_transcription=types.AudioTranscriptionConfig(),
)
async with client.aio.live.connect(model="gemini-3.5-transcribe-live", config=config) as session:
# Stream audio
await session.send_realtime_input(audio=types.Blob(data=chunk, mime_type="audio/pcm;rate=16000"))
# Hybrid VAD: notify turn end on client-detected silence for zero latency
await session.send_realtime_input(audio_stream_end=True)const session = await ai.live.connect({
model: 'gemini-3.5-transcribe-live',
config: {
responseModalities: ['text'],
inputAudioTranscription: { mode: 'smart' }
},
callbacks: {
onmessage: (msg) => {
if (msg.serverContent?.interimInputTranscription) {
console.log('Interim:', msg.serverContent.interimInputTranscription.text);
}
if (msg.serverContent?.inputTranscription) {
console.log('Final:', msg.serverContent.inputTranscription.text);
}
}
}
});
session.sendRealtimeInput({ audio: { data: chunkBase64, mimeType: 'audio/pcm;rate=16000' } });
session.sendRealtimeInput({ audioStreamEnd: true }); // Hybrid VAD{
"setup": {
"model": "models/gemini-3.5-transcribe-live",
"generationConfig": {
"responseModalities": ["TEXT"],
"speechConfig": {
"voiceConfig": {}
}
},
"inputAudioTranscription": {
"mode": "smart"
}
}
}TEXT or AUDIO per session, not both. Native audio models output audio (response_modalities=["AUDIO"]); enable output_audio_transcription if you need text transcripts.For step-by-step migration checklists and protocol deltas when upgrading from gemini-3.1-flash-live-preview, gemini-2.5-flash-native-audio-*, or gemini-2.0-flash-live-001 to Gemini 3.8 Live or Gemini 3.8 Live Extended Thinking, read references/migration.md.
send_realtime_input for real-time user input (audio, video, text). Use send_client_content with explicit user/model roles to inject context turns mid-streamaudioStreamEnd / audio_stream_end (Hybrid VAD) when the mic is paused or user finishes speakinginterrupted: true)interaction_status (IN_PROGRESS vs IDLE) when using gemini-3.8-live-extended-thinking rather than relying on turn_complete aloneIf the search_docs tool (from the Google MCP server) is available, use it as your only documentation source:
search_docs with your query[!IMPORTANT] When MCP tools are present, never fetch URLs manually. MCP provides up-to-date, indexed documentation that is more accurate and token-efficient than URL fetching.
If no MCP documentation tools are available, fetch from the official docs index:
llms.txt URL: https://ai.google.dev/gemini-api/docs/llms.txt
This index contains links to all documentation pages in .md.txt format. Use web fetch tools to:
llms.txt to discover available documentation pageshttps://ai.google.dev/gemini-api/docs/live-session.md.txt)[!IMPORTANT] Those are not all the documentation pages. Use the
llms.txtindex to discover available documentation pages
The Live API supports 70 languages including: English, Spanish, French, German, Italian, Portuguese, Chinese, Japanese, Korean, Hindi, Arabic, Russian, and many more. Native audio models automatically detect and switch languages.
© google-gemini, Apache-2.0. Rendered from Markdown: HTML in the file is shown as text, images as links, and headings moved down two levels. Raw file
SKILL.md and 1 other file (references) in skills/gemini-live-api-dev of google-gemini/gemini-skills.
Open the folder on GitHubat commit 832c8f9
Gemini Live API Dev next to the 5 skills that share the most tags, products or categories with it. Stars are the repository's; “used in” counts other GitHub owners with a copy.
| Skill | Stars | Used in | Tokens | Auto-check | Licence | Repo updated |
|---|---|---|---|---|---|---|
| Gemini Live API Dev this skillgoogle-gemini/gemini-skills | 4.3k | — | ~4.6k | Automated safety check: Pass | Apache-2.0 | |
| Gemini Live API DevJetBrains/skills | 363 | — | ~2.6k | Automated safety check: Pass | None | |
| Gemini Interactions APIAyuilos/Miffan | 182 | — | ~4.6k | Automated safety check: Pass | AGPL-3.0 | |
| Gemini API DevAyuilos/Miffan | 182 | 1 repos | ~1.4k | Automated safety check: Pass | AGPL-3.0 | |
| Azure AI Voicelive Pymicrosoft/skills | 3.1k | 6 repos | ~2.9k | Automated safety check: Pass | MIT | |
| Gemini API DevJetBrains/skills | 363 | — | ~1.6k | Automated safety check: Pass | None |
JetBrains/skills
A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API.
Ayuilos/Miffan
A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, streaming responses…
Ayuilos/Miffan
A skill your agent uses when building applications with Gemini API hosted models, including Gemini and Gemma 4, working with multimodal content (text, images, audio, video), implementing function…
microsoft/skills
Build real-time voice AI applications using Azure AI Voice Live SDK (azure-ai-voicelive).
JetBrains/skills
A skill your agent uses when building applications with Gemini models, Gemini API, working with multimodal content (text, images, audio, video), implementing function calling, using structured…
JetBrains/skills
A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, streaming responses, background research tasks…
google-gemini/gemini-skills
A skill your agent uses when writing code that calls the Gemini API for text generation, multi-turn chat, multimodal understanding, image generation, video generation, speech generation (TTS), voice…
google-gemini/gemini-skills
A skill your agent uses for generative video editing, text-to-video, image-referenced video generation, first-frame-to-video, first-and-last-frame transitions, and video extensions using Gemini Omni…
Works with
A skill your agent uses when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live. Gemini Live API Dev is an agent skill from google-gemini/gemini-skills, published by the product's own GitHub organization.8 Live.
Gemini Live API Dev fits situations like: building real-time; bidirectional streaming applications with the Gemini Live API; migrating legacy Live models (2.0/2.5/3; gemini 3.8 Live.
Run `npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a claude-code`. Or copy the skill folder (skills/gemini-live-api-dev in google-gemini/gemini-skills) into .claude/skills/gemini-live-api-dev in your project. Claude Code loads it when a task matches its description.
Run `npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a codex`. Or copy the skill folder (skills/gemini-live-api-dev in google-gemini/gemini-skills) into .agents/skills/gemini-live-api-dev in your project. Codex loads it when a task matches its description.
Cursor, Gemini CLI, GitHub Copilot and OpenCode also load SKILL.md folders. With the skills CLI, run `npx skills add google-gemini/gemini-skills --skill gemini-live-api-dev -a cursor` (or -a gemini-cli, github-copilot or opencode for the others). To copy it by hand, put the folder in .cursor/skills/gemini-live-api-dev, .gemini/skills/gemini-live-api-dev, .github/skills/gemini-live-api-dev and .opencode/skills/gemini-live-api-dev in your project.
Going by SKILL.md and its folder, Gemini Live API Dev needs the command-line tools its instructions call (pip and npm). Our summary lists: Python 3; Node.js; A credential in YOUR_API_KEY.
SKILL.md names 8 domains. In commands or code: ai.google.dev; the agent is likely to contact it when it follows the instructions. As links in the text: docs.livekit.io, docs.pipecat.ai, docs.fishjam.io, visionagents.ai, voximplant.com, firebase.google.com and colab.research.google.com. This is read from the text; nothing was executed.
Our automated static check of SKILL.md found no risky patterns, such as piping downloads into a shell, reading credential files or hidden Unicode. It is not a guarantee. Review the folder before installing.
Gemini Live API Dev is published under the Apache-2.0 licence (the repository's licence). It allows redistribution, so the full SKILL.md is shown on this page.
About 4.6k tokens (SKILL.md is roughly 19k characters). Agents keep only the skill's name and description in context until a task matches; then they load SKILL.md in full. Its references folder adds about 2.2k tokens, read only when the agent opens those files.
Skills that share tags, products or a category with Gemini Live API Dev: Gemini Live API Dev (JetBrains/skills, 363 stars), Gemini Interactions API (Ayuilos/Miffan, 182 stars), Gemini API Dev (Ayuilos/Miffan, 182 stars) and Azure AI Voicelive Py (microsoft/skills, 3.1k stars). The comparison table on this page puts their stars, adoption, token cost, safety result and licence side by side.
google-gemini (a GitHub organization, an official publisher) maintains it in google-gemini/gemini-skills, which has 4,252 GitHub stars. The repository holds 3 skills in this directory. The repository was last updated on October 6, 2026.
Source: google-gemini/gemini-skills on GitHub. Facts on this page come from the repository at the commit we read; the author's words are quoted as theirs.