# Gemini Audio Upload

> Use this tool when you need to analyze audio files using advanced multimodal models, and guide the model's behavior with additional context and system instructions. It solves problems such as speech recognition, audio classification, and sentiment analysis by providing detailed insights into audio data. The tool accepts audio files as input and outputs analyzed data, making it ideal for applications where audio understanding is crucial.

Canonical page: https://skillsregistry.net/skills/unscene-gemini-audio-upload  
JSON: https://api.skillsregistry.net/v1/skills/unscene-gemini-audio-upload

## Description

Enables audio file analysis using Google's Gemini multimodal models with support for additional context and system instructions to guide the model's behavior.

## Trust

- **Trust score (0–1):** 0.99
- **Verification tier:** verified
- **Last scanned:** 2026-09-28

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-09-28

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/lkwwcb1vph)
- **Repository:** <https://github.com/unscene/gemini-audio-upload>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "unscene-gemini-audio-upload"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/unscene-gemini-audio-upload` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/unscene-gemini-audio-upload/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
