# Gemini TTS MCP Server

> Use this tool when you need to convert text into natural-sounding speech, especially for long texts or multi-speaker dialogues. It solves problems such as enhancing audio content, facilitating accessibility, and improving user experience through its text-to-speech capabilities. It takes text input and outputs audio playback via Windows Media Player, supporting multiple voices and automatic chunking.

Canonical page: https://skillsregistry.net/skills/bsmi021-mcp-gemini-tts  
JSON: https://api.skillsregistry.net/v1/skills/bsmi021-mcp-gemini-tts

## Description

Provides text-to-speech capabilities using Google's Gemini TTS API with support for multiple voices, automatic chunking of long text, multi-speaker dialogue, and audio playback via Windows Media Player.

## Trust

- **Trust score (0–1):** 0.69
- **Verification tier:** scanned
- **Last scanned:** 2026-09-03

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-09-03

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/gy77960tqi)
- **Repository:** <https://github.com/bsmi021/mcp-gemini-tts>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "bsmi021-mcp-gemini-tts"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/bsmi021-mcp-gemini-tts` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/bsmi021-mcp-gemini-tts/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
