# vibevoice-asr

> Use this tool when you need to transcribe spoken audio into text locally, with capabilities for speaker diarization, and integrate it into AI tools like Claude Code, Cursor, and OpenCode. It solves problems of manual transcription and enables efficient audio analysis, providing text outputs from audio inputs. Ideal for use cases requiring accurate and private speech-to-text transcription within AI workflows.

Canonical page: https://skillsregistry.net/skills/tjameswilliams-vibevoice-server  
JSON: https://api.skillsregistry.net/v1/skills/tjameswilliams-vibevoice-server

## Description

Local speech-to-text transcription using Microsoft's VibeVoice-ASR model with speaker diarization, enabling audio transcription directly in AI tools like Claude Code, Cursor, and OpenCode.

## Trust

- **Trust score (0–1):** 0.70
- **Verification tier:** verified
- **Last scanned:** 2026-08-31

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-08-31

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/aalldj30b0)
- **Repository:** <https://github.com/tjameswilliams/vibevoice-server>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "tjameswilliams-vibevoice-server"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/tjameswilliams-vibevoice-server` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/tjameswilliams-vibevoice-server/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
