Gemini Media Analysis
MCP Video Recognition Server provides tools for image, audio, and video analysis using Google's Gemini AI. Built with TypeScript and the MCP SDK, it offers three main tools: image recognition for describing visual content, audio recognition for transcription and analysis, and video recognition for understanding video content. The server supports both stdio and SSE transport methods, includes file caching to improve performance, and handles the complexities of waiting for video processing to complete. It's particularly useful for applications requiring media content analysis, automated descriptions, or accessibility features without requiring direct integration with Google's APIs.
Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately. Last scanned 2026-09-28.
Scan details: Circle-IR · 2026-09-28 · Appeal
View full trust & usage report →Metadata
- Version
- 1.0.0
- Skill type
- atomic
- Execution layer
- mcp-remote
- Category
- media
- Source
- PulseMCP
- Author type
- human
- Last scanned
- 2026-09-28
- Updated
- 2026-09-28
Use via MCP
Resolve Gemini Media Analysis from your agent
Streamable HTTP transport at https://api.skillsregistry.net/mcp. No auth for read tools. Discovery: .well-known/mcp.json.
One command in your shell — Claude Code wires it up and verifies the connection. Run /mcp in any session to confirm.
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp --scope user for --scope project to commit it to .mcp.json.