# Tika MCP Server

> Use this tool when you need to extract text and metadata from various file formats, such as PDF, DOCX, and images, to enable AI assistants to understand file contents. It solves problems of file content analysis and information retrieval, providing inputs of file formats and outputs of extracted text and metadata. Ideal for use cases where AI agents require access to file content for further processing or analysis.

Canonical page: https://skillsregistry.net/skills/daveyproctor-tika-mcp  
JSON: https://api.skillsregistry.net/v1/skills/daveyproctor-tika-mcp

## Description

Extracts text and metadata from various file formats (PDF, DOCX, images with OCR) using Apache Tika, enabling AI assistants to understand file contents.

## Trust

- **Trust score (0–1):** 0.68
- **Verification tier:** scanned
- **Last scanned:** 2026-09-01

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** media
- **Updated:** 2026-09-01

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/qgs65fzatp)
- **Repository:** <https://github.com/daveyproctor/tika-mcp>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "daveyproctor-tika-mcp"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/daveyproctor-tika-mcp` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/daveyproctor-tika-mcp/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
