# Tesseract PDF MCP Server

> Use this tool when you need to extract text from PDF documents using Optical Character Recognition (OCR) technology, solving problems such as data extraction and document indexing. It takes PDF files as input and outputs extracted text, supporting multiple languages including English and Simplified Chinese. Ideal for use cases where manual data entry is impractical, such as large-scale document processing and automated data capture.

Canonical page: https://skillsregistry.net/skills/maximdx-tesseract-mcp-server  
JSON: https://api.skillsregistry.net/v1/skills/maximdx-tesseract-mcp-server

## Description

Provides OCR capabilities to extract text from PDF documents using Tesseract, with support for multiple languages including English and Simplified Chinese.

## Trust

- **Trust score (0–1):** 0.70
- **Verification tier:** verified
- **Last scanned:** 2026-09-02

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** file-system
- **Updated:** 2026-09-02

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/h5l194aym9)
- **Repository:** <https://github.com/maximdx/tesseract-mcp-server>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "maximdx-tesseract-mcp-server"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/maximdx-tesseract-mcp-server` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/maximdx-tesseract-mcp-server/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
