Processes PDF documents with intelligent type detection, extracting text from both text-based and scanned documents. Integrates pdf-parse, pdf-lib, and pdf-to-png-converter with Baidu OCR API for multilingual document recognition. Provides text extraction with page range support, full-text search with regex capabilities, metadata extraction, and automatic fallback from text extraction to OCR when needed. Useful for document processing pipelines, content analysis systems, and AI-assisted document review workflows.
Cognium trust score
85%
Tier
Verified
Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately.
Last scanned 2026-09-19.
Returns 7 tools: search_skills, get_skill, list_leaderboard, get_trust_breakdown, resolve_composition, plus the ChatGPT-connector search and fetch. Every tool is annotated read-only.
Resolve this skill directly via MCP tools/call get_skill.