Markdown Web Extractor
Built by Vishwajeet Dabholkar, this server extracts clean markdown content from web pages using Playwright's headless Chrome browser. It intelligently identifies main content areas while filtering out navigation, headers, footers, and advertisements, with special optimizations for platforms like Confluence. The implementation supports configurable options including image and link inclusion, custom CSS selectors for waiting, timeout controls, and handles JavaScript-heavy sites by waiting for content to load. This makes it valuable for content analysis, documentation processing, and web scraping workflows where clean, readable text extraction is needed.
Composite of vulnerability cleanliness, spec conformance, provenance, stability, and usage signals — scanned and weighted by Cognium. Human and agent signals are tracked separately. Last scanned 2026-09-19.
Scan details: Circle-IR · 2026-09-19 · Appeal
View full trust & usage report →Metadata
- Version
- 1.0.0
- Skill type
- atomic
- Execution layer
- mcp-remote
- Category
- browser-automation
- Source
- PulseMCP
- Repository
- github.com/vishwajeetdabholkar/markdown-mcp
- Author type
- human
- Last scanned
- 2026-09-19
- Updated
- 2026-09-19
Use via MCP
Resolve Markdown Web Extractor from your agent
Streamable HTTP transport at https://api.skillsregistry.net/mcp. No auth for read tools. Discovery: .well-known/mcp.json.
One command in your shell — Claude Code wires it up and verifies the connection. Run /mcp in any session to confirm.
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp --scope user for --scope project to commit it to .mcp.json.