# Website Scraper MCP Server

> Use this tool when you need to extract and organize website content for search and retrieval purposes. It solves problems related to data collection, preprocessing, and indexing by scraping, crawling, and cleaning website data, then outputting it to Azure AI Search. Ideal for use cases requiring full-text retrieval, such as information gathering, research, or data enrichment, with input of website URLs and output of indexed content in Azure AI Search.

Canonical page: https://skillsregistry.net/skills/lalit9168-web-scrapping  
JSON: https://api.skillsregistry.net/v1/skills/lalit9168-web-scrapping

## Description

Enables AI agents to scrape, crawl, clean, chunk, and index website content into Azure AI Search for full-text retrieval.

## Trust

- **Trust score (0–1):** 0.69
- **Verification tier:** scanned
- **Last scanned:** 2026-08-30

## Facts

- **Version:** 1.0.0
- **Skill type:** atomic
- **Execution layer:** mcp-remote
- **Runtime environment:** api
- **Category:** cloud-infra
- **Updated:** 2026-08-30

## Source

- **Source listing:** [Glama](https://glama.ai/mcp/servers/da2pquiq7h)
- **Repository:** <https://github.com/lalit9168/web-scrapping>

## Use it

Resolve this record through the SkillsRegistry MCP server (no auth, read-only):

```
claude mcp add --transport http --scope user skillsregistry https://api.skillsregistry.net/mcp
```

```json
{
  "jsonrpc": "2.0",
  "id": 1,
  "method": "tools/call",
  "params": {
    "name": "get_skill",
    "arguments": {
      "slug": "lalit9168-web-scrapping"
    }
  }
}
```

REST: `GET https://api.skillsregistry.net/v1/skills/lalit9168-web-scrapping` · pull for local use: `GET https://api.skillsregistry.net/v1/skills/lalit9168-web-scrapping/pull`

---
SkillsRegistry indexes agent skills from public registries and GitHub. Skills we have analysed are scanned with Circle-IR and scored on six dimensions; each listing states its scan coverage. More: https://skillsregistry.net/llms.txt
