by kwp-lab
Web Fetch is a web scraping tool that converts web pages to markdown, extracts images, and works with proxies for secure
Fetches web pages and converts them to markdown format while extracting image URLs. Includes proxy support for corporate networks and restricted environments.
Web Fetch is a community-built MCP server published by kwp-lab that provides AI assistants with tools and capabilities via the Model Context Protocol. Web Fetch is a web scraping tool that converts web pages to markdown, extracts images, and works with proxies for secure It is categorized under search web.
You can install Web Fetch in your AI client of choice. Use the install panel on this page to get one-click setup for Cursor, Claude Desktop, VS Code, and other MCP-compatible clients. This server runs locally on your machine via the stdio transport.
MIT
Web Fetch is released under the MIT license. This is a permissive open-source license, meaning you can freely use, modify, and distribute the software.
Fetch and extract information from websites automatically
Example
Research competitor pricing, scrape product reviews, monitor news mentions
Automate 5-10 hours/week of manual web research
Track website changes, new content, price updates
Example
Monitor competitor blog for new posts, track stock availability, watch for pricing changes
Stay informed without manual checking, never miss important updates
Extract structured data from multiple websites
Example
Compile product listings from 10 e-commerce sites, aggregate job postings, collect real estate data
Build datasets 100x faster than manual copying
Share your MCP server with the developer community
We wired Web Fetch into a staging workspace; the listing’s GitHub and npm pointers saved time versus hunting across READMEs.
Web Fetch reduced integration guesswork — categories and install configs on the listing matched the upstream repo.
I recommend Web Fetch for teams standardizing on MCP; the explainx.ai page compares cleanly with sibling servers.
Web Fetch is among the better-indexed MCP projects we tried; the explainx.ai summary tracks the official description.
According to our notes, Web Fetch benefits from clear Model Context Protocol framing — fewer ambiguous “AI plugin” claims.
Web Fetch has been reliable for tool-calling workflows; the MCP profile page is a good permalink for internal docs.
Strong directory entry: Web Fetch surfaces stars and publisher context so we could sanity-check maintenance before adopting.
We evaluated Web Fetch against two servers with overlapping tools; this profile had the clearer scope statement.
According to our notes, Web Fetch benefits from clear Model Context Protocol framing — fewer ambiguous “AI plugin” claims.
Web Fetch is among the better-indexed MCP projects we tried; the explainx.ai summary tracks the official description.
showing 1-10 of 54
Model Context Protocol server for fetching web content with custom http proxy. This allows Claude Desktop (or any MCP client) to fetch web content and handle images appropriately.
<a href="https://glama.ai/mcp/servers/@kwp-lab/mcp-fetch"> <img width="380" height="200" src="https://glama.ai/mcp/servers/@kwp-lab/mcp-fetch/badge" /> </a>This repository forks from the @smithery/mcp-fetch and replaces the node-fetch implementation with the library node-fetch-native.
The server will use the http_proxy and https_proxy environment variables to route requests through the proxy server by default if they are set.
You also can set the MCP_HTTP_PROXY environment variable to use a different proxy server.
fetch: Retrieves URLs from the Internet and extracts their content as markdown. If images are found, their URLs will be included in the response.Image Processing Specifications:
Only extract image urls from the article content, and append them to the tool result:
{
"params": {
"url": "https://www.example.com/articles/123"
},
"response": {
"content": [
{
"type": "text",
"text": "Contents of https://www.example.com/articles/123:
Here is the article content
Images found in article:
- https://www.example.com/1.jpg.webp
- https://www.example.com/2.jpg.webp
- https://www.example.com/3.webp"
}
]
}
}
To use this tool with Claude Desktop, simply add the following to your Claude Desktop configuration (~/Library/Application Support/Claude/claude_desktop_config.json):
{
"tools": {
"fetch": {
"command": "npx",
"args": ["-y", "@kwp-lab/mcp-fetch"],
"env": {
"MCP_HTTP_PROXY": "https://example.com:10890" // Optional, remove if not needed
}
}
}
}
This will automatically download and run the latest version of the tool when needed.
The following sections are for those who want to develop or modify the tool.
npm install -g tsx)To install MCP Fetch for Claude Desktop automatically via Smithery:
npx -y @smithery/cli install @kwp-lab/mcp-fetch --client claude
git clone https://github.com/kwp-lab/mcp-fetch.git
cd mcp-fetch
npm install
npm run build
Make sure Claude Desktop is installed and running.
Install tsx globally if you haven't:
npm install -g tsx
# or
pnpm add -g tsx
Modify your Claude Desktop config located at:
~/Library/Application Support/Claude/claude_desktop_config.json
You can easily find this through the Claude Desktop menu:
Add the following to your MCP client's configuration:
{
"tools": {
"fetch": {
"args": ["tsx", "/path/to/mcp-fetch/index.ts"]
}
}
}
Interact with services that don't offer APIs
Example
Check form submissions, validate website functionality, test user flows
Automate interactions with any website, even without API
Prerequisites
Time Estimate
20-40 minutes including configuration and testing
Steps
Troubleshooting
✓ Do
✗ Don't
💡 Pro Tips
Architecture
MCP server handles HTTP requests, HTML parsing, JavaScript rendering (if headless browser), and returns structured data to Claude.
Protocols
Compatibility
✓ Use when
Use for research automation, content monitoring, data aggregation from multiple sources, and when official APIs don't exist. Best for read-only information gathering.
✗ Avoid when
Avoid for sites with APIs (use API instead), sites that explicitly forbid scraping, when data is copyrighted, or for login-required content without proper authorization.