Skip to main content

Open-source MCP Server thay thế Tavily - Web search, extract, crawl với SearXNG

Project description

WET - Web ExTract MCP Server

PyPI version License: MIT

Open-source MCP Server replacing Tavily for web scraping & multimodal extraction

Zero-install experience: just uvx wet-mcp - automatically setups and manages SearXNG container.

Features

Feature Description
Web Search Search via SearXNG (metasearch: Google, Bing, DuckDuckGo, Brave)
Content Extract Extract clean content (Markdown/Text/HTML)
Deep Crawl Crawl multiple pages from a root URL with depth control
Site Map Discover website URL structure
Media List and download images, videos, audio files
Anti-bot Stealth mode bypasses Cloudflare, Medium, LinkedIn, Twitter

Quick Start

Prerequisites

  • Docker daemon running (for SearXNG)
  • Python 3.13+ (or use uvx)

MCP Client Configuration

Claude Desktop / Cursor / Windsurf / Antigravity:

{
  "mcpServers": {
    "wet": {
      "command": "uvx",
      "args": ["wet-mcp"]
    }
  }
}

That's it! When the MCP client calls wet-mcp for the first time:

  1. Automatically installs Playwright chromium
  2. Automatically pulls SearXNG Docker image
  3. Starts wet-searxng container
  4. Runs the MCP server

Without uvx

pip install wet-mcp
wet-mcp

Tools

Tool Actions Description
web search, extract, crawl, map Web operations
media list, download Media discovery & download
help - Full documentation

Examples

# Search
{"action": "search", "query": "python web scraping", "max_results": 10}

# Extract content
{"action": "extract", "urls": ["https://example.com"]}

# Crawl with depth
{"action": "crawl", "urls": ["https://docs.python.org"], "depth": 2}

# Map site structure
{"action": "map", "urls": ["https://example.com"]}

# List media
{"action": "list", "url": "https://github.com/python/cpython"}

# Download media
{"action": "download", "media_urls": ["https://example.com/image.png"]}

Tech Stack

Component Technology
Language Python 3.13
MCP Framework FastMCP
Web Search SearXNG (auto-managed Docker)
Web Crawling Crawl4AI
Docker Management python-on-whales

How It Works

┌─────────────────────────────────────────────────────────┐
│                    MCP Client                           │
│            (Claude, Cursor, Windsurf)                   │
└─────────────────────┬───────────────────────────────────┘
                      │ MCP Protocol
                      ▼
┌─────────────────────────────────────────────────────────┐
│                   WET MCP Server                        │
│  ┌──────────┐  ┌──────────┐  ┌──────────────────────┐   │
│  │   web    │  │  media   │  │        help          │   │
│  │ (search, │  │ (list,   │  │  (full documentation)│   │
│  │ extract, │  │ crawl,   │  └──────────────────────┘   │
│  │ crawl,   │  │ download)│                             │
│  │ map)     │  └────┬─────┘                             │
│  └────┬─────┘       │                                   │
│       │             │                                   │
│       ▼             ▼                                   │
│  ┌──────────┐  ┌──────────┐                             │
│  │ SearXNG  │  │ Crawl4AI │                             │
│  │ (Docker) │  │(Playwright)│                           │
│  └──────────┘  └──────────┘                             │
└─────────────────────────────────────────────────────────┘

Configuration

Environment variables:

Variable Default Description
WET_AUTO_DOCKER true Auto-manage SearXNG container
WET_SEARXNG_PORT 8080 SearXNG container port
SEARXNG_URL http://localhost:8080 External SearXNG URL
LOG_LEVEL INFO Logging level

Container Management

# View SearXNG logs
docker logs wet-searxng

# Stop SearXNG
docker stop wet-searxng

# Remove container (will be recreated on next run)
docker rm wet-searxng

# Reset auto-setup (forces re-install Playwright)
rm ~/.wet-mcp/.setup-complete

License

MIT License

Project details


Release history Release notifications | RSS feed

This version

1.3.0

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

wet_mcp-1.3.0.tar.gz (13.5 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

wet_mcp-1.3.0-py3-none-any.whl (18.5 kB view details)

Uploaded Python 3

File details

Details for the file wet_mcp-1.3.0.tar.gz.

File metadata

  • Download URL: wet_mcp-1.3.0.tar.gz
  • Upload date:
  • Size: 13.5 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: uv/0.9.28 {"installer":{"name":"uv","version":"0.9.28","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"Ubuntu","version":"24.04","id":"noble","libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":true}

File hashes

Hashes for wet_mcp-1.3.0.tar.gz
Algorithm Hash digest
SHA256 97ec23ea681556f5c516cc335640b9e2239b95ada58dc1a97c7a64e6631f5d54
MD5 9b2f96f5fe9460a1dc358cbf5f367aeb
BLAKE2b-256 206e9c136bb8597fb31d9d315e37361459cea71339f32bba9d89f447de366cf7

See more details on using hashes here.

File details

Details for the file wet_mcp-1.3.0-py3-none-any.whl.

File metadata

  • Download URL: wet_mcp-1.3.0-py3-none-any.whl
  • Upload date:
  • Size: 18.5 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: uv/0.9.28 {"installer":{"name":"uv","version":"0.9.28","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"Ubuntu","version":"24.04","id":"noble","libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":true}

File hashes

Hashes for wet_mcp-1.3.0-py3-none-any.whl
Algorithm Hash digest
SHA256 ff185162830893001828f3c4f95ca395c0a6e0923334bff617eb96f9ea4ab90a
MD5 de9b6ce60cff9e9e1a54ed3020936691
BLAKE2b-256 c7a2adfb5be95672708672faa6eda73b0ba51eb08412ada15d0a91bfd2ebef2f

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page