MCP Deep Web Research Server (v0.3.0)

by qpd-v

86 429 downloads Not rated yet MIT
GitHub

About

It is an MCP server that brings real‑time web research into Claude Desktop. It provides intelligent search queuing, enhanced content extraction, and deep research capabilities for users who need structured, sourced information.

Details

License
MIT

Explore

- Intelligent Search Queue System
- Batch search operations with rate limiting
- Queue management with progress tracking
- Error recovery and automatic retries
- Search result deduplication

- Enhanced Content Extraction
- TF-IDF based relevance scoring
- Keyword proximity analysis
- Content section weighting
- Readability scoring
- Improved HTML structure parsing
- Structured data extraction
- Better content cleaning and formatting

- Core Features
- Google search integration
- Webpage content extraction
- Research session tracking
- Markdown conversion with improved formatting

Setting up with Highlight

This MCP is not yet compatible with Highlight’s one-click setup. However, you can still use it with Highlight by following these steps:

  1. Download and install Highlight from highlightai.com/download
  2. Navigate to the plugins tab and select "Add Custom Plugin"
  3. Configure the plugin with the settings below
    Plugin Name MCP Deep Web Research Server (v0.3.0)
    Command (node, npx, python, etc.)

    Please refer to the README for specific instructions on how to obtain API keys or other required environment variables.

  4. Enable "Start Automatically" if you want the plugin to start when Highlight launches

From the repository

- Node.js >= 18 (includes npm and npx)
- Claude Desktop app


npm install -g mcp-deepwebresearch

bash

After installation, run this command to install required browser dependencies:

npx playwright install chromium

The server can be configured through environment variables:

- MAX_PARALLEL_SEARCHES: Maximum number of concurrent searches (default: 5)
- SEARCH_DELAY_MS: Delay between searches in milliseconds (default: 200)
- MAX_RETRIES: Number of retry attempts for failed requests (default: 3)
- TIMEOUT_MS: Request timeout in milliseconds (default: 55000)
- LOG_LEVEL: Logging level (default: 'info')

pnpm install

1. deep_research
- Performs comprehensive research with content analysis
- Arguments:

     {
topic: string;
maxDepth?: number; // default: 2
maxBranching?: number; // default: 3
timeout?: number; // default: 55000 (55 seconds)
minRelevanceScore?: number; // default: 0.7
}

- Returns:
     {
findings: {
mainTopics: Array<{name: string, importance: number}>;
keyInsights: Array<{text: string, confidence: number}>;
sources: Array<{url: string, credibilityScore: number}>;
};
progress: {
completedSteps: number;
totalSteps: number;
processedUrls: number;
};
timing: {
started: string;
completed?: string;
duration?: number;
operations?: {
parallelSearch?: number;
deduplication?: number;
topResultsProcessing?: number;
remainingResultsProcessing?: number;
total?: number;
};
};
}

2. parallel_search
- Performs multiple Google searches in parallel with intelligent queuing
- Arguments: { queries: string[], maxParallel?: number }
- Note: maxParallel is limited to 5 to ensure reliable performance

3. visit_page
- Visit a webpage and extract its content
- Arguments: { url: string }
- Returns:

     {
url: string;
title: string;
content: string; // Markdown formatted content
}

Claude Desktop / Cursor

Paste into your MCP client config file to install this server.

{
    "mcpServers": {
        "mcp deep web research server (v0.3.0)": {
            "mcp-DEEPwebresearch": {
                "command": "npx",
                "args": [
                    "playwright",
                    "install",
                    "chromium"
                ]
            }
        }
    }
}

McpServers

{
    "mcp-DEEPwebresearch": {
        "command": "npx",
        "args": [
            "playwright",
            "install",
            "chromium"
        ]
    }
}

Node.js Version
TypeScript
License: MIT

A Model Context Protocol (MCP) server for advanced web research.

<a href="https://glama.ai/mcp/servers/5afpizjl6x">Web Research Server MCP server</a>

Latest Changes

- Added visit_page tool for direct webpage content extraction
- Optimized performance to work within MCP timeout limits
Reduced default maxDepth and maxBranching parameters
Improved page loading efficiency
Added timeout checks throughout the process
Enhanced error handling for timeouts

> This project is a fork of mcp-webresearch by mzxrai, enhanced with additional features for deep web research capabilities. We're grateful to the original creators for their foundational work.

Bring real-time info into Claude with intelligent search queuing, enhanced content extraction, and deep research capabilities.

Features

- Intelligent Search Queue System
- Batch search operations with rate limiting
- Queue management with progress tracking
- Error recovery and automatic retries
- Search result deduplication

- Enhanced Content Extraction
- TF-IDF based relevance scoring
- Keyword proximity analysis
- Content section weighting
- Readability scoring
- Improved HTML structure parsing
- Structured data extraction
- Better content cleaning and formatting

- Core Features
- Google search integration
- Webpage content extraction
- Research session tracking
- Markdown conversion with improved formatting

Prerequisites

- Node.js >= 18 (includes npm and npx)
- Claude Desktop app

Installation

Global Installation (Recommended)

```bash

No reviews yet — be the first

Sign in to leave a review

Use Google, GitHub, or an email account so ratings stay tied to real people.

Email sign in

No reviews posted yet.