MCP server intelligence profile

MCP Web Scrape MCP Server

A comprehensive web scraping server that transforms web content into clean, agent-ready Markdown with automatic citations and efficient caching.

Local Onlymukul975
Awaiting current scanNpm · 1.0.7

The selected current version does not yet have completed public verification. Unknown does not mean clean or vulnerable.

1Distribution channel
48Independently observed tools
0Linked remote endpoints
AvailableVersion intelligence

Detailed security scan evidence is not public for this MCP yet. Public identity, registry metadata, and independently observed protocol inventory remain available.

Install and connect

Installation and connection instructions are shown only when supported by retained package, repository, or endpoint evidence.

Install mcp-web-scrape from npm

Version 1.0.7 declares 1 executable entrypoint.

npm install --save-exact mcp-web-scrape@1.0.7
npx -y -p mcp-web-scrape@1.0.7 mcp-web-scrape
MCP client configuration example
{
  "mcpServers": {
    "mcp-web-scrape": {
      "command": "npx",
      "args": [
        "-y",
        "-p",
        "mcp-web-scrape@1.0.7",
        "mcp-web-scrape"
      ]
    }
  }
}

Identity

Canonical slugmcp-web-scrape-8f341becDeploymentLocal Only
Canonical packagenpm:mcp-web-scrapeRepositorymukul975/mcp-web-scrape
First publishedLatest release
Last security verificationClassification confidence90%
PublicationDraftOfficial distributionNot verified

Distributions

ChannelIdentifierCurrent versionVersionsSource
npmmcp-web-scrape1.0.77Repository

Current release

PackageVersionPublished / observedInventorySecurity scan
npmmcp-web-scrape1.0.7CurrentSep 5, 202648 toolsPartial · 0 resources · 0 promptsEvidence restricted
Enterprise protection

Continuously monitor this MCP for security risk

Independently scan the exact version your agents use, receive alerts when its risk changes, and investigate every finding with retained version evidence.

  • Independent exact-version security scans
  • Continuous release and vulnerability monitoring
  • Risk-change alerts with capability context
  • Historical evidence and API exports
Custom pricingContact salesTailored to your organization, integrations, data needs, and support requirements.

Current version evidence

No public current-version evidence is available yet.

Current protocol inventory

2024-11-05Negotiated protocol
mcp-web-scrapeServer-reported name
2Capability groups
Aug 15, 2026Observed

Tools 48

ToolCategoryAnnotationsRisk
analyze_competitorsAnalyze competitor websites for SEO and content insights
Input schema
{
  "type": "object",
  "properties": {
    "urls": {
      "type": "array",
      "items": {
        "type": "string"
      },
      "description": "Array of competitor URLs to analyze"
    },
    "metrics": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "keywords",
          "meta-tags",
          "headings",
          "links",
          "performance",
          "all"
        ]
      },
      "description": "Metrics to compare (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "urls"
  ]
}
analyze_cookiesAnalyze cookies set by web pages for privacy and security
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze cookies for"
    },
    "includeThirdParty": {
      "type": "boolean",
      "description": "Whether to include third-party cookies (default: true)",
      "default": true
    },
    "checkSecurity": {
      "type": "boolean",
      "description": "Whether to check cookie security flags (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
analyze_page_speedAnalyze page loading speed and performance metrics
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze page speed for"
    },
    "device": {
      "type": "string",
      "enum": [
        "desktop",
        "mobile",
        "both"
      ],
      "description": "Device type for analysis (default: both)",
      "default": "both"
    },
    "metrics": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "fcp",
          "lcp",
          "cls",
          "fid",
          "ttfb",
          "all"
        ]
      },
      "description": "Performance metrics to analyze (default: all)",
      "default": [
        "all"
      ]
    }
  },
  "required": [
    "url"
  ]
}
analyze_performanceAnalyze web page performance metrics
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze performance for"
    },
    "metrics": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "size",
          "resources",
          "seo",
          "accessibility",
          "all"
        ]
      },
      "description": "Performance metrics to analyze (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
analyze_readabilityAnalyze text readability using various metrics
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze readability for"
    },
    "metrics": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "flesch",
          "gunning-fog",
          "coleman-liau",
          "ari",
          "all"
        ]
      },
      "description": "Readability metrics to calculate (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
analyze_traffic_patternsAnalyze traffic patterns and user behavior indicators
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze traffic patterns for"
    },
    "timeframe": {
      "type": "string",
      "enum": [
        "1h",
        "24h",
        "7d",
        "30d"
      ],
      "description": "Analysis timeframe (default: 24h)",
      "default": "24h"
    },
    "metrics": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "pageviews",
          "bounce-rate",
          "session-duration",
          "referrers",
          "all"
        ]
      },
      "description": "Traffic metrics to analyze (default: all)",
      "default": [
        "all"
      ]
    }
  },
  "required": [
    "url"
  ]
}
batch_extractExtract content from multiple URLs in a single operation
Input schema
{
  "type": "object",
  "properties": {
    "urls": {
      "type": "array",
      "items": {
        "type": "string"
      },
      "description": "Array of URLs to extract content from"
    },
    "format": {
      "type": "string",
      "enum": [
        "markdown",
        "text",
        "json"
      ],
      "description": "Output format (default: markdown)",
      "default": "markdown"
    },
    "maxConcurrent": {
      "type": "number",
      "description": "Maximum concurrent requests (default: 3)",
      "default": 3
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "urls"
  ]
}
benchmark_performanceBenchmark website performance against competitors and industry standards
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to benchmark performance for"
    },
    "competitors": {
      "type": "array",
      "items": {
        "type": "string"
      },
      "description": "Competitor URLs for comparison"
    },
    "metrics": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "speed",
          "seo",
          "accessibility",
          "security",
          "all"
        ]
      },
      "description": "Performance metrics to benchmark (default: all)",
      "default": [
        "all"
      ]
    }
  },
  "required": [
    "url"
  ]
}
check_broken_linksCheck for broken links and redirects on web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to check for broken links"
    },
    "checkExternal": {
      "type": "boolean",
      "description": "Whether to check external links (default: true)",
      "default": true
    },
    "timeout": {
      "type": "number",
      "description": "Timeout for link checks in seconds (default: 10)",
      "default": 10
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
check_privacy_policyAnalyze privacy policy content and compliance
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to check privacy policy for"
    },
    "regulations": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "gdpr",
          "ccpa",
          "coppa",
          "pipeda",
          "all"
        ]
      },
      "description": "Privacy regulations to check compliance for (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
check_ssl_certificateCheck SSL certificate validity and security details
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to check SSL certificate for"
    },
    "includeChain": {
      "type": "boolean",
      "description": "Whether to include certificate chain details (default: false)",
      "default": false
    }
  },
  "required": [
    "url"
  ]
}
check_url_statusCheck if URL is accessible and get HTTP status codes
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to check status for"
    },
    "followRedirects": {
      "type": "boolean",
      "description": "Whether to follow redirects (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
classify_contentClassify web content into categories and topics
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to classify content for"
    },
    "categories": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "news",
          "blog",
          "ecommerce",
          "education",
          "entertainment",
          "technology",
          "business",
          "health",
          "sports",
          "all"
        ]
      },
      "description": "Content categories to classify into (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
clear_cacheClear cached content entries
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "Specific URL to clear from cache (if not provided, clears all)"
    }
  },
  "required": []
}
compare_contentCompare content between two URLs or cached versions
Input schema
{
  "type": "object",
  "properties": {
    "url1": {
      "type": "string",
      "description": "First URL to compare"
    },
    "url2": {
      "type": "string",
      "description": "Second URL to compare"
    },
    "compareType": {
      "type": "string",
      "enum": [
        "text",
        "structure",
        "metadata"
      ],
      "description": "Type of comparison to perform (default: text)",
      "default": "text"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url1",
    "url2"
  ]
}
convert_to_pdfConvert web page content to PDF format
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to convert to PDF"
    },
    "format": {
      "type": "string",
      "enum": [
        "A4",
        "Letter",
        "Legal"
      ],
      "description": "PDF page format (default: A4)",
      "default": "A4"
    },
    "includeImages": {
      "type": "boolean",
      "description": "Whether to include images in PDF (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
detect_languageDetect the primary language of web page content
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to detect language for"
    },
    "confidence": {
      "type": "boolean",
      "description": "Whether to include confidence scores (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
detect_trackingDetect tracking scripts and privacy-related elements
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to detect tracking on"
    },
    "trackerTypes": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "analytics",
          "advertising",
          "social",
          "fingerprinting",
          "all"
        ]
      },
      "description": "Types of trackers to detect (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_contact_infoExtract contact information like emails, phones, addresses from web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract contact information from"
    },
    "types": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "email",
          "phone",
          "address",
          "all"
        ]
      },
      "description": "Types of contact information to extract (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_contentExtract and clean content from a web page, returning Markdown with citation
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to fetch and extract content from"
    },
    "format": {
      "type": "string",
      "enum": [
        "markdown",
        "text",
        "json"
      ],
      "description": "Output format (default: markdown)",
      "default": "markdown"
    },
    "includeImages": {
      "type": "boolean",
      "description": "Whether to include images in the output (default: true)",
      "default": true
    },
    "includeLinks": {
      "type": "boolean",
      "description": "Whether to include links in the output (default: true)",
      "default": true
    },
    "bypassRobots": {
      "type": "boolean",
      "description": "Whether to bypass robots.txt restrictions (default: false)",
      "default": false
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_entitiesExtract named entities (people, places, organizations) from web content
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract entities from"
    },
    "entityTypes": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "person",
          "organization",
          "location",
          "date",
          "money",
          "all"
        ]
      },
      "description": "Types of entities to extract (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_feedsDiscover and parse RSS/Atom feeds from web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to discover feeds from"
    },
    "maxItems": {
      "type": "number",
      "description": "Maximum number of feed items to return (default: 10)",
      "default": 10
    },
    "includeContent": {
      "type": "boolean",
      "description": "Whether to include full content of feed items (default: false)",
      "default": false
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_formsExtract form elements and their structure from web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract forms from"
    },
    "includeHidden": {
      "type": "boolean",
      "description": "Whether to include hidden form fields (default: false)",
      "default": false
    },
    "includeDisabled": {
      "type": "boolean",
      "description": "Whether to include disabled form fields (default: false)",
      "default": false
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_headingsExtract document structure and heading hierarchy from web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract headings from"
    },
    "levels": {
      "type": "array",
      "items": {
        "type": "number",
        "minimum": 1,
        "maximum": 6
      },
      "description": "Heading levels to extract (1-6, default: all)",
      "default": [
        1,
        2,
        3,
        4,
        5,
        6
      ]
    },
    "includeText": {
      "type": "boolean",
      "description": "Whether to include heading text content (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_imagesExtract all images from a web page with metadata
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract images from"
    },
    "includeAltText": {
      "type": "boolean",
      "description": "Whether to include alt text (default: true)",
      "default": true
    },
    "includeDimensions": {
      "type": "boolean",
      "description": "Whether to include image dimensions if available (default: false)",
      "default": false
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_keywordsExtract important keywords and phrases from web content
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract keywords from"
    },
    "maxKeywords": {
      "type": "number",
      "description": "Maximum number of keywords to extract (default: 20)",
      "default": 20
    },
    "includePhrases": {
      "type": "boolean",
      "description": "Whether to include multi-word phrases (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_linksExtract all links from a web page with filtering options
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract links from"
    },
    "linkType": {
      "type": "string",
      "enum": [
        "all",
        "internal",
        "external"
      ],
      "description": "Type of links to extract (default: all)",
      "default": "all"
    },
    "includeAnchorText": {
      "type": "boolean",
      "description": "Whether to include anchor text (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_schema_markupExtract and validate schema.org structured data markup
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract schema markup from"
    },
    "schemaTypes": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "Article",
          "Product",
          "Organization",
          "Person",
          "Event",
          "Recipe",
          "all"
        ]
      },
      "description": "Schema types to extract (default: all)",
      "default": [
        "all"
      ]
    },
    "validate": {
      "type": "boolean",
      "description": "Whether to validate schema markup (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_social_mediaExtract social media links and metadata from web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract social media links from"
    },
    "platforms": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "twitter",
          "facebook",
          "instagram",
          "linkedin",
          "youtube",
          "tiktok",
          "all"
        ]
      },
      "description": "Social media platforms to extract (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_structured_dataExtract JSON-LD, microdata, and schema.org data
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract structured data from"
    },
    "dataTypes": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "json-ld",
          "microdata",
          "rdfa",
          "opengraph"
        ]
      },
      "description": "Types of structured data to extract (default: all)",
      "default": [
        "json-ld",
        "microdata",
        "rdfa",
        "opengraph"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_tablesExtract and parse HTML tables with optional CSV export
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract tables from"
    },
    "format": {
      "type": "string",
      "enum": [
        "json",
        "csv",
        "markdown"
      ],
      "description": "Output format for tables (default: json)",
      "default": "json"
    },
    "includeHeaders": {
      "type": "boolean",
      "description": "Whether to include table headers (default: true)",
      "default": true
    },
    "minRows": {
      "type": "number",
      "description": "Minimum number of rows to include table (default: 1)",
      "default": 1
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
extract_text_onlyExtract plain text content without any formatting or HTML
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract text from"
    },
    "removeWhitespace": {
      "type": "boolean",
      "description": "Whether to remove extra whitespace (default: true)",
      "default": true
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
generate_meta_tagsGenerate optimized meta tags for SEO based on content analysis
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to generate meta tags for"
    },
    "targetKeywords": {
      "type": "array",
      "items": {
        "type": "string"
      },
      "description": "Target keywords for optimization"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
generate_reportsGenerate comprehensive reports combining multiple analysis tools
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to generate report for"
    },
    "reportType": {
      "type": "string",
      "enum": [
        "seo",
        "performance",
        "security",
        "accessibility",
        "comprehensive"
      ],
      "description": "Type of report to generate (default: comprehensive)",
      "default": "comprehensive"
    },
    "format": {
      "type": "string",
      "enum": [
        "json",
        "html",
        "markdown"
      ],
      "description": "Report output format (default: json)",
      "default": "json"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
generate_sitemapGenerate sitemap by crawling website pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The base URL to start crawling from"
    },
    "maxDepth": {
      "type": "number",
      "description": "Maximum crawl depth (default: 2)",
      "default": 2
    },
    "maxPages": {
      "type": "number",
      "description": "Maximum number of pages to crawl (default: 50)",
      "default": 50
    },
    "includeExternal": {
      "type": "boolean",
      "description": "Whether to include external links (default: false)",
      "default": false
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
generate_word_cloudGenerate word frequency analysis and word cloud data from web content
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze for word frequency"
    },
    "maxWords": {
      "type": "number",
      "description": "Maximum number of words to include (default: 100)",
      "default": 100
    },
    "minLength": {
      "type": "number",
      "description": "Minimum word length to include (default: 3)",
      "default": 3
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
get_cache_statsGet detailed cache statistics and usage information
Input schema
{
  "type": "object",
  "properties": {
    "includeEntries": {
      "type": "boolean",
      "description": "Whether to include list of cached entries (default: false)",
      "default": false
    }
  },
  "required": []
}
get_page_metadataExtract meta tags, title, description, keywords from web pages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to extract metadata from"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
monitor_changesMonitor web page content changes over time
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to monitor for changes"
    },
    "interval": {
      "type": "number",
      "description": "Check interval in seconds (default: 3600)",
      "default": 3600
    },
    "threshold": {
      "type": "number",
      "description": "Change detection threshold 0-1 (default: 0.1)",
      "default": 0.1
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
monitor_uptimeMonitor website uptime and availability
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to monitor uptime for"
    },
    "interval": {
      "type": "number",
      "description": "Check interval in seconds (default: 300)",
      "default": 300
    },
    "timeout": {
      "type": "number",
      "description": "Request timeout in seconds (default: 30)",
      "default": 30
    },
    "expectedStatus": {
      "type": "number",
      "description": "Expected HTTP status code (default: 200)",
      "default": 200
    }
  },
  "required": [
    "url"
  ]
}
scan_vulnerabilitiesScan web pages for common security vulnerabilities
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to scan for vulnerabilities"
    },
    "scanTypes": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "xss",
          "csrf",
          "headers",
          "forms",
          "cookies",
          "all"
        ]
      },
      "description": "Types of vulnerability scans to perform (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
search_contentSearch for specific text patterns within extracted content
Input schema
{
  "type": "object",
  "properties": {
    "content": {
      "type": "string",
      "description": "The content to search within"
    },
    "query": {
      "type": "string",
      "description": "The search query or pattern"
    },
    "caseSensitive": {
      "type": "boolean",
      "description": "Whether search should be case sensitive (default: false)",
      "default": false
    },
    "useRegex": {
      "type": "boolean",
      "description": "Whether to treat query as regex pattern (default: false)",
      "default": false
    },
    "maxResults": {
      "type": "number",
      "description": "Maximum number of results to return (default: 10)",
      "default": 10
    }
  },
  "required": [
    "content",
    "query"
  ]
}
sentiment_analysisAnalyze sentiment and emotional tone of web content
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to analyze sentiment for"
    },
    "granularity": {
      "type": "string",
      "enum": [
        "document",
        "paragraph",
        "sentence"
      ],
      "description": "Level of sentiment analysis (default: document)",
      "default": "document"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
summarize_contentGenerate a summary of already extracted content
Input schema
{
  "type": "object",
  "properties": {
    "content": {
      "type": "string",
      "description": "The content to summarize"
    },
    "maxLength": {
      "type": "number",
      "description": "Maximum length of the summary (default: 500)",
      "default": 500
    },
    "format": {
      "type": "string",
      "enum": [
        "paragraph",
        "bullets"
      ],
      "description": "Summary format (default: paragraph)",
      "default": "paragraph"
    }
  },
  "required": [
    "content"
  ]
}
track_changes_detailedTrack detailed changes in web page content with diff analysis
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to track changes for"
    },
    "sections": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "title",
          "headings",
          "content",
          "links",
          "images",
          "all"
        ]
      },
      "description": "Page sections to track changes for (default: all)",
      "default": [
        "all"
      ]
    },
    "sensitivity": {
      "type": "string",
      "enum": [
        "low",
        "medium",
        "high"
      ],
      "description": "Change detection sensitivity (default: medium)",
      "default": "medium"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
translate_contentTranslate web page content to different languages
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to translate"
    },
    "targetLanguage": {
      "type": "string",
      "description": "Target language code (e.g., es, fr, de, zh)"
    },
    "sourceLanguage": {
      "type": "string",
      "description": "Source language code (auto-detect if not provided)"
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url",
    "targetLanguage"
  ]
}
validate_htmlValidate HTML structure, accessibility, and SEO
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to validate"
    },
    "checks": {
      "type": "array",
      "items": {
        "type": "string",
        "enum": [
          "structure",
          "accessibility",
          "seo",
          "performance",
          "all"
        ]
      },
      "description": "Validation checks to perform (default: all)",
      "default": [
        "all"
      ]
    },
    "useCache": {
      "type": "boolean",
      "description": "Whether to use cached content if available (default: true)",
      "default": true
    }
  },
  "required": [
    "url"
  ]
}
validate_robotsCheck robots.txt compliance for specific URLs
Input schema
{
  "type": "object",
  "properties": {
    "url": {
      "type": "string",
      "description": "The URL to check robots.txt compliance for"
    },
    "userAgent": {
      "type": "string",
      "description": "User agent to check against (default: *)",
      "default": "*"
    }
  },
  "required": [
    "url"
  ]
}

Resources 0

  • None observed.

Resource templates 0

  • None observed.

Prompts 0

  • None observed.

Remote endpoints

EndpointTransportAuthenticationHealthObserved
No verified remote endpoint is linked.

MCP Web Scrape MCP Server questions

How do I install MCP Web Scrape MCP Server?

Install the selected package version with: npm install --save-exact mcp-web-scrape@1.0.7

What tools does MCP Web Scrape MCP Server provide?

MCP Web Scrape MCP Server exposed 48 tools during independent protocol observation, including analyze_competitors, analyze_cookies, analyze_page_speed, analyze_performance, analyze_readability, analyze_traffic_patterns, batch_extract, benchmark_performance, and others.

Is MCP Web Scrape MCP Server secure?

The selected current version does not yet have completed public verification. Unknown does not mean clean or vulnerable.

Explore related MCP server guides

Curated product and capability guides containing this catalog record.

Official vs Community MCP Servers

Let’s talk about MCP security.

Share your details and our security team will contact you.