AgentMarketMCP / SKILL 资产档案馆

目录 / crawlforge-mcp

MCP 鉴权未知 未评级 已上架

crawlforge-mcp

CrawlForge MCP is a production-ready MCP server that gives AI agents the power to scrape websites, extract structured data, run deep research, bypass anti-bot detection, and process documents. It packages 27 specialized web scraping tools into a single MCP server with credit-based pricing and a free tier.

该来源不提供完整文件导出(国内平台多为平台内托管),仅存元数据与原链

接入信息

传输形态
http
鉴权方式
鉴权未知
端点
https://crawlforge-mcp--crawlforgedev.run.tools
鉴权方式未标注,请核对官方文档后再接入——不要直接使用以下片段
{
  "mcpServers": {
    "crawlforge-mcp": {
      "url": "https://crawlforge-mcp--crawlforgedev.run.tools"
    }
  }
}

能力清单

工具说明
fetch_urlFetch content from a URL with optional headers and timeout
extract_textExtract clean text content from a webpage
extract_linksExtract all links from a webpage with optional filtering
extract_metadataExtract metadata from a webpage (title, description, keywords, etc.)
scrape_structuredExtract structured data from a webpage using CSS selectors
search_webSearch the web using Google Search API (proxied through CrawlForge)
crawl_deepCrawl websites deeply using breadth-first search
map_siteDiscover and map website structure
extract_contentExtract and analyze main content from web pages with enhanced readability detection
process_documentProcess documents from multiple sources and formats including PDFs and web pages
summarize_contentGenerate intelligent summaries of text content with configurable options
analyze_contentPerform comprehensive content analysis including language detection and topic extraction
extract_structuredExtract structured data from a webpage using LLM-powered analysis and a JSON Schema. Falls back to CSS selector extraction when no LLM provider is configured.
batch_scrapeProcess multiple URLs simultaneously with support for async job management and webhook notifications
scrape_with_actionsExecute browser action chains before scraping content, with form auto-fill and intermediate state capture
deep_researchConduct comprehensive multi-stage research with intelligent query expansion, source verification, and conflict detection
track_changesEnhanced content change tracking with baseline capture, comparison, scheduled monitoring, advanced comparison engine, alert system, and historical analysis
generate_llms_txtAnalyze websites and generate standard-compliant LLMs.txt and LLMs-full.txt files defining AI model interaction guidelines
stealth_modeAdvanced anti-detection browser management with stealth features, fingerprint randomization, and human behavior simulation
localizationMulti-language and geo-location management with country-specific settings, browser locale emulation, timezone spoofing, and geo-blocked content handling
纠错与举报(发现条目失效、署名有误或涉及侵权?)
提交举报 / 纠错

侵权举报经核验成立后,我们会即时下线该条目并删除已存的内容副本。