服务器详情
Web content extraction API for AI agents. Scrape any URL and get clean, structured Markdown content with navigation, ads, and scripts stripped. Full JavaScript rendering via headless Chromium. Single and batch (10 URLs) modes. Built for RAG pipelines and AI research. Tools: web_scrape_to_markdown (single), web_scrape_batch (up to 10 URLs). Use this for RAG ingestion, research, content analysis, data extraction, or competitive intelligence. IMPORTANT: For screenshots/PDFs of pages, use capture_screenshot instead. For SEO analysis, use seo_audit_page. Returns: {markdown, title, wordCount, links[]}. No API key required — x402 micropayment $0.005/call on Base L2.
一个提供网页内容抓取功能的模型上下文协议服务器,使大型语言模型能够检索网页并将 HTML 转换为 Markdown,便于后续阅读、总结和处理。
收录本 MCP 服务的合集
工具测试
fetch
从互联网上抓取一个 URL,并将 HTML 转换为 Markdown 内容。
url字符串,必需。要抓取的网页地址。max_length整数,可选。返回的最大字符数,默认 5000。start_index整数,可选。从指定字符索引开始提取,便于分块读取。raw布尔值,可选。是否返回未经 Markdown 转换的原始内容。连接方式
{
"mcpServers": {
"Web Scraper — Clean Markdown from Any URL": {
"url": "部署后由服务提供方生成 Remote 地址"
}
}
}Remote 地址需要在部署后由服务提供方生成,页面不会伪造不可用的端点。
{
"mcpServers": {
"Web Scraper — Clean Markdown from Any URL": {
"command": "uvx",
"args": [
"mcp-server-fetch"
]
}
}
}Fetch 可直接使用 uvx 运行 mcp-server-fetch。
如何使用
- 01步骤 1
先检查服务器能力和权限范围。
- 02步骤 2
复制安装命令或 JSON 配置。
- 03步骤 3
在客户端中进行小范围连接测试。
- 04步骤 4
确认数据访问和维护状态后再长期使用。
讨论与反馈
这里保留讨论入口,方便继续核对来源、使用体验和维护状态。
前往来源页面