{"repo":"weidwonder/crawl4ai-mcp-server","free":true,"listed":false,"github":"https://github.com/weidwonder/crawl4ai-mcp-server","clone":"git clone https://github.com/weidwonder/crawl4ai-mcp-server.git","description":"用于提供给本地开发者的 LLM的高效互联网搜索&内容获取的MCP Server， 节省你的token","language":"Python","stars":150,"topics":[],"license":"MIT","category":"mcp-servers","readme_excerpt":"Crawl4AI MCP Server 这是一个基于MCP (Model Context Protocol)的智能信息获取服务器,为AI助手系统提供强大的搜索能力和面向LLM优化的网页内容理解功能。通过多引擎搜索和智能内容提取,帮助AI系统高效获取和理解互联网信息,将网页内容转换为最适合LLM处理的格式。 特性 - 🔍 强大的多引擎搜索能力,支持DuckDuckGo和Google - 📚 面向LLM优化的网页内容提取,智能过滤非核心内容 - 🎯 专注信息价值,自动识别和保留关键内容 - 📝 多种输出格式,支持引用溯源 - 🚀 基于FastMCP的高性能异步设计 安装 方式1: 大部分的安装场景 1. 确保您的系统满足以下要求: - Python = 3.9 - 建议使用专门的虚拟环境 2. 克隆仓库: 3. 创建并激活虚拟环境: 4. 安装依赖: 5. 安装playwright浏览器: 方式2: 安装到Claude桌面客户端 via Smithery 通过 Smithery 将 Crawl4AI MCP 的 Claude 桌面端服务安装自动配置至您本地的 Claude 伸展中心 : 使用方法 服务器提供以下工具: search 强大的网络搜索工具,支持多个搜索引擎: - DuckDuckGo搜索(默认): 无需API密钥,全面处理AbstractText、Results和RelatedTopics - Google搜索: 需要配置API密钥,提供精准搜索结果 - 支持同时使用多个引擎获取更全面的结果 参数说明: - query : 搜索查询字符串 - num results : 返回结果数量(默认10) - engine : 搜索引擎选择 - \"duckduckgo\": DuckDuckGo搜索(默认) - \"google\": Google搜索(需要API密钥) - \"all\": 同时使用所有可用的搜索引擎 示例: read url 面向LLM优化的网页内容理解工具,提供智能内容提取和格式转换: - markdown with citations : 包含内联引用的Markdown(默认),保持信息溯源 - fit markdown : 经过LLM优化的精简内容,去除冗余信息 - raw markdown : 基础HTML→Markdown转换 - references markdown : 单独的引用/参考文献部分 - fit html : 生成fit markdown的过滤后HTML - markdown : 默认Markdown格式 示例: 示例: 如需使用Google搜索,需要在config.json中配置API密钥: LLM内容优化 服务器采用了一系列针对LLM的内容优化策略: - 智能内容识别: 自动识别并保留文章主体、关键信息段落 - 噪音过滤: 自动过滤导航栏、广告、页脚等对理解无帮助的内容 - 信息完整性: 保留URL引用,支持信息溯源 - 长度优化: 使用最小词数阈值(10)过滤无效片段 - 格式优化: 默认输出markdown with citations格式,便于LLM理解和引用 开发说明 项目结构: 配置说明 1. 复制配置示例文件: 2. 如需使用Google搜索,在config.json中配置API密钥: 更新日志 - 2025.02.08: 添加搜索功能,支持DuckDuckGo(默认)和Google搜索 - 2025.02.07: 重构项目结构,使用FastMCP实现,优化依赖管理 - 2025.02.","default_branch":null,"files":null,"tree":[],"storefront":"/r/weidwonder","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/weidwonder/crawl4ai-mcp-server/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}