{"repo":"dev-chenxing/jjwxc-crawler","free":true,"listed":false,"github":"https://github.com/dev-chenxing/jjwxc-crawler","clone":"git clone https://github.com/dev-chenxing/jjwxc-crawler.git","description":"基于Scrapy开发的晋江爬虫，根据书号下载小说非V章节，生成可编辑的Word文档 | A simple tool to scrape and download non-V chapters of any novel from jjwxc.net in .docx format, built with Python and Scrapy","language":"Python","stars":15,"topics":["docx","download","jjwxc","python","scrapy","word","chinese","crawler","scraping","cli"],"license":null,"category":"scrapers-browser-automation","readme_excerpt":"《重生之我在绿江爪爪巴》 一键下载 晋江文学城 (https://www.jjwxc.net) 网站小说非 V 章节 简体中文 English 特点功能 - 命令行界面 - 支持输出 DOCX 和 TXT 格式 - 可自定义输出路径 - ................... 有建议或 bug 可以提 issue. 命令行界面使用命令行 UI 库Rich编写。 界面样例： 安装文档 下载文件 点击 Code - Download ZIP，下载后解压缩得到文件夹，建议重命名为 jjwxc-crawler 环境配置 - Python 3.9.15 - Windows 安装 Python 后，第一步，打开所在目录的命令行，输入以下命令创建并激活虚拟环境 在Linux系统下， 此时命令行前应显示有 (venv) ，表示当前已激活虚拟环境 venv 第二步，在虚拟环境内安装 Scrapy 和其他依赖 运行小程序 下载章节将保存至根目录下的 novels 文件夹 默认输出格式为.docx，如果要更改为.txt 格式输出，可编辑 \\jjcrawler\\jjcrawler\\spiders\\config.py 中参数 下载一整页的小说 ⬆ 回到顶部","default_branch":null,"files":null,"tree":[],"storefront":"/r/dev-chenxing","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/dev-chenxing/jjwxc-crawler/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}