{"repo":"apache/nutch-webapp","free":true,"listed":false,"github":"https://github.com/apache/nutch-webapp","clone":"git clone https://github.com/apache/nutch-webapp.git","description":"Apache Nutch is an extensible and scalable web crawler","language":"Java","stars":10,"topics":["web-crawler","crawling","java","nutch","hadoop","apache"],"license":"Apache-2.0","category":"scrapers-browser-automation","readme_excerpt":"Apache Nutch WebApp README ========================== For the latest information about Nutch, please visit our website at: https://nutch.apache.org/ and our wiki, at: https://cwiki.apache.org/confluence/display/NUTCH/Home Introduction ------------ The Nutch WebApp is built using the Apache Wicket Java web framework and Spring. Running locally --------------- N.B. Currently, you must have a running Nutch REST Server on the same host. You can easily run the WebApp by executing the following If you want to run the WebApp in a Jakarta Servlet container i.e. Apache Tomcat, then run the following You can then access the WebApp on the Tomcat host on port 8080. Contributing ------------ To contribute a patch, follow these instructions (note that installing Hub is not strictly required, but is recommended). IDE setup --------- Generate Eclipse project files and follow the instructions in Importing existing projects. IntelliJ IDEA users can also import Eclipse projects using the \"Eclipser\" pluginhttps://plugins.jetbrains.com/plugin/7153-eclipser), see also Importing Eclipse Projects into IntelliJ IDEA.","default_branch":null,"files":null,"tree":[],"storefront":"/r/apache","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/apache/nutch-webapp/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}