Tags
Language
Tags
August 2025
Su Mo Tu We Th Fr Sa
27 28 29 30 31 1 2
3 4 5 6 7 8 9
10 11 12 13 14 15 16
17 18 19 20 21 22 23
24 25 26 27 28 29 30
31 1 2 3 4 5 6
    Attention❗ To save your time, in order to download anything on this site, you must be registered 👉 HERE. If you do not have a registration yet, it is better to do it right away. ✌

    ( • )( • ) ( ͡⚆ ͜ʖ ͡⚆ ) (‿ˠ‿)
    SpicyMags.xyz

    Web Crawling and Data Mining with Apache Nutch

    Posted By: ksveta6
    Web Crawling and Data Mining with Apache Nutch

    Web Crawling and Data Mining with Apache Nutch by Dr. Zakir Laliwala, Abdulbasit Shaikh
    2013 | ISBN: 1783286857 | English | 136 Pages | PDF | 3 MB

    Perform web crawling and apply data mining in your application

    Overview

    Learn to run your application on single as well as multiple machines
    Customize search in your application as per your requirements
    Acquaint yourself with storing crawled webpages in a database and use them according to your needs
    In Detail

    Apache Nutch helps you to create your own search engine and customize it according to your needs. You can integrate Apache Nutch very easily with your existing application and get the maximum benefit from it. It can be easily integrated with different components like Apache Hadoop, Eclipse, and MySQL.

    "Web Crawling and Data Mining with Apache Nutch" shows you all the necessary steps to help you in crawling webpages for your application and using them to make your application searching more efficient. You will create your own search engine and will be able to improve your application page rank in searching.

    "Web Crawling and Data Mining with Apache Nutch" starts with the basics of crawling webpages for your application. You will learn to deploy Apache Solr on server containing data crawled by Apache Nutch and perform Sharding with Apache Nutch using Apache Solr.

    You will integrate your application with databases such as MySQL, Hbase, and Accumulo, and also with Apache Solr, which is used as a searcher.

    With this book, you will gain the necessary skills to create your own search engine. You will also perform link analysis and scoring that are helpful in improving the rank of your application page.

    What you will learn from this book

    Carry out web crawling for your application
    Make your application searching efficient by integrating it with Apache Solr
    Integrate your application with different databases for data storage purposes
    Run your application in a cluster environment by integrating it with Apache Hadoop
    Perform crawling operations with Eclipse, which is used as an IDE instead of the command line
    Create your own plugin in Apache Nutch
    Integrate Apache Solr with Apache Nutch, and deploy Apache Solr on Apache Tomcat
    Apply Sharding on Apache Tomcat for getting good results from Apache Solr while searching
    Approach

    This book is a user-friendly guide that covers all the necessary steps and examples related to web crawling and data mining using Apache Nutch.

    Who this book is written for

    "Web Crawling and Data Mining with Apache Nutch" is aimed at data analysts, application developers, web mining engineers, and data scientists. It is a good start for those who want to learn how web crawling and data mining is applied in the current business world. It would be an added benefit for those who have some knowledge of web crawling and data mining.