Image from Google Jackets

Feature Selection and Enhanced Krill Herd Algorithm for Text Document Clustering / by Laith Mohammad Qasim Abualigah.

By: Abualigah, Laith Mohammad Qasim, autor
Contributor(s): SpringerLink (Online service)
Series: (Studies in Computational Intelligence, 1860-949X; 816); (Intelligent Technologies and Robotics (Springer-42732)).Publisher: Cham : Imprint: Springer, 2019Description: 1 recurso en línea (XXVII, 165 páginas).ISBN: 9783030106744.Subject: Análisis ClusterOnline resources: Acceso a este recurso digital (usuarios Universidad Europea de Madrid)Digital Resources
Contents:
Chapter 1. Introduction -- Chapter 2. Krill Herd Algorithm -- Chapter 3. Literature Review -- Chapter 4. Proposed Methodology -- Chapter 5. Experimental Results -- Chapter 6. Conclusion and Future Work -- References -- List Of Publications.
Abstract: This book puts forward a new method for solving the text document (TD) clustering problem, which is established in two main stages: (i) A new feature selection method based on a particle swarm optimization algorithm with a novel weighting scheme is proposed, as well as a detailed dimension reduction technique, in order to obtain a new subset of more informative features with low-dimensional space. This new subset is subsequently used to improve the performance of the text clustering (TC) algorithm and reduce its computation time. The k-mean clustering algorithm is used to evaluate the effectiveness of the obtained subsets. (ii) Four krill herd algorithms (KHAs), namely, the (a) basic KHA, (b) modified KHA, (c) hybrid KHA, and (d) multi-objective hybrid KHA, are proposed to solve the TC problem; each algorithm represents an incremental improvement on its predecessor. For the evaluation process, seven benchmark text datasets are used with different characterizations and complexities. Text document (TD) clustering is a new trend in text mining in which the TDs are separated into several coherent clusters, where all documents in the same cluster are similar. The findings presented here confirm that the proposed methods and algorithms delivered the best results in comparison with other, similar methods to be found in the literature.
Tags from this library: No tags from this library for this title. Log in to add tags.
Star ratings
    Average rating: 0.0 (0 votes)
Holdings
Item type Current library Collection Call number Status Date due Barcode Item holds
LIBRO-E NO PRÉSTAMO LIBRO-E NO PRÉSTAMO Madrid Digital Acceso Electrónico (UEM) Ciencias e Ingeniería QA278.55 2019 EB (Browse shelf(Opens below)) Acceso electrónico eBooks26062314
Total holds: 0

Chapter 1. Introduction -- Chapter 2. Krill Herd Algorithm -- Chapter 3. Literature Review -- Chapter 4. Proposed Methodology -- Chapter 5. Experimental Results -- Chapter 6. Conclusion and Future Work -- References -- List Of Publications.

This book puts forward a new method for solving the text document (TD) clustering problem, which is established in two main stages: (i) A new feature selection method based on a particle swarm optimization algorithm with a novel weighting scheme is proposed, as well as a detailed dimension reduction technique, in order to obtain a new subset of more informative features with low-dimensional space. This new subset is subsequently used to improve the performance of the text clustering (TC) algorithm and reduce its computation time. The k-mean clustering algorithm is used to evaluate the effectiveness of the obtained subsets. (ii) Four krill herd algorithms (KHAs), namely, the (a) basic KHA, (b) modified KHA, (c) hybrid KHA, and (d) multi-objective hybrid KHA, are proposed to solve the TC problem; each algorithm represents an incremental improvement on its predecessor. For the evaluation process, seven benchmark text datasets are used with different characterizations and complexities. Text document (TD) clustering is a new trend in text mining in which the TDs are separated into several coherent clusters, where all documents in the same cluster are similar. The findings presented here confirm that the proposed methods and algorithms delivered the best results in comparison with other, similar methods to be found in the literature.

There are no comments on this title.

to post a comment.