Performance analysis of elastic search technique in identification and removalof duplicate data

Subhani Shaik; Nallamothu Naga; Malleswara Rao

Journal ArticleOPEN ACCESS

Performance analysis of elastic search technique in identification and removalof duplicate data

International Journal of Innovative Technology and Exploring Engineering (2019) 8(10) 2401-2405

DOI: 10.35940/ijitee.H6579.0881019

1Citations

9Readers

Get full text

Abstract

Elastic search is a way to organize the data and make it easily accessible. It is a server based search on Lucene. It is a highly scalable, distributed and full-text search engine. Elastic search is developed in Java. It is published as open source under the terms of the Apache License. Elastic search is the most popular enterprise search engine. Elastic search includes all advances in speed, security, scalability, and hardware efficiency. Elastic search is a tool for querying written words. It can perform some other smart tasks, but its principal is returning text similar to a given query and statistical analyses of a quantity of text. Elasticsearch is a standalone database server, which is written in Java and using HTTP/JSON protocol,it’s takes data and optimized the data according to language based searches and stores it in a sophisticated format. Elastic search is very convenient, supporting clustering and leader selection out of the box. Whether it’s searching a database of trade products by description, finding similar text in a body of crawled web pages. In this manuscript elastic search capability of copied data identification and its removing techniques performance are analyzed.

Author supplied keywords

Cite

CITATION STYLE

APA

Shaik, S., Naga, N., & Rao, M. (2019). Performance analysis of elastic search technique in identification and removalof duplicate data. International Journal of Innovative Technology and Exploring Engineering, 8(10), 2401–2405. https://doi.org/10.35940/ijitee.H6579.0881019

Performance analysis of elastic search technique in identification and removalof duplicate data

Abstract

Author supplied keywords

Cite

Register to see more suggestions