Performance analysis of elastic search technique in identification and removalof duplicate data

0Citations
Citations of this article
6Readers
Mendeley users who have this article in their library.
Get full text

Abstract

Elastic search is a way to organize the data and make it easily accessible. It is a server based search on Lucene. It is a highly scalable, distributed and full-text search engine. Elastic search is developed in Java. It is published as open source under the terms of the Apache License. Elastic search is the most popular enterprise search engine. Elastic search includes all advances in speed, security, scalability, and hardware efficiency. Elastic search is a tool for querying written words. It can perform some other smart tasks, but its principal is returning text similar to a given query and statistical analyses of a quantity of text. Elasticsearch is a standalone database server, which is written in Java and using HTTP/JSON protocol,it’s takes data and optimized the data according to language based searches and stores it in a sophisticated format. Elastic search is very convenient, supporting clustering and leader selection out of the box. Whether it’s searching a database of trade products by description, finding similar text in a body of crawled web pages. In this manuscript elastic search capability of copied data identification and its removing techniques performance are analyzed.

Cite

CITATION STYLE

APA

Shaik, S., Naga, N., & Rao, M. (2019). Performance analysis of elastic search technique in identification and removalof duplicate data. International Journal of Innovative Technology and Exploring Engineering, 8(10), 2401–2405. https://doi.org/10.35940/ijitee.H6579.0881019

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free