The performance of text similarity algorithms

Open

Didik Dwi Prasetya, Aji Prasetya Wibawa, Tsukasa Hirashima

2018 International Journal of Advances in Intelligent Informatics Vol. 4 Issue 1 Article Cited by 72 SDG 17SDG 4SDG 16 Quartile

Abstract

Text similarity measurement compares text with available references to indicate the degree of similarity between those objects. There have been many studies of text similarity and resulting in various approaches and algorithms. This paper investigates four majors text similarity measurements, which include String-based, Corpus-based, Knowledge-based, and Hybrid similarities. The results of the investigation showed that the semantic similarity approach is more rational in finding substantial relationship between texts. © 2018, Universitas Ahmad Dahlan. All rights reserved.

Affiliations

Department of Electrical Engineering, State University of Malang, Indonesia; Graduate School of Engineering, Hiroshima University, Japan

Research at a Glance

Premium content — register to unlock

Research at a Glance

Register to unlock

Topics & SDG Alignment

Premium content — register to unlock

Topics & SDG Alignment

Register to unlock

Collaboration

Premium content — register to unlock

Collaboration

Register to unlock

Author Profile (Selected)

Premium content — register to unlock

Author Profile (Selected)

Register to unlock

References Overview

Premium content — register to unlock

References Overview

Register to unlock

Journal & Source

Premium content — register to unlock

Journal & Source

Register to unlock

Metadata & Integrity

Premium content — register to unlock

Metadata & Integrity

Register to unlock