arrow
Return

Faster algorithm of string comparison

delete2003-06-01
delete51
PRE
AI
Y
Yang, QX
S
Sung Sam Yuan
L
Liping Zhao
C
Chun, L
S
Sheng‐Lung Peng
DOI:10.1007/s10044-002-0180-8delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In many applications, it is necessary to determine the field similarity. Our paper introduces a package of substring-based new algorithms to determine Field Similarity. Combined together, our new algorithms not only achieves higher accuracy, but also gains the time complexity O(knm) (k<0.75) for the worst case, O(beta*n) where beta<6 for the average case and O(1) for the best case. Throughout the paper, we use the approach of comparative examples to show the higher accuracy of our algorithms compared to that proposed in Lee et al. [1]. Theoretical analysis, concrete examples and experimental results show that our algorithms can significantly improve the accuracy and time complexity of the calculation of field similarity.
Keywords:
data cleaning
data mining
field similarity
pattern recognition
record similarity
string similarity

Journal

Pattern Analysis and Applications cover
Pattern Analysis and Applications
IF:
2
Papers:
1.9K
Citations:
1.9K

Organization

No organization information available