How to compare uncertain data types: towards robust similarity measures
Publication Date
July 21, 2020
Creator
Abstract
In view of the importance of effective comparison of uncertain data in real-world applications, this thesis focuses on developing new similarity measures with high accuracy. As it identifies and articulates an inherent limitation of the popular set-theoretic similarity measures for continuous intervals where they return the same similarity value for very different sets of intervals (termed as aliasing), this thesis first underpins a new axiomatic definition of a robust similarity measure and then proposes a new similarity measure for continuous intervals based on their bidirectional subsethood. Beyond establishing theoretical foundation of the new measure, the thesis also demonstrates its robust results vis-a-vis existing measures and suitability for real world applications. In the next stage, it develops a generalized framework to assess similarity between discontinuous intervals as current approaches involve loss of discontinuity information and are also affected by aliasing of the popular measures— these weaknesses impact the accuracy of similarity results. This thesis further integrates Allen’s theory with the new generalized framework to make the latter more efficient.
Moving beyond intervals, this thesis extends the new similarity measure both vertically and horizontally (α-cut based) for comparing type-1 (T1) fuzzy sets as the shortcoming of popular similarity measures persists with their extension to T1 fuzzy sets. The empirical evaluation of the extended new measures with respect to key existing fuzzy set-theoretic similarity measures shows that the vertically extended new measure behaves intuitively for various types of fuzzy sets, except for non-normal fuzzy sets; however, the α-cut based extended new measure meets expectation in all cases. At the final stage, the utility of the new similarity measure is explored to improve the robustness of fuzzy integral (FI) based uncertain-data aggregation. As existing approaches to generate fuzzy measures (FMs) rely on popular similarity measures to capture the degree of similarity among individual sources (and their combinations), they are also impacted by their aforesaid limitation. Therefore, this thesis develops a new FM based on the new similarity measure which can generate intuitive aggregation outcome when used in combination with an FI.
Item Type
ethesis
Thesis Type
PhD
Supervisors
Subjects (LC)
Associated Schools / Departments
School of Computer Science (UK)
eprints ID
59792
UoN Repository URI
Except where otherwise noted, this item's license is described as
File(s)![Thumbnail Image]()
Name
How to Compare Uncertain Data Types - Towards Robust Similarity Measures.pdf
Type
Full-text
Description
Examined
Size
2.26 MB
Format
Adobe PDF
Checksum (MD5)
35d9d51bdd1035301dfbcebaad6387e6