返回
Improving high-dimensional data fusion by exploiting the multivariate advantage
DOI:10.1016/j.chemolab.2016.05.010.png)
摘要
En 中文
As no analytical chemical platform exists that is able to characterize the full chemical composition of a sample, often multiple platforms are used to measure the same sample. The chemometric analysis of the resulting data then requires the data to be 'fused'. The more comprehensive view on each sample should enhance understanding of the underlying chemistry, and/or increase predictive accuracy of the resulting model. Different data fusion approaches have been proposed for this purpose; each has its own drawbacks and advantages. In this paper we propose a new strategy for data fusion by combining the advantages of low-level fusion with those of mid and high-level data fusion. We argue that the information that is usually discarded in the latter fusion approaches can still benefit both classification and regression when multiple data blocks are considered together. This information may be recovered by a regression employing the intraclass correlation between the discarded and retained data. A comprehensive simulation study shows that, for classification, the resulting data fusion method outperforms the conventional data fusion approaches in many scenarios of communal information between data blocks. A real-life example on predicting the bitterness of different beers shows that the method also has great potential for regression. (C) 2016 Elsevier B.V. All rights reserved.
Keyword:
Data fusion
Partial correlation
Classification
Mid-level data fusion
High-level data fusion
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
IF:
3.8
论文数:
4.6K
被引数:
1.2W
机构
引用论文
A Hybrid Agent-based Design Methodology for Dynamic Cross-layer Reliability in Heterogeneous Embedded Systems异构嵌入式系统中基于混合代理的动态跨层可靠性设计方法
MSClust: a tool for unsupervised mass spectra extraction of chromatography-mass spectrometry ion-wise aligned data
METABOLOMICS
IF3.3

