返回
Selecting statistical models and variable combinations for optimal classification using otolith microchemistry
DOI:10.1890/09-1887.1.png)
摘要
En 中文
Reliable assessment of fish origin is of critical importance for exploited species, since nursery areas must be identified and protected to maintain recruitment to the adult stock. During the last two decades, otolith chemical signatures (or fingerprints) have been increasingly used as tools to discriminate between coastal habitats. However, correct assessment of fish origin from otolith fingerprints depends on various environmental and methodological parameters, including the choice of the statistical method used to assign fish to unknown origin. Among the available methods of classification, Linear Discriminant Analysis (LDA) is the most frequently used, although it assumes data are multivariate normal with homogeneous within-group dispersions, conditions that are not always met by otolith chemical data, even after transformation. Other less constrained classification methods are available, but there is a current lack of comparative analysis in applications to otolith microchemistry. Here, we assessed stock identification accuracy for four classification methods (LDA, Quadratic Discriminant Analysis [QDA], Random Forests [RF], and Artificial Neural Networks [ANN]), through the use of three distinct data sets. In each case, all possible combinations of chemical elements were examined to identify the elements to be used for optimal accuracy in fish assignment to their actual origin. Our study shows that accuracy differs according to the model and the number of elements considered. Best combinations did not include all the elements measured, and it was not possible to define an ad hoc multielement combination for accurate site discrimination. Among all the models tested, RF and ANN performed best, especially for complex data sets (e. g., with numerous fish species and/or chemical elements involved). However, for these data, RF was less time-consuming and more interpretable than ANN, and far more efficient and less demanding in terms of assumptions than LDA or QDA. Therefore, when LDA and QDA assumptions cannot be reached, the use of machine learning methods, such as RF, should be preferred for stock assessment and nursery identification based on otolith microchemistry, especially when data set include multispecific otolith signatures and/or many chemical elements.
Keyword:
Artificial Neural Networks
classification method
fish origin
habitat discrimination
induced coupled plasma-mass spectrometer (SB ICP-MS)
Linear Discriminant Analysis
nursery habitats
otolith microchemistry
Quadratic Discriminant Analysis
Random Forest
trace elements
期刊
IF:
4.3
论文数:
5.0K
被引数:
2.2W
机构
引用论文
Heavy metals research in Nigeria: a review of studies and prioritization of research needs尼日利亚的重金属研究: 研究回顾和研究需求的优先次序

