返回
Semi-random subspace method for writeprint identification
DOI:10.1016/j.neucom.2012.11.015.png)
摘要
En 中文
The anonymous nature of online messages distribution causes a series of moral and legal issues. By analyzing identity cues people leave behind their texts, i.e., writeprint, potential authors can be identified individually. But writeprint identification is a difficult learning task, because of the high redundancy in stylistic feature set and high similarity of some authors' writing-style. In this paper, we propose a novel method, called semi-random subspace (Semi-RS), to simultaneously address the two problems. Different from the conventional random subspace method (RSM) which samples features from the whole feature set in a completely random way, the proposed Semi-RS randomly samples features on each individual-author feature set (IAFS) partitioned from the whole feature set. More specifically, we first divide the whole feature set into several IAFSs in a deterministic way, then construct a set of base classifiers on different randomly sampled feature sets from each IAFS, and finally combine all base classifiers for the final decision. Experimental results on the benchmark dataset demonstrate the effectiveness of the proposed method which improves previously reported results. In addition, we analyze the diversity of algorithm, reveals that Semi-RS constructs more diverse base classifiers than conventional RSMs. (C) 2012 Elsevier B.V. All rights reserved.
Keyword:
Writeprint
Individual-author feature set (IAFS)
Random subspace method (RSM)
Principal component analysis (PCA)
Diversity
期刊
IF:
6.5
论文数:
2.5W
被引数:
6.5W
机构
暂无机构信息
引用论文
Adaptive Impedance Control to Enhance Human Skill on a Haptic Interface System自适应阻抗控制以增强触觉接口系统上的人类技能
A Computational Framework for the Automated Construction of Glycosylation Reaction Networks
PLoS ONE
IF0

