arrow
Return

Grasp Pattern Recognition Using Convolutional Vision Transformer

delete2026-01-01
delete0
PRE
AI
J
Johnson, Jeffy
R
Ravi, Vidharshana
S
Swagat Kumar Samantaray
S
Sekar Anup Chander
S
Srikanth Vasamsetti *
DOI:10.1007/978-3-031-93691-3_27delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
The goal of grasp pattern classification is to establish the grasp type for an object to be grasped, which can be used to improve prosthetic hand control and reduce the strain on amputees. We present a Convolutional Vision Transformer with two components: Convolutional Neural Network (CNN) and Vision Transformer (ViT). Convolutional operations within CNN are skilful at extracting local features but struggle to capture global representations. In contrast, cascaded self-attention modules in visual transformers capture long-distance feature dependencies but degrade local detail. The CNN extracts learnable features, while the ViT uses an attention mechanism to categorize the learnt features. The model was trained on two household object datasets: the RGBD Object Dataset and the Hit-GPRec dataset, and has achieved global accuracy of 78.57% and 84.02%, respectively, under BOC and 99% and 99.59% under WWC. Our contribution involves incorporating a CNN module into the ViT architecture, resulting in competitive results on the datasets.
Keywords:
Grasp pattern classification
convolutional neural network
vision transformer
deep learning
computer vision

Journal

C
COMPUTER VISION AND IMAGE PROCESSING, CVIP 2024, PT II
IF:
0
Papers:
33
Citations:
0

Organization

C
College of Engineering Guindy
Scholars:
193
Papers: 184
Citations: 0
A
Anna University Chennai
Scholars:
2.4K
Papers: 2.5K
Citations: 2
V
VIT Bhopal University
Scholars:
119
Papers: 84
Citations: 0
A
Anna University
Scholars:
7.0K
Papers: 6.4K
Citations: 32
researcher View more organizations