Publicación:
Hybrid Convolutional Vision Transformer for Robust Low-Channel sEMG Hand Gesture Recognition: A Comparative Study with CNNs

dc.contributor.authorRodríguez Serrezuela, Ruthber
dc.date.accessioned2026-08-26T20:04:06Z
dc.date.available2026-08-26T20:04:06Z
dc.date.issued2025-12-12
dc.description.abstractHand gesture classification using surface electromyography (sEMG) is fundamental for prosthetic control and human–machine interaction. However, most existing studies focus on high-density recordings or large gesture sets, leaving limited evidence on performance in low-channel, reduced-gesture configurations. This study addresses this gap by comparing a classical convolutional neural network (CNN), inspired by Atzori’s design, with a Convolutional Vision Transformer (CViT) tailored for compact sEMG systems. Two datasets were evaluated: a proprietary Myo-based collection (10 subjects, 8 channels, six gestures) and a subset of NinaPro DB3 (11 transradial amputees, 12 channels, same gestures). Both models were trained using standardized preprocessing, segmentation, and balanced windowing procedures. Results show that the CNN performs robustly on homogeneous signals (Myo:94.2% accuracy) but exhibits increased variability in amputee recordings (NinaPro: 92.0%). In contrast, the CViT consistently matches or surpasses the CNN, reaching 96.6% accuracy on Myo and 94.2% on NinaPro. Statistical analyses confirm significant differences in the Myo dataset. The objective of this work is to determine whether hybrid CNN–ViT architectures provide superior robustness and generalization under low-channel sEMG conditions. Rather than proposing a new architecture, this study delivers the first systematic benchmark of CNN and CViT models across amputee and non-amputee subjects using short windows, heterogeneous signals, and identical protocols, highlighting their suitability for compact prosthetic–control systems.
dc.identifier.urihttps://dspace.corhuila.edu.co/handle/123456789/295
dc.language.isoen
dc.rightsAttribution-NonCommercial-NoDerivs 2.5 Colombiaen
dc.rights.urihttp://creativecommons.org/licenses/by-nc-nd/2.5/co/
dc.subjectconvolutional neural network (CNN)
dc.subjecthand gesture recognition
dc.subjecthybrid deep learning
dc.subjectmyoelectric pattern recognition
dc.subjectsurface electromyography (sEMG)
dc.subjectVision Transformer (ViT)
dc.titleHybrid Convolutional Vision Transformer for Robust Low-Channel sEMG Hand Gesture Recognition: A Comparative Study with CNNs
dc.typeArtículo
dspace.entity.typePublication
relation.isAuthorOfPublicationbc225318-4cd6-449c-b8dd-b5859fb32d45
relation.isAuthorOfPublication.latestForDiscoverybc225318-4cd6-449c-b8dd-b5859fb32d45

Archivos

Bloque original

Mostrando 1 - 1 de 1
Miniatura
Nombre:
Hybrid_Convolutional_Vision_Transformer_for_Robust.pdf
Tamaño:
3.65 MB
Formato:
Adobe Portable Document Format

Bloque de licencias

Mostrando 1 - 1 de 1
Miniatura
Nombre:
license.txt
Tamaño:
1.14 KB
Formato:
Item-specific license agreed to upon submission
Descripción: