Skip to content

From text to sign language pose sequences: a Transformer-based approach for the automatic generation of gestures

2026 · CITA-DW · pp. 71-84 · 0 citations · 13 references
Computer Science

TL;DR

A Transformer-based approach for the automatic generation of sign language pose sequences directly from text without relying on gloss-based intermediate representations is proposed, demonstrating the potential of attention-based architectures for sign language generation and highlighting their relevance for improving digital accessibility.

View source

Similar papers

Conference Aug 2026

A Greedy Skeleton Retrieval Framework for Vietnamese Text-to-Sign Generation

Sign Language Production (SLP) plays a crucial role in bridging the communication gap between the Deaf community and broader society, functioning alongside Sign Language Translation (SLT) and Recognition (SLR). In addition to the limited scale of available data, research on Vietnamese Sign Language (VSL) is further hin...

D. Thanh, Thang Cap · 0 citations
Aug 2026

Variational Sign Language Translation

A novel framework based on conditional Variational autoencoder for SLT (VSLT) that facilitates direct and sufficient cross-modal alignment between sign language videos and spoken language text is proposed, and a shared Attention Residual Gaussian Distribution (ARGD) which considers the textual information as a residual...

Rui Zhao, Liang Zhang, Biao Fu et al. · 0 citations
Conference Open access Sep 2026

A Gloss-driven Indian Sign Language Production System Using Learned Pose Representations

A scalable and modular SLP framework based on Sign-Pose-VQ-VAE model, designed for low-resource settings, achieves state-of-the-art performance among keypoint-based methods on the PHOENIX14T benchmark, attaining a BLEU-4 score of 10.03 and surpassing the previous best method by 0.67 points.

Suvajit Patra, Arkadip Maitra, Swami Punyeshwarananda et al. · 0 citations
2026

Bridging Text-to-Sign Translation via Codebook-Oriented Pretraining

This work proposes a novel text-to-sign translation based on model pretraining, which enhances semantic alignment by inheriting codebook-oriented prior knowledge from masked self-supervised models.

Ninlawat Phuangchoke, C. Polprasert · 0 citations
#computer vision Preprint Sep 2026

Investigating Temporal Motion Features for Pose-to-Text Indian Sign Language Translation

We investigate the effect of pretrained T5 model scale and explicit motion features on pose-to-text Indian Sign Language Translation (SLT) for the WSLP 2026 Shared Task. Pose sequences are projected into the embedding space of T5 through a lightweight pose encoder, with the complete model fine-tuned to generate English...

Manav Dhamecha, Praveen Kumar Chandaliya, Pruthwik Mishra · 0 citations
#computer vision Preprint Sep 2026

Zero-Shot Cross-Lingual Recognition of Sign Language Handshapes

This work presents the first zero-shot cross-lingual framework for handshape recognition, transferring from ASL to Catalan Sign Language (LSC), and leverages the decomposition of handshapes into five phonological features shared across both languages, to decode LSC handshapes from predicted features via a composite pho...

Marcel Granero-Moya, Carolina del Corral Farrarós, G. Haro et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.