Open Access iconOpen Access

ARTICLE

A Spatial-Temporal Normalized Contrastive Embedding for Robust Motion Similarity Retrieval

Seung-su Lee1, Young-Been Noh1, HwaYoung Jeong2, Kwang-il Hwang1,*

1 Department of Embedded Systems Engineering, Incheon National University, Incheon, Republic of Korea
2 Humanitas College, Kyung Hee University, Seoul, Republic of Korea

* Corresponding Author: Kwang-il Hwang. Email: email

Computers, Materials & Continua 2026, 88(3), 5 https://doi.org/10.32604/cmc.2026.081251

Abstract

Robust motion similarity retrieval from monocular 2D pose sequences is challenged by body-scale variation, viewpoint inconsistency, translation drift, and temporal misalignment. Existing contrastive skeleton learning methods primarily address action recognition and rarely integrate explicit geometric canonicalization for retrieval-oriented metric learning. This paper proposes a spatial-temporal normalized contrastive embedding framework that unifies structured nuisance suppression with scalable similarity representation learning. A four-stage normalization pipeline—torso-scale normalization, pelvis-centered alignment, posture-axis alignment, and phase-synchronized temporal resampling—removes geometric and temporal distortions prior to embedding. The normalized sequences are encoded using an acausal dilated temporal convolutional network trained with a hybrid contrastive objective combining NT-Xent and semi-hard triplet loss, enabling both global separation and fine-grained stylistic discrimination. A prototype-based representation further supports interpretable amateur-to-professional style mapping. Experiments on a golf swing benchmark achieve a Top-1 accuracy of 91.3%, outperforming BiLSTM and Dynamic Time Warping baselines. The framework establishes an invariant and interpretable paradigm for motion similarity retrieval applicable to broader human movement analysis tasks.

Keywords

Motion similarity retrieval; spatial-temporal normalization; contrastive representation learning; temporal convolutional networks (TCN); prototype-based embedding

Cite This Article

APA Style
Lee, S., Noh, Y., Jeong, H., Hwang, K. (2026). A Spatial-Temporal Normalized Contrastive Embedding for Robust Motion Similarity Retrieval. Computers, Materials & Continua, 88(3), 5. https://doi.org/10.32604/cmc.2026.081251
Vancouver Style
Lee S, Noh Y, Jeong H, Hwang K. A Spatial-Temporal Normalized Contrastive Embedding for Robust Motion Similarity Retrieval. Comput Mater Contin. 2026;88(3):5. https://doi.org/10.32604/cmc.2026.081251
IEEE Style
S. Lee, Y. Noh, H. Jeong, and K. Hwang, “A Spatial-Temporal Normalized Contrastive Embedding for Robust Motion Similarity Retrieval,” Comput. Mater. Contin., vol. 88, no. 3, pp. 5, 2026. https://doi.org/10.32604/cmc.2026.081251



cc Copyright © 2026 The Author(s). Published by Tech Science Press.
This work is licensed under a Creative Commons Attribution 4.0 International License , which permits unrestricted use, distribution, and reproduction in any medium, provided the original work is properly cited.
  • 128

    View

  • 22

    Download

  • 0

    Like

Share Link