TMM Volume 25 | 2023 | IEEE Signal Processing Society

2023

TMM Volume 25 | 2023

Unified Adaptive Relevance Distinguishable Attention Network for Image-Text Matching

TMM Volume 25 | 2023

Image-text matching, as a fundamental cross-modal task, bridges the gap between vision and language. The core is to accurately learn semantic alignment to find relevant shared semantics in image and text. Existing methods typically attend to all fragments with word-region similarity greater than empirical threshold zero as relevant shared semantics, e.g. , via a ReLU operation that forces the negative to zero and maintains the positive.

Decompose to Adapt: Cross-Domain Object Detection Via Feature Disentanglement

TMM Volume 25 | 2023

TMM Articles

Recent advances in unsupervised domain adaptation (UDA) techniques have witnessed great success in cross-domain computer vision tasks, enhancing the generalization ability of data-driven deep learning architectures by bridging the domain distribution gaps.

Block Division Convolutional Network With Implicit Deep Features Augmentation for Micro-Expression Recognition

TMM Volume 25 | 2023

TMM Articles

Despite the development of computer vision techniques, the micro-expression (ME) recognition task still remains a great challenge because MEs have very low intensity and short duration. However, the ME recognition is of great significance since it provides important clues for real affective states detection. This paper proposes a novel Block Division Convolutional Network (BDCNN) with the implicit deep features augmentation.

Deep Margin-Sensitive Representation Learning for Cross-Domain Facial Expression Recognition

TMM Volume 25 | 2023

TMM Articles

Cross-domain Facial Expression Recognition (FER) aims to safely transfer the learned knowledge from labeled source data to unlabeled target data, which is challenging due to the subtle difference between various expressions and the large discrepancy between domains. Existing methods mainly focus on reducing the domain shift for transferable features but fail to learn discriminative representations for recognizing facial expression, which may result in negative transfer under cross-domain settings.

3D Holoscopic Image Compression Based on Gaussian Mixture Model

TMM Volume 25 | 2023

TMM Articles

We introduce a Gaussian Mixture Model (GMM) framework for 3D holoscopic image compression in this paper. The elemental-images of the 3D holoscopic image are predicted using GMM and the parameters of GMM are estimated using the common Expectation-Maximization (EM) algorithm. GMM Model Optimization (GMO) is used in this framework to select the optimal number of distributions and avoid local optimum of EM at the same time.

Fast Human Pose Estimation in Compressed Videos

TMM Volume 25 | 2023

TMM Articles

Current approaches for human pose estimation in videos can be categorized into per-frame and warping-based methods. Both approaches have their pros and cons. For example, per-frame methods are generally more accurate, but they are often slow. Warping-based approaches are more efficient, but the performance is usually not good. To bridge the gap, in this paper, we propose a novel fast framework for human pose estimation to meet the real-time inference with controllable accuracy degradation in compressed video domain.

success.jpg

New Society Officer Elected

EUSIPCO_2025.jpg

(EUSIPCO 2025) 2025 European Signal Processing Conference

webinar_ASI.jpg

SPS-DSI (DEGAS) Webinar: Low Distortion Embedding with Bottom-up Manifold Learning

What is Signal Processing?

Popular Pages

Today's:

All time:

Last viewed:

TMM Volume 25 | 2023

TMM Menu

Publications & Resources

For Authors

award_nomination_article_2023_new.jpg

success.jpg

stars_general.jpg

Top Reasons to Join SPS Today!

Unified Adaptive Relevance Distinguishable Attention Network for Image-Text Matching

Decompose to Adapt: Cross-Domain Object Detection Via Feature Disentanglement

Block Division Convolutional Network With Implicit Deep Features Augmentation for Micro-Expression Recognition

Deep Margin-Sensitive Representation Learning for Cross-Domain Facial Expression Recognition

3D Holoscopic Image Compression Based on Gaussian Mixture Model

Fast Human Pose Estimation in Compressed Videos

SPS on Twitter

IEEE SPS Educational Resources

What is Signal Processing?

Popular Pages

Today's:

All time:

Last viewed:

TMM Volume 25 | 2023

Search form

You are here

TMM Menu

Publications & Resources

For Authors

Top Reasons to Join SPS Today!

SPS on Twitter

IEEE SPS Educational Resources