IEEE TMM Article

Zero-shot learning (ZSL) has received extensive attention recently especially in areas of fine-grained object recognition, retrieval, and image captioning. Due to the complete lack of training samples and high requirement of defense transferability, the ZSL model learned is particularly vulnerable against adversarial attacks. Recent work also showed adversarially robust generalization requires more data.

Federated Adversarial Domain Hallucination for Privacy-Preserving Domain Generalization

TMM Volume 26 | 2024

TMM Articles

Domain generalization aims to reduce the vulnerability of deep neural networks in the out-of-domain distribution scenario. With the recent and increasing data privacy concerns, federated domain generalization, where multiple domains are distributed on different local clients, has become an important research problem and brings new challenges for learning domain-invariant information from separated domains.

Unified Adaptive Relevance Distinguishable Attention Network for Image-Text Matching

TMM Volume 25 | 2023

TMM Articles

Image-text matching, as a fundamental cross-modal task, bridges the gap between vision and language. The core is to accurately learn semantic alignment to find relevant shared semantics in image and text. Existing methods typically attend to all fragments with word-region similarity greater than empirical threshold zero as relevant shared semantics, e.g. , via a ReLU operation that forces the negative to zero and maintains the positive.

Decompose to Adapt: Cross-Domain Object Detection Via Feature Disentanglement

TMM Volume 25 | 2023

TMM Articles

Recent advances in unsupervised domain adaptation (UDA) techniques have witnessed great success in cross-domain computer vision tasks, enhancing the generalization ability of data-driven deep learning architectures by bridging the domain distribution gaps.

Block Division Convolutional Network With Implicit Deep Features Augmentation for Micro-Expression Recognition

TMM Volume 25 | 2023

TMM Articles

Despite the development of computer vision techniques, the micro-expression (ME) recognition task still remains a great challenge because MEs have very low intensity and short duration. However, the ME recognition is of great significance since it provides important clues for real affective states detection. This paper proposes a novel Block Division Convolutional Network (BDCNN) with the implicit deep features augmentation.

Deep Margin-Sensitive Representation Learning for Cross-Domain Facial Expression Recognition

TMM Volume 25 | 2023

TMM Articles

Cross-domain Facial Expression Recognition (FER) aims to safely transfer the learned knowledge from labeled source data to unlabeled target data, which is challenging due to the subtle difference between various expressions and the large discrepancy between domains. Existing methods mainly focus on reducing the domain shift for transferable features but fail to learn discriminative representations for recognizing facial expression, which may result in negative transfer under cross-domain settings.

3D Holoscopic Image Compression Based on Gaussian Mixture Model

TMM Volume 25 | 2023

TMM Articles

We introduce a Gaussian Mixture Model (GMM) framework for 3D holoscopic image compression in this paper. The elemental-images of the 3D holoscopic image are predicted using GMM and the parameters of GMM are estimated using the common Expectation-Maximization (EM) algorithm. GMM Model Optimization (GMO) is used in this framework to select the optimal number of distributions and avoid local optimum of EM at the same time.

Fast Human Pose Estimation in Compressed Videos

TMM Volume 25 | 2023

TMM Articles

Current approaches for human pose estimation in videos can be categorized into per-frame and warping-based methods. Both approaches have their pros and cons. For example, per-frame methods are generally more accurate, but they are often slow. Warping-based approaches are more efficient, but the performance is usually not good. To bridge the gap, in this paper, we propose a novel fast framework for human pose estimation to meet the real-time inference with controllable accuracy degradation in compressed video domain.

general_get_involved_tc_article_full.jpg

Get Involved with Technical Committees by becoming a TC Affiliate

lecture_mic_general2.jpg

Call for Project Proposals: IEEE SPS SigMA Program - Signal Processing Mentorship Academy

nominate.jpg

Call for Nominations for Fellow Evaluation Committee Member Extended

What is Signal Processing?

Popular Pages

Today's:

All time:

Last viewed:

IEEE TMM Article

Top Reasons to Join SPS Today!

IEEE TMM Article

ATZSL: Defensive Zero-Shot Recognition in the Presence of Adversaries

Federated Adversarial Domain Hallucination for Privacy-Preserving Domain Generalization

Unified Adaptive Relevance Distinguishable Attention Network for Image-Text Matching

Decompose to Adapt: Cross-Domain Object Detection Via Feature Disentanglement

Block Division Convolutional Network With Implicit Deep Features Augmentation for Micro-Expression Recognition

Deep Margin-Sensitive Representation Learning for Cross-Domain Facial Expression Recognition

3D Holoscopic Image Compression Based on Gaussian Mixture Model

Fast Human Pose Estimation in Compressed Videos

Call for Nominations: IEEE T-MM 2023 Multimedia Prize Paper Award

nominate.jpg

Call for Nominations: IEEE T-MM 2023 Multimedia Prize Paper Award

nominate_2_general.jpg

Pages

SPS Social Media

IEEE SPS Educational Resources

What is Signal Processing?

Popular Pages

Today's:

All time:

Last viewed:

IEEE TMM Article

Search form

You are here

Top Reasons to Join SPS Today!

IEEE TMM Article

Pages

SPS Social Media

IEEE SPS Educational Resources