Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios

Facial expressions convey massive information and play a crucial role in emotional expression. Deep neural network (DNN) accompanied by deep metric learning (DML) techniques boost the discriminative ability of the model in facial expression recognition (FER) applications. DNN, equipped with only cla...

Full description

Bibliographic Details
Main Authors: Rajoli, Hossein, Lotfi, Fatemeh, Atyabi, Adham, Afghah, Fatemeh
Format: Text
Language:unknown
Published: 2023
Subjects:
DML
Online Access:http://arxiv.org/abs/2302.04108
id ftarxivpreprints:oai:arXiv.org:2302.04108
record_format openpolar
spelling ftarxivpreprints:oai:arXiv.org:2302.04108 2023-09-05T13:19:06+02:00 Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios Rajoli, Hossein Lotfi, Fatemeh Atyabi, Adham Afghah, Fatemeh 2023-02-08 http://arxiv.org/abs/2302.04108 unknown http://arxiv.org/abs/2302.04108 Computer Science - Computer Vision and Pattern Recognition Computer Science - Artificial Intelligence Computer Science - Computational Complexity Computer Science - Computer Science and Game Theory text 2023 ftarxivpreprints 2023-08-16T17:31:58Z Facial expressions convey massive information and play a crucial role in emotional expression. Deep neural network (DNN) accompanied by deep metric learning (DML) techniques boost the discriminative ability of the model in facial expression recognition (FER) applications. DNN, equipped with only classification loss functions such as Cross-Entropy cannot compact intra-class feature variation or separate inter-class feature distance as well as when it gets fortified by a DML supporting loss item. The triplet center loss (TCL) function is applied on all dimensions of the sample's embedding in the embedding space. In our work, we developed three strategies: fully-synthesized, semi-synthesized, and prediction-based negative sample selection strategies. To achieve better results, we introduce a selective attention module that provides a combination of pixel-wise and element-wise attention coefficients using high-semantic deep features of input samples. We evaluated the proposed method on the RAF-DB, a highly imbalanced dataset. The experimental results reveal significant improvements in comparison to the baseline for all three negative sample selection strategies. Comment: The paper has been accepted in the CISS 2023 and will be published very soon Text DML ArXiv.org (Cornell University Library)
institution Open Polar
collection ArXiv.org (Cornell University Library)
op_collection_id ftarxivpreprints
language unknown
topic Computer Science - Computer Vision and Pattern Recognition
Computer Science - Artificial Intelligence
Computer Science - Computational Complexity
Computer Science - Computer Science and Game Theory
spellingShingle Computer Science - Computer Vision and Pattern Recognition
Computer Science - Artificial Intelligence
Computer Science - Computational Complexity
Computer Science - Computer Science and Game Theory
Rajoli, Hossein
Lotfi, Fatemeh
Atyabi, Adham
Afghah, Fatemeh
Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
topic_facet Computer Science - Computer Vision and Pattern Recognition
Computer Science - Artificial Intelligence
Computer Science - Computational Complexity
Computer Science - Computer Science and Game Theory
description Facial expressions convey massive information and play a crucial role in emotional expression. Deep neural network (DNN) accompanied by deep metric learning (DML) techniques boost the discriminative ability of the model in facial expression recognition (FER) applications. DNN, equipped with only classification loss functions such as Cross-Entropy cannot compact intra-class feature variation or separate inter-class feature distance as well as when it gets fortified by a DML supporting loss item. The triplet center loss (TCL) function is applied on all dimensions of the sample's embedding in the embedding space. In our work, we developed three strategies: fully-synthesized, semi-synthesized, and prediction-based negative sample selection strategies. To achieve better results, we introduce a selective attention module that provides a combination of pixel-wise and element-wise attention coefficients using high-semantic deep features of input samples. We evaluated the proposed method on the RAF-DB, a highly imbalanced dataset. The experimental results reveal significant improvements in comparison to the baseline for all three negative sample selection strategies. Comment: The paper has been accepted in the CISS 2023 and will be published very soon
format Text
author Rajoli, Hossein
Lotfi, Fatemeh
Atyabi, Adham
Afghah, Fatemeh
author_facet Rajoli, Hossein
Lotfi, Fatemeh
Atyabi, Adham
Afghah, Fatemeh
author_sort Rajoli, Hossein
title Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
title_short Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
title_full Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
title_fullStr Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
title_full_unstemmed Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
title_sort triplet loss-less center loss sampling strategies in facial expression recognition scenarios
publishDate 2023
url http://arxiv.org/abs/2302.04108
genre DML
genre_facet DML
op_relation http://arxiv.org/abs/2302.04108
_version_ 1776199910604931072