Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios
Facial expressions convey massive information and play a crucial role in emotional expression. Deep neural network (DNN) accompanied by deep metric learning (DML) techniques boost the discriminative ability of the model in facial expression recognition (FER) applications. DNN, equipped with only cla...
Main Authors: | , , , |
---|---|
Format: | Text |
Language: | unknown |
Published: |
2023
|
Subjects: | |
Online Access: | http://arxiv.org/abs/2302.04108 |
id |
ftarxivpreprints:oai:arXiv.org:2302.04108 |
---|---|
record_format |
openpolar |
spelling |
ftarxivpreprints:oai:arXiv.org:2302.04108 2023-09-05T13:19:06+02:00 Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios Rajoli, Hossein Lotfi, Fatemeh Atyabi, Adham Afghah, Fatemeh 2023-02-08 http://arxiv.org/abs/2302.04108 unknown http://arxiv.org/abs/2302.04108 Computer Science - Computer Vision and Pattern Recognition Computer Science - Artificial Intelligence Computer Science - Computational Complexity Computer Science - Computer Science and Game Theory text 2023 ftarxivpreprints 2023-08-16T17:31:58Z Facial expressions convey massive information and play a crucial role in emotional expression. Deep neural network (DNN) accompanied by deep metric learning (DML) techniques boost the discriminative ability of the model in facial expression recognition (FER) applications. DNN, equipped with only classification loss functions such as Cross-Entropy cannot compact intra-class feature variation or separate inter-class feature distance as well as when it gets fortified by a DML supporting loss item. The triplet center loss (TCL) function is applied on all dimensions of the sample's embedding in the embedding space. In our work, we developed three strategies: fully-synthesized, semi-synthesized, and prediction-based negative sample selection strategies. To achieve better results, we introduce a selective attention module that provides a combination of pixel-wise and element-wise attention coefficients using high-semantic deep features of input samples. We evaluated the proposed method on the RAF-DB, a highly imbalanced dataset. The experimental results reveal significant improvements in comparison to the baseline for all three negative sample selection strategies. Comment: The paper has been accepted in the CISS 2023 and will be published very soon Text DML ArXiv.org (Cornell University Library) |
institution |
Open Polar |
collection |
ArXiv.org (Cornell University Library) |
op_collection_id |
ftarxivpreprints |
language |
unknown |
topic |
Computer Science - Computer Vision and Pattern Recognition Computer Science - Artificial Intelligence Computer Science - Computational Complexity Computer Science - Computer Science and Game Theory |
spellingShingle |
Computer Science - Computer Vision and Pattern Recognition Computer Science - Artificial Intelligence Computer Science - Computational Complexity Computer Science - Computer Science and Game Theory Rajoli, Hossein Lotfi, Fatemeh Atyabi, Adham Afghah, Fatemeh Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios |
topic_facet |
Computer Science - Computer Vision and Pattern Recognition Computer Science - Artificial Intelligence Computer Science - Computational Complexity Computer Science - Computer Science and Game Theory |
description |
Facial expressions convey massive information and play a crucial role in emotional expression. Deep neural network (DNN) accompanied by deep metric learning (DML) techniques boost the discriminative ability of the model in facial expression recognition (FER) applications. DNN, equipped with only classification loss functions such as Cross-Entropy cannot compact intra-class feature variation or separate inter-class feature distance as well as when it gets fortified by a DML supporting loss item. The triplet center loss (TCL) function is applied on all dimensions of the sample's embedding in the embedding space. In our work, we developed three strategies: fully-synthesized, semi-synthesized, and prediction-based negative sample selection strategies. To achieve better results, we introduce a selective attention module that provides a combination of pixel-wise and element-wise attention coefficients using high-semantic deep features of input samples. We evaluated the proposed method on the RAF-DB, a highly imbalanced dataset. The experimental results reveal significant improvements in comparison to the baseline for all three negative sample selection strategies. Comment: The paper has been accepted in the CISS 2023 and will be published very soon |
format |
Text |
author |
Rajoli, Hossein Lotfi, Fatemeh Atyabi, Adham Afghah, Fatemeh |
author_facet |
Rajoli, Hossein Lotfi, Fatemeh Atyabi, Adham Afghah, Fatemeh |
author_sort |
Rajoli, Hossein |
title |
Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios |
title_short |
Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios |
title_full |
Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios |
title_fullStr |
Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios |
title_full_unstemmed |
Triplet Loss-less Center Loss Sampling Strategies in Facial Expression Recognition Scenarios |
title_sort |
triplet loss-less center loss sampling strategies in facial expression recognition scenarios |
publishDate |
2023 |
url |
http://arxiv.org/abs/2302.04108 |
genre |
DML |
genre_facet |
DML |
op_relation |
http://arxiv.org/abs/2302.04108 |
_version_ |
1776199910604931072 |