ARECA-Lite: A lightweight modified ArecaNet with reduced complexity for real-time robust facial emotion recognition
Article Sidebar
Issue Vol. 22 No. 3 (2026)
-
Enforcing label consistency and lowering labelling time in infants’ pose data via semi-automatic annotation
Greta DI MARINO, Emanuele CARDINALE, Alessio CORREANI, Lucia MIGLIORELLI, Sara MOCCIA1-14
-
A data-driven framework for AI adoption efficiency assessment using hybrid DEA and machine learning methods
Ewa CHODAKOWSKA15-29
-
ARECA-Lite: A lightweight modified ArecaNet with reduced complexity for real-time robust facial emotion recognition
Mustapha Abdelkader LAOUMIR, Amina KINANE DAOUADJI, Fatima BENDELLA30-51
-
Designing vehicle structures using materials with a low carbon footprint
Bartosz ŁATA, Jacek CZARNIGOWSKI, Wiktor ISKRA, Miłosz CHWIEJCZAK, Iga KOPEĆ52-61
-
Parallelogram-mode: A novel clustering method for categorical data
Ashuza KUDERHA, Olamma IHEANETU62-81
-
Digitalisation of relay protection and implementation of main protections in digital form
Dmytro DANYLCHENKO, Vladyslav TSIUPA, Oleksandr MIROSHNYK, Taras SHCHUR, Katarzyna PIOTROWSKA82-96
-
Modified snake optimizer algorithm for solving the permutation flow shop scheduling problem
Hassan ALMAZINI, Salah MORTADA, Hussein Fouad ALMAZINI97-107
-
Computer-based data processing approaches to production scrap management
Łukasz WÓJCIK, Arkadiusz GOLA, Jakub PIZOŃ108-120
-
A hybrid parameter-tuning for adaptive variable-length particle swarm optimisation in cancer feature selection
Shir Li WANG, Siti RAMADHANI, Muhammad FIKRY, Haldi BUDIMAN, Theam Foo NG, Sumayyah DZULKIFLY, Roziana ARIFFIN121-147
-
Beyond classical optimisation: Toward feasibility-aware computational architectures for synchronised systems
Grzegorz BOCEWICZ, Czesław SMUTNICKI, Zbigniew BANASZAK148-167
-
A composite latency model for evaluating hybrid OLTP/OLAP information systems
Volodymyr SOLOHUB, Volodymyr PASHKEVYCH168-180
-
Automatic methods for 3D motion trajectories gap filling: Custom-based Kalman vs. BiLSTM
Kamil ŻELAZOWSKI, Wojciech WOJCIECHEWICZ, Maria SKUBLEWSKA-PASZKOWSKA, Paweł POWROŹNIK181-195
-
Implementation of an IEC 61215-oriented photovoltaic module test emulator with integrated predictive maintenance capabilities
Aristide TOLOK NELEM, Yannick Antoine ABANDA, Steyve Samson NYATTE, Mathieu Jean Pierre PESDJOCK, Achille MELINGUI, Pierre ELE196-218
-
Modelling the predictive reliability of rotating machines using Artificial Intelligence.
Fernand Joseph TOUKAP NONO, Tokoue Ngatcha DIANORRÉ, Offole FLORENC, Mouzong Pemi MARCELIN219-243
-
Anomaly detection in vibroarthrographic signals using handcrafted signal features and one-class methods
Robert KARPIŃSKI, Arkadiusz SYTA244–261
Archives
-
Vol. 22 No. 3
2026-09-30 15
-
Vol. 22 No. 2
2026-06-30 15
-
Vol. 22 No. 1
2026-03-31 15
-
Vol. 21 No. 4
2025-12-31 12
-
Vol. 21 No. 3
2025-09-30 12
-
Vol. 21 No. 2
2025-06-30 12
-
Vol. 21 No. 1
2025-03-31 12
-
Vol. 20 No. 4
2024-12-31 12
-
Vol. 20 No. 3
2024-09-30 12
-
Vol. 20 No. 2
2024-06-30 12
-
Vol. 20 No. 1
2024-03-30 12
-
Vol. 19 No. 4
2023-12-31 10
-
Vol. 19 No. 3
2023-09-30 10
-
Vol. 19 No. 2
2023-06-30 10
-
Vol. 19 No. 1
2023-03-31 10
-
Vol. 18 No. 4
2022-12-30 8
-
Vol. 18 No. 3
2022-09-30 8
-
Vol. 18 No. 2
2022-06-30 8
-
Vol. 18 No. 1
2022-03-31 8
Main Article Content
Authors
mustaphaabdelkader.laoumir@univ-usto.dz
Abstract
Facial emotion recognition (FER) in e-learning environments faces a persistent tension between recognition performance and deployment efficiency; high-performing architectures typically impose prohibitive computational costs that preclude real-time use on commodity hardware. This paper presents ARECA-Lite, a lightweight variant of ArecaNet designed to resolve this tension in resource-constrained deployments. ARECA-Lite integrates a truncated MobileNetV3-Small backbone with dual VGGFace2-pretrained InceptionResNetV1 sub-branches, a wavelet pooling module for structured frequency-domain decomposition, and a modified Assembled Residual Enhanced Cross-Attention (A.R.E.C.A.) module composed of two parallel RECA blocks. A FACS-guided offline preprocessing pipeline, targeted disgust oversampling, and a composite Dice-BCE loss collectively address class imbalance and low-contrast action unit (AU) evidence. Experiments are conducted on FER2013. ARECA-Lite achieves 82.41% accuracy and an 82.50% macro F1-score, outperforming a retrained ArecaNet baseline under the same experimental conditions by 7.10 and 9.43 percentage points, respectively, while reducing parameter count from 24.95 M to 4.69 M, GPU inference latency from 21.83 ms to 5.44 ms, and computational cost from 6.57 GFLOPs to 1.64 GFLOPs. The preprocessing and augmentation pipeline adds 14.88 percentage points in accuracy relative to raw-data training. Evaluation is limited to FER2013; broader validation is left to future work. These results demonstrate that competitive, class-balanced facial emotion recognition is achievable without the substantial computational overhead of larger attention-based models, making ARECA-Lite well-suited to low-latency affective computing on commodity hardware.
Keywords:
Sustainable Development Goal (SDG)
- Quality education
- Industry, Innovation, Technology and Infrastructure
References
Ba, J. L., Kiros, J. R., & Hinton, G. E. (2016). Layer normalization. ArXiv, abs/1607.06450. https://doi.org/10.48550/arXiv.1607.06450
Cao, Q., Shen, L., Xie, W., Parkhi, O. M., & Zisserman, A. (2018). VGGFace2: A dataset for recognising faces across pose and age. In 2018 13th IEEE International Conference on Automatic Face & Gesture Recognition (FG 2018) (pp. 67–74). IEEE. https://doi.org/10.1109/FG.2018.00020 DOI: https://doi.org/10.1109/FG.2018.00020
Dosovitskiy, A., Beyer, L., Kolesnikov, A., Weissenborn, D., Zhai, X., Unterthiner, T., Dehghani, M., Minderer, M., Heigold, G., Gelly, S., Uszkoreit, J., & Houlsby, N. (2021). An image is worth 16 × 16 words: Transformers for image recognition at scale. International Conference on Learning Representations (ICLR 2021).
Ekman, P., & Friesen, W. V. (1978). Facial Action Coding System: A technique for the measurement of facial movement. Consulting Psychologists Press. DOI: https://doi.org/10.1037/t27734-000
Goodfellow, I. J., Erhan, D., Carrier, P. L., Courville, A., Mirza, M., Hamner, B., Cukierski, W., Tang, Y., Thaler, D., Lee, D.-H., Zhou, Y., Ramaiah, C., Feng, F., Li, R., Wang, X., Athanasakis, D., Shawe-Taylor, J., Milakov, M., Park, J., … Bengio, Y. (2013). Challenges in representation learning: A report on three machine learning contests. In M. Lee, A. Hirose, Z.-G. Hou, & R. M. Kil (Eds.), Neural information processing: ICONIP 2013 (Lecture Notes in Computer Science, Vol. 8228, pp. 117–124). Springer. https://doi.org/10.1007/978-3-642-42051-1_16 DOI: https://doi.org/10.1007/978-3-642-42051-1_16
Gursesli, M. C., Lombardi, S., Duradoni, M., Bocchi, L., Guazzini, A., & Lanatà, A. (2024). Facial emotion recognition (FER) through custom lightweight CNN model: Performance evaluation in public datasets. IEEE Access, 12, 45543–45559. https://doi.org/10.1109/ACCESS.2024.3380847 DOI: https://doi.org/10.1109/ACCESS.2024.3380847
Hendrycks, D., & Gimpel, K. (2016). Gaussian error linear units (GELUs). ArXiv, abs/1606.08415. https://doi.org/10.48550/arXiv.1606.08415
Howard, A. G., Zhu, M., Chen, B., Kalenichenko, D., Wang, W., Weyand, T., Andreetto, M., & Adam, H. (2017). MobileNets: Efficient convolutional neural networks for mobile vision applications. ArXiv, abs/1704.04861. https://doi.org/10.48550/arXiv.1704.04861
Howard, A., Sandler, M., Chu, G., Chen, L.-C., Chen, B., Tan, M., Wang, W., Zhu, Y., Pang, R., Vasudevan, V., Le, Q. V., & Adam, H. (2019). Searching for MobileNetV3. In 2019 IEEE/CVF International Conference on Computer Vision (ICCV) (pp. 1314–1324). IEEE. https://doi.org/10.1109/ICCV.2019.00140 DOI: https://doi.org/10.1109/ICCV.2019.00140
Hu, J., Shen, L., & Sun, G. (2018). Squeeze-and-excitation networks. In 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (pp. 7132–7141). IEEE. https://doi.org/10.1109/CVPR.2018.00745 DOI: https://doi.org/10.1109/CVPR.2018.00745
Khan, T., Yasir, M., & Choi, C. (2025). Attention-enhanced optimized deep ensemble network for effective facial emotion recognition. Alexandria Engineering Journal, 119, 111–123. https://doi.org/10.1016/j.aej.2025.01.078 DOI: https://doi.org/10.1016/j.aej.2025.01.078
Kim, J., & Choi, G. (2025). ArecaNet: Robust facial emotion recognition via assembled residual enhanced cross-attention networks for emotion-aware human–computer interaction. Sensors, 25(23), Article 7375. https://doi.org/10.3390/s25237375 DOI: https://doi.org/10.3390/s25237375
Kinane Daouadji, A., & Bendella, F. (2024). Improving e-learning by facial expression analysis. Applied Computer Science, 20(2), 126–137. https://doi.org/10.35784/acs-2024-20 DOI: https://doi.org/10.35784/acs-2024-20
Li, S., Deng, W., & Du, J. (2017). Reliable crowdsourcing and deep locality-preserving learning for expression recognition in the wild. In 2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) (pp. 2852–2861). IEEE. https://doi.org/10.1109/CVPR.2017.277 DOI: https://doi.org/10.1109/CVPR.2017.277
Liao, X., Zhang, Y., & Liu, Z. (2024). RS-Xception: A lightweight facial emotion recognition network using depthwise separable convolutions. Journal of Real-Time Image Processing, 21(3), Article 89. https://doi.org/10.1007/s11554-024-01452-1
Loshchilov, I., & Hutter, F. (2019). Decoupled weight decay regularization. In International Conference on Learning Representations (ICLR 2019). OpenReview.
Milletari, F., Navab, N., & Ahmadi, S. A. (2016). V-net: Fully convolutional neural networks for volumetric medical image segmentation. In 2016 Fourth International Conference on 3D Vision (3DV) (pp. 565–571). IEEE. https://doi.org/10.1109/3DV.2016.79 DOI: https://doi.org/10.1109/3DV.2016.79
Mollahosseini, A., Hasani, B., & Mahoor, M. H. (2019). AffectNet: A database for facial expression, valence, and arousal computing in the wild. IEEE Transactions on Affective Computing, 10(1), 18–31. https://doi.org/10.1109/TAFFC.2017.2740923 DOI: https://doi.org/10.1109/TAFFC.2017.2740923
Naik, S., Bagayatkar, S., & Singh, P. (2026). Facial emotion recognition on FER-2013 using an EfficientNetB2-based approach. ArXiv, abs/2601.18228. https://doi.org/10.48550/arXiv.2601.18228
Ramachandran, P., Zoph, B., & Le, Q. V. (2017). Searching for activation functions. ArXiv, abs/arXiv.1710.05941. https://doi.org/10.48550/arXiv.1710.05941
Ramirez-Quintana, J. A., Muñoz-Pacheco, J. J., Ramirez-Alonso, G., Medrano-Hermosillo, J. A., & Corral-Saenz, A. D. (2025). Lightweight convolutional neural network with efficient channel attention mechanism for real-time facial emotion recognition in embedded systems. Sensors, 25(23), Article 7264. https://doi.org/10.3390/s25237264 DOI: https://doi.org/10.3390/s25237264
Roy, A. K., Kathania, H. K., Sharma, A., Dey, A., & Ansari, M. S. A. (2024). ResEmoteNet: Bridging accuracy and loss reduction in facial emotion recognition. ArXiv, abs/2409.10545. https://doi.org/10.48550/arXiv.2409.10545 DOI: https://doi.org/10.36227/techrxiv.172651476.62062165/v1
Saurav, S., Saini, R., & Singh, S. (2025). An integrated attention-guided deep convolutional neural network for facial expression recognition in the wild. Multimedia Tools and Applications, 84(12), 10027–10069. https://doi.org/10.1007/s11042-024-19012-2 DOI: https://doi.org/10.1007/s11042-024-19012-2
Sun, X., Yang, J., & Zhou, Y. (2026). Research on a lightweight real-time facial expression recognition system based on an improved Mini-Xception algorithm. Information, 17(1), Article 111. https://doi.org/10.3390/info17010111 DOI: https://doi.org/10.3390/info17010111
Syabil, M., Rahman, A., & Ahmad, F. (2024). Baseline convolutional neural networks for static emotion recognition on FER2013. Journal of Artificial Intelligence and Computer Science, 12(2), 45–58.
Szegedy, C., Ioffe, S., Vanhoucke, V., & Alemi, A. A. (2017). Inception-v4, Inception-ResNet and the impact of residual connections on learning. In Proceedings of the AAAI Conference on Artificial Intelligence (Vol. 31, No. 1, pp. 4278–4284). AAAI Press. https://doi.org/10.1609/aaai.v31i1.11231 DOI: https://doi.org/10.1609/aaai.v31i1.11231
Vaswani, A., Shazeer, N., Parmar, N., Uszkoreit, J., Jones, L., Gomez, A. N., Kaiser, Ł., & Polosukhin, I. (2017). Attention is all you need. In I. Guyon, U. von Luxburg, S. Bengio, H. Wallach, R. Fergus, S. Vishwanathan, & R. Garnett (Eds.), Advances in Neural Information Processing Systems 30 (pp. 5998–6008). Curran Associates, Inc.
Wang, H., Zhang, Y., Liu, X., Chen, M., & Zhao, L. (2025). Enhancing real-time facial emotion recognition in classrooms via attention-ResNet optimization. ArXiv, abs/2512.08970. https://doi.org/10.48550/arXiv.2512.08970 DOI: https://doi.org/10.1007/s00371-025-04265-1
Williams, T., & Li, R. (2018). Wavelet pooling for convolutional neural networks. In 6th International Conference on Learning Representations (ICLR 2018). OpenReview.
Yang, Q., He, Y., Chen, H., Wu, Y., & Rao, Z. (2025). A novel lightweight facial expression recognition network based on deep shallow network fusion and attention mechanism. Algorithms, 18(8), Article 473. https://doi.org/10.3390/a18080473 DOI: https://doi.org/10.3390/a18080473
Yun, S., Han, D., Oh, S. J., Chun, S., Choe, J., & Yoo, Y. (2019). CutMix: Regularization strategy to train strong classifiers with localizable features. In 2019 IEEE/CVF International Conference on Computer Vision (ICCV) (pp. 6023–6032). IEEE. https://doi.org/10.1109/ICCV.2019.00612 DOI: https://doi.org/10.1109/ICCV.2019.00612
Zhang, H., Cissé, M., Dauphin, Y. N., & Lopez-Paz, D. (2018). mixup: Beyond empirical risk minimization. In 6th International Conference on Learning Representations (ICLR 2018). OpenReview.
Zuiderveld, K. (1994). Contrast limited adaptive histogram equalization. In P. S. Heckbert (Ed.), Graphics gems IV (pp. 474–485). Academic Press. DOI: https://doi.org/10.1016/B978-0-12-336156-1.50061-6
Article Details
License

This work is licensed under a Creative Commons Attribution 4.0 International License.
All articles published in Applied Computer Science are open-access and distributed under the terms of the Creative Commons Attribution 4.0 International License.
