Enforcing label consistency and lowering labelling time in infants’ pose data via semi-automatic annotation

Main Article Content

Greta DI MARINO

greta.dimarino@phd.unich.it

https://orcid.org/0009-0004-5264-258X
Emanuele CARDINALE

emanuele.cardinale@phd.unich.it

https://orcid.org/0009-0000-2677-2904
Alessio CORREANI

a.correani@staff.univpm.it

https://orcid.org/0000-0001-5666-7806
Lucia MIGLIORELLI

lmigliorelli@unite.it

https://orcid.org/0000-0003-2388-1501
Sara MOCCIA

sara.moccia@unich.it

https://orcid.org/0000-0002-4494-8907

Abstract

Infants’ spontaneous movements provide clinically relevant information about neurodevelopment, but their assessment still relies on qualitative visual inspection, which is prone to variability. Video-based systems with automatic pose estimation algorithms have been proposed to address this, yet annotating data to train these algorithms is time-consuming and prone to intra- and inter-annotator variability. To improve the efficiency and consistency of annotation, semi-automatic annotation has been proposed in the broader human pose estimation literature. To assess the extent to which semi-automatic annotation may also be beneficial for infants’ pose estimation, we benchmarked six state-of-the-art adult-trained models on a new dataset comprising 46 videos of preterm infants recorded in a neonatal unit. The predicted joints from the best-performing model were presented to human annotators for optional refinement, who were also asked to label the same joints manually for comparison. ViTPose and Sapiens achieved the highest joint localisation accuracy, with ViTPose requiring markedly lower computational cost. Semi-automatic annotation reduced inter-annotator variability by 7.82% and decreased the time to label a frame by 35.9% compared to full manual labelling. These findings support the use of semi-automatic annotation as an effective strategy to enhance the labelling process of infant pose estimation data from real clinical settings.

Keywords:

Human pose estimation, semi-automatic annotation, preterm infants, Sapiens, ViTPose

Sustainable Development Goal (SDG)

  • Good health and well-being

References

Article Details

DI MARINO, G., CARDINALE, E., CORREANI, A., MIGLIORELLI, L., & MOCCIA, S. (2026). Enforcing label consistency and lowering labelling time in infants’ pose data via semi-automatic annotation. Applied Computer Science, 22(3), 1-14. https://doi.org/10.35784/acs_9670