A text-guided vision model for enhanced recognition of small instances. Applied Computer Science, v. 22, n. 1, p. 35–46, 31 Mar.2026.