[1]
X. Y. Yang and K. Wang, “Spatial Transformer Networks for Facial Action Unit Detection: Modeling Inter-AU Dependencies via Structured Attention”, AEM, vol. 15, no. 3, pp. 9761–9771, Aug. 2026.