Research on Automatic Segmentation and Layout of Poster Theme Visual Elements Using Text-Prompt-Driven 3D Base Model (SAT) and Contrastive Learning

Main Article Content

G. Wang

Abstract

Automatic segmentation and layout of poster visual elements require joint modeling of visual structure, theme semantics, and spatial configuration. To overcome weak theme understanding, unclear element boundaries, and limited semantic relevance in existing poster-design automation methods, this study proposes a text-prompt-driven segmentation and layout framework based on a SAT spatial-semantic foundation model and contrastive learning. A text-image-layout triplet dataset is constructed, containing poster theme descriptions, pixel-level masks, bounding boxes, and element category labels. The SAT encoder extracts multi-scale visual features, while a BERT-based prompt encoder projects theme semantics into a high-dimensional representation to guide visual attention toward theme-relevant regions. A contrastive-learning module constructs semantic association matrices among elements, and an attention fusion module outputs element masks and layout coordinates. Experiments on a self-constructed poster dataset show that the proposed method achieves an average IoU of 0.870, a layout rationality score of 8.6/10, and an inference time of 0.30 s per image. Ablation experiments confirm the contribution of text prompting, semantic guidance, and contrastive learning. The framework supports visual signal processing, multimodal feature alignment, and spatial layout optimization in intelligent design systems.

Downloads

Download data is not yet available.

Article Details

How to Cite
Wang, G. (2026). Research on Automatic Segmentation and Layout of Poster Theme Visual Elements Using Text-Prompt-Driven 3D Base Model (SAT) and Contrastive Learning. Advanced Electromagnetics, 15(3), 8212–8216. https://doi.org/10.7716/aem.v15i3.3939
Section
Research Articles

References

H. Liu, “The influence of the composition system of Chinese painting on poster design,” Design Art Research, vol. (1), pp. 4, 2020, doi: CNKI:SUN:JCGS.0.2020-01-027.

A. Ma, “A brief analysis of the application of creative programming in poster design,” Footwear Technology and Design, vol. 2, no. 20, pp. 64-66, 2022, doi: 10.3969/j.issn.2096-3793.2022-20-020.

View Article

Z. Deng W and Wu Y. Application Research of AR (Augmented Reality) Technology in Creative Interaction of Poster Design[J], “2021,”, doi: 10.2991/ASSEHR.K.210106.131.

View Article

L. Zhang, “A study on the impact of visual design of League of Legends global event posters on communication effect: focusing on composition, color and cultural symbols,” China Design, pp. 8(10), 2025, doi: 10.12428/zgsj2025.10.100.

View Article

N. Wang and N. Sun, “Visual grammar analysis of the movie poster “The Water Gate Bridge of Changjin Lake”,” Modern Linguistics, vol. 11, no. 7, pp. 3171-3176, 2023, doi: 10.12677/ML.2023.117430.

View Article

Y. Pyo, H. Cho, T. Lee, et al., “P-236: Late-News Poster: Layout Optimization of AMOLED Pixel Circuits based on Deep Reinforcement Learning,” SID International Symposium Digest of Technical Papers, vol. 56, no. 1, pp. 1635-1638, 2025, doi: 10.1002/sdtp.18514.

View Article

Z. Long, G. Wu, Y. Zhang, et al., “Poster: Repair Cross Browser Layout Issues by Combining Learning and Search-based technique,” IEEE, 2021, doi: 10.1109/ICST49551.2021.00062.

View Article

C. Xu, K. Han, and W. Xu, “Image-aware layout generation with user constraints for poster design,” The Visual Computer, vol. 41, no. 6, pp. 4239-4252, 2025, doi: 10.1007/s00371-024-03657-z.

View Article

M. Zhang, F. Liu, and M. C. Ran, “CrePoster: Leveraging multi-level features for cultural relic poster generation via attention-based framework,” Expert Systems with Application, vol. 245, no. Jul., pp. 123136.1-123136.14, 2024, doi: 10.1016/j.eswa.2024.123136.

View Article

E. Barker and V. Phillips, “Creating conference posters: Structure, form and content,” Journal of Perioperative Practice, vol. 31, no. 7-8, pp. 296-299, 2021, doi: 10.1177/1750458921996254.

View Article