CVPR Poster SKE-Layout: Spatial Knowledge Enhanced Layout Generation with LLMs

Poster

SKE-Layout: Spatial Knowledge Enhanced Layout Generation with LLMs

Junsheng Wang · Nieqing Cao · Yan Ding · Mengying Xie · Fuqiang Gu · Chao Chen

ExHall D Poster #344

[ Abstract ] [ Paper PDF ]

Sat 14 Jun 3 p.m. PDT — 5 p.m. PDT

Abstract:

Generating layouts from textual descriptions by large language models (LLMs) plays a crucial role in precise spatial reasoning-induced domains such as robotic object rearrangement and text-to-image generation. However, current methods face challenges in limited real-world examples, handling diverse layout descriptions and varying levels of granularity. To address these issues, a novel framework named Spatial Knowledge Enhanced Layout (SKE-Layout), is introduced. SKE-Layout integrates mixed spatial knowledge sources, leveraging both real and synthetic data to enhance spatial contexts. It utilizes diverse representations tailored to specific tasks and employs contrastive learning and multitask learning techniques for accurate spatial knowledge retrieval. This framework generates more accurate and fine-grained visual layouts for object rearrangement and text-to-image generation tasks, achieving improvements of 5\%-30\% compared to existing methods.

Live content is unavailable. Log in and register to view live content