Layout Generation Agents with Large Language Models (2405.08037v1)

Published 13 May 2024 in cs.HC and cs.AI

Abstract: In recent years, there has been an increasing demand for customizable 3D virtual spaces. Due to the significant human effort required to create these virtual spaces, there is a need for efficiency in virtual space creation. While existing studies have proposed methods for automatically generating layouts such as floor plans and furniture arrangements, these methods only generate text indicating the layout structure based on user instructions, without utilizing the information obtained during the generation process. In this study, we propose an agent-driven layout generation system using the GPT-4V multimodal LLM and validate its effectiveness. Specifically, the LLM manipulates agents to sequentially place objects in the virtual space, thus generating layouts that reflect user instructions. Experimental results confirm that our proposed method can generate virtual spaces reflecting user instructions with a high success rate. Additionally, we successfully identified elements contributing to the improvement in behavior generation performance through ablation study.

PDF HTML Abstract

Summarize Bookmark Chat (Pro)

References (11)

Authors (2)

Yuichi Sasazawa (4 papers)
Yasuhiro Sogawa (13 papers)

Tweets

https://twitter.com/gastronomy/status/1790595074553729285

HackerNews

Layout Generation Agents with Large Language Models (2 points, 0 comments)

Layout Generation Agents with Large Language Models (2405.08037v1)

Related Papers

Tweets

HackerNews