Learning Spatial Knowledge for Text to 3D Scene Generation
Anne Lynn S. Chang, Manolis Savva, Christopher D. Manning · 2014
We address the grounding of natural language to concrete spatial constraints, and inference of implicit pragmatics in 3D environments.We apply our approach to the task of text-to-3D scene generation.We present a representation for common sense spatial knowledge and an approach to extract it from 3D scene data.In text-to-3D scene generation, a user provides as input natural language text from which we extract explicit constraints on the objects that should appear in the scene.The main innovation of this work is to show how to augment these explicit constraints with learned spatial knowledge to infer missing objects and likely layouts for the objects in the scene.We demonstrate that spatial knowledge is useful for interpreting natural language and show examples of learned knowledge and generated 3D scenes.