InseRF: Text-Driven Generative Object Insertion in Neural 3D Scenes
We introduce InseRF, a novel method for generative object insertion in the NeRF reconstructions of 3D scenes. Based on a user-provided textual description and a 2D bounding box in a reference viewpoint, InseRF generates new objects in 3D scenes. Recently, methods for 3D scene editing have been profoundly transformed, owing to the use of strong priors of text-to-image diffusion models in 3D generative modeling. Existing methods are mostly effective in editing 3D scenes via style and appearance changes or removing existing objects. Generating new objects, however, remains a challenge for such methods, which we address in this study. Specifically, we propose grounding the 3D object insertion to a 2D object insertion in a reference view of the scene. The 2D edit is then lifted to 3D using a single-view object reconstruction method. The reconstructed object is then inserted into the scene, guided by the priors of monocular depth estimation methods. We evaluate our method on various 3D scene
1ETH Zurich, 2Google Zurich *Equal Contribution InseRF generates an object in a 3D scene via a text prompt and one 2D bounding box. Abstract We introduce InseRF, a novel method for generative object insertion in the NeRF reconstructions of 3D scenes. Based on a user-provided textual description and a 2D bounding box in a reference viewpoint, InseRF generates new objects in 3D scenes. Recently, methods for 3D scene editing have been profoundly transformed, owing to the use of strong priors of text-to-image diffusion models in 3D generative modeling. Existing methods are mostly effective…
saved by
related reading
- InseRF: Text-Driven Generative Object Insertion in Neural 3D Scenesarxiv.org
- An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversiontextual-inversion.github.io
- DreamGaussian: Generative Gaussian Splatting for Efficient 3D Content Creation | PDFarxiv.org
- NeRF: Neural Radiance Fieldsmatthewtancik.com
- One-2-3-45++: Fast Single Image to 3D Objects with Consistent Multi-View Generation and 3D Diffusionarxiv.org
- AI 3D Model Generator: Create 3D from Text & Images | Meshymeshy.ai
- Building NeRF at City Scaleneuralfields.cs.brown.edu
- [2211.10440] Magic3D: High-Resolution Text-to-3D Content Creationarxiv.org
- 2311.15980arxiv.org
- ReferIt3D: Neural Listeners for Fine-Grained 3D Object Identification in Real-World Scenesecva.net
- D-NeRF: Neural Radiance Fields for Dynamic Scenesopenaccess.thecvf.com
- 3DMV3dmv2023.github.io