Understanding Pure CLIP Guidance for Voxel Grid NeRF Models

09/30/2022
by   Han-Hung Lee, et al.
0

We explore the task of text to 3D object generation using CLIP. Specifically, we use CLIP for guidance without access to any datasets, a setting we refer to as pure CLIP guidance. While prior work has adopted this setting, there is no systematic study of mechanics for preventing adversarial generations within CLIP. We illustrate how different image-based augmentations prevent the adversarial generation problem, and how the generated results are impacted. We test different CLIP model architectures and show that ensembling different models for guidance can prevent adversarial generations within bigger models and generate sharper results. Furthermore, we implement an implicit voxel grid model to show how neural networks provide an additional layer of regularization, resulting in better geometrical structure and coherency of generated objects. Compared to prior work, we achieve more coherent results with higher memory efficiency and faster training speeds.

READ FULL TEXT

page 1

page 15

research
12/02/2022

3D-TOGO: Towards Text-Guided Cross-Category 3D Object Generation

Text-guided 3D object generation aims to generate 3D objects described b...
research
06/01/2023

Diffusion Self-Guidance for Controllable Image Generation

Large-scale generative models are capable of producing high-quality imag...
research
10/23/2022

Compressing Explicit Voxel Grid Representations: fast NeRFs become also small

NeRFs have revolutionized the world of per-scene radiance field reconstr...
research
10/18/2017

Amending the Characterization of Guidance in Visual Analytics

At VAST 2016, a characterization of guidance has been presented. It incl...
research
06/04/2023

Detector Guidance for Multi-Object Text-to-Image Generation

Diffusion models have demonstrated impressive performance in text-to-ima...
research
07/08/2022

Guiding the retraining of convolutional neural networks against adversarial inputs

Background: When using deep learning models, there are many possible vul...

Please sign up or login with your details

Forgot password? Click here to reset