Search papers, labs, and topics across Lattice.
This paper introduces TGRHuman, a novel method for generating realistic 3D human models from text, addressing the limitations of current NeRF-based approaches in geometry and texture quality. By decoupling geometry and texture generation, TGRHuman employs explicit multi-view observation generation for efficient 3D synthesis, utilizing a high-resolution generative module and a diffusion renderer for detailed texture synthesis. Experimental results demonstrate that TGRHuman significantly outperforms existing text-to-3D human generation methods in both geometry and texture fidelity while maintaining 3D consistency and inference efficiency.
TGRHuman achieves unprecedented realism in 3D human generation from text, outperforming existing methods in both geometry and texture quality.
Realistic 3D human generation plays a crucial role in many graphics applications. However, current methods still struggle to generate high-quality human geometry and texture while maintaining 3D consistency and inference efficiency. In this work, we address these limitations by introducing TGRHuman, a novel approach for generating realistic 3D humans from text. Our method decouples geometry and texture generation to alleviate the issues commonly encountered in NeRF-based methods. Instead of relying on slow, implicit score-distillation-based optimization, we directly use explicit multi-view observation generation and optimization for efficient 3D synthesis. For geometry generation, we propose a high-resolution generative module for multi-view normals together with a geometry-carving strategy that preserves view consistency and supports loose clothing. For texture generation, we produce spatially consistent RGB observations from densely sampled surrounding views using a carefully designed texture-prior acquisition strategy and a diffusion renderer, enabling detailed human texture synthesis. Experiments show that our method can generate high-quality and consistent 3D human geometry and texture efficiently. TGRHuman outperforms existing text-to-3D human methods in both geometry and texture quality.