HumanMaterial: Human Material Estimation from a Single Image via Progressive Training

📅 2025-07-24
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This paper addresses the ill-posed problem of full-body material estimation from a single human image. We propose HumanMaterial: first, we introduce OpenHumanBRDF—a high-fidelity, open-source dataset—marking the first explicit modeling of displacement maps and subsurface scattering (SSS) in human inverse rendering; second, we design a progressive multi-stage training framework integrated with controllable physically based rendering (PBR) loss (CPR), jointly optimizing six physically grounded material maps—normal, albedo, roughness, specular, displacement, and SSS. To mitigate multi-task imbalance and underfitting, we incorporate physics-driven prior prediction and fine-tuning. Our method achieves significant improvements over state-of-the-art approaches on both OpenHumanBRDF and real-world images, particularly enhancing geometric fidelity and optical realism in skin regions, enabling high-quality, photorealistic relighting under arbitrary illumination.

Technology Category

Application Category

📝 Abstract
Full-body Human inverse rendering based on physically-based rendering aims to acquire high-quality materials, which helps achieve photo-realistic rendering under arbitrary illuminations. This task requires estimating multiple material maps and usually relies on the constraint of rendering result. The absence of constraints on the material maps makes inverse rendering an ill-posed task. Previous works alleviated this problem by building material dataset for training, but their simplified material data and rendering equation lead to rendering results with limited realism, especially that of skin. To further alleviate this problem, we construct a higher-quality dataset (OpenHumanBRDF) based on scanned real data and statistical material data. In addition to the normal, diffuse albedo, roughness, specular albedo, we produce displacement and subsurface scattering to enhance the realism of rendering results, especially for the skin. With the increase in prediction tasks for more materials, using an end-to-end model as in the previous work struggles to balance the importance among various material maps, and leads to model underfitting. Therefore, we design a model (HumanMaterial) with progressive training strategy to make full use of the supervision information of the material maps and improve the performance of material estimation. HumanMaterial first obtain the initial material results via three prior models, and then refine the results by a finetuning model. Prior models estimate different material maps, and each map has different significance for rendering results. Thus, we design a Controlled PBR Rendering (CPR) loss, which enhances the importance of the materials to be optimized during the training of prior models. Extensive experiments on OpenHumanBRDF dataset and real data demonstrate that our method achieves state-of-the-art performance.
Problem

Research questions and friction points this paper is trying to address.

Estimating high-quality human materials from single images
Overcoming ill-posed inverse rendering with realistic material constraints
Balancing multi-material prediction via progressive training strategy
Innovation

Methods, ideas, or system contributions that make the work stand out.

Constructed high-quality dataset OpenHumanBRDF for realistic rendering
Designed HumanMaterial model with progressive training strategy
Introduced Controlled PBR Rendering loss for material optimization
🔎 Similar Papers
No similar papers found.
Y
Yu Jiang
School of Computer Science, Wuhan University, Wuhan, China
Jiahao Xia
Jiahao Xia
Research Fellow, University of Technology Sydney
Deep Learning
J
Jiongming Qin
School of Computer Science, Wuhan University, Wuhan, China
Y
Yusen Wang
School of Computer Science, Wuhan University, Wuhan, China
T
Tuo Cao
School of Computer Science, Wuhan University, Wuhan, China
Chunxia Xiao
Chunxia Xiao
Professor of Computer Science, Wuhan University
Computer VisionComputer GraphicsMachine learning