Video-driven Neural Physically-based Facial Asset for Production
2022-12-01
发表期刊ACM TRANSACTIONS ON GRAPHICS
ISSN0730-0301
EISSN1557-7368
卷号41期号:6
发表状态已发表
DOI10.1145/3550454.3555445
摘要Production-level workflows for producing convincing 3D dynamic human faces have long relied on an assortment of labor-intensive tools for geometry and texture generation, motion capture and rigging, and expression synthesis. Recent neural approaches automate individual components but the corresponding latent representations cannot provide artists with explicit controls as in conventional tools. In this paper, we present a new learning-based, video-driven approach for generating dynamic facial geometries with high-quality physically-based assets. For data collection, we construct a hybrid multiview-photometric capture stage, coupling with ultra-fast video cameras to obtain raw 3D facial assets. We then set out to model the facial expression, geometry and physically-based textures using separate VAEs where we impose a global MLP based expression mapping across the latent spaces of respective networks, to preserve characteristics across respective attributes. We also model the delta information as wrinkle maps for the physically-based textures, achieving high-quality 4K dynamic textures. We demonstrate our approach in high-fidelity performer-specific facial capture and cross-identity facial motion retargeting. In addition, our multi-VAE-based neural asset, along with the fast adaptation schemes, can also be deployed to handle in-the-wild videos. Besides, we motivate the utility of our explicit facial disentangling strategy by providing various promising physically-based editing results with high realism. Comprehensive experiments show that our technique provides higher accuracy and visual fidelity than previous video-driven facial reconstruction and animation methods.
关键词Physically-Based Face Rendering Facial Modeling Digital Human Video-Driven Animation
URL查看原文
收录类别SCI ; EI ; CPCI-S
语种英语
资助项目Shanghai YangFan Program[21YF1429500] ; Shanghai Local College Capacity Building Program[22010502800] ; NSFC programs["61976138","61977047"] ; National Key Research and Development Program[2018YFB2100500] ; STCSM[2015F0203000-06] ; SHMEC[2019-01-07-00-01-E00003]
WOS研究方向Computer Science
WOS类目Computer Science, Software Engineering
WOS记录号WOS:000891651900028
出版者ASSOC COMPUTING MACHINERY
引用统计
文献类型期刊论文
条目标识符https://kms.shanghaitech.edu.cn/handle/2MSLDSTB/266605
专题信息科学与技术学院_本科生
信息科学与技术学院_PI研究组_虞晶怡组
信息科学与技术学院_硕士生
信息科学与技术学院_博士生
信息科学与技术学院_PI研究组_许岚组
通讯作者Xu, Lan; Yu, Jingyi
作者单位
1.ShanghaiTech Univ, Shanghai, Peoples R China
2.Deemos Technol Co Ltd, Shanghai, Peoples R China
3.Huazhong Univ Sci & Technol, Wuhan, Peoples R China
4.Shanghai Engn Res Ctr Intelligent Vis & Imaging, Shanghai, Peoples R China
第一作者单位上海科技大学
通讯作者单位上海科技大学
第一作者的第一单位上海科技大学
推荐引用方式
GB/T 7714
Zhang, Longwen,Zeng, Chuxiao,Zhang, Qixuan,et al. Video-driven Neural Physically-based Facial Asset for Production[J]. ACM TRANSACTIONS ON GRAPHICS,2022,41(6).
APA Zhang, Longwen.,Zeng, Chuxiao.,Zhang, Qixuan.,Lin, Hongyang.,Cao, Ruixiang.,...&Yu, Jingyi.(2022).Video-driven Neural Physically-based Facial Asset for Production.ACM TRANSACTIONS ON GRAPHICS,41(6).
MLA Zhang, Longwen,et al."Video-driven Neural Physically-based Facial Asset for Production".ACM TRANSACTIONS ON GRAPHICS 41.6(2022).
条目包含的文件 下载所有文件
文件名称/大小 文献类型 版本类型 开放类型 使用许可
个性服务
查看访问统计
谷歌学术
谷歌学术中相似的文章
[Zhang, Longwen]的文章
[Zeng, Chuxiao]的文章
[Zhang, Qixuan]的文章
百度学术
百度学术中相似的文章
[Zhang, Longwen]的文章
[Zeng, Chuxiao]的文章
[Zhang, Qixuan]的文章
必应学术
必应学术中相似的文章
[Zhang, Longwen]的文章
[Zeng, Chuxiao]的文章
[Zhang, Qixuan]的文章
相关权益政策
暂无数据
收藏/分享
文件名: 10.1145@3550454.3555445.pdf
格式: Adobe PDF
此文件暂不支持浏览
所有评论 (0)
暂无评论
 

除非特别说明,本系统中所有内容都受版权保护,并保留所有权利。