Faces and Expressions

Institute Homepage

Institute Homepage Sign In

Back

Research Overview

Intrinsically Motivated Learning

Regularity as Intrinsic Reward for Free Play

SENSEI: Semantic Exploration Guided by Foundation Models to Learn Versatile World Models

Curious Exploration via Structured World Models Yields Zero-Shot Object Manipulation

Learning with Muscles

Natural and Robust Walking from Generic Rewards

The effect of muscles in Learning Behavior

Scaling RL to Large Musculoskeletal Systems

Reinforcement Learning for Diverse Solutions

Offline Diversity Under Imitation Constraints

Learning Diverse Skills for Local Navigation

Learning Agile Skills via Adversarial Imitation of Rough Partial Demonstrations

Reinforcement Learning and Control

Model-based Reinforcement Learning and Planning

Object-centric Self-supervised Reinforcement Learning

Self-exploration of Behavior

Causal Reasoning in RL

Equation Learner for Extrapolation and Control

Intrinsically Motivated Hierarchical Learner

Regularity as Intrinsic Reward for Free Play

Curious Exploration via Structured World Models Yields Zero-Shot Object Manipulation

Natural and Robust Walking from Generic Rewards

Goal-conditioned Offline Planning

Offline Diversity Under Imitation Constraints

Learning Diverse Skills for Local Navigation

Learning Agile Skills via Adversarial Imitation of Rough Partial Demonstrations

Deep Learning

Combinatorial Optimization as a Layer / Blackbox Differentiation

Object-centric Self-supervised Reinforcement Learning

Symbolic Regression and Equation Learning

Representation Learning

Stepsize adaptation for stochastic optimization

Probabilistic Neural Networks

Learning with 3D rotations: A hitchhiker’s guide to SO(3)

Haptic Sensing

Super-resolution Sensing for Haptics

Insight: a Haptic Sensor Powered by Vision and Machine Learning

Minsight: Learning-based tactile sensing for robotics

ML for Science

Predicting brain activity (fMRI)

Equation Learning for Statistical Physics

Machine Learning for Understanding Quantum Systems

Symbolic Regression and Equation Learning

Previous Research Projects

The Playful Machine

Robust and Affordable Haptic Sensation with Sparse Sensor Configuration

Perzeptive Systeme Members Publications

Faces and Expressions

Sab face and expression — Analysis and Synthesis of 3D Faces: Top left: FLAME [] captures 3D face shape, pose, and expression shape variations. Top right: DECA [] reconstructs animatable 3D faces from 2D images with FLAME’s parameter control. Bottom left: ToFu [] reconstructs high-fidelity 3D faces from calibrated multi-view images. Bottom right: GIF [] generates realistic face images with FLAME’s parameter control.

Facial shape and motion are essential to communication. They are also fundamentally three dimensional. Consequently, we need a 3D model of the face that can capture the full range of face shapes and expressions. Such a model should be realistic, easy to animate, and easy to fit to data. See [] for a comprehensive overview of different facial representations.

To that end, we train an expressive 3D head model called FLAME from over 33,000 3D scans. Because it is learned from large-scale, expressive data, it is more realistic than previous models. To capture non-linear expression shape variations, we introduce CoMA [], a versatile autoencoder framework for meshes with hierarchical mesh up- and down-sampling operations. Models like FLAME and CoMA require large datasets of 3D faces in dense semantic correspondence across different identities and expressions. ToFu [], a geometry inference framework that facilitates a hierarchical volumetric feature aggregation scheme, predicts facial meshes in a consistent mesh topology directly from calibrated multi-view images three orders of magnitude faster than traditional techniques.

To capture, model, and understand facial expressions, we need to estimate the parameters of our face models from images and videos. Training a neural network to regress model parameters from image pixels is difficult because we lack paired training data of images and the true 3D face shape. To address this, RingNet [] directly learns this mapping using only 2D image features. DECA [] additionally learns an animatable detailed displacement model from in-the-wild images. This enables important applications such as creation of animatable avatars from a single image. Our NoW benchmark enables the field to quantitatively compare such methods for the first time.

Classical rendering methods can be used to generate images using FLAME but these look unrealistic due to the lack of hair, eyes, and the mouth cavity (i.e., teeth or tongue). To address this, we are developing new neural rendering methods. GIF [] combines a generative adversarial network (GAN) with FLAME’s parameter control to generate realistic looking face images.

Members

Perzeptive Systeme

Timo Bolkart

Research Scientist

Guest Scientist

Perzeptive Systeme

Anurag Ranjan

Doctoral Researcher

Affiliated Researcher

Guest Scientist

Perzeptive Systeme

Haiwen Feng

Doctoral Researcher

Empirische Inferenz

Partha Ghosh

Postdoctoral Researcher

Publications

Perceiving Systems Conference Paper Topologically Consistent Multi-View Face Inference Using Volumetric Sampling Li, T., Liu, S., Bolkart, T., Liu, J., Li, H., Zhao, Y. In Proc. International Conference on Computer Vision (ICCV), :3804-3814, IEEE, Piscataway, NJ, International Conference on Computer Vision, October 2021 (Published) project paper DOI BibTeX

Perceiving Systems Article Learning an Animatable Detailed 3D Face Model from In-the-Wild Images Feng, Y., Feng, H., Black, M. J., Bolkart, T. ACM Transactions on Graphics, 40(4):88:1-88:13, August 2021 (Published) pdf Sup Mat code video talk DOI BibTeX

Perceiving Systems Conference Paper GIF: Generative Interpretable Faces Ghosh, P., Gupta, P. S., Uziel, R., Ranjan, A., Black, M. J., Bolkart, T. In 2020 International Conference on 3D Vision (3DV 2020), 1:868-878, IEEE, Piscataway, NJ, International Conference on 3D Vision (3DV 2020), November 2020 (Published) pdf project code video DOI BibTeX

Perceiving Systems Article 3D Morphable Face Models - Past, Present and Future Egger, B., Smith, W. A. P., Tewari, A., Wuhrer, S., Zollhoefer, M., Beeler, T., Bernard, F., Bolkart, T., Kortylewski, A., Romdhani, S., et al. ACM Transactions on Graphics, 39(5):157, October 2020 (Published) project page pdf preprint DOI BibTeX

Perceiving Systems Conference Paper Capture, Learning, and Synthesis of 3D Speaking Styles Cudeiro, D., Bolkart, T., Laidlaw, C., Ranjan, A., Black, M. J. In Proceedings IEEE Conf. on Computer Vision and Pattern Recognition (CVPR), :10101-10111, IEEE International Conference on Computer Vision and Pattern Recognition (CVPR), June 2019 () code Project Page video paper BibTeX

Perceiving Systems Conference Paper Learning to Regress 3D Face Shape and Expression from an Image without 3D Supervision Sanyal, S., Bolkart, T., Feng, H., Black, M. J. In Proceedings IEEE Conf. on Computer Vision and Pattern Recognition (CVPR), :7763-7772, IEEE International Conference on Computer Vision and Pattern Recognition (CVPR), June 2019 () code pdf preprint URL BibTeX

Perceiving Systems Conference Paper Generating 3D Faces using Convolutional Mesh Autoencoders Ranjan, A., Bolkart, T., Sanyal, S., Black, M. J. In European Conference on Computer Vision (ECCV), Lecture Notes in Computer Science, vol 11207:725-741, Springer, Cham, September 2018 () Code (tensorflow) Code (pytorch) Project Page paper supplementary DOI BibTeX

Perceiving Systems Article Learning a model of facial shape and expression from 4D scans Li, T., Bolkart, T., Black, M. J., Li, H., Romero, J. ACM Transactions on Graphics, 36(6):194:1-194:17, November 2017, Two first authors contributed equally () data/model video code chumpy code tensorflow paper supplemental BibTeX