3D-CariGAN: An End-to-End Solution to 3D Caricature Generation from Normal Face Photos

Last update: Oct 09, 2022

Related tags

Overview

3D-CariGAN: An End-to-End Solution to 3D Caricature Generation from Normal Face Photos

This repository contains the source code and dataset for the paper 3D-CariGAN: An End-to-End Solution to 3D Caricature Generation from Normal Face Photos by Zipeng Ye, Mengfei Xia, Yanan Sun, Ran Yi, Minjing Yu, Juyong Zhang, Yu-Kun Lai and Yong-Jin Liu, which is accepted by IEEE Transactions on Visualization and Computer Graphics (TVCG).

This repository contains two parts: dataset and source code.

2D and 3D Caricature Dataset

2D Caricature Dataset

We collect 5,343 hand-drawn portrait caricature images from Pinterest.com and WebCaricature dataset with facial landmarks extracted by a landmark detector, followed by human interaction for correction if needed.

The 2D dataset is in cari_2D_dataset.zip file.

3D Caricature Dataset

We use the method to generate 5,343 3D caricature meshes of the same topology. We align the pose of the generated 3D caricature meshes with the pose of a template 3D head using an ICP method, where we use 5 key landmarks in eyes, nose and mouth as the landmarks for ICP. We normalize the coordinates of the 3D caricature mesh vertices by translating the center of meshes to the origin and scaling them to the same size.

The 3D dataset is in cari_3D_dataset.zip file.

3DCariPCA

We use the 3D caricature dataset to build a PCA model. We use sklearn.decomposition.PCA to build 3DCariPCA. The PCA model is pca200_icp.model file. You could use joblib to load the model and use it.

Download

You can download the two datasets and PCA in google drive and BaiduYun (code: 3kz8).

Source Code

Running Environment

Ubuntu 16.04 + Python3.7

You can install the environment directly by using conda env create -f env.yml in conda.

Training

We use our 3D caricature dataset and CelebA-Mask-HQ dataset to train 3D-CariGAN. You could download CelebA-Mask-HQ dataset and then reconstruct their 3D normal heads of all images. The 3D normal heads are for calculating loss.

Inferring

The inferring code is cari_pipeline.py file in pipeline folder. You could train your model or use our pre-trained model.

The pipeline includes two optional sub-program eye_complete and color_complete, which are implemented by C++. You should compile them and then use them. The eye_complete is for completing the eye part of mesh and the color_complete is for texture completion.

Pre-trained Model

You can download pre-trained model latest.pth in google drive and BaiduYun (code: 3kz8). You should put it into ./checkpoints.

Additional notes

Please cite the following paper if the dataset and code help your research:

Citation:

@article{ye2021caricature,
 author = {Ye, Zipeng and Xia, Mengfei and Sun, Yanan and Yi, Ran and Yu, Minjing and Zhang, Juyong and Lai, Yu-Kun and Liu, Yong-Jin},
 title = {3D-CariGAN: An End-to-End Solution to 3D Caricature Generation from Normal Face Photos},
 journal = {IEEE Transactions on Visualization and Computer Graphics},
 year = {2021},
 doi={10.1109/TVCG.2021.3126659},
}

The paper will be published.

3D-CariGAN: An End-to-End Solution to 3D Caricature Generation from Normal Face Photos

Related tags

Overview

3D-CariGAN: An End-to-End Solution to 3D Caricature Generation from Normal Face Photos

2D and 3D Caricature Dataset

2D Caricature Dataset

3D Caricature Dataset

3DCariPCA

Download

Source Code

Running Environment

Training

Inferring

Pre-trained Model

Additional notes

Owner

TCTrack: Temporal Contexts for Aerial Tracking (CVPR2022)

A real world application of a Recurrent Neural Network on a binary classification of time series data

Nested Graph Neural Network (NGNN) is a general framework to improve a base GNN's expressive power and performance

An example showing how to use jax to train resnet50 on multi-node multi-GPU

Code for the preprint "Well-classified Examples are Underestimated in Classification with Deep Neural Networks"

Python package for covariance matrices manipulation and Biosignal classification with application in Brain Computer interface

Extracts data from the database for a graph-node and stores it in parquet files

a Pytorch easy re-implement of "YOLOX: Exceeding YOLO Series in 2021"

All the code and files related to the MI-Lab of UE19CS305 course in sem 5

A curated list of awesome papers for Semantic Retrieval (TOIS Accepted: Semantic Models for the First-stage Retrieval: A Comprehensive Review).

Conversational text Analysis using various NLP techniques

Benchmark library for high-dimensional HPO of black-box models based on Weighted Lasso regression

[ACM MM 2021] Joint Implicit Image Function for Guided Depth Super-Resolution

An open source machine learning library for performing regression tasks using RVM technique.

[ICCV'21] Neural Radiance Flow for 4D View Synthesis and Video Processing

A Free and Open Source Python Library for Multiobjective Optimization

The implementation of 'Image synthesis via semantic composition'.

Benchmarks for Model-Based Optimization

PyTorch inference for "Progressive Growing of GANs" with CelebA snapshot

A rough implementation of the paper "A Steering Algorithm for Redirected Walking Using Reinforcement Learning"