InverseRenderNet: Learning single image inverse rendering, CVPR 2019.

Last update: Dec 20, 2022

Related tags

Overview

InverseRenderNet: Learning single image inverse rendering

!! Check out our new work InverseRenderNet++ paper and code, which improves the inverse rendering results and shadow handling.

This is the implementation of the paper "InverseRenderNet: Learning single image inverse rendering". The model is implemented in tensorflow.

If you use our code, please cite the following paper:

@inproceedings{yu19inverserendernet,
    title={InverseRenderNet: Learning single image inverse rendering},
    author={Yu, Ye and Smith, William AP},
    booktitle={Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)},
    year={2019}
}

Evaluation

Dependencies

To run our evaluation code, please create your environment based on following dependencies:

tensorflow 1.12.0
python 3.6
skimage
cv2
numpy

Pretrained model

Download our pretrained model from: Link
Unzip the downloaded file
Make sure the model files are placed in a folder named "irn_model"

Test on demo image

You can perform inverse rendering on random RGB image by our pretrained model. To run the demo code, you need to specify the path to pretrained model, path to RGB image and corresponding mask which masked out sky in the image. The mask can be generated by PSPNet, which you can find on https://github.com/hszhao/PSPNet. Finally inverse rendering results will be saved to the output folder named by your argument.

python3 test_demo.py --model /PATH/TO/irn_model --image demo.jpg --mask demo_mask.jpg --output test_results

Test on IIW

IIW dataset should be downloaded firstly from http://opensurfaces.cs.cornell.edu/publications/intrinsic/#download
Run testing code where you need to specify the path to model and IIW data:

python3 test_iiw.py --model /PATH/TO/irn_model --iiw /PATH/TO/iiw-dataset

Training

Train from scratch

The training for InverseRenderNet contains two stages: pre-train and self-train.

To begin with pre-train stage, you need to use training command specifying option -m to pre-train.
After finishing pre-train stage, you can run self-train by specifying option -m to self-train.

In addition, you can control the size of batch in training, and the path to training data should be specified.

An example for training command:

python3 train.py -n 2 -p Data -m pre-train

Data for training

To directly use our code for training, you need to pre-process the training data to match the data format as shown in examples in Data folder.

In particular, we pre-process the data before training, such that five images with great overlaps are bundled up into one mini-batch, and images are resized and cropped to a shape of 200 * 200 pixels. Along with input images associated depth maps, camera parameters, sky masks and normal maps are stored in the same mini-batch. For efficiency, every mini-batch containing all training elements for 5 involved images are saved as a pickle file. While training the data feeding thread directly load each mini-batch from corresponding pickle file.

InverseRenderNet: Learning single image inverse rendering, CVPR 2019.

Related tags

Overview

InverseRenderNet: Learning single image inverse rendering

Evaluation

Dependencies

Pretrained model

Test on demo image

Test on IIW

Training

Train from scratch

Data for training

Owner

Ye Yu

computer vision, image processing and machine learning on the web browser or node.

This is used to convert a string to an Image with Handwritten Characters.

A webcam-based 3x3x3 rubik's cube solver written in Python 3 and OpenCV.

ARU-Net - Deep Learning Chinese Word Segment

In this project we will be using the live feed coming from the webcam to create a virtual mouse with complete functionalities.

Computer vision applications project (Flask and OpenCV)

Deskewing images with slanted content

An Implementation of the seglink alogrithm in paper Detecting Oriented Text in Natural Images by Linking Segments

An expandable and scalable OCR pipeline

LEARN OPENCV IN 3 HOURS USING PYTHON - INCLUDING EXAMPLE PROJECTS

Discord QR Scam Code Generator + Token grab mobile device.

scene-linear test images

Text Detection from images using OpenCV

Automatically download multiple papers by keywords in CVPR

Sign Language Recognition service utilizing a deep learning model with Long Short-Term Memory to perform sign language recognition.

Automatic Number Plate Recognition (ANPR) is a highly accurate system capable of reading vehicle number plates without human intervention

python ocr using tesseract/ with EAST opencv detector

WACV 2022 Paper - Is An Image Worth Five Sentences? A New Look into Semantics for Image-Text Matching

document image degradation