Single-view 3D Reconstruction Computational Photography Derek Hoiem, University of Illinois 10/18/12 Some slides from Alyosha Efros, Steve Seitz.

Slides:



Advertisements
Similar presentations
Single-view Metrology and Camera Calibration
Advertisements

3D reconstruction.
CS 691 Computational Photography Instructor: Gianfranco Doretto 3D to 2D Projections.
Paris town hall.
Photo Stitching Panoramas from Multiple Images Computer Vision CS 543 / ECE 549 University of Illinois Derek Hoiem 02/28/12.
Lecture 8: Stereo.
Camera calibration and epipolar geometry
Structure from motion.
Photo Stitching Panoramas from Multiple Images Computer Vision CS 543 / ECE 549 University of Illinois Derek Hoiem 03/3/15.
Single-view metrology
Lecture 19: Single-view modeling CS6670: Computer Vision Noah Snavely.
Homographies and Mosaics : Computational Photography Alexei Efros, CMU, Fall 2005 © Jeffrey Martin (jeffrey-martin.com) with a lot of slides stolen.
Epipolar geometry. (i)Correspondence geometry: Given an image point x in the first view, how does this constrain the position of the corresponding point.
Structure from motion. Multiple-view geometry questions Scene geometry (structure): Given 2D point matches in two or more images, where are the corresponding.
Uncalibrated Geometry & Stratification Sastry and Yang
Scene Modeling for a Single View : Computational Photography Alexei Efros, CMU, Fall 2005 René MAGRITTE Portrait d'Edward James …with a lot of slides.
Announcements Project 2 questions Midterm out on Thursday
Lecture 7: Image Alignment and Panoramas CS6670: Computer Vision Noah Snavely What’s inside your fridge?
CS485/685 Computer Vision Prof. George Bebis
Projective geometry Slides from Steve Seitz and Daniel DeMenthon
Panorama signups Panorama project issues? Nalwa handout Announcements.
1Jana Kosecka, CS 223b Cylindrical panoramas Cylindrical panoramas with some slides from R. Szeliski, S. Seitz, D. Lowe, A. Efros,
Previously Two view geometry: epipolar geometry Stereo vision: 3D reconstruction epipolar lines Baseline O O’ epipolar plane.
Automatic Photo Popup Derek Hoiem Alexei A. Efros Martial Hebert Carnegie Mellon University.
Single-view Metrology and Camera Calibration Computer Vision Derek Hoiem, University of Illinois 02/26/15 1.
CSE 576, Spring 2008Projective Geometry1 Lecture slides by Steve Seitz (mostly) Lecture presented by Rick Szeliski.
Camera parameters Extrinisic parameters define location and orientation of camera reference frame with respect to world frame Intrinsic parameters define.
Scene Modeling for a Single View : Computational Photography Alexei Efros, CMU, Spring 2010 René MAGRITTE Portrait d'Edward James …with a lot of.
Cameras, lenses, and calibration
Image Warping Computational Photography Derek Hoiem, University of Illinois 09/27/11 Many slides from Alyosha Efros + Steve Seitz Photo by Sean Carroll.
3-D Scene u u’u’ Study the mathematical relations between corresponding image points. “Corresponding” means originated from the same 3D point. Objective.
3D Scene Models Object recognition and scene understanding Krista Ehinger.
Single-view Metrology and Cameras Computational Photography Derek Hoiem, University of Illinois 10/8/13.
Single-view metrology
Projective Geometry and Single View Modeling CSE 455, Winter 2010 January 29, 2010 Ames Room.
Single-view modeling, Part 2 CS4670/5670: Computer Vision Noah Snavely.
Image Stitching Ali Farhadi CSE 455
CSC 589 Lecture 22 Image Alignment and least square methods Bei Xiao American University April 13.
Image-based modeling Digital Visual Effects, Spring 2008 Yung-Yu Chuang 2008/5/6 with slides by Richard Szeliski, Steve Seitz and Alexei Efros.
02/26/02 (c) 2002 University of Wisconsin, CS 559 Last Time Canonical view pipeline Orthographic projection –There was an error in the matrix for taking.
Single-view Metrology and Camera Calibration Computer Vision Derek Hoiem, University of Illinois 01/25/11 1.
Image Warping Computational Photography Derek Hoiem, University of Illinois 09/23/10 Many slides from Alyosha Efros + Steve Seitz Photo by Sean Carroll.
Single View Geometry Course web page: vision.cis.udel.edu/cv April 9, 2003  Lecture 20.
Single-view 3D Reconstruction Computational Photography Derek Hoiem, University of Illinois 10/12/10 Some slides from Alyosha Efros, Steve Seitz.
Project 2 –contact your partner to coordinate ASAP –more work than the last project (> 1 week) –you should have signed up for panorama kitssigned up –sign.
Computational Photography Derek Hoiem, University of Illinois
Single-view Metrology and Cameras Computational Photography Derek Hoiem, University of Illinois 10/06/15.
Basic Perspective Projection Watt Section 5.2, some typos Define a focal distance, d, and shift the origin to be at that distance (note d is negative)
Single-view 3D Reconstruction
stereo Outline : Remind class of 3d geometry Introduction
More with Pinhole + Single-view Metrology
12/15/ Fall 2006 Two-Point Perspective Single View Geometry Final Project Faustinus Kevin Gozali an extended tour.
Image-Based Rendering Geometry and light interaction may be difficult and expensive to model –Think of how hard radiosity is –Imagine the complexity of.
Image Warping Many slides from Alyosha Efros + Steve Seitz + Derek oeim Photo by Sean Carroll.
Project 2 went out on Tuesday You should have a partner You should have signed up for a panorama kit Announcements.
Homographies and Mosaics : Computational Photography Alexei Efros, CMU, Fall 2006 © Jeffrey Martin (jeffrey-martin.com) with a lot of slides stolen.
55:148 Digital Image Processing Chapter 11 3D Vision, Geometry
Single-view metrology
Prof. Adriana Kovashka University of Pittsburgh October 3, 2016
CSCE 441 Computer Graphics 3-D Viewing
Scene Modeling for a Single View
Two-view geometry Computer Vision Spring 2018, Lecture 10
Epipolar geometry.
Projective geometry Readings
Viewing (Projections)
Viewing (Projections)
Announcements list – let me know if you have NOT been getting mail
Filtering An image as a function Digital vs. continuous images
Announcements Project 2 extension: Friday, Feb 8
Presentation transcript:

Single-view 3D Reconstruction Computational Photography Derek Hoiem, University of Illinois 10/18/12 Some slides from Alyosha Efros, Steve Seitz

Take-home question Suppose you have estimated three vanishing points corresponding to orthogonal directions. How can you recover the rotation matrix that is aligned with the 3D axes defined by these points? – Assume that intrinsic matrix K has three parameters – Remember, in homogeneous coordinates, we can write a 3d point at infinity as (X, Y, Z, 0) Photo from online Tate collection VP x VP z. VP y

Take-home question Assume that the camera height is 5 ft. – What is the height of the man? – What is the height of the building?

Difficulty in macro (close-up) photography For close objects, we have a small relative DOF Can only shrink aperture so far How to get both bugs in focus?

Solution: Focus stacking 1.Take pictures with varying focal length Example from

Solution: Focus stacking 1.Take pictures with varying focal length 2.Combine

Focus stacking

Focus stacking How to combine? Web answer: With software (Photoshop, CombineZM) How to do it automatically?

Focus stacking How to combine? 1.Align images (e.g., using corresponding points) 2.Two ideas a)Mask regions by hand and combine with pyramid blend b)Gradient domain fusion (mixed gradient) without masking y/Workflow.htm#Focus%20Stacking Automatic solution would make a very interesting final project school.com/an-introduction-to-focus- stacking Recommended Reading:

Relation between field of view and focal length Field of view (angle width) Film/Sensor Width Focal length

Dolly Zoom or “Vertigo Effect” Zoom in while moving away How is this done?

Dolly zoom (or “Vertigo effect”) Distance between object and camera width of object Field of view (angle width) Film/Sensor Width Focal length

Today’s class: 3D Reconstruction

The challenge One 2D image could be generated by an infinite number of 3D geometries ? ? ?

The solution Make simplifying assumptions about 3D geometry Unlikely Likely

Today’s class: Two Models Box + frontal billboards Ground plane + non-frontal billboards

“Tour into the Picture” (Horry et al. SIGGRAPH ’97) Create a 3D “theatre stage” of five billboards Specify foreground objects through bounding polygons Use camera transformations to navigate through the scene Following slides modified from Efros

The idea Many scenes can be represented as an axis-aligned box volume (i.e. a stage) Key assumptions All walls are orthogonal Camera view plane is parallel to back of volume How many vanishing points does the box have? Three, but two at infinity Single-point perspective Can use the vanishing point to fit the box to the particular scene

Fitting the box volume User controls the inner box and the vanishing point placement (# of DOF???) Q: What’s the significance of the vanishing point location? A: It’s at eye (camera) level: ray from COP to VP is perpendicular to image plane – Under single-point perspective assumptions, the VP should be the principal point of the image

High Camera Example of user input: vanishing point and back face of view volume are defined

High Camera Example of user input: vanishing point and back face of view volume are defined

Low Camera

Example of user input: vanishing point and back face of view volume are defined

High CameraLow Camera Comparison of how image is subdivided based on two different camera positions. You should see how moving the box corresponds to moving the eyepoint in the 3D world.

Left Camera Another example of user input: vanishing point and back face of view volume are defined

Left Camera Another example of user input: vanishing point and back face of view volume are defined

Right Camera Another example of user input: vanishing point and back face of view volume are defined

Right Camera Another example of user input: vanishing point and back face of view volume are defined

Left CameraRight Camera Comparison of two camera placements – left and right. Corresponding subdivisions match view you would see if you looked down a hallway.

Question Think about the camera center and image plane… – What happens when we move the box? – What happens when we move the vanishing point?

2D to 3D conversion First, we can get ratios leftright top bottom vanishing point back plane

Size of user-defined back plane determines width/height throughout box (orthogonal sides) Use top versus side ratio to determine relative height and width dimensions of box Left/right and top/bot ratios determine part of 3D camera placement leftright top bottom camera pos 2D to 3D conversion

Depth of the box Can compute by similar triangles (CVA vs. CV’A’) Need to know focal length f (or FOV) Note: can compute position on any object on the ground – Simple unprojection – What about things off the ground?

Homography A B C D A’ B’ C’D’ 2d coordinates 3d plane coordinates

Image rectification To unwarp (rectify) an image solve for homography H given p and p’: wp’=Hp p p’

Computing homography Assume we have four matched points: How do we compute homography H? Direct Linear Transformation (DLT)

Computing homography Direct Linear Transform Apply SVD: UDV T = A h = V smallest (column of V corr. to smallest singular value) Matlab [U, S, V] = svd(A); h = V(:, end); Explanations of SVD and solving homogeneous linear systemsSVDsolving homogeneous linear systems

Solving for homographies

Ah0 Defines a least squares problem: 2n × 9 92n Since h is only defined up to scale, solve for unit vector ĥ Solution: ĥ = eigenvector of A T A with smallest eigenvalue Works with 4 or more points

Tour into the picture algorithm 1.Set the box corners

Tour into the picture algorithm 1.Set the box corners 2.Set the VP 3.Get 3D coordinates – Compute height, width, and depth of box 4.Get texture maps – homographies for each face x

Result Render from new views

Foreground Objects Use separate billboard for each For this to work, three separate images used: – Original image. – Mask to isolate desired foreground images. – Background with objects removed

Foreground Objects Add vertical rectangles for each foreground object Can compute 3D coordinates P0, P1 since they are on known plane. P2, P3 can be computed as before (similar triangles)

Foreground Result Video from CMU class: mGwcuM mGwcuM

Automatic Photo Pop-up Input Ground Vertical Sky Geometric LabelsCut’n’Fold3D Model Image Learned Models Hoiem et al. 2005

Cutting and Folding Fit ground-vertical boundary – Iterative Hough transform

Cutting and Folding Form polylines from boundary segments – Join segments that intersect at slight angles – Remove small overlapping polylines Estimate horizon position from perspective cues

Cutting and Folding ``Fold’’ along polylines and at corners ``Cut’’ at ends of polylines and along vertical-sky boundary

Cutting and Folding Construct 3D model Texture map

Results Automatic Photo Pop-up Input Image Cut and Fold

Results Automatic Photo Pop-up Input Image

Comparison with Manual Method Input Image Automatic Photo Pop-up (15 sec)! [Liebowitz et al. 1999]

Failures Labeling Errors

Failures Foreground Objects

Adding Foreground Labels Recovered Surface Labels + Ground-Vertical Boundary Fit Object Boundaries + Horizon

Final project ideas If a one-person project: – Interactive program to make 3D model from an image (e.g., output in VRML, or draw path for animation) If a two-person team, 2 nd person: – Add tools for cutting out foreground objects and automatic hole-filling

Summary 2D  3D is mathematically impossible Need right assumptions about the world geometry Important tools – Vanishing points – Camera matrix – Homography

Next Week Monday: Project 3 is due – Office hours at 3pm today, 10am Monday Next Tuesday: start of new section – Finding correspondences automatically – Image stitching – Various forms of recognition