New: I am co-organizing Three questions about virtual try-on @ CVPR 2025 with Aayush Bansal. I am the Head of Engineering at SpreeAI. Within my 2.5 years tenure, I transformed the SpreeAI from a struggling $12M valuation company with poor product-market-fit to a $1.5B valuation AI e-commerce company with my virtual tryon brain child. Beyond overseeing all R&D, my greatest pride lies in building and leading a world-class team of passionate machine learning researchers and engineers.Previously, I was a Research Scientist at Meta Reality Labs Research, where I tech-led a group of researchers to develop 3D perception and human sensing algorithms for Meta Aria glasses. Before that, I was a Ph.D. student at The Robotics Institute, Carnegie Mellon University where I worked with Prof. Srinivasa Narasimhan and Prof. Yaser Sheikh on novel methods to capture dense and accurate 3D shape of human bodies. I also worked with Prof. Zhaoyang Wang at the Catholic University of America, where I got B.E. degree in Electrical Engineering, on camera calibration, structured light system, and tracking algorithms. I am also working part-time with a consulting company, LeCON LLC, to help startups and companies to further develop capabilities in large vision language model (VLM), including data curation, model training, and deployment, and various core components for intelligent systems such as mapping, perception, and prediction. Reach out to us if you are interested in collaborating. ResearchI am very interested in various aspects of 3D vision, physics-based vision, and generative models for photorealistic digitial avatar creation and human scene understanding. The goal is to develop holistic and end-to-end machine learning systems that understand and recreate virtual environments that are perceptually indistinguishable from reality. Jobs opportuninty: I am hiring full time CV&ML&Graphics researchers. I strike to balance between core and applied research with patents, papers, and product as outputs. Send me an email if you are interested in working with me. Award
Patent
Publication
|
|
ODAM: Object Detection, Association, and Mapping using Posed RGB Video
|
|
ContactOpt: Optimizing Contact to Improve Grasps
|
|
ANR: Articulated Neural Rendering for Virtual Avatars
|
|
TexMesh: Reconstructing Detailed Human Texture and Geometry from Monocular Video
|
|
Long-term Human Motion Prediction with Scene Context
|
|
4D Visualization of Dynamic Events from Unconstrained Multi-View Videos
|
|
Spatiotemporal Bundle Adjustment for Dynamic 3D Human Reconstruction in the Wild
|
|
Self-supervised Multi-view Person Association and Its Applications
|
|
|
Occlusion-Net: 2D/3D Occluded Keypoint Localization Using Graph Networks
|
|
CarFusion: Combining Part Detection and Point Tracking for Dynamic 3D Reconstruction of
Vehicles
|
|
Texture Illumination Separation for Single-shot Structured Light Reconstruction
|
|
Passive Tomography of Turbulance Strength
|
|
Automated fast initial guess in digital image correlation
|
|
Hyper-accurate flexible calibration technique for fringe-projection-based
three-dimensional imaging
|
|
Three-dimensional phantoms for curvature correction in spatial frequency domain
imaging
|
|
Advanced geometric camera calibration for machine vision
|
|
|
Accuracy enhancement of digital image correlation with B-spline interpolation
|
|
Phase extraction from optical interferograms in presence of intensity nonlinearity and
arbitrary phase shifts
|
|
Flexible calibration technique for fringe-projection-based three-dimensional
imaging
|
|
Exploiting Point Motion, Shape Deformation, and Semantic Priors for Dynamic 3D
Reconstruction in the Wild
|