Han Lin
555
posts
PhD
@UNC| MS
@Columbia| Research Intern
@AIatMeta| Working on generative models, multimodal learning, and LLMs
Chapel Hill, NC
Joined December 2021
Glad to share that V-Co is accepted to #ECCV2026! 🇸🇪 ✨ V-Co introduces a practical representation learning recipe for image generation by jointly denoising pixels and pretrained semantic features (e.g., DINOv2). Its design combines a fully dual-stream JiT architecture,
Spatial reasoning is not just about getting the answer right — it is also about knowing when the image does not provide enough evidence to answer, and what additional viewpoint is needed to resolve the uncertainty. ✨Excited to share SpatialUncertain, a controlled framework for
🌟Glad to introduce PhyMotion: a structured 3D motion reward for physics-grounded human video generation. Realistic human motion remains a major challenge in video generation. Existing rewards often stay in 2D pixel space, missing failures such as floating feet,




