I research computer vision systems for cameras and sensors that need to operate in the real world. My work spans computational photography, synthetic data, and vision-language models for specialized image domains.
At Sony, I train vision and multimodal models for semiconductor imaging and yield optimization, combining sensor knowledge with machine learning and deploying agents that support high-volume manufacturing workflows.
I am deeply interested in
- World models and sensing How richer learned representations can reshape cameras, sensors, and physical perception.
- Robotics and embodied AI How world models, visual reasoning, and memory can help robots understand and operate in real environments.
- Creative computational photography How optics, sensors, algorithms, and learning can be combined to build new ways of seeing.
- VLM reasoning for specialized domains How vision-language models can be adapted to reason over technical visual data, from sensor evaluation to manufacturing.
work
2024 – Present
Sony
Senior AI Researcher · Pentas Vision (SSS)
I lead a small team working on computer vision across sensors, manufacturing, and image quality. Lately I've been building an agentic VLM platform for visual reasoning, post-training the models and figuring out how to evaluate and reward them well. Before that, synthetic data for edge vision models and optimization systems for defect detection.
2021 – 2023
SenseBrain · SenseTime Research
Research Engineer · Computational photography & imaging
Novel sensors, ML-based artifact removal and image enhancement, and generative models for mobile imaging systems. Exploring the limits of smartphone cameras.
2016 – 2020
Northwestern University
B.S & M.S in Computer science · computational photography lab
My M.S. research was with Northwestern’s Computational Photography Lab, advised by Ollie Cossairt and Aggelos Katsaggelos, where I worked on event-camera 3D reconstruction. I also spent a lot of time at The Garage, Northwestern’s startup incubator, as both a builder and mentor.