Chi-Pin Huang

Chi-Pin Huang is a Research Scientist at NVIDIA Research Taiwan. His research focuses on Vision-Language Generative Models and Vision-Language-Action Models (VLAs), with particular interest in bridging perception, generation, and decision-making. He received his Ph.D. degree from National Taiwan University in 2026 under the supervision of Prof. Yu-Chiang Frank Wang, and earned his B.S.

Sameer Dharur

Sameer Dharur is a research scientist on the Cosmos team at NVIDIA, helping to build vision-language-models (VLMs) that reason better about the world. Prior to that, he spent ~4.5 years as a researcher and engineer at Apple specializing in computer vision and natural language processing to solve problems in image and video understanding, question answering, and robotics.

Wei-Cheng Tseng

Wei-Cheng Tseng is a research scientist at NVIDIA Research. He is also Ph.D. student in University of Toronto. His research interests are computer vision, generative AI applications in physical AI. He received his M.S. and B.S. in Electrical Engineering from National Tsing Hua University.

Website: https://weichengtseng.github.io/