Postdoctoral Researcher · University of Pennsylvania
My research studies multimodal AI, from vision-language foundation models to agents that reason and act. My current focus is multimodal AI for healthcare, including vision-language pretraining for 3D medical images, multimodal LLMs, and evidence-grounded clinical agents. This direction builds on my Ph.D. research in embodied and autonomous systems, where I studied multi-sensor perception, vision-language reasoning, planning, and reinforcement learning. I remain involved in collaborative work on embodied AI, including vision-language-action learning and language-guided robotics.
I received my Ph.D. in Computer Technology from the University of Science and Technology of China (USTC), where I worked in the LINKE Lab with Prof. Yanyong Zhang and Assoc. Prof. Jianmin Ji. I earned my B.E. in Applied Physics from Anhui University of Science and Technology.
My current work has two primary directions in healthcare, alongside continuing collaborations in embodied AI:
Medical vision-language learning · 3D imaging, anatomy-grounded pretraining, clinical report generation
Clinical reasoning and agents · multimodal LLMs, diagnostic QA, traceable evidence, tool use and verification
Embodied AI collaborations · vision-language-action learning and language-guided robotics, building on earlier work in sensor fusion, planning and reinforcement learning; exploring applications in healthcare
Publications
Authors are listed in publication order; Guoliang You is in bold and * marks equal contribution. Preprints are identified separately from peer-reviewed publications.
Click an illustration to view the hand-drawn overview or the original paper figure.
Medical AI
Semantically Calibrated Evidence Composition for CT Vision-Language Learning
Combines image-based hazard detection with 3D scene priors to estimate power-line clearance.
Experience
University of PennsylvaniaPostdoctoral Researcher · Philadelphia, PA
LimX DynamicsResearch Intern, multimodal foundation models and vision-language-action learning · Beijing
University of Science and Technology of China (USTC)Ph.D. in Computer Technology, LINKE Lab · HefeiAdvisors: Prof. Yanyong Zhang and Assoc. Prof. Jianmin Ji
AI Research Institute, Hefei Comprehensive National Science CenterResearch Intern, reinforcement learning and sim-to-real autonomous driving · Hefei
Anhui University of Science and TechnologyBachelor's degree in Applied Physics · Huainan
Beyond papers
Some of the most memorable parts of my research happened away from a desk: building platforms, testing them outside, and learning what changes when a model meets a physical system. These photos are from that part of my journey.
First-generation autonomous vehicle platformSecond-generation platform with upgraded sensorsAutonomous vehicle testing outdoorsSingle-arm robotic platformDual-arm robotic platformDrone-robotic arm grasping platform