Fei Pan

Fei Pan潘飞

Associate Research Fellow, Zhejiang Lab 副研究员,之江实验室

CV Google Scholar GitHub
About简介

I am an Associate Research Fellow at Zhejiang Lab. My research focuses on multimodal foundation vision-language models, with an emphasis on their applications in Earth science. My work spans visual perception and representation learning, aiming to build visual and embodied intelligence systems with strong generalization capability.

I received my Ph.D. from KAIST in 2023, during which I was awarded the Robert Bosch Ph.D. Fellowship and the Qualcomm Innovation Fellowship. After my Ph.D., I conducted postdoctoral research at the University of Michigan, working with Prof. Stella X. Yu on large-scale vision-language models.

目前于之江实验室任副研究员,主要研究多模态基础视觉语言模型,并探索其在地球科学领域的应用。 研究方向包括视觉感知与表征学习,致力于构建强泛化能力的视觉及具身智能系统。

2023 年于韩国科学技术院(KAIST)取得博士学位,博士在读期间曾获 罗伯特‑博世博士奖学金(Robert Bosch PhD Fellowship)与 高通创新奖学金(Qualcomm Innovation Fellowship)。 博士毕业后,于密歇根大学(University of Michigan)从事博士后研究,与 Stella X. Yu 教授合作开展大规模视觉语言模型相关的工作。

Multi-modal Vision-Language Models 多模态视觉语言模型
Embodied AI & Vision-Language-Action (VLA) Models 具身智能与视觉-语言-动作(VLA)模型
Domain Adaptation & Generalization 域自适应与泛化
Self-supervised & Unsupervised Representation Learning 自监督与无监督表征学习
Semantic Segmentation & Open-Vocabulary Perception 语义分割与开放词汇视觉感知
Model Compression, Quantization & Edge Deployment 模型压缩、量化与端侧部署
Experience工作履历
2026.09 – Present至今
Zhejiang Lab之江实验室
Associate Research Fellow副研究员

Multi-modal Vision-Language Models and their applications to Earth science.

多模态视觉语言模型及其在地球科学中的应用。

2025.04 – 2026.08
Vibe Inc.
AI Algorithm ResearcherAI 算法研发

Embodied AI and robotics — 3D perception, robotic grasping, VLA model fine-tuning, and on-device AI model training and deployment.

具身智能与机器人方向 — 3D 视觉感知、机械臂抓取、VLA 模型微调、端侧AI模型训练与部署。

2023.09 – 2025.03
University of Michigan, Ann Arbor 密西根大学安娜堡分校
Research Fellow, EECS博士后研究员,EECS

Large-scale vision-language models, diffusion models, and ViTs, with Prof. Stella X. Yu.

与 Stella X. Yu 教授合作,研究大规模视觉语言模型、扩散模型与 ViTs。

2021.05 – 2021.11
Robert Bosch GmbH
R&D Intern研发实习

YOLO-based multi-camera vehicle detection with GAN-based synthetic data augmentation.

基于 YOLO 的多路摄像头车辆检测,结合 GAN 合成数据增强。

2016.03 – 2023.08
KAIST
Graduate Research Assistant研究生助理,GSA

Domain-adaptive and self-supervised semantic segmentation using GANs, Cut & Mixing, and geometry-guided motion cues.

基于 GAN、Cut & Mixing 、3D视觉运动信息的域自适应与自监督语义分割研究。

Education教育背景
Ph.D., Electrical Engineering博士,电子信息工程 — KAIST (QS 65)(QS 世界排名 65)
2018.03 – 2023.08 · Advisor: Prof. In So Kweon导师:In So Kweon 教授
Thesis: Geometric-guided Domain Adaptation for Semantic Segmentation. 博士论文:Geometric-guided Domain Adaptation for Semantic Segmentation。
M.S., Electrical Engineering硕士,电子信息工程 — KAIST
2016.02 – 2018.02
B.Eng., Communication Engineering学士,通信工程Xidian University西安电子科技大学
2011.09 – 2015.07
Ranked 5/825 (top 0.6%) · GPA 3.94/4.0 学院排名 5/825(前 0.6%)· 成绩 3.94/4.0
Publications发表著作
Selected Projects软件著作
Zero-shot Classification & Segmentation
Zero-shot recognition built on large-scale vision-language models, covering classification, detection, segmentation, and captioning without task-specific training data. 基于大规模视觉语言模型的零样本识别工具,涵盖分类、检测、分割与图像描述,无需特定任务训练数据。
github.com/BuildingInfoSys/zeroshot_attribute_extraction
ImageNet-D — Diffusion Robustness Benchmark
Diffusion-synthesized objects for evaluating robustness of vision models from ResNet to ViT-based foundation models like CLIP and MiniGPT-4. 基于扩散模型生成合成对象,用于评估从 ResNet 到 CLIP、MiniGPT-4 等 ViT 基础模型的视觉模型鲁棒性。
github.com/chenshuang-zhang/imagenet_d
IntraDA — Domain Adaptive Semantic Segmentation
Domain-adaptive semantic segmentation framework improving generalization to real-world driving scenes. 具备域自适应能力的语义分割框架,旨在提升模型在真实驾驶场景图像上的泛化能力。
github.com/feipanir/IntraDA
Awards & Honors获奖与荣誉
2024.06
Best Paper Award最佳论文奖

CVPR Workshop on Learning with Limited Labelled Data for Image and Video Understanding. CVPR「有限标注数据学习在图像与视频理解中的应用」研讨会。

2020.12
Qualcomm Innovation Fellowship高通公司创新研发奖

Awarded for outstanding research achievements during graduate study in Korea. 因在韩国攻读研究生期间的杰出研究成果而获得。

2019.09
Robert Bosch Ph.D. Scholarship (2 years)博世博士奖学金(两年)

Supported research on intelligent transportation systems. 支持在智能交通系统方面的创新研究。

2015.05
Goodix Technology Scholarship汇顶科技奖学金

Outstanding graduate, Xidian University. 西安电子科技大学本科优秀毕业生。

2014.05
National Scholarship国家奖学金

For outstanding academic performance, Xidian University. 因本科期间的优异学术表现而获得。

Skills技能
PyTorch TensorFlow Caffe Python C / C++ NumPy OpenCV Git Docker SciPy CNNs GANs Diffusion Models ControlNet ViTs LLMs VLMs
Academic Service学术服务