YiFan Cai ☕️
YiFan Cai
(he/him)

Graduate student

I am a graduate student in Systems Engineering at the University of Pennsylvania, with a B.S. in Computer Science from ShanghaiTech University. I have research experience in diffusion models, computer vision, robotic manipulation, and molecular drug design. My current interests focus on world models, computer vision, and artificial intelligence.
Download CV

Education

ShanghaiTech University

Bachelor of Science in Computer Science

ShanghaiTech logo
Major: Computer Science

Key focus: 2D Diffusion

University of Wisconsin-Madison

Exchange Student

UW-Madison logo
Major: Computer Science

Key focus: Intro to Cryptography, Software Security, HCI

University of Pennsylvania

Master of Science in Systems Engineering

UPenn logo
Major: System Engineering

Key Focus: World Models, Video Diffusion

Leadership & Service
Language
Python
C/C++
Matlab
R
JAVA
AI / ML
Diffusion Models
Computer Vision
Reinforcement Learning
World Models
Probabilistic Models
100%
Chinese Native
90%
English Fluent
📚 My Research

My research explores diffusion-based generative models as world models for perception, prediction, and decision-making. I focus on three closely related directions:

  • Diffusion-based image restoration, including controllable inpainting and completion.
  • Generative diffusion models for small-molecule drug design, guided by structured and non-differentiable objectives.
  • World models for robotic manipulation, learning latent dynamics that support planning and control.

Across these settings, my broader goal is to understand how generative models can bridge perception and action by providing rich, structured representations of the environment.

Recent Publications
Industry Experience And Research
ByteDance - Douyin Ads featured image

ByteDance - Douyin Ads

Ads ranking and recommendation systems internship.

Alibaba Group - Tmall Campus featured image

Alibaba Group - Tmall Campus

Development and optimization of generative image stylization models based on GAN and diffusion architectures for portrait-to-anime image generation.

GRASP Lab, University of Pennsylvania featured image

GRASP Lab, University of Pennsylvania

Research on 3D scene understanding, temporal modeling, and world models for robotic manipulation.

YesAI Lab, ShanghaiTech University featured image

YesAI Lab, ShanghaiTech University

Research on diffusion models for image generation, computer vision, and molecular generation.

Selected Projects
PERoKF: Physics-Enhanced Super-Resolution of Kolmogorov Flow featured image

PERoKF: Physics-Enhanced Super-Resolution of Kolmogorov Flow

Single-frame fluid super-resolution with physics-consistency loss and physics-derived features

FEM for PDE-Constrained Optimization in Heat Conduction featured image

FEM for PDE-Constrained Optimization in Heat Conduction

Applying finite element methods and convex optimization to sequential PDE-constrained problems in heat conduction (SAND-style formulation).

HeartAI: Heart Rate Game AI featured image

HeartAI: Heart Rate Game AI

An AI agent for Heart Rate Game combining reinforcement learning and physiological control strategies.

Object Detection in Counter-Strike 2 featured image

Object Detection in Counter-Strike 2

Applying and comparing object detection algorithms in Counter-Strike 2 for reliable player/weapon detection under complex in-game scenarios.

Building an Intelligent Pac-Man Agent featured image

Building an Intelligent Pac-Man Agent

Incrementally upgrading a Pac-Man agent from rule-based planning to probabilistic reasoning and reinforcement learning, forming a complete AI decision-making pipeline.