Yuhai Deng 邓宇海

I am currently pursuing an M.S. degree at the College of Computer Science, Nankai University, under the supervision of Prof. Xiang Li, whose insightful guidance has greatly shaped my research interests and approach.
Prior to that, I received my B.E. degree from Central South University in 2024, where I was honored as an Outstanding Graduate. During my undergraduate studies, I had the privilege of working with Prof. Shichao Kan, who taught me how to approach research problems and introduced me to the world of computer vision.

I believe intelligence is inherently multimodal. Models should be able not only to understand the world, but also to think in vision. Intelligent models should unify visual representation, visual reasoning, and visual generation, bridging understanding, thinking, and creation.

Research Experience

My research experience spans visual retrieval, reference-guided image tone editing, and expressive talking-face generation. Across these projects, I have primarily explored diffusion-based generation and representation learning.

Reference-Guided Image Tone Editing

Developed ICTone to explore semantic-aware image tone editing from a reference image through diffusion-based generation.

Object Retrieval

Studied object retrieval for outside-knowledge VQA and developed MS-GCEL for long-tailed visual retrieval.

Expressive Talking-Face Generation

Developed VAExpress to explore expressive talking-face generation with audio-aligned facial expressions.

Education

M.S. Student in Computer Science and Technology | Nankai University

Time: 2024.09 — Present

Honors & Awards

Outstanding Graduate

Central South University

Second Prize

The 16th Chinese Collegiate Computing Competition

National Scholarship

Central South University