showSidebars ==
showTitleBreadcrumbs == 1
node.field_disable_title_breadcrumbs.value ==

PhD Dissertation Proposal by XU Chenshu | Towards User-Controllable Visual Media Generation and Editing

Please click here if you are unable to view this page.

 
Towards User-Controllable Visual Media Generation and Editing

XU Chenshu

PhD Candidate
School of Computing and Information Systems
Singapore Management University
 

FULL PROFILE 

Research Area

  • Human-Machine Collaborative Systems
    • Multimedia Systems

Dissertation Committee

Advisor:
Members:
 
 

Date

29 Jul y 2026 (Wednesday)

Time

1:00pm – 2:00pm

Venue

Meeting room 5.1, Level 5
School of Computing and Information Systems 1, 
Singapore Management University, 
80 Stamford Road, 
Singapore 178902

Please register by 27 July 2026.

We look forward to seeing you at this research seminar.

 

ABOUT THE TALK

Recent generative models have substantially advanced visual content creation, yet practical creative workflows require users to control generated content precisely, intuitively, and predictably. We investigate user-controllable visual media generation and editing across raster images and vector graphics through four progressively forms of control. First, we address identity–style entanglement in personalized image generation, enabling a character’s identity and artistic style to be learned and manipulated independently. Second, we introduce DragNoise, an interactive point-based image editing method that improves editing accuracy, semantic consistency, and computational efficiency. Third, we present DragSVG, which extends drag-based interaction to vector graphics and produces temporally coherent animations that follow user-defined trajectories while preserving the structure and editability of the original SVG. Finally, we propose DragPhysSVG, a framework for scene-aware physical animation that adapts physics-grounded motion to the geometry of a target scene, allowing animated vector objects to interact more naturally with their surrounding environments. Together, these studies demonstrate how suitable representations, interaction mechanisms, and optimization strategies can translate increasingly sophisticated forms of user intent into faithful and predictable visual outcomes, contributing toward generative visual systems that serve as controllable and expressive tools for human creativity.

ABOUT THE SPEAKER

Chenshu XU is a PhD candidate in the School of Computing and Information Systems (SCIS) at Singapore Management University (SMU). She is a member of Visual Understanding and Generation Lab (VUG) @ SMU supervised by Prof. HE Shengfeng. Her research interests lie at the intersection of Computer Vision and Computer Graphics.