CSD722 Project: Depth-Conditioned Video Generation using ControlNet & AnimateDiff

Year: Jan – May 2026

Supervisor: Dr. Sumit Shekhar & Dr. Saurabh Shigwan (CSE Dept, SNU)
ControlNetAnimateDiffVideoGenerationDepthConditioningDiffusionModels
  • Extended ControlNet to text-to-video diffusion models using AnimateDiff and Motion LoRA.
  • Modified model architecture to prevent training collapse where generated videos ignored depth-map conditioning; stabilized training by adding auxiliary supervision losses.
  • Conducted ablation experiments on prompts, depth conditioning, and text guidance.

Resources & Links

Continue Exploring

Discover other related projects or articles.

Back to All Projects