• Home
  • Search
  • Motion-Zero: A Zero-Shot Trajectory Control Framework of Moving Object for Diffusion-Based Video Generation
  • Cite Icon1
  • https://doi.org/10.1609/aaai.v39i2.32198Copy DOI Icon

Motion-Zero: A Zero-Shot Trajectory Control Framework of Moving Object for Diffusion-Based Video Generation

Show More
  • Abstract
  • Literature Map
  • Citations
  • Similar Papers
Abstract

Recent large-scale pre-trained diffusion models have demonstrated a powerful generative ability to produce high-quality videos from detailed text descriptions. However, exerting control over the motion of objects in videos generated by any video diffusion model remains a challenging problem. In this paper, we propose a novel zero-shot moving object trajectory control framework, Motion-Zero, to enable arbitrary single-object-trajectory control for the text-to-video diffusion model. To this end, an initial noise prior module is designed to provide a position-based prior to improve the stability of the appearance of the moving object and the accuracy of position. In addition, based on the attention map of the U-Net, spatial constraints are directly applied to the denoising process of diffusion models, which further ensures the positional consistency of moving objects during the inference. Furthermore, temporal consistency is guaranteed with a proposed shift temporal attention mechanism. Our method can be flexibly applied to various state-of-the-art video diffusion models without any training process. Extensive experiments demonstrate our proposed method can control the motion trajectories of arbitrary objects while preserving the original ability to generate high-quality videos.

Similar Papers
  • Conference Article
  • Citations4

Extracting video object's motion trajectory by velocity voting

  • Dec 15, 2003
  • Chang Yuan +2
  • Conference Article

Video object trajectory perturbation based data hiding satisfying statistical and perceptual invisibility

  • Mar 01, 2011
  • Abdullah Cay +2
  • Research Article

Talking-DiSSM: Enhancing Temporal Consistency in Talking Face Video Generation with Bidirectional SSMs

  • Jul 22, 2025
  • ACM Transactions on Intelligent Systems and Technology
  • Zhen Xiao +5
  • Conference Article
  • Citations12

Global Coding of Multi-source Surveillance Video Data

  • Apr 01, 2015
  • Jing Xiao +4
  • Research Article
  • Citations27

Transparent Image Layer Diffusion using Latent Transparency

  • Jul 19, 2024
  • ACM Transactions on Graphics
  • Lvmin Zhang +1
  • Research Article

УДОСКОНАЛЕННЯ БОРТОВИХ НАВІГАЦІЙНИХ КОМПЛЕКСІВ АВТОТРАНСПОРТНИХ ЗАСОБІВ З ВИКОРИСТАННЯМ ІНЕРЦІАЛЬНИХ ДАТЧИКІВ ТА ІМОВІРНІСНО-ГЕОМЕТРИЧНИХ МЕТОДІВ ОБРОБКИ ІНФОРМАЦІЇ

  • Jan 01, 2020
  • The National Transport University Bulletin
  • Topolskov E +2
  • Research Article
  • Citations13

Unified Spatio-Temporal Dynamic Routing for Efficient Video Object Segmentation

  • May 01, 2024
  • IEEE Transactions on Intelligent Transportation Systems
  • Jisheng Dang +3
  • PDF
  • Research Article
  • Citations37

Feature point based 3D tracking of multiple fish from multi-view images

  • Jun 30, 2017
  • PLoS ONE
  • Zhi-Ming Qian +1
  • Conference Article

InstantAS: Minimum Coverage Sampling for Arbitrary-Size Image Generation

  • Oct 28, 2024
  • Changshuo Wang +5
  • Conference Article

OMS: One More Step Noise Searching to Enhance Membership Inference Attacks for Diffusion Models

  • Sep 01, 2025
  • Xiaomeng Fu +6
  • Research Article
  • Citations10

Stabilizing Diffusion Model for Robotic Control With Dynamic Programming and Transition Feasibility

  • Sep 01, 2024
  • IEEE Transactions on Artificial Intelligence
  • Haoran Li +4
  • Research Article
  • Citations9

Diffusion Models for 3D Generation: A Survey

  • Feb 01, 2025
  • Computational Visual Media
  • Chen Wang +4
  • Conference Article
  • Citations27

Direct-a-Video: Customized Video Generation with User-Directed Camera Movement and Object Motion

  • Jul 13, 2024
  • Shiyuan Yang +7
  • Conference Article

RePaint-Enhanced Conditional Diffusion Model for Generating Designs Under Performance Constraints

  • Aug 17, 2025
  • Ke Wang +4
  • Research Article

InstantPainting: Expanding GANs for Efficient Text-Conditioned Image Generation Platform

  • Apr 11, 2025
  • Proceedings of the AAAI Conference on Artificial Intelligence
  • Bing-Kun Bao +4
Cactus Communications logo

Copyright 2026 Cactus Communications. All rights reserved.