Motion Beyond Morphology: Bootstrapping Cross-Category Motion Transfer from Abstract Motion Representations
Abstract
Video motion transfer aims to animate a target object using dynamics from a reference video. Existing formulations largely rely on fixed structural correspondence, which becomes ill-defined when reference and target objects differ substantially in morphology, articulation, or deformation mechanisms. We introduce Motion Beyond Morphology, a perspective that seeks to transfer motion beyond fixed structural correspondence, by preserving dynamics that remain meaningful across different target morphologies. To realize this, we propose a two-stage framework. Stage~I learns complementary multi-granularity abstract motion views and uses them to bootstrap cross-category video pairs that preserve transferable dynamics across diverse morphologies. Stage~II internalizes this supervision into direct reference-video-conditioned generation, removing the need for explicit motion extraction at inference. We further introduce OpenVMT-Dataset and OpenVMT-Bench for training and evaluating image- and text-conditioned motion transfer across Same, Near, and Far category gaps, and plan to release both upon acceptance. Extensive experiments demonstrate state-of-the-art motion fidelity and target preservation. Project page: https://miniz233.github.io/MotionBeyondMorphology/
Community
Project Page: https://miniz233.github.io/MotionBeyondMorphology/
Github: https://github.com/miniz233/MBM
Morphology-Adaptive Motion Transfer (Image2Video)
๐ We introduce Motion Beyond Morphology (MBM), a framework for transferring motion across objects with substantially different morphologies, without relying on fixed structural correspondence.
MBM learns complementary multi-granularity abstract motion representations to bootstrap cross-category supervision, and then enables direct reference-video-conditioned generation at inference. Experiments across Same, Near, and Far category gaps demonstrate state-of-the-art motion fidelity and target preservation.
๐ Project page: https://miniz233.github.io/MotionBeyondMorphology/
๐ป Code: https://github.com/miniz233/MBM
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- Motion4Motion: Motion Transfer Across Subjects at Inference (2026)
- VideoWeave: Unlocking Geometric Consistency in Video Generation via Joint Geometry-Video Modeling (2026)
- UniMoCa: Unifying Motion and Camera Controls as Visual Proxies for Faithful Human Video Generation (2026)
- TriMotion: Modality-Agnostic Camera Control for Video Generation (2026)
- ProxyUp: Training-Free Proxy-Conditioned Video Generation for Controllable Dynamics (2026)
- SCAIL-2: Unifying Controlled Character Animation with End-to-end In-Context Conditioning (2026)
- Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2608.01628 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper