BLARM: Animating 3D Objects from Video via Blending Latent Rigid Motion Primitives
Abstract
BLARM predicts temporally coherent 3D mesh animations from monocular video using learned rigid motion components and skinning weights without explicit rigs.
We introduce BLARM, a feed-forward method for video-driven 3D mesh animation. Given a monocular video and a static object mesh, BLARM predicts a temporally coherent animated mesh whose motion follows the video. Rather than relying on explicit rigs or directly regressing high-dimensional vertex motion, we represent animation using a compact set of learned, time-varying rigid motion components and time-invariant vertex-to-component skinning weights. This yields a low-dimensional deformation space without requiring skeletons, cages, skinning weights, or rig annotations. Our architecture conditions geometry-derived deformation latents on video features through factorized spatial-temporal attention, then decodes rigid transformations blended by predicted skinning weights. Trained with trajectory reconstruction, entropy regularization, and motion-aware contrastive learning, BLARM produces accurate and temporally stable animations while recovering compact, interpretable motion structure from monocular video.
Community
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- RegHead: Non-Humanoid Head Blendshapes via Feed-Forward Registration (2026)
- ScaleVid: Geometry-Aware Video Object Scaling with Mesh-Free Inference (2026)
- SM4RT: Learning Structured Motion Geometry for 4D Reconstruction (2026)
- ViP-Rig: Visual-Prompted Controllable Rigging (2026)
- OASIS: Occlusion-aware Single-image Hand Avatar Reconstruction via 3D Gaussian Splatting (2026)
- DSAR: Dual-Stream Autoregressive Modeling of Temporal Cloth Dynamics for Photorealistic Animatable Avatars (2026)
- HarmoHOI: Harmonizing Appearance and 3D Motion for Multi-view Hand-Object Interaction Synthesis (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper