FIRE3D: Feed-forward Interactive 3D Scene Reconstruction Within A Minute
Abstract
FIRE3D is a feed-forward framework that rapidly converts single images or casual videos into interactive, simulation-ready 3D scenes with complete object-level geometry, poses, and textures.
We present FIRE3D, a unified framework that takes a single RGB image or casual RGB video and transforms it into simulation-ready 3D scene assets for games and interactive applications in under a minute. At the core of FIRE3D is a feed-forward, end-to-end network that predicts a compositional scene representation from posed RGB-D observations estimated from the RGB capture, including the 6-DoF pose, bounding box, mesh, and texture for every object. By modeling the scene as a collection of discrete entities, FIRE3D produces amodally complete and simulation-ready environments where objects are physically decoupled and ready for interaction. Our framework requires no test-time optimization, runs orders of magnitude faster than prior interaction-ready methods, and provides object-level completeness beyond existing feed-forward 3D approaches. We demonstrate competitive or state-of-the-art results across pose accuracy, geometry completeness, and texture quality across various datasets while being orders of magnitudes faster. Project page: https://xiahongchi.github.io/Fire3D/
Get this paper in your agent:
hf papers read 2609.08848 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 1
Datasets citing this paper 2
hongchi/Fire3D_examples
hongchi/Fire3D
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper