view article Article iFAN: Training Plain Mask Transformers for the Way They Actually Infer xxlucas • 3 days ago • 2
view article Article iFAN: Training Plain Mask Transformers for the Way They Actually Infer xxlucas • 3 days ago • 2
view article Article AnchorDiT: Semantic Anchor-Guided Transformers for Pixel-Space Image Generation xxlucas • 4 days ago • 3
view article Article AnchorDiT: Semantic Anchor-Guided Transformers for Pixel-Space Image Generation xxlucas • 4 days ago • 3
HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion Paper • 2605.15741 • Published May 15 • 2
Dynamic-TreeRPO: Breaking the Independent Trajectory Bottleneck with Structured Sampling Paper • 2509.23352 • Published Sep 27, 2025 • 10
RePainter: Empowering E-commerce Object Removal via Spatial-matting Reinforcement Learning Paper • 2510.07721 • Published Oct 9, 2025 • 9
UM-Text: A Unified Multimodal Model for Image Understanding Paper • 2601.08321 • Published Jan 13 • 21
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition Paper • 2605.10127 • Published May 13 • 11