arxiv:2602.02493
ZehongMa
zehongma
AI & ML interests
MLLMs, Image/Video Generation, Multi-modal Representation Learning
Recent Activity
upvoted a paper 11 days ago
MiniWorld: Democratizing the Training of Video World Models from Scratch authored a paper 28 days ago
PixelGen: Pixel Diffusion Beats Latent Diffusion with Perceptual LossOrganizations
None yet