Harnessing CLIP and DINO: An Uncertainty-Aware Cascaded Fusion Network for Generalizable Deepfake Image Detection Paper • 2609.07670 • Published 3 days ago • 15
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 7 days ago • 25
Listening Forward: Next Patch Embedding Prediction Enables Scalable Audio Learners Paper • 2608.19863 • Published 21 days ago • 6
OmniScientist: An Omni-Modal Omni-Discipline AI Scientist Paper • 2608.13558 • Published 28 days ago • 94
SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models Paper • 2608.10538 • Published 30 days ago • 15
CAPEval: A Decoupled Caption Evaluation across Understanding and Generation Paper • 2608.02589 • Published Aug 3 • 25