AgentHazard: A Benchmark for Evaluating Harmful Behavior in Computer-Use Agents Paper • 2604.02947 • Published 4 days ago • 12
PixelSmile: Toward Fine-Grained Facial Expression Editing Paper • 2603.25728 • Published 12 days ago • 116