Submitted by taesiri 3 Code-Space Response Oracles: Generating Interpretable Multi-Agent Policies with Large Language Models Deepmind 1
Submitted by Junyi Zhang 63 LoGeR: Long-Context Geometric Reconstruction with Hybrid Memory Deepmind 609 7
Submitted by taesiri 19 Fantastic Reasoning Behaviors and Where to Find Them: Unsupervised Discovery of the Reasoning Process Deepmind 3
Submitted by taesiri 8 The FACTS Leaderboard: A Comprehensive Benchmark for Large Language Model Factuality Deepmind 2
Submitted by Tyler Zhu 4 Dynamic Reflections: Probing Video Representations with Text Alignment Deepmind 2