AlignBench: Benchmarking Chinese Alignment of Large Language Models Paper • 2311.18743 • Published Nov 30, 2023 • 1
UrbanLLaVA: A Multi-modal Large Language Model for Urban Intelligence with Spatial Reasoning and Understanding Paper • 2506.23219 • Published Jun 29, 2025 • 7
Mitigating Geospatial Knowledge Hallucination in Large Language Models: Benchmarking and Dynamic Factuality Aligning Paper • 2507.19586 • Published Jul 25, 2025
RAISECity: A Multimodal Agent Framework for Reality-Aligned 3D World Generation at City-Scale Paper • 2511.18005 • Published Nov 22, 2025 • 1
SpatialAct: Probing Spatial Reasoning-to-Action Capabilities of VLM Agents in 3D Scenes Paper • 2605.31148 • Published May 29 • 4
A Survey of Large Language Model-Powered Spatial Intelligence Across Scales: Advances in Embodied Agents, Smart Cities, and Earth Science Paper • 2504.09848 • Published Apr 14, 2025 • 1
CritiqueLLM: Scaling LLM-as-Critic for Effective and Explainable Evaluation of Large Language Model Generation Paper • 2311.18702 • Published Nov 30, 2023
Can AI Agents Learn Their Way to the Top? Evaluating Heuristic Learning in a Long-Running Game Agent Competition Paper • 2610.12341 • Published 3 days ago • 13