When Many Answers Are Valid, Voting Fails: Symbolic Verification for Best-of-K Causal Reasoning in LLMs
Paper • 2608.03506 • Published • 3
We're creating a causal atlas of global well-being---combining satellite imagery, AI agent swarms, traditional machine learning, and causal methods to track and test human development at scale
Temporal Dynamics of Development Aid in Africa: Evidence from a Staggered Difference-in-Differences Study of China and World Bank Projects in Africa
Remote Auditing: Design-based Tests of Randomization, Selection, and Missingness with Broadly Accessible Satellite Imagery