Running Repro - Questioning the Coverage-Length Metric in Conformal Prediction: When Shorter Intervals Are Not Better ๐ฏ Browse and collaborate on a research logbook with an AI agent
Running Repro - Demystifying LLM-as-a-Judge: Analytically Tractable Model for Inference-Time Scaling ๐ฏ Explore and share experiment logs with an AI assistant
Running Repro - The Catastrophic Failure of the k-Means Algorithm in High Dimensions ๐ฏ View and collaborate on experiment logbooks online
Running 1 Repro - Efficient Test-Time Scaling via Hierarchical Search and Self-Verification for Discrete Diffusion Language Models ๐ฏ Track and share research notes with an AI coding assistant
Running 1 Repro - Esoteric Language Models: A Family of Any-Order Diffusion LLMs ๐ฏ Explore and sync experiment logs with an AI agent