perplexity-ai/pplx-decider-v1.1-27b Text Classification • 26B • Updated about 7 hours ago • 599 • 74
SpIDER: Spatially Informed Dense Embedding Retrieval for Software Issue Localization Paper • 2512.16956 • Published Feb 5 • 1
Hermes: Learning Contextual Reasoning Unlocks Test-Time Scaling Paper • 2609.38332 • Published 10 days ago • 10
MILO: Automated Harness Discovery via Orchestrated Multi-Agent Evolution Paper • 2609.38349 • Published 10 days ago • 22
On the Off-Policy Teacher in On-Policy Distillation Paper • 2609.38360 • Published 10 days ago • 23
AQuA: A Benchmarking Tool for Label Quality Assessment Paper • 2306.09467 • Published Jun 15, 2023 • 1
Investigating Compositional Reasoning in Time Series Foundation Models Paper • 2502.06037 • Published Feb 9, 2025 • 2
Towards Long-Context Time Series Foundation Models Paper • 2409.13530 • Published Sep 20, 2024 • 2
ARFBench: Benchmarking Time Series Question Answering Ability for Software Incident Response Paper • 2604.21199 • Published Apr 23 • 1
Chronos-2: From Univariate to Universal Forecasting Paper • 2510.15821 • Published Oct 17, 2025 • 28
SpIDER: Spatially Informed Dense Embedding Retrieval for Software Issue Localization Paper • 2512.16956 • Published Feb 5 • 1
STAMP: Spatial-Temporal Adapter with Multi-Head Pooling Paper • 2511.10848 • Published Nov 13, 2025 • 2
MOMENT: A Family of Open Time-series Foundation Models Paper • 2402.03885 • Published Feb 6, 2024 • 9
TimeSeriesGym: A Scalable Benchmark for (Time Series) Machine Learning Engineering Agents Paper • 2505.13291 • Published May 19, 2025 • 3
MILO: Automated Harness Discovery via Orchestrated Multi-Agent Evolution Paper • 2609.38349 • Published 10 days ago • 22
Hermes: Learning Contextual Reasoning Unlocks Test-Time Scaling Paper • 2609.38332 • Published 10 days ago • 10
On the Off-Policy Teacher in On-Policy Distillation Paper • 2609.38360 • Published 10 days ago • 23
Running 68 Don't Train the Model, Evolve the Harness 🌿 68 Evolving an agent's harness, not its model, on Harvey's LAB