The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 4 days ago • 137 • 6
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 4 days ago • 137
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 4 days ago • 137
The Tasteful Agent: Measuring and Improving Taste in Long-Horizon Tasks Paper • 2609.25804 • Published 4 days ago • 137
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts Paper • 2606.05922 • Published Jun 4 • 37
The Hidden Dimensions of LLM Alignment: A Multi-Dimensional Analysis of Orthogonal Safety Directions Paper • 2502.09674 • Published May 27, 2025
World Pilot: Steering Vision-Language-Action Models with World-Action Priors Paper • 2606.12403 • Published Jun 10 • 27
ICA Lens: Interpreting Language Models Without Training Another Dictionary Paper • 2606.11722 • Published Jun 10 • 18
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts Paper • 2606.05922 • Published Jun 4 • 37
Retrospective Harness Optimization: Improving LLM Agents via Self-Preference over Trajectory Rollouts Paper • 2606.05922 • Published Jun 4 • 37
Breaking Language Barriers: Cross-Lingual Continual Pre-Training at Scale Paper • 2407.02118 • Published Jul 2, 2024 • 1
Can LLMs Refuse Questions They Do Not Know? Measuring Knowledge-Aware Refusal in Factual Tasks Paper • 2510.01782 • Published Oct 2, 2025
Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs Paper • 2602.01914 • Published Feb 2 • 1
Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs Paper • 2602.01914 • Published Feb 2 • 1