ASCIIEval: Benchmarking Models' Visual Perception in Text Strings via ASCII Art Paper • 2410.01733 • Published Oct 2, 2024
ResearchGPT: Benchmarking and Training LLMs for End-to-End Computer Science Research Workflows Paper • 2510.20279 • Published Oct 23, 2025
Info-Coevolution: An Efficient Framework for Data Model Coevolution Paper • 2506.08070 • Published Jun 9, 2025
InfoBatch: Lossless Training Speed Up by Unbiased Dynamic Data Pruning Paper • 2303.04947 • Published Mar 8, 2023
Preventing Zero-Shot Transfer Degradation in Continual Learning of Vision-Language Models Paper • 2303.06628 • Published Mar 12, 2023
StateM: Reaching 95.3% Raw Accuracy, or a \$15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling Paper • 2608.15089 • Published 4 days ago • 196