More Choices, Fewer Decisions: Ordinal-Scale Bias in JEV-like Direct-Decision Models Paper • 2609.38827 • Published 2 days ago • 49
LoRA-GGPO: Mitigating Double Descent in LoRA Fine-Tuning via Gradient-Guided Perturbation Optimization Paper • 2502.14538 • Published Feb 20, 2025
XTRUST: On the Multilingual Trustworthiness of Large Language Models Paper • 2409.15762 • Published Sep 24, 2024
OCR-MetaReasoning Benchmark: Evaluating the Meta-Reasoning Ability of MLLMs in Text-Rich Image Understanding Paper • 2608.30678 • Published Aug 31 • 1
CaSKG: Counterfactual-Causal Skill Graphs for Scalable Agent Skill Retrieval Paper • 2608.25500 • Published Aug 26 • 7
Thinking vs. NoThinking: Towards Interpreting Reasoning Mechanisms of Large Language Models via Sparse Autoencoders Paper • 2608.08168 • Published Aug 8
AGGC: Adaptive Group Gradient Clipping for Stabilizing Large Language Model Training Paper • 2601.11864 • Published Jan 17 • 1
Align, Don't Divide: Revisiting the LoRA Architecture in Multi-Task Learning Paper • 2508.05078 • Published Aug 7, 2025
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability Paper • 2508.04017 • Published Aug 6, 2025 • 12
Don't Take the Premise for Granted: Evaluating the Premise Critique Ability of Large Language Models Paper • 2505.23715 • Published May 29, 2025 • 2
THINK-Bench: Evaluating Thinking Efficiency and Chain-of-Thought Quality of Large Reasoning Models Paper • 2505.22113 • Published May 28, 2025 • 2
More Choices, Fewer Decisions: Ordinal-Scale Bias in JEV-like Direct-Decision Models Paper • 2609.38827 • Published 2 days ago • 49
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability Paper • 2508.04017 • Published Aug 6, 2025 • 12
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability Paper • 2508.04017 • Published Aug 6, 2025 • 12 • 2
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens Paper • 2508.01191 • Published Aug 2, 2025 • 240
NLoRA: Nyström-Initiated Low-Rank Adaptation for Large Language Models Paper • 2502.14482 • Published Feb 20, 2025
NLoRA: Nyström-Initiated Low-Rank Adaptation for Large Language Models Paper • 2502.14482 • Published Feb 20, 2025 • 1
StructFlowBench: A Structured Flow Benchmark for Multi-turn Instruction Following Paper • 2502.14494 • Published Feb 20, 2025 • 15