EfficientRollout: System-Aware Self-Speculative Decoding for RL Rollouts Paper • 2606.18967 • Published Jun 17 • 24
ParallelBench: Understanding the Trade-offs of Parallel Decoding in Diffusion LLMs Paper • 2510.04767 • Published Oct 6, 2025 • 28
Optimized Feature Generation for Tabular Data via LLMs with Decision Tree Reasoning Paper • 2406.08527 • Published Jun 12, 2024 • 1
EfficientRollout: System-Aware Self-Speculative Decoding for RL Rollouts Paper • 2606.18967 • Published Jun 17 • 24
State-offset Tuning: State-based Parameter-Efficient Fine-Tuning for State Space Models Paper • 2503.03499 • Published Mar 5, 2025 • 5
VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data Paper • 2502.06737 • Published Feb 10, 2025
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation Paper • 2508.05399 • Published Aug 7, 2025 • 17
XQuant: Breaking the Memory Wall for LLM Inference with KV Cache Rematerialization Paper • 2508.10395 • Published Aug 14, 2025 • 42
XQuant: Breaking the Memory Wall for LLM Inference with KV Cache Rematerialization Paper • 2508.10395 • Published Aug 14, 2025 • 42
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation Paper • 2508.05399 • Published Aug 7, 2025 • 17
Sparsified State-Space Models are Efficient Highway Networks Paper • 2505.20698 • Published May 27, 2025 • 2
Hierarchical Context Merging: Better Long Context Understanding for Pre-trained LLMs Paper • 2404.10308 • Published Apr 16, 2024
UNCAGE: Contrastive Attention Guidance for Masked Generative Transformers in Text-to-Image Generation Paper • 2508.05399 • Published Aug 7, 2025 • 17
VersaPRM: Multi-Domain Process Reward Model via Synthetic Reasoning Data Paper • 2502.06737 • Published Feb 10, 2025
ReVISE: Learning to Refine at Test-Time via Intrinsic Self-Verification Paper • 2502.14565 • Published Feb 20, 2025
Counting Guidance for High Fidelity Text-to-Image Synthesis Paper • 2306.17567 • Published Jun 30, 2023