-
L4Q: Parameter Efficient Quantization-Aware Training on Large Language Models via LoRA-wise LSQ
Paper • 2402.04902 • Published • 5 -
QWHA: Quantization-Aware Walsh-Hadamard Adaptation for Parameter-Efficient Fine-Tuning on Large Language Models
Paper • 2509.17428 • Published • 9 -
LRAgent: Efficient KV Cache Sharing for Multi-LoRA LLM Agents
Paper • 2602.01053 • Published • 8 -
PReCache: Efficient KV Cache Sharing for Multi-LoRA Agents via Low-Rank Precomputation and Neutral Reconstruction
Paper • 2609.34054 • Published • 4
Hyesung Jeon
hjeon2k
AI & ML interests
None yet
Recent Activity
updated a collection about 13 hours ago
Authored Paper updated a collection about 13 hours ago
Authored Paper submitted a paper about 13 hours ago
KVCMAS: Efficient KV cache Correction for Shared Context in Multi-Agent Systems