Yixuan Wang
LuckyOrz
ยท
AI & ML interests
LLM pretraining, hierarchical model, efficient LLM
Recent Activity
upvoted a collection 3 days ago
NCP_ArchPreview upvoted a paper 28 days ago
MemSFT: Mitigating Alignment Tax with an External Parametric Memory upvoted a paper about 1 month ago
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory