Jingzhi Wang
jzwang666
ยท
AI & ML interests
LLM pretrain & finetuning, develop high effiency attention architecture, agentic RL system
Recent Activity
upvoted a paper 1 day ago
Memory Decoder at Scale: A Pretrained, Parametric Long-Term Memory upvoted a paper about 1 month ago
Information-Aware KV Cache Compression for Long Reasoning