Zhi Zheng
zz1358m
AI & ML interests
LLM reasoning, Trustworthy LLM, LLM application, Neural combinatorial optimization.
Recent Activity
liked a model about 19 hours ago
zz1358m/Qwen3.5-4B-MATH-ReAct-Agentic-ESOpt upvoted a paper about 23 hours ago
Agentic ESOpt: Fine-Tuning Long-Horizon LLM Agents with Minimal GPU Requirements updated a collection 1 day ago
Agentic-ESOpt Checkpoints Collection