-
starriver030515/hapo_data
Viewer • Updated • 1.59k • 32 -
starriver030515/Qwen2.5-Math-1.5B-16k
Text Generation • 2B • Updated • 14 -
starriver030515/Qwen2.5-Math-7B-32k
Text Generation • 8B • Updated • 20 -
From Uniform to Heterogeneous: Tailoring Policy Optimization to Every Token's Nature
Paper • 2509.16591 • Published • 2
Zheng Liu
starriver030515
AI & ML interests
None yet
Recent Activity
upvoted a paper 13 days ago
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression liked a model 20 days ago
deepseek-ai/DeepSeek-V4.1-Flash liked a model about 2 months ago
Qwen/Qwen3.8-27BOrganizations
FUSION-Data
FUSION-Stage1
FUSION-Model
-
starriver030515/FUSION-LLaMA3.1-8B
Image-Text-to-Text • 10B • Updated • 13 -
starriver030515/FUSION-X-LLaMA3.1-8B
Image-Text-to-Text • 11B • Updated • 12 -
starriver030515/FUSION-X-Phi3.5-3B
Image-Text-to-Text • 6B • Updated • 20 • 2 -
starriver030515/FUSION-Phi3.5-3B
Image-Text-to-Text • 5B • Updated • 7
FUSION-Stage1.5
HAPO
-
starriver030515/hapo_data
Viewer • Updated • 1.59k • 32 -
starriver030515/Qwen2.5-Math-1.5B-16k
Text Generation • 2B • Updated • 14 -
starriver030515/Qwen2.5-Math-7B-32k
Text Generation • 8B • Updated • 20 -
From Uniform to Heterogeneous: Tailoring Policy Optimization to Every Token's Nature
Paper • 2509.16591 • Published • 2
FUSION-Model
-
starriver030515/FUSION-LLaMA3.1-8B
Image-Text-to-Text • 10B • Updated • 13 -
starriver030515/FUSION-X-LLaMA3.1-8B
Image-Text-to-Text • 11B • Updated • 12 -
starriver030515/FUSION-X-Phi3.5-3B
Image-Text-to-Text • 6B • Updated • 20 • 2 -
starriver030515/FUSION-Phi3.5-3B
Image-Text-to-Text • 5B • Updated • 7
FUSION-Data
FUSION-Stage1.5
FUSION-Stage1