Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Siyuan Li's picture

Siyuan Li

Lupin1998
16 60 15
Xin1118's profile picture syjian's profile picture JackyWangAI's profile picture
·
https://lupin1998.github.io/
  • LupinLSY
  • Lupin1998
  • siyuan-li-lupin1998

AI & ML interests

Network Design, Self-supervised Learning, Computer Vision, Data-centric ML, AI for Science

Organizations

MogaNet's profile picture ICML2023's profile picture OpenSTL's profile picture odl-raiser's profile picture OpenRaiser's profile picture

Collections 2

LLMS
  • Taming LLMs by Scaling Learning Rates with Gradient Grouping

    Paper • 2506.01049 • Published Jun 1, 2025 • 40
AIGC
  • MergeVQ: A Unified Framework for Visual Generation and Representation with Disentangled Token Merging and Quantization

    Paper • 2504.00999 • Published Apr 1, 2025 • 98
  • Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

    Paper • 2409.12191 • Published Sep 18, 2024 • 80
LLMS
  • Taming LLMs by Scaling Learning Rates with Gradient Grouping

    Paper • 2506.01049 • Published Jun 1, 2025 • 40
AIGC
  • MergeVQ: A Unified Framework for Visual Generation and Representation with Disentangled Token Merging and Quantization

    Paper • 2504.00999 • Published Apr 1, 2025 • 98
  • Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

    Paper • 2409.12191 • Published Sep 18, 2024 • 80

Papers 39

arxiv:2607.04033
arxiv:2606.07454
arxiv:2605.21195
arxiv:2605.15963
View 39 papers

models 1

Lupin1998/DeepSeek-R1-Distill-Qwen-1.5B-GRPO

Updated Aug 29, 2025

datasets 0

None public yet
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs