Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
zhaohanlin's picture

zhaohanlin

ultrazhl
7
·

AI & ML interests

None yet

Organizations

None yet

upvoted an article 3 months ago
view article
Article

Is using a validation set useful for end-to-end learning in robotics?

m1b
•
Dec 1, 2024
• 17
upvoted a paper 3 months ago

Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

Paper • 2605.30280 • Published May 28 • 146
upvoted a paper about 1 year ago

UI-Venus Technical Report: Building High-performance UI Agents with RFT

Paper • 2508.10833 • Published Aug 14, 2025 • 46
upvoted 2 collections almost 2 years ago

Molmo

Collection
Artifacts for open multimodal language models. • 5 items • Updated Dec 23, 2025 • 310

UI Agent

Collection
a collection of algorithmic agents for user interfaces/interactions, program synthesis, and robotics • 509 items • Updated 24 days ago • 69
upvoted a paper almost 2 years ago

MM1.5: Methods, Analysis & Insights from Multimodal LLM Fine-tuning

Paper • 2409.20566 • Published Sep 30, 2024 • 54
upvoted a paper about 2 years ago

VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents

Paper • 2408.06327 • Published Aug 12, 2024 • 17
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs