Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

alibaba

company
https://www.alibabagroup.com/
https://github.com/alibaba
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

xiaoying0505  submitted a paper about 10 hours ago
Business Arena: Benchmarking LLM Agents in a Realistic Marketplace
KhanCold  submitted a paper 7 days ago
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations
KuanCao320  published a dataset 3 months ago
alibabagroup/OmniDoc-TokenBench
View all activity

Papers

Business Arena: Benchmarking LLM Agents in a Realistic Marketplace

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

View all Papers

lzj's profile picture jinran's profile picture boxin's profile picture Wensheng's profile picture li's profile picture FengYi's profile picture Di Yang's profile picture
alibabagroup 's papers 6
Submitted by
Xiaoying Xing
12

Business Arena: Benchmarking LLM Agents in a Realistic Marketplace

alibabagroup alibaba
9 5
Submitted by
Qiming Shi
96

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

alibabagroup alibaba
92 4
Submitted by
Yang Li (SJTU & SII)
8

How Does Reasoning Flow? Tracing Attention-Induced Information Flow for Targeted RL in LLMs

alibabagroup alibaba
2
Submitted by
Ningyu Zhang
13

Why Steering Works: Toward a Unified View of Language Model Parameter Dynamics

alibabagroup alibaba
3
Submitted by
Yang Li (SJTU & SII)
59

Attention Illuminates LLM Reasoning: The Preplan-and-Anchor Rhythm Enables Fine-Grained Policy Optimization

alibabagroup alibaba
2
Submitted by
Jiaming Wang
68

Winning the Pruning Gamble: A Unified Approach to Joint Sample and Token Pruning for Efficient Supervised Fine-Tuning

alibabagroup alibaba
3 3
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs