Toolkits (DUPLICATE them, never use the public ones) A set of tools to enable finetuning, evaluations, prototyping, agentic workflows etc. ATTENTION: ALWAYS DUPLICATE THESE SPACES ON OUR INFRA!!! Running 135 AutoTrain Advanced π 135 Create powerful AI models without code Runtime error Agents 40 LLM Merge Adapter π’ 40 Runtime error Agents Featured 288 mergekit-gui π 288 Merge AI models using a YAML configuration file
Runtime error Agents Featured 288 mergekit-gui π 288 Merge AI models using a YAML configuration file
Benchmarks Most commonly used leaderboards to check model capabilities Running on CPU Upgrade 14.1k Open LLM Leaderboard π 14.1k Track, rank and evaluate open LLMs and chatbots Running Featured 488 LLM Performance Leaderboard π¨ 488 View the LLM leaderboard rankings Running 5.01k Arena Leaderboard π 5.01k View the LMArena model performance leaderboard Running on CPU Upgrade 7.7k MTEB Leaderboard π 7.7k Embedding Leaderboard
Running on CPU Upgrade 14.1k Open LLM Leaderboard π 14.1k Track, rank and evaluate open LLMs and chatbots
Toolkits (DUPLICATE them, never use the public ones) A set of tools to enable finetuning, evaluations, prototyping, agentic workflows etc. ATTENTION: ALWAYS DUPLICATE THESE SPACES ON OUR INFRA!!! Running 135 AutoTrain Advanced π 135 Create powerful AI models without code Runtime error Agents 40 LLM Merge Adapter π’ 40 Runtime error Agents Featured 288 mergekit-gui π 288 Merge AI models using a YAML configuration file
Runtime error Agents Featured 288 mergekit-gui π 288 Merge AI models using a YAML configuration file
Benchmarks Most commonly used leaderboards to check model capabilities Running on CPU Upgrade 14.1k Open LLM Leaderboard π 14.1k Track, rank and evaluate open LLMs and chatbots Running Featured 488 LLM Performance Leaderboard π¨ 488 View the LLM leaderboard rankings Running 5.01k Arena Leaderboard π 5.01k View the LMArena model performance leaderboard Running on CPU Upgrade 7.7k MTEB Leaderboard π 7.7k Embedding Leaderboard
Running on CPU Upgrade 14.1k Open LLM Leaderboard π 14.1k Track, rank and evaluate open LLMs and chatbots