Swift 1.5 27B Collection Swift 1.5 on Qwen3.8-27B: stronger than Swift 1.0 on agentic and coding tasks, with fewer thinking tokens. BF16 weights and every quant. • 11 items • Updated about 19 hours ago • 10
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 Paper • 2608.27370 • Published 29 days ago • 40
Muse Glimmer Collection Muse Glimmer 30B: multimodal agentic model for local deployment. BF16 weights, GGUF k-quants, ExecuTorch builds, DFlash drafter. • 4 items • Updated Aug 10 • 109
Granite 4.2 Language Models Collection Efficient reasoning and thinking language models for multilingual generation, coding, and AI assistant workflows. • 24 items • Updated 11 days ago • 41
MoE-SpAc: Efficient MoE Inference Based on Speculative Activation Utility in Heterogeneous Edge Scenarios Paper • 2603.09983 • Published Feb 12 • 4
FreeToken: Efficient Edge-Native MoE Serving with Bandwidth-Adaptive Execution Paper • 2608.16157 • Published Aug 17 • 112
From RLVR to RLSVR: Task Transformation Induces Self-Verifiable Rewards for Open-Ended LLM Self-Improvement Paper • 2607.23802 • Published Jul 26 • 96
Agents-A1 Collection Agents-A1 is a Long-horizon Agentic Model that reaches trillion-parameter-level performance by scaling the agent horizon. • 12 items • Updated Jul 16 • 44
DSpark: Confidence-Scheduled Speculative Decoding with Semi-Autoregressive Generation Paper • 2607.05147 • Published Jul 6 • 51
Laguna XS 2.1 Collection Designed for agentic coding and long-horizon work on a local machine. Licensed under OpenMDW-1.1. • 9 items • Updated Jul 2 • 25
Tmax Collection Data and models associated with "Tmax: A simple recipe for terminal agents". paper: https://arxiv.org/abs/2606.23321 • 23 items • Updated Jun 23 • 20
DFlash Collection Block Diffusion for Flash Speculative Decoding • 23 items • Updated 13 days ago • 156