ggml-org/Qwen3-Reranker-0.6B-Q8_0-GGUF Text Ranking • 0.6B • Updated Oct 3, 2025 • 50.4k • 28
Alibaba-NLP/gte-reranker-modernbert-base Text Ranking • 0.1B • Updated Jul 4, 2025 • 2.22M • 97
Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 519
XQuant: Breaking the Memory Wall for LLM Inference with KV Cache Rematerialization Paper • 2508.10395 • Published Aug 14, 2025 • 42
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL Paper • 2508.13167 • Published Aug 6, 2025 • 129