Qwen3.8 Collection Qwen3.8 Unsloth quants including Qwen3.8-27B! Run and train Qwen3.8 with the Unsloth Desktop app. • 9 items • Updated 29 days ago • 71
Quiet-STaR: Language Models Can Teach Themselves to Think Before Speaking Paper • 2403.09629 • Published Mar 14, 2024 • 81
Breaking the Modality Barrier: Universal Embedding Learning with Multimodal LLMs Paper • 2504.17432 • Published Apr 24, 2025 • 41
Helping or Herding? Reward Model Ensembles Mitigate but do not Eliminate Reward Hacking Paper • 2312.09244 • Published Dec 14, 2023 • 9