view article Article Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers tomaarsen • 27 days ago • 144
Autodata: An agentic data scientist to create high quality synthetic data Paper • 2606.25996 • Published Jun 24 • 21
Running 122 Unlocking On-Policy Distillation for Any Model Family 📝 122 Explore on-policy distillation visualization for any model
Running Featured 93 Distilling 100B+ Models 40x Faster with TRL 📝 93 TRL distillation for 100B+ teachers, 40x faster
unsloth/Mistral-Small-3.2-24B-Instruct-2506-unsloth-bnb-4bit Image-Text-to-Text • 25B • Updated Jun 23, 2025 • 2k • 13
Enhancing Retrieval for ESGLLM via ESG-CID -- A Disclosure Content Index Finetuning Dataset for Mapping GRI and ESRS Paper • 2503.10674 • Published Mar 10, 2025 • 3
Beyond Semantic Similarity: Rethinking Retrieval for Agentic Search via Direct Corpus Interaction Paper • 2605.05242 • Published May 3 • 126
view article Article 🛡️ Nemotron PII: Synthesized Data for Privacy-Preserving AI nvidia • Oct 28, 2025 • 37
Next-Embedding Prediction Makes Strong Vision Learners Paper • 2512.16922 • Published Dec 18, 2025 • 91
LLM-JEPA: Large Language Models Meet Joint Embedding Predictive Architectures Paper • 2509.14252 • Published Sep 11, 2025 • 10