NEW Articles from Team or Enterprise organizations will get promoted to the main section. Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community
ResterChed
• • 40
J-Space: Yet Another LLM Mind Reader?
dlouapre
• • 31
Aether-7B-5Attn: A 100% Open-Source Sovereign Foundation Model — and a Controlled Experiment in Heterogeneous Attention
FINAL-Bench
• • 19
One Adapter, Both Modalities: Field Notes from Building and Serving a Multimodal Reranker
lightonai
• • 15
KV Caching Explained: Optimizing Transformer Inference Efficiency
not-lain
• • 374
The state-of-the-art in open-source AI for Swiss legal tasks
joelniklaus
• • 5
When will language models be good enough?
Uncensor any LLM with abliteration
mlabonne
• • 876
The Best Open Source and Open-Weight LLM Models to Run Locally in 2026
daya-shankar
• • 10
Data for Agents
nvidia
• • 25
Distillation in 2026 (so far): which frontier models use it and how
sergiopaniego
• • 18
VKUE: No GPU? Runs Anyway — a 34.7B Reasoner on a Laptop and on Bare CPU
FINAL-Bench
• • 17
Giving AI Agents 3D Bodies, Real Jobs, and Wallets on three.ws
three-ws
• • 19
Deploy GLM-5.2-FP8 as your open, frontier-level agent
juanjucm
• • 7
makeMoE: Implement a Sparse Mixture of Experts Language Model from Scratch
AviSoori1x
• • 124
Code a simple RAG from scratch
ngxson
• • 363
Small Language Models (SLM): A Comprehensive Overview
jjokah
• • 167
Arabic LLM Models
silma-ai
• • 49