NEW Articles from Team or Enterprise organizations will get promoted to the main section. Magpie Speech — Applying an LLM Data Synthesis Method to an LLM-Based TTS Model to Synthesize a Speech Dataset
Aratako
• • 15
GitChameleon 2.0: Evaluating AI Code Generation Against Python Library Version Incompatibilities
cabbage972
• • 1
Automating Airline IROPS Re-accommodation with KaibanJS
darielnoel
• Byte Pair Encoding (BPE) — From Banana to Bandana
sweatSmile
• • 1
High-Confidence NLP on CPU: A Hybrid BERT and Phrase-List Architecture
tlogandesigns
• Decoding the Shift and Diffusion Models Training Like Qwen Image, FLUX, SDXL, and More
MonsterMMORPG
• • 2
EMOTRON-3B 🤬🤢😨😀😐😭😲 - Experiments in GRPO
dleemiller
• The AI Evaluation Chart Crisis
andrewtran117
• • 4
Announcing the Synthetic Online Conversations Dataset (SOC)
marcodsn
• • 13
How I Built 7 Custom Gradio Components in Just 12 Days!
elismasilva
• • 7
<p style="text-align:center;"> Bridging the Gap: Making Robotics Feel Like Machine Learning </p>
hba123
• • 12
How to Run a Hugging Face Model in JAX (Part 3)
qihqi
• • 10
NVIDIA Releases 3 Million Sample Dataset for OCR, Visual Question Answering, and Captioning Tasks
nvidia
• • 76
RynnVLA-001: Using Human Demonstrations to Improve Robot Manipulation
Alibaba-DAMO-Academy
• • 28
Luth: Efficient French Specialization for Small Language Models
MaxLSB
• • 21
Re-understanding KL Approximation from an RL-for-LLM Lens: Notes on “Approximating KL Divergence”
NormalUhr
• • 15
Introducing : 🤏🏻🏭SmolFactory
AIG 1.0: Revolutionary AI-Optimized Image Format with Multi-Center Radial Compression
AI Voice Actor: Personalized Voice and Conversation Pattern Replication
sadpig70
• • 1
OpenAI just dropped two massive open-weight models — *but how do we separate the reality from the hype?*
stefanwebb
• • 11