view article Article Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident +2 hlarcher, XciD, raphael-gl, chris-rannou • 14 days ago • 452
GigaChat 3.5 Collection GigaChat 3.5 is a large-scale Mixture-of-Experts (MoE) Hybrid model with 432B total parameters • 5 items • Updated 27 days ago • 19
ReAligned-Qwen3.5 Collection Lazarus AI's ReAligned finetune of Qwen 3.5 alters the alignment of the model, eliminating unwanted behaviors like propaganda, lying, & gaslighting. • 18 items • Updated Jun 3 • 5
Granite 4.1 Language Models Collection Efficient language models for multilingual generation, coding, RAG, and AI assistant workflows. • 6 items • Updated Apr 29 • 63
view article Article How Long Prompts Block Other Requests - Optimizing LLM Performance tngtech • Jun 12, 2025 • 14
view article Article Prefill and Decode for Concurrent Requests - Optimizing LLM Performance tngtech • Apr 16, 2025 • 87
EXAONE 4.5 Collection LG's First Open-Weight Vision-Language Model for Industrial Intelligence • 5 items • Updated Apr 22 • 47
Gemma 4 Collection Gemma 4 is Google's new model family including including E2B, E4B, 26B-A4B, and 31B. • 43 items • Updated 6 days ago • 254
Nemotron-Pre-Training-Datasets Collection Large scale pre-training datasets used in the Nemotron family of models. • 15 items • Updated 11 days ago • 183
NVIDIA Nemotron v3 Collection Open, Production-ready Enterprise Models • 23 items • Updated 24 days ago • 347