Replacing Large Language Models with Jev Decision Models for Low-Latency Edge Service Orchestration Paper • 2609.22753 • Published 10 days ago • 5
FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observations Paper • 2609.20817 • Published 19 days ago • 40
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16 Text Generation • 32B • Updated Aug 24 • 572k • • 227
OasisKV: Scaling In-Decode KV Cache Beyond HBM with Lookahead Sparse Prefetching Paper • 2608.08097 • Published Aug 8 • 26
DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF Image-Text-to-Text • 9B • Updated 4 days ago • 1.96M • 943
ScrambleToolBench: Agents Search Exhaustively Even When Their Own Map Points to the Next Step Paper • 2608.02358 • Published Aug 3 • 12