ChronoLens: Measuring Language Change Across Time, Languages, and Linguistic Levels Paper • 2608.03507 • Published 4 days ago • 3
ChronoLens: Measuring Language Change Across Time, Languages, and Linguistic Levels Paper • 2608.03507 • Published 4 days ago • 3
A Frozen Pixel-Space Diffusion Model Can Guide Itself with Its Own Samples Paper • 2607.29122 • Published 8 days ago • 4
WorldDiT: A Unified Diffusion Architecture for World and Action Modeling Paper • 2607.23909 • Published 12 days ago • 7
view post Post 770 How does billing work for Hugging Face Inference Providers? If I already have Hugging Face GPU credits, can I use them, or do I need to add a credit card and pay separately? See translation 3 replies · ➕ 3 3 + Reply
MonkeyOCRv2: A Visual-Text Foundation Model for Document AI Paper • 2607.11562 • Published 26 days ago • 77
Xiaomi-Robotics-U0: Unified Embodied Synthesis with World Foundation Model Paper • 2607.11643 • Published 26 days ago • 45
Single-Rollout Asynchronous Optimization for Agentic Reinforcement Learning Paper • 2607.07508 • Published about 1 month ago • 29
Sparse Delta Memory: Scaling the State of Linear RNNs through Sparsity Paper • 2607.07386 • Published about 1 month ago • 13
HunyuanOCR-1.5: Making Lightweight OCR VLMs Faster and Better Paper • 2607.04884 • Published Jul 6 • 9
Unified Audio Intelligence Without Regressing on Text Intelligence Paper • 2607.05196 • Published Jul 6 • 23
Duration Aware Scheduling for ASR Serving Under Workload Drift Paper • 2603.11273 • Published Mar 11 • 3
Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models Paper • 2606.03748 • Published Jun 2 • 22
Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025 Paper • 2606.02255 • Published Jun 1
Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025 Paper • 2606.02255 • Published Jun 1
Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini Paper • 2605.27295 • Published May 26 • 23