NEW Articles from Team or Enterprise organizations will get promoted to the main section. OCR Processing and Text in Image Analysis with DeepSeek Janus-1.3B
PandorAI1995
• • 3
Navigating Korean LLM Research #1: Models
amphora
• • 26
Aria: First Open Multimodal Native MoE Model
RhymesAI
• • 9
Allegro: Advanced Video Generation Model
RhymesAI
• • 59
🇮🇹🇯🇵🇧🇷 Generating multilingual instruction datasets with Magpie 🐦⬛
anakin87
• • 20
Advanced Flux Dreambooth LoRA Training with 🧨 diffusers
linoyts
• • 42
Turn your newsletters into a Podcast with NotebookLM
MedEmbed: Fine-Tuned Embedding Models for Medical / Clinical IR
abhinand
• • 54
AI is turning nuclear: a review
as-cle-bert
• • 10
LLM ChatBots 3.0: Merging LLMs with Dynamic UI Elements
airabbitX
• • 7
Occam’s Sheath: A Simpler Approach to AI Safety Guardrails
daniel-de-leon
• • 8
Mamba Out
rwightman
• • 11
OCR Processing and Text in Image Analysis with Florence-2-base and Qwen2-VL-2B
PandorAI1995
• • 17
EmbeddingAlign RAG: Boosting QA Systems
ColFlor: Towards BERT-Size Vision-Language Document Retrieval Models
ahmed-masry
• • 22
Unlocking the Power of Large Language Models (LLMs) for Business Applications
ImranzamanML
• • 1
How to build a custom text classifier without days of human labeling
sdiazlor
• • 57
Organizing a Privacy-preserving Hackathon
binoua
• • 9
Image Search with Text Prompt
tonyassi
• • 6
bismillah
amrvictim
• • 1