·
AI & ML interests
NLP
Organizations
upvoted a paper about 1 year ago upvoted an article about 1 year ago view article DeepSeek-R1 Dissection: Understanding PPO & GRPO Without Any Prior Reinforcement Learning Knowledge
NormalUhr
• • 297
view article Deploy LLMs with Hugging Face Inference Endpoints
philschmid
• • 18