Rajdeep Haldar
rhaldar97
AI & ML interests
Adversarial Robustness
Computer Vision
LLM Human Alignment
Recent Activity
updated a model 1 day ago
phaseMHD/PHASE-Turbulence-MR submitted a paper 8 months ago
f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment liked a dataset over 1 year ago
argilla/distilabel-math-preference-dpo