Karthik Iyer
kiyer97
ยท
AI & ML interests
Reinforcement learning, reward modeling, RLHF, policy optimization, offline RL
Recent Activity
upvoted a paper about 6 hours ago
Flash-dLLM: IO-Aware KV Caching and Parallel Decoding for Fast, Memory-Efficient Diffusion LLMs liked a dataset 1 day ago
youinwww/reinforcement_learning upvoted a paper 1 day ago
Towards Full Pipeline FP8 Reinforcement Learning for LLMsOrganizations
None yet