🔄 In a Training Loop
Md Ismail Sojal
0xSojalSec
AI & ML interests
Ai-ML-RL-LLM-MoE-SLMs-MLMs-LAMs-VLMs
Re-searcher/ post-training / reasoning models / RAG /
Recent Activity
liked a model about 4 hours ago
XiaomiMiMo/MiMo-V2.6-Flash-RL liked a Space 9 days ago
Pepe104/MiniMax-H3-Turbo-Lora-UNCENSORED upvoted an article 16 days ago
Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps