Leo Raphael Rodrigues
LeoRodrigues05
AI & ML interests
Mechanistic Interpretability, AI Safety, Alignment, Generative Recommendation Systems, Machine Translation, AI Trustworthiness, AI Policy.
Recent Activity
published a dataset about 2 months ago
LeoRodrigues05/safety-cot-interventions-v5 updated a dataset about 2 months ago
LeoRodrigues05/safety-cot-interventions-v5Organizations
None yet