Leo Raphael Rodrigues's picture

Leo Raphael Rodrigues

LeoRodrigues05

AI & ML interests

Mechanistic Interpretability, AI Safety, Alignment, Generative Recommendation Systems, Machine Translation, AI Trustworthiness, AI Policy.

Recent Activity

published a dataset about 2 months ago
LeoRodrigues05/safety-cot-interventions-v5
updated a dataset about 2 months ago
LeoRodrigues05/safety-cot-interventions-v5
View all activity

Organizations

None yet