AI & ML interests

Deterministic, falsifiable evaluation of LLM agents: multi-agent reasoning, theory of mind, game theory and causal spatial reasoning (iXentBench)

Recent Activity

antoniguasch  updated a collection about 6 hours ago
iXentBench · Falsifiable reasoning
antoniguasch  updated a collection about 6 hours ago
iXentBench · Falsifiable reasoning
antoniguasch  updated a collection about 6 hours ago
iXentBench · Falsifiable reasoning
View all activity