arxiv:2509.17978
Antoni Guasch
antoniguasch
·
AI & ML interests
Deterministic evaluation of LLM agents, multi-agent reasoning, theory of mind, game theory, falsifiable benchmarks (iXentBench)
Recent Activity
updated a collection about 23 hours ago
iXentBench · Falsifiable reasoning updated a collection about 23 hours ago
iXentBench · Falsifiable reasoning updated a collection about 23 hours ago
iXentBench · Falsifiable reasoning