Publications
- A benchmark of expert-level academic questions to assess AI capabilities (Humanity’s Last Exam)Center for AI Safety, Scale AI, HLE Contributors ConsortiumContributor to the expert-level AI benchmark published in Nature.
- SMoSE: Sparse Mixture of Shallow Experts for Interpretable Reinforcement Learning in Continuous Control TasksM. Vincze, L. Ferrarotti, L. L. Custode, B. Lepri, G. IaccaFirst-author work on sparse, interpretable reinforcement-learning policies with only 108–672 active actor parameters.@ AAAI · 2025