Mustafa Mert Çelikok
Assistant Professor in Multi-Agent Reinforcement Learning, University of Southern Denmark
Department of Mathematics and Computer Science
University of Southern Denmark
Odense, Denmark
I am a tenure-track Assistant Professor at the University of Southern Denmark, in the Data Science Section of the Department of Mathematics and Computer Science (IMADA).
My research focuses on multi-agent reinforcement learning (MARL) and cooperative AI. I study how autonomous agents learn to cooperate with unfamiliar partners, including humans and other AI agents, particularly in complex and changing environments. My broader interests include reinforcement learning, multi-agent learning, continual learning, and human-AI interaction.
Before joining SDU, I was a postdoctoral researcher at Delft University of Technology, where I worked on multi-agent reinforcement learning for collaborative AI systems as part of the Dutch Hybrid Intelligence Centre. I completed my PhD in Computer Science at Aalto University under the supervision of Professor Samuel Kaski, and was a member of the first cohort of the European Laboratory for Learning and Intelligent System’s (ELLIS) PhD programme. I am currently an ELLIS member.
news
| Jun 01, 2026 | Our paper “Distributional Active Inference” was accepted to ICML 2026. Also co-organizing the ICLR 2026 Workshop on Multi-Agent Learning and Its Opportunities in the Era of Generative AI. |
|---|---|
| Jan 01, 2026 | Became Co-director of the Danish Pioneer Centre for AI Programme on Learning and Optimal Control in Dynamical Systems (2026–2028). |
| Sep 01, 2025 | Our paper “Value Improved Actor Critic Algorithms” was accepted to NeurIPS 2025. |
| Sep 01, 2025 | Started as a tenure-track Assistant Professor at the University of Southern Denmark, Data Science Section, Department of Mathematics and Computer Science (IMADA). |
| May 01, 2025 | Our paper “On the Complexity of Learning to Cooperate with Populations of Socially Rational Agents” was accepted to AAMAS 2025. |
selected publications
- Distributional Active InferenceIn Forty-third International Conference on Machine Learning, 2026
- The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent MisalignmentarXiv preprint arXiv:2606.10747, 2026
- On the Complexity of Learning to Cooperate with Populations of Socially Rational AgentsIn Proceedings of the 24th International Conference on Autonomous Agents and Multiagent Systems, 2025
- Value Improved Actor Critic AlgorithmsIn Advances in Neural Information Processing Systems (NeurIPS), 2025
- Uncoupled Learning of Differential Stackelberg Equilibria with CommitmentsIn Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems, 2024
- Best-Response Bayesian Reinforcement Learning with Bayes-adaptive POMDPs for CentaursIn Proceedings of the 21st International Conference on Autonomous Agents and Multiagent Systems, 2022