Skip to main content

Online Learning for Strategic Resource Allocation: From Policies to Mechanisms

Vincent Leon

Abstract:
Modern multi-agent systems increasingly operate in environments where agents are strategic, information is incomplete, and system dynamics evolve over time. This talk presents recent frameworks that address this challenge by developing learning-based methods for dynamic games, unified by the broader theme of strategic resource allocation in evolving and uncertain environments. The two frameworks progress from learning strategies in a two-player game to designing mechanisms for multi-agent systems, with performance measured in both cases by regret against an offline benchmark. The first studies a dynamic Colonel Blotto game in which a budget-constrained learner repeatedly allocates troops across battlefields against an opponent with an unknown strategy. Combining Lagrangian relaxation for the budget constraint with an efficient combinatorial bandits algorithm that exploits the graph structure of the game yields regret that is sublinear in the time horizon and polynomial in the problem parameters. The second addresses dynamic mechanism design for sequential auctions modeled as an infinite-horizon average-reward Markov decision process with unknown transition kernel and reward functions. The Vickrey-Clarke-Groves mechanism is extended to this dynamic setting, together with a reinforcement learning algorithm that achieves sublinear regret while approximately preserving efficiency, truthfulness, and individual rationality.

Time: Fri 2026-10-02 11.00 - 12.00

Location: KTH, Seminar room 3721

Participating: Vincent Leon

Export to calendar

Bio:

Vincent Leon is a postdoctoral researcher at the Department of Decision and Control Systems, KTH Royal Institute of Technology, supervised by Professors Henrik Sandberg and Karl Henrik Johansson. He received his Ph.D. in Industrial Engineering from the University of Illinois Urbana-Champaign in May 2026 and his B.Eng. in Civil Engineering with First Class Honours from The University of Hong Kong. From August 2024 to January 2025, he was a visiting scholar at the Singapore University of Technology and Design. His research interests include equilibrium analysis of games, learning algorithm design for multi-agent systems, online control of networked systems, and their applications to cyber-physical systems. 

Page responsible:Per Enqvist
Belongs to: Stockholm Mathematics Centre
Last changed: Sep 29, 2026