Distributed Control by Lagrangian Steepest Descent
David H. Wolpert, Stefan Bieniawski
Abstract
Often adaptive, distributed control can be viewed as an iterated game between independent players. The coupling between the players' mixed strategies, arising as the system evolves from one instant to the next, is determined by the system designer. Information theory tells us that the most likely joint strategy of the players, given a value of the expectation of the overall control objective function, is the minimizer of a Lagrangian function of the joint strategy. So the goal of the system designer is to speed evolution of the joint strategy to that Lagrangian minimizing point, lower the expectated value of the control objective function, and repeat. Here we elaborate the theory of algorithms that do this using local descent procedures, and that thereby achieve efficient, adaptive, distributed control.
Create a lesson
Related papers
Reputation as Community Memory for the Agentic Web
Ryan Chard, Gus Ellerm, Alexander Brace et al.
CC-OPI: Online Distributed Task Allocation for UAV Swarms under Communication Constraints
Biao Liu, Tong Zhang
Social Laws for Multi-agent Coordination in Stochastic Environments
Rolando Fernandez, Caleb Probine, Tyler Lee et al.
ABM-SIRTEM: A Hybrid Agent-Based and Epidemiological Model for Pandemic Response
Sheryl Paul, Samuel Williams, Preetom K. Biswas et al.
Message capacity and claim wording set the transition points of collective truth-finding in language-model networks
Makoto Fukushima
Agentic Societies Need a Social Harness
Tapan Chugh, Vidushi Singh, Krish Jain et al.