arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:1810.00135v2 [eess.SY] 22 Dec 2018

A Simple Framework for Stability Analysis of State-Dependent Networks of Heterogeneous Agents

S. Rasoul Etesami Affiliation: Department of Industrial and Enterprise Systems Engineering
University of Illinois at Urbana-Champaign, Urbana, IL 61801
Email: etesami1@illinois.edu
Abstract

Stability and analysis of multi-agent network systems with state-dependent switching typologies have been a fundamental and longstanding challenge in control, social sciences, and many other related fields. These already complex systems become further complicated once one accounts for asymmetry or heterogeneity of the underlying agents/dynamics. Despite extensive progress in analysis of conventional networked decision systems where the network evolution and state dynamics are driven by independent or weakly coupled processes, most of the existing results fail to address multi-agent systems where the network and state dynamics are highly coupled and evolve based on status of heterogeneous agents. Motivated by numerous applications of such dynamics in social sciences, in this paper we provide a new direction toward analysis of dynamic networks of heterogeneous agents under complex time-varying environments. As a result we show how Lyapunov stability and convergence of several challenging problems from opinion dynamics can be established using a simple application of our framework. Moreover, we introduce a new class of asymmetric opinion dynamics, namely nearest neighbor dynamics, and show how our approach can be used to analyze their behavior. In particular, we extend our results to game-theoretic settings and provide new insights toward analysis of complex networked multi-agent systems using exciting field of sequential optimization.

Index Terms: 
Lyapunov stability; multi-agent decision systems; state-dependent dynamics; switching network dynamics; opinion dynamics, block coordinate descent, game theory.

I Introduction

Researchers in a number of fields are currently finding a variety of applications for complex networks, and distributed multi-agent network systems are currently the focal point of many new applications. Such applications relate to the growing popularity of online social networks, the analysis of large network data sets, the problems that arise from interactions among agents in complex networks such as formation control, smart grids, political, economic, and biological systems, and the expansion of power and wireless networks in our daily life.

There is ample evidence that decision making is often guided by heterogeneous agents interacting in a complex time-varying environment. Perhaps one simple example is when a set of heterogeneous robots with different communication capabilities want to rendezvous despite the fact that they are simultaneously moving and yet have to maintain communication connectivity. However, it is often observed that in practice the behavior of multi-agent decision systems under static symmetric/homogeneous setting is fundamentally different from its dynamic asymmetric/heterogeneous counterpart. Unlike the static homogeneous case, often any comprehensive analysis of dynamic heterogeneous multi-agent systems is quite challenging, particularly in dynamic environments, and this class of problems has eluded researchers for many years. New ideas and methodologies need to be developed to address such shortcomings in which any progress can impact numerous applications in variety of domains including opinion formation in social networks, formation control, cyber-physical security, dynamic clustering, among many others. Since better understanding of such complex systems will allow us to design novel or perhaps fundamentally different mechanisms, in this work we take some initial steps toward extending the existing results on multi-agent network systems from the static homogeneous setting to highly dynamic and heterogeneous environments.

I-A Motivation

There are many motivating examples of relationships in political, social, and engineering applications, which are governed by complex networks of heterogeneous agents. Agents may possibly belong to multiple groups and be connected by multi-layered networks. The networks can also be dynamic in the sense that they can vary over time depending on the agents’ status. As a few illustrative examples one can consider (Figure 1):

  • In social networks, there are often clear affinities among people based on shared political or cultural beliefs. However, on specific issues, alliances form among people from different groups. Almost every congressional vote provides an example of this phenomenon, where some representatives break away from their respective parties to vote with the other party.

  • In formation control a basic task is to design a distributed protocol so that a set of robots collectively form a certain structure. Robots have different communication capabilities and can only communicate with those in their local neighborhood. Consequently, the communication network between them may vary depending on their relative distances from each other.

  • Relationships between countries in the Middle East, and their ties to the US and Russia are nuanced, with affiliations changing over time, and depending on context. Similarly, relationships among terrorist organizations, such as ISIS, Al Qaeda, Taliban, and LeT, often change due to their battle for supremacy, and their fight for the allegiance of their extremist followers.

  • In opinion systems such as political election polls, individuals initially have different opinions about a certain topic/candidate. Individuals frequently interact and become friend/unfriend depending on how close their opinions are from each other. In particular, through such interactions their opinions gradually form and a collective opinion eventually emerges.

Refer to caption
Fig. 1: The left figure depicts a complex network of global conflicts between countries (agents) due to their political or religious ties. The middle figure represents a social network of individuals’ opinions for buying different cell-phone products. The right figure shows a specific network formation taken by UAVs and drones. In all the figures the networks may dynamically change over time and the agents can be heterogeneous (e.g., drones and UAVs have different communication capabilities).

Motivated by the above, and many other similar examples, our objective in this work is to provide a simple framework to understand and analyze the behavior of networks of heterogeneous agents with a rich dynamic network structure which may evolve or vary based on agents’ states. To this end, we provide new connections between analysis of multi-agent network systems and some developed methods in the mature fields of successive optimization and game theory. Utilizing such connections, we establish Lyapunov stability and convergence of several classes of heterogeneous multi-agent network systems with switching state-dependent dynamics.

I-B Literature Review and Organization

There has been a rich body of literature on analysis of distributed multi-agent network systems, mainly from the static point of view, in which a set of agents iteratively interact over a fixed communication network so as to achieve a certain goal such as consensus or optimizing a global objective function. The classical models of Degroot [1] and Friedkin& Johnsen [2] in social science are two special types of such systems. As the literature on this area is quite vast, we refer interested readers for an overview to [3] and [4]. However, a crucial assumption in almost all of these works is that the communication network among the agents is fixed. Often these results can be generalized to time-varying networks by assuming a certain “independency” between the network process and the state dynamics. For instance, one of the commonly used assumptions is that the network dynamics are governed by an exogenous process which is uncoupled from the state dynamics [5, 6, 7, 8, 9, 10]. A generalization of this idea is to consider multi-agent network dynamics where the network and state dynamics are allowed to be coupled, however a certain condition such as network connectivity or communication symmetry must be satisfied at each time instance [11, 12, 13, 14]. A further extension of these results is the line of works in [15, 16] which shows that any sequence of stochastic matrices viewed as update matrices of multi-agent network dynamics admits a “Lyapunov” type function. Unfortunately, such Lyapunov functions are based on another so-called adjoint dynamics which depend on the future instances of the original dynamics. Thus, unless there is a strong inherent property in the underlying stochastic matrices [17, 14], it seems unlikely to leverage the co-evolution of original and adjoint dynamics so as to establish meaningful convergence results.

While the above techniques can properly address a large class of multi-agent network systems, there are still many examples which do not fit into any of the aforementioned frameworks (or the application of the above techniques provides very poor results on the behavior of the multi-agent system). One of the main reasons is that most of the developed frameworks for analysis of multi-agent network systems aim to isolate the evolution of network dynamics from those of state dynamics in which case one derives a conclusion about local time instances (e.g. one time slot or a window of time instances) and then generalizes this behavior to the entire trajectories. But the main question here is that what if the network and state dynamics are completely determined endogenously so that no local property can be assumed or checked a priori? In other words, if we do not take into account the actual correlation between network and state dynamics (or how they evolve in terms of each other), it seems hopeless to have a good understanding of the overall behavior of the dynamics. This shortcoming is even more pronounced once we take into account agents heterogeneity or asymmetric communication among them. In this work we show that despite these challenges it is still possible to capture the co-evolution of network-state dynamics for several (and perhaps many other) multi-agent network dynamics even under asymmetric or heterogeneous environment. While the evolution of state dynamics are commonly captured by difference/differential equations, one of the key novelties of our work is to explicitly derive analogous equations for network dynamics by explicitly introducing network variables and couple them with the state variables. This allows us to derive network topology from state dynamics and vice versa and even incorporate other constraints into the system behavior.

The rest of the paper is organized as follows: In Section II we provide a general framework which we use frequently throughout the paper. In Section III we apply this framework to a well-known state-dependent switching dynamics from social sciences, known as Hegselmann-Krause (HK) opinion dynamics, and extend it to an edge-heterogeneous setting with additional communication constraints. In particular, we show how this approach can handle up to some extent asymmetric communication among the agents in the HK model. In section IV we introduce a new class of asymmetric opinion dynamics, namely nearest neighbor dynamics, and show that our framework can also be applied suitably to these models despite the fact that the underlying communication networks suffer from asymmetry. In Section V, we provide an application of our framework to game-theoretic settings and show how the existing results in game theory can be leveraged to analyze heterogeneous multi-agent network dynamics. We conclude the paper by identifying some future directions of research in Section VI.

Notations

We adopt the following notations throughput the paper: We use bold symbols for vectors. Given a vector 𝒗\boldsymbol{v} we let diag(𝒗)diag(\boldsymbol{v}) be a diagonal matrix with vector 𝒗\boldsymbol{v} as its diagonal elements and zero, everywhere else. Moreover, we denote the transpose of 𝒗\boldsymbol{v} by 𝒗T\boldsymbol{v}^{T}. For a positive integer n+n\in\mathbb{Z}^{+} we set [n]:={1,2,,n}[n]:=\{1,2,\ldots,n\}. We let 𝟏\boldsymbol{1} be a column vector of all ones. Given a positive-definite matrix QQ and a vector 𝒗n\boldsymbol{v}\in\mathbb{R}^{n}, we let 𝒗Q2=𝒗TQ𝒗\|\boldsymbol{v}\|^{2}_{Q}=\boldsymbol{v}^{T}Q\boldsymbol{v}, and 𝒗\|\boldsymbol{v}\| to be the Euclidean norm of 𝒗\boldsymbol{v}. Given a real number xx\in\mathbb{R} we set (x):=min{0,x}(x)^{-}:=\min\{0,x\}. We denote the cardinality of a finite set SS by |S||S|.

II A Sequential Optimization Framework

Often multi-agent dynamical systems which commonly arise in control or social sciences are images of optimization algorithms which are commonly used in machine learning literature. To make this connection more clear, let us first consider the following iterated block successive minimization process which is frequently used in the machine learning literature for minimizing a smooth/non-smooth function [18]. Consider the optimization problem:

minf(𝒚1,𝒚2,,𝒚n),𝒚iYi,i,\displaystyle\min f(\boldsymbol{y}_{1},\boldsymbol{y}_{2},\ldots,\boldsymbol{y}_{n}),\ \boldsymbol{y}_{i}\in Y_{i},\forall i,

where YimiY_{i}\subseteq\mathbb{R}^{m_{i}} is a closed convex set, and f:i=1nYif:\prod_{i=1}^{n}Y_{i}\to\mathbb{R} is a continuous function. A popular approach to solve the above optimization problem is the block coordinate descent (BCD) method. At each iteration of this method, the objective function is minimized with respect to a single block of variables while the rest of the blocks are held fixed. More specifically, at iteration t=0,1,t=0,1,\ldots of the algorithm, the block variable 𝒚i\boldsymbol{y}_{i} is updated by solving the following subproblem:

𝒚it=argmin𝒛iYif(𝒚1t,,𝒚i1t,𝒛i,𝒚i+1t,,𝒚nt),i[n].\displaystyle\boldsymbol{y}_{i}^{t}=\arg\min_{\boldsymbol{z}_{i}\in Y_{i}}f(\boldsymbol{y}^{t}_{1},\ldots,\boldsymbol{y}_{i-1}^{t},\boldsymbol{z}_{i},\boldsymbol{y}_{i+1}^{t},\ldots,\boldsymbol{y}^{t}_{n}),\ i\in[n].

In particular, an important question here is whethere or not the generated sequence {𝒚t}t=0\{\boldsymbol{y}^{t}\}_{t=0}^{\infty} where 𝒚t=(𝒚1t,,𝒚nt)\boldsymbol{y}^{t}=(\boldsymbol{y}^{t}_{1},\ldots,\boldsymbol{y}^{t}_{n}) will converge to a local/global minimizer of the objective function f()f(\cdot). Due to its particular simple implementation, the BCD method has been widely used for solving problems such as power allocation in image denoising and image reconstruction, wireless communication systems, clustering, and dynamic programming [19, 20, 18]. On the other hand, since in practice finding the exact minimum in each iteration with respect to a block variable might be expensive, one can consider different variants of the BCD method, such as inexact BCD method, where one adds a smooth regularizer to the objective function or approximates it by a smooth upper bound function. In either case, and under some mild assumptions, it can be shown that the BCD method will converge to a stationary point of the objective function f()f(\cdot) [18].

To see how BCD method can be used toward stability analysis of multi-agent network systems, let us consider a special case of the above minimization where there are only two block variables, namely a state block variable 𝒚:=(𝒚1,,𝒚n)n×d\boldsymbol{y}:=(\boldsymbol{y}_{1},\ldots,\boldsymbol{y}_{n})\in\mathbb{R}^{n\times d}, where 𝒚id\boldsymbol{y}_{i}\in\mathbb{R}^{d} denotes the state of agent ii, and a network block variable 𝝀:=(λij)Λ[0,1]n×n\boldsymbol{\lambda}:=(\lambda_{ij})\in\Lambda\subseteq[0,1]^{n\times n}, where λij=1\lambda_{ij}=1 if there is a directed edge from agent (node) ii to agent jj so that 𝒚i\boldsymbol{y}_{i} can be influenced by 𝒚j\boldsymbol{y}_{j}, and λij=0\lambda_{ij}=0 if no such an edge exists. In other words, an integral block variable 𝝀\boldsymbol{\lambda} encodes the adjacency matrix of the communication network among agents. Note that we also allow the network to contain self-loops whenever λii=1\lambda_{ii}=1 for some ii.

Now let us consider a class of multi-agent network dynamics with nn agents evolving over discrete time instances t=0,1,2,t=0,1,2,\ldots as:

𝒙(t+1)=g1(𝒙(t),𝝀(t)),\displaystyle\boldsymbol{x}(t+1)=g_{1}(\boldsymbol{x}(t),\boldsymbol{\lambda}(t)), (1)
𝝀(t+1)=g2(𝒙(t+1)),\displaystyle\boldsymbol{\lambda}(t+1)=g_{2}(\boldsymbol{x}(t+1)), (2)

where 𝒙(t):=(x1(t),,xn(t))Xn×d\boldsymbol{x}(t):=(x_{1}(t),\ldots,x_{n}(t))\in X\subseteq\mathbb{R}^{n\times d} denotes agent ii’s state at time tt, and 𝝀(t)=(λij(t))ijΛ[0,1]n×n\boldsymbol{\lambda}(t)=(\lambda_{ij}(t))_{ij}\in\Lambda\subseteq[0,1]^{n\times n} identifies the communication links between each pair of agents at that time. Here g1()g_{1}(\cdot) and g2()g_{2}(\cdot) are two functions which capture the update rule of the underlying dynamics. As it can be seen from (1), the state of the dynamics at the next time instance 𝒙(t+1)\boldsymbol{x}(t+1) is determined by the joint pair of state-network at time tt, i.e., (𝒙(t),𝝀(t))(\boldsymbol{x}(t),\boldsymbol{\lambda}(t)), while the network structure at time t+1t+1 is determined by the state variable at time tt.

Proposition 1

Let (S,S)(S,\leq_{S}) be a totally ordered set.11 1 In most applications of this paper (with an exception of Section IV) we set (S,S)(S,\leq_{S}) to be the set of real numbers endowed by its natural order. Moreover, assume that there exists a function Φ(𝐲,𝛌):X×ΛS\Phi(\boldsymbol{y},\boldsymbol{\lambda}):X\times\Lambda\to S, such that given any fixed network 𝛌Λ\boldsymbol{\lambda}^{*}\in\Lambda, Φ(𝐲,𝛌):XS\Phi(\boldsymbol{y},\boldsymbol{\lambda}^{*}):X\to S is nonincreasing (with respect to S\leq_{S}) along the image of g1()g_{1}(\cdot), i.e.,

Φ(g1(𝒚,𝝀),𝝀)SΦ(𝒚,𝝀),𝒚X.\displaystyle\Phi\big(g_{1}(\boldsymbol{y},\boldsymbol{\lambda}^{*}),\boldsymbol{\lambda}^{*}\big)\leq_{S}\Phi\big(\boldsymbol{y},\boldsymbol{\lambda}^{*}\big),\ \forall\boldsymbol{y}\in X. (3)

If there exists a function f1:ΛSf_{1}:\Lambda\to S such that for any fixed state 𝐲X\boldsymbol{y}\in X,

g2(𝒚)argmin𝝀Λ{f(𝒚,𝝀):=Φ(𝒚,𝝀)+f1(𝝀)},\displaystyle g_{2}(\boldsymbol{y})\in\arg\min_{\boldsymbol{\lambda}\in\Lambda}\{f(\boldsymbol{y},\boldsymbol{\lambda}):=\Phi(\boldsymbol{y},\boldsymbol{\lambda})+f_{1}(\boldsymbol{\lambda})\}, (4)

then f(𝐲,𝛌)f(\boldsymbol{y},\boldsymbol{\lambda}) is nonincreasing (with respect to S\leq_{S}) along the trajectories of (1).

Proof:

Using the definition of joint dynamics (1) we can write:

f(𝒙(t+1),𝝀(t+1))\displaystyle f(\boldsymbol{x}(t+1),\boldsymbol{\lambda}(t+1)) =f(𝒙(t+1),g2(𝒙(t+1)))\displaystyle=f(\boldsymbol{x}(t+1),g_{2}(\boldsymbol{x}(t+1)))
=min𝝀Λf(𝒙(t+1),𝝀)\displaystyle=\min_{\boldsymbol{\lambda}\in\Lambda}f(\boldsymbol{x}(t+1),\boldsymbol{\lambda}) (5)
Sf(𝒙(t+1),𝝀(t))\displaystyle\leq_{S}f(\boldsymbol{x}(t+1),\boldsymbol{\lambda}(t)) (6)
=Φ(g1(𝒙(t),𝝀(t)),𝝀(t))+f1(𝝀(t))\displaystyle=\Phi\big(g_{1}(\boldsymbol{x}(t),\boldsymbol{\lambda}(t)),\boldsymbol{\lambda}(t)\big)+f_{1}(\boldsymbol{\lambda}(t)) (7)
SΦ(𝒙(t),𝝀(t))+f1(𝝀(t))\displaystyle\leq_{S}\Phi\big(\boldsymbol{x}(t),\boldsymbol{\lambda}(t)\big)+f_{1}(\boldsymbol{\lambda}(t)) (8)
=f(𝒙(t),𝝀(t)),\displaystyle=f(\boldsymbol{x}(t),\boldsymbol{\lambda}(t)), (9)

where the second equality is due to (4), and the last inequality holds by (3) given the fixed 𝝀=𝝀(t)\boldsymbol{\lambda}^{*}=\boldsymbol{\lambda}(t). As a result, the coupled dynamics in (1) can be replicated by applying the BCD method to the objective function f(𝒚,𝝀)=Φ(𝒚,𝝀)+f1(𝝀)f(\boldsymbol{y},\boldsymbol{\lambda})=\Phi(\boldsymbol{y},\boldsymbol{\lambda})+f_{1}(\boldsymbol{\lambda}) with constraint sets 𝝀Λ\boldsymbol{\lambda}\in\Lambda and 𝒙X\boldsymbol{x}\in X, where at each iteration we fix either network or state variable and optimize the objective function ff with respect to the other variable. Q.E.D.

Definition 1

Let (S,S)(S,\leq_{S}) be a totally ordered set. A function V:mSV:\mathbb{R}^{m}\to S is called a Lyapunov function for the discrete time dynamical system 𝐳(t+1)=h(𝐳(t))\boldsymbol{z}(t+1)=h(\boldsymbol{z}(t)), if it is nonincreasing along the trajectories of the dynamics, i.e., V(𝐳(t+1))SV(𝐳(t))V(\boldsymbol{z}(t+1))\leq_{S}V(\boldsymbol{z}(t)). We refer to a dynamical system which admits a Lyapunov function as Lyapunov stable.

Intuitively, Proposition 1 states that if trajectories of the projected state dynamics in (1) over a fixed network 𝝀\boldsymbol{\lambda}^{*} admit a Lyapunov function Φ(𝒚,𝝀)\Phi(\boldsymbol{y},\boldsymbol{\lambda}^{*}), while minimizing f(𝒚,𝝀)=Φ(𝒚,𝝀)+f1(𝝀)f(\boldsymbol{y},\boldsymbol{\lambda})=\Phi(\boldsymbol{y},\boldsymbol{\lambda})+f_{1}(\boldsymbol{\lambda}) with respect to 𝝀Λ\boldsymbol{\lambda}\in\Lambda accurately captures the network associated to the fixed state 𝒚\boldsymbol{y}, then the joint state-network dynamics (1) admit a Lyapunov function. In particular, V(𝒚)=min𝝀Λf(𝒚,𝝀)V(\boldsymbol{y})=\min_{\boldsymbol{\lambda}\in\Lambda}f(\boldsymbol{y},\boldsymbol{\lambda}) serves as a Lyapunov function for the dynamics {𝒙(t),t=0,1,}\{\boldsymbol{x}(t),t=0,1,\ldots\} generated by (1). In what follows next we show how this simple framework can be used to establish Lyapunov stability and convergence of several important state-dependent multi-agent network dynamics.

III Hegselmann-Krause Opinion Dynamics

To show effectiveness of the proposed framework in Section II toward stability and convergence analysis of multi-agent network systems, in this section we consider a well-known model from opinion dynamics known as Hegselmann-Krause (HK) model [21]. A natural question that commonly arises in social sciences is the extent to which one can predict the outcome of the opinion formation of entities under some complex interaction process running among these social actors [1, 21, 22, 23, 24, 25]. In this regard, one of the first studies was undertaken by Hegselmann and Krause in [21] with many applications in the robotics rendezvous [26, 27], linguistic formation [28], social networks [29], trust and marketing [30], among many others [25]. In the HK model, a finite number of agents frequently update their opinions where the opinion of each agent is captured by a scalar (or vector) quantity in one (or higher) dimension.22 2 For simplicity of presentation, in this section we only consider one dimensional HK model. However, all the results can be extended in a straightforward manner to higher dimensions. Because of the conservative nature of social entities, each agent in this model communicates only with those whose opinions are closer to him and lie within a certain level of his confidence.

In the homogeneous HK model, there are a set of [n][n] agents. It is assumed that at each time instance t=0,1,2,t=0,1,2,\ldots, the opinion (state) of agent i[n]i\in[n] can be represented by a scalar xi(t)x_{i}(t)\in\mathbb{R}. Each agent ii updates its value at time tt by taking the arithmetic average of its own value and those of all the others that are in its ϵ\epsilon-neighborhood at time tt. Here the parameter ϵ>0\epsilon>0 is a constant which captures the confidence bound. More precisely, the evolution of opinion vectors 𝒙(t):=(x1(t),,xn(t))n\boldsymbol{x}(t):=(x_{1}(t),\ldots,x_{n}(t))\in\mathbb{R}^{n} can be modeled by the following discrete-time dynamics:

𝒙(t+1)=A(t)𝒙(t),\displaystyle\boldsymbol{x}(t+1)=A(t)\boldsymbol{x}(t), (10)
Aij(t)={1|Ni(t)|if jNi(t),0 else,\displaystyle A_{ij}(t)=\begin{cases}\frac{1}{|N_{i}(t)|}&\mbox{if }j\in N_{i}(t),\\ 0&\mbox{ else},\end{cases} (11)

where Ni(t)N_{i}(t) is the set of neighbors of agent ii, i.e.,

Ni(t)={j[n]:|xi(t)xj(t)|ϵ}.\displaystyle N_{i}(t)=\{j\in[n]:|x_{i}(t)-x_{j}(t)|\leq\epsilon\}.

In the node heterogeneous HK dynamics everything remains as above except that different agents can have different confidence bounds ϵi,i[n]\epsilon_{i},i\in[n]. The node heterogeneous model reflects the fact that some agents are very open minded (large ϵi\epsilon_{i}) and are willing to communicate with many others before updating their opinions, while some agents are closed-minded (small ϵi\epsilon_{i}) and are biased towards their own opinions. For instance ϵi=0\epsilon_{i}=0 means that agent ii is stubborn and will not change its opinion at all. Although at first glance the differences between homogeneous and node-heterogeneous HK dynamics may seem negligible, their outcomes are substantially different, such that most of the results from one cannot be carried over to the other [31, 32, 33, 34]. In this regard, one of the fundamental questions concerning HK dynamics is whether or not they eventually converge to a final outcome.

III-A Homogeneous HK Model

Let us first focus on the homogeneous HK model. Although convergence and detailed analysis of the homogeneous HK model have been established and studied extensively in the past literature (see, e.g., [25] for a comprehensive survey), in this subsection we provide a simple argument to show why this model fits into our framework. This will allow us to generalize stability of homogeneous HK model to account for higher degrees of heterogeneity or asymmetry among the agents.

Let us define 𝒢t=([n],t)\mathcal{G}_{t}=([n],\mathcal{E}_{t}) to be the communication graph at a generic time tt such that there is an edge between agents ii and jj at time tt, i.e., (i,j)t(i,j)\in\mathcal{E}_{t} if and only if jNi(t)j\in N_{i}(t). Note that in the homogeneous HK model the communication graph is undirected as if agent jj is a neighbor of agent ii, the converse is also true. In fact, what makes the analysis of HK dynamics challenging is the strong coupling between the evolution of the network 𝒢t\mathcal{G}_{t} and the state 𝒙(t)\boldsymbol{x}(t). This is because at any time tt the state vector 𝒙(t)\boldsymbol{x}(t) determines the network topology 𝒢t\mathcal{G}_{t}, and this new network determines the state vector at the next time step 𝒙(t+1)\boldsymbol{x}(t+1). This puts HK dynamics to the class of complex time-dependent and state-dependent network dynamics [35, 36, 37, 38, 39, 40]. In particular, the communication network 𝒢t\mathcal{G}_{t} may switch many times depending on how the opinion vectors evolve which brings additional complication to the analysis.

Now let us consider the following objective function comprised of two block variables, namely 𝒚n\boldsymbol{y}\in\mathbb{R}^{n} and 𝝀=(λij)[0,1]n×n\boldsymbol{\lambda}=(\lambda_{ij})\in[0,1]^{n\times n},

f(𝒚,𝝀):=i,jλij((yiyj)2ϵ2).\displaystyle f(\boldsymbol{y},\boldsymbol{\lambda}):=\sum_{i,j}\lambda_{ij}\Big((y_{i}-y_{j})^{2}-\epsilon^{2}\Big). (12)

This function can also be written in a compact form as f(𝒚,𝝀)=𝒚T𝒚tr()ϵ2f(\boldsymbol{y},\boldsymbol{\lambda})=\boldsymbol{y}^{T}\mathcal{L}\boldsymbol{y}-tr(\mathcal{L})\epsilon^{2}, where :=diag(𝝀𝟏)𝝀\mathcal{L}:=diag(\boldsymbol{\lambda}\boldsymbol{1})-\boldsymbol{\lambda}, and tr()tr(\cdot) denotes the trace function. Intuitively, the block variable 𝝀\boldsymbol{\lambda} is meant to capture the communication network 𝒢t\mathcal{G}_{t}, and the block variable 𝒚\boldsymbol{y} captures the opinion states. Note that if we restrict λij\lambda_{ij}s to binary variables in {0,1}\{0,1\}, then 𝝀\boldsymbol{\lambda} simply represents the adjacency matrix of a network of nn agents where λij=1\lambda_{ij}=1 if there is a directed edge (i,j)(i,j) from node ii to node jj, and λij=0\lambda_{ij}=0 otherwise. Moreover, for such a binary block variable 𝝀\boldsymbol{\lambda}, the matrix \mathcal{L} is precisely the Laplacian matrix of the communication network associated with 𝝀\boldsymbol{\lambda}. Although we still need to assume that 𝝀{0,1}n×n\boldsymbol{\lambda}\in\{0,1\}^{n\times n}, to avoid complication of handling integral variables, for now we allow λij\lambda_{ij}s to vary continuously in the interval [0,1][0,1]. As we shall see soon the integrality of network variables will be automatically achieved during iterations of the BCD method.

Now let us consider the BCD method applied to the objective function (12) with block variables 𝒙\boldsymbol{x} and 𝝀\boldsymbol{\lambda}. For a generic time tt, let us fix the state variable to 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t). Minimizing (12) with respect to the network variable 𝝀Λ=[0,1]n×n\boldsymbol{\lambda}\in\Lambda=[0,1]^{n\times n} we obtain,

𝝀t:=argmini,j𝝀[0,1]n2λij((xi(t)xj(t))2ϵ2),\displaystyle\boldsymbol{\lambda}_{t}:=\arg\min_{\boldsymbol{\lambda}\in[0,1]^{n^{2}}}\sum_{i,j}\lambda_{ij}\Big((x_{i}(t)-x_{j}(t))^{2}-\epsilon^{2}\Big), (13)

where

(𝝀t)ij=(𝝀t)ji={1if |xi(t)xj(t)|ϵ,0else.\displaystyle(\boldsymbol{\lambda}_{t})_{ij}=(\boldsymbol{\lambda}_{t})_{ji}=\begin{cases}1&\mbox{if }|x_{i}(t)-x_{j}(t)|\leq\epsilon,\\ 0&\mbox{else}.\end{cases} (14)

This simply follows because the objective function f(𝒙(t),𝝀)f(\boldsymbol{x}(t),\boldsymbol{\lambda}) is a linear function of the network variable 𝝀\boldsymbol{\lambda} and achieves its minimum in an extreme point of [0,1]n×n[0,1]^{n\times n}. In particular, the optimal extreme point can be found by an easy inspection as given in (14). But note that 𝝀t\boldsymbol{\lambda}_{t} is precisely the adjacency matrix of the communication graph 𝒢t\mathcal{G}_{t} in the homogeneous HK model (recall that in the homogeneous HK model two agents are each others’ neighbors at time tt if and only if their distance is at most ϵ\epsilon). Therefore, fixing the state variable to 𝒙(t)\boldsymbol{x}(t) and minimizing (12) with respect to the network variable exactly delivers the adjacency matrix of the communication network 𝒢t\mathcal{G}_{t} in the homogeneous HK model at that time.

Now let us fix the network variable to the minimizer 𝝀t\boldsymbol{\lambda}_{t} which represents an undirected graph with associated Laplacian matrix t\mathcal{L}_{t}. It is well-known that given a fixed undirected graph with Laplacian matrix t\mathcal{L}_{t} and arbitrary values {yi,i[n]}\{y_{i},i\in[n]\} at its nn verticies, the quadratic function Φ(𝒚,𝝀t):=𝒚Tt𝒚\Phi(\boldsymbol{y},\boldsymbol{\lambda}_{t}):=\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y} is nonincreasing if each node updates its value to the average value of its neighbors [41, 42]. In other words, defining 𝒚=At𝒚\boldsymbol{y}^{\prime}=A_{t}\boldsymbol{y} where At:=(diag(𝝀t𝟏))1𝝀tA_{t}:=(diag(\boldsymbol{\lambda}_{t}\boldsymbol{1}))^{-1}\boldsymbol{\lambda}_{t}, we have (𝒚)Tt𝒚𝒚t𝒚(\boldsymbol{y}^{\prime})^{T}\mathcal{L}_{t}\boldsymbol{y}^{\prime}\leq\boldsymbol{y}\mathcal{L}_{t}\boldsymbol{y}. But note that AtA_{t} is precisely the update matrix of the homogeneous HK model given in (10), which means that for the specific choice of 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t) we have 𝒚=𝒙(t+1)\boldsymbol{y}^{\prime}=\boldsymbol{x}(t+1). Thus the evolution of the homogeneous HK dynamics is governed by the application of BCD method to the objective function (12). Appealing to Proposition 1, this shows that f(𝒙(t+1),𝝀(t+1))f(𝒙(t),𝝀(t))f(\boldsymbol{x}(t+1),\boldsymbol{\lambda}(t+1))\leq f(\boldsymbol{x}(t),\boldsymbol{\lambda}(t)), implying that homogeneous HK dynamics admit a Lyapunov function.

Remark 1

Adapting the same notation as in Section II we have

g1(𝒚,𝝀t)=At𝒚,g2(𝒚)=(𝟏{|yiyj|ϵ})ij,\displaystyle g_{1}(\boldsymbol{y},\boldsymbol{\lambda}_{t})=A_{t}\boldsymbol{y},\ \ \ \ \ \ \ \ \ g_{2}(\boldsymbol{y})=(\boldsymbol{1}_{\{|y_{i}-y_{j}|\leq\epsilon\}})_{ij},
Φ(𝒚,𝝀):=𝒚T𝒚,f1(𝝀)=tr()ϵ2,\displaystyle\Phi(\boldsymbol{y},\boldsymbol{\lambda}):=\boldsymbol{y}^{T}\mathcal{L}\boldsymbol{y},\ \ \ \ \ \ \ f_{1}(\boldsymbol{\lambda})=-tr(\mathcal{L})\epsilon^{2}, (15)

and f(𝐲,𝛌)=𝐲T𝐲tr()ϵ2f(\boldsymbol{y},\boldsymbol{\lambda})=\boldsymbol{y}^{T}\mathcal{L}\boldsymbol{y}-tr(\mathcal{L})\epsilon^{2} so that the first term 𝐲T𝐲\boldsymbol{y}^{T}\mathcal{L}\boldsymbol{y} captures the internal coupling between network and opinion states in the HK model.

III-B Restricted Edge Heterogeneous HK Dynamics

In this part we show how the framework of Section II can extend the analysis of homogeneous HK model by capturing higher degrees of heterogeneity or constraints. For this purpose, we consider restricted edge-heterogeneous HK model which is a variant of the homogeneous HK model with the following two additional changes:

1) Edge Heterogeneity: Consider the same dynamical system as in the homogeneous HK model (10), except that the distance between every pair of nodes is measured based on possibly a different confidence bound. More precisely, let {ϵij=ϵji>0,ij}\{\epsilon_{ij}=\epsilon_{ji}>0,\forall i\neq j\} be a set of fixed thresholds (one for each pair of agents) so that agents ii and jj at time instance tt can communicate if and only if their distance at that time is less than ϵij\epsilon_{ij}, and thus Ni(t)={j[n]:|xi(t)xj(t)|ϵij}N_{i}(t)=\{j\in[n]:|x_{i}(t)-x_{j}(t)|\leq\epsilon_{ij}\}. This captures the heterogeneous relationship among individuals due to family or other social ties. For instance two family members will still continue to communicate even if their opinions are relatively far from each other, while two strangers are more likely to terminate their interactions as soon as their opinions slightly deviate from each other. As before we assume that at each time instance agents update their opinions by taking the arithmetic average of their neighbors’ opinions, determined based on the heterogeneous thresholds ϵij\epsilon_{ij}. Note that the homogeneous HK dynamics is a special case of the edge heterogeneous setting where ϵij=ϵ,ij\epsilon_{ij}=\epsilon,\forall i\neq j.

2) Communication Restrictions: Other than closeness in opinion, often there are other important factors which determine whether or not two agents should communicate. For instance, individuals often communicate with those who have closer opinion to them and are within small geographic distance from them. One direct way of handling such restriction is to add an extra component to each agent’s opinion where this new component encodes the geographic location of that agent. As a result, in this higher dimensional opinion space two agents are each others’ neighbors if each component of their opinion vectors (and in particular the geographic component of their opinion vector) are close to each other. This implies that two agents are eligible to communicate if they are close to each other both opinionwise and geographic-wise.33 3 Although this approach increases the dimension of the opinion space, yet most of the results such as convergence of the dynamics can be extended to this higher dimensional setting [43]. An alternative approach however for imposing new constraints is to consider a predefined underlying network 𝒢\mathcal{G} which restricts the agents’ interactions to only those who are both connected through the edges of 𝒢\mathcal{G} and have closer opinion to each other. Intuitively, one can imagine running HK dynamics over the graph 𝒢\mathcal{G}. Thus, denoting the communication network of the edge-heterogeneous HK dynamics at time tt by 𝒢(t)\mathcal{G}(t), the actual communication graph at time tt is given by the intersection of edges which appear in both 𝒢\mathcal{G} and 𝒢(t)\mathcal{G}(t). This second approach is particularly more suitable when there are certain hard communication constraints among individuals due to age gaps, gender restrictions, or other social laws.

Theorem 2

The restricted edge-heterogeneous HK model over an undirected graph 𝒢=([n],)\mathcal{G}=([n],\mathcal{E}) is Lyapunov stable. Moreover, in the absence of edge heterogeneity (i.e., when all pairs have the same confidence bound), the restricted HK model converges to an equilibrium point geometrically fast.

Proof:

Given a set of pairwise thresholds {ϵij=ϵji>0:i,j[n]}\{\epsilon_{ij}=\epsilon_{ji}>0:i,j\in[n]\}, and an undirected restricting graph 𝒢=([n],)\mathcal{G}=([n],\mathcal{E}), let us consider the BCD method applied to the following minimization problem:

minf(𝒚,𝝀):=i,jλij((yiyj)2ϵij2)\displaystyle\min f(\boldsymbol{y},\boldsymbol{\lambda}):=\sum_{i,j}\lambda_{ij}\Big((y_{i}-y_{j})^{2}-\epsilon_{ij}^{2}\Big)
𝝀Λ,𝒚n,\displaystyle\ \ \ \ \ \ \ \ \boldsymbol{\lambda}\in\Lambda,\ \ \boldsymbol{y}\in\mathbb{R}^{n}, (16)

where Λ={(λij)[0,1]n×n:λij=0,{i,j},λij=λjii,j}\Lambda=\{(\lambda_{ij})\in[0,1]^{n\times n}:\lambda_{ij}=0,\forall\{i,j\}\notin\mathcal{E},\lambda_{ij}=\lambda_{ji}\ \forall i,j\} is the constraint set for the block variable 𝝀\boldsymbol{\lambda}. The above minimization problem can be rewritten as

min\displaystyle\min {i,j}λij((yiyj)2ϵij2),\displaystyle\sum_{\{i,j\}\in\mathcal{E}}\lambda_{ij}\Big((y_{i}-y_{j})^{2}-\epsilon_{ij}^{2}\Big), (17)
λij[0,1],{i,j},𝒚n.\displaystyle\ \lambda_{ij}\in[0,1],\forall\{i,j\}\in\mathcal{E},\ \ \boldsymbol{y}\in\mathbb{R}^{n}. (18)

Now given a fixed state variable 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t) at a generic time tt, minimizing (17) with respect to block variable 𝝀\boldsymbol{\lambda} gives us 𝝀t\boldsymbol{\lambda}_{t} where

(𝝀t)ij=(𝝀t)ji={1if |xi(t)xj(t)|ϵij,{i,j}0else.\displaystyle(\boldsymbol{\lambda}_{t})_{ij}=(\boldsymbol{\lambda}_{t})_{ji}=\begin{cases}1&\mbox{if }|x_{i}(t)-x_{j}(t)|\leq\epsilon_{ij},\{i,j\}\in\mathcal{E}\\ 0&\mbox{else}.\end{cases}

In other words, minimizing (17) with respect to 𝝀Λ\boldsymbol{\lambda}\in\Lambda precisely captures the communication structure in the restricted edge-heterogeneous HK model. Now let us fix the network variable to 𝝀t\boldsymbol{\lambda}_{t}. Since 𝝀t\boldsymbol{\lambda}_{t} represents the adjacency matrix of an undirected graph with corresponding Laplacian t\mathcal{L}_{t}, as before Φ(𝒚,𝝀t):=𝒚Tt𝒚\Phi(\boldsymbol{y},\boldsymbol{\lambda}_{t}):=\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y} is nonincreasing for the HK update rule over this fixed network 𝝀t\boldsymbol{\lambda}_{t}. Thus for any 𝒚n\boldsymbol{y}\in\mathbb{R}^{n}, if we denote the update matrix of the restricted edge-heterogeneous HK model on the undirected graph 𝝀t\boldsymbol{\lambda}_{t} by AtA_{t}, we have (At𝒚)Tt(At𝒚)𝒚T𝒚(A_{t}\boldsymbol{y})^{T}\mathcal{L}_{t}(A_{t}\boldsymbol{y})\leq\boldsymbol{y}^{T}\mathcal{L}\boldsymbol{y}. In particular, by choosing 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t), where 𝒙(t)\boldsymbol{x}(t) denotes the state vector of the restricted edge-heterogeneous HK model at time tt, we have 𝒙(t+1)Tt𝒙(t+1)𝒙T(t)t𝒙(t)\boldsymbol{x}(t+1)^{T}\mathcal{L}_{t}\boldsymbol{x}(t+1)\leq\boldsymbol{x}^{T}(t)\mathcal{L}_{t}\boldsymbol{x}(t). Therefore, BCD method applied to (17) replicates the dynamics of the restricted edge-heterogeneous HK model. In particular, this shows that V(𝒚):=min𝝀Λf(𝒚,𝝀)={i,j}((yiyj)2ϵij2)V(\boldsymbol{y}):=\min_{\boldsymbol{\lambda}\in\Lambda}f(\boldsymbol{y},\boldsymbol{\lambda})=\sum_{\{i,j\}\in\mathcal{E}}\Big((y_{i}-y_{j})^{2}-\epsilon_{ij}^{2}\Big)^{-} serves as a Lyapunov function for the restricted edge-heterogeneous HK model.

In what follows next, we use the above Lyapunov function to establish asymptotic convergence of the restricted HK model to an equilibrium point in the absence of edge heterogeneity (i.e., when all ϵij\epsilon_{ij} are the same which by rescaling from now we may assume ϵij=1,i,j\epsilon_{ij}=1,\forall i,j). To lower bound the decrease of Lyapunov function V()V(\cdot) at a given time step tt, we note that this decrease is lower bounded by the decrease amount which is achieved due to the state update. Thus

V(𝒙(t))V(𝒙(t+1))\displaystyle V(\boldsymbol{x}(t))-V(\boldsymbol{x}(t+1)) 𝒙T(t)t𝒙(t)𝒙(t+1)Tt𝒙(t+1)\displaystyle\geq\boldsymbol{x}^{T}(t)\mathcal{L}_{t}\boldsymbol{x}(t)-\boldsymbol{x}(t+1)^{T}\mathcal{L}_{t}\boldsymbol{x}(t+1) (19)
=𝒙T(t)(tAtTtAt)𝒙(t)\displaystyle=\boldsymbol{x}^{T}(t)(\mathcal{L}_{t}-A_{t}^{T}\mathcal{L}_{t}A_{t})\boldsymbol{x}(t) (20)
=𝒙T(t)(IAt)T(diag(𝝀𝒕𝟏)+𝝀𝒕)(IAt)𝒙(t)\displaystyle=\boldsymbol{x}^{T}(t)(I-A_{t})^{T}(diag(\boldsymbol{\lambda_{t}}\boldsymbol{1})+\boldsymbol{\lambda_{t}})(I-A_{t})\boldsymbol{x}(t) (21)
=(𝒙(t)𝒙(t+1))T(diag(𝝀𝒕𝟏)+𝝀𝒕)(𝒙(t)𝒙(t+1))\displaystyle=(\boldsymbol{x}(t)-\boldsymbol{x}(t+1))^{T}(diag(\boldsymbol{\lambda_{t}}\boldsymbol{1})+\boldsymbol{\lambda_{t}})(\boldsymbol{x}(t)-\boldsymbol{x}(t+1)) (22)
𝒙(t)𝒙(t+1)2,\displaystyle\geq\|\boldsymbol{x}(t)-\boldsymbol{x}(t+1)\|^{2}, (23)

where the first equality is obtained by using 𝒙(t+1)=At𝒙(t)\boldsymbol{x}(t+1)=A_{t}\boldsymbol{x}(t), and the second equality is valid by a simple matrix multiplication and noting that diag(𝝀t𝟏)At=𝝀tdiag(\boldsymbol{\lambda}_{t}\boldsymbol{1})A_{t}=\boldsymbol{\lambda}_{t}. Finally the last inequality holds because λt\lambda_{t} is the adjacency matrix of a connected undirected graph,44 4 Here, without loss of generality we may assume λt\lambda_{t} to be connected, otherwise for the rest of analysis we can restrict our attention to one of its connected components. and hence diag(𝝀𝒕𝟏)+𝝀𝒕diag(\boldsymbol{\lambda_{t}}\boldsymbol{1})+\boldsymbol{\lambda_{t}} is a positive definite matrix whose eigenvalues are greater than or equal to 1. By summing (19) for all τt\tau\leq t, and rearranging the terms we get

τ=0t𝒙(τ)𝒙(τ+1)2V(𝒙(0))V(𝒙(t))V(𝒙(0))+n2,\displaystyle\sum_{\tau=0}^{t}\|\boldsymbol{x}(\tau)-\boldsymbol{x}(\tau+1)\|^{2}\leq V(\boldsymbol{x}(0))-V(\boldsymbol{x}(t))\leq V(\boldsymbol{x}(0))+n^{2},

where the second inequality is valid since by the definition of V()V(\cdot) we always have V()n2V(\cdot)\geq-n^{2}. Thus τ=0𝒙(τ)𝒙(τ+1)2=τ=0i=1n(xi(τ)xi(τ+1))2\sum_{\tau=0}^{\infty}\|\boldsymbol{x}(\tau)-\boldsymbol{x}(\tau+1)\|^{2}=\sum_{\tau=0}^{\infty}\sum_{i=1}^{n}(x_{i}(\tau)-x_{i}(\tau+1))^{2} is a convergent series, and hence for any δ2>0\delta^{2}>0, there exists a sufficiently large time tδt_{\delta} such that τ=tδi=1n(xi(τ)xi(τ+1))2<δ2\sum_{\tau=t_{\delta}}^{\infty}\sum_{i=1}^{n}(x_{i}(\tau)-x_{i}(\tau+1))^{2}<\delta^{2}. On the other hand, it is shown in Lemma 1 that if δ<1n2\delta<\frac{1}{n^{2}}, no switch in the communication network can occur after time tδt_{\delta}. Thus after at most finite time tδt_{\delta} the communication network of the restricted HK model remains unchanged. This implies that from time tδt_{\delta} onward the evolution of the dynamics is governed by powers of a fixed stochastic matrix which is well-known to converge to an equilibrium point geometrically fast. Q.E.D.

Remark 2

The update matrix AtA_{t} of the HK model is the transition matrix of a lazy simple random walk on its underlying network 𝛌t\boldsymbol{\lambda}_{t}. However, Φ(𝐲,𝛌t):=𝐲Tt𝐲\Phi(\boldsymbol{y},\boldsymbol{\lambda}_{t}):=\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y} serves as a Lyapunov function for any irreducible random walk on the undirected graph 𝛌t\boldsymbol{\lambda}_{t}. As a result, one can allow more general update weights {wij>0,i,j[n]}\{w_{ij}>0,i,j\in[n]\} than original weights {1k,k[n]}\{\frac{1}{k},k\in[n]\} appearing in the update matrices of the HK model, and still use the above analysis to show that the generated dynamics are Lyapunov stable. This can be done by replacing variables λij\lambda_{ij} by wijλijw_{ij}\lambda_{ij} in the above proof.

III-C Asymmetric 0-1 HK Dynamics

Finally, in this subsection we take one step further and consider a special case of the node-heterogeneous HK model. As we mentioned earlier, the dynamics of node-heterogeneous HK model follow exactly the same update rule as homogeneous HK model given in (10) except that different agents might have different confidence bounds ϵi,i[n]\epsilon_{i},i\in[n]. Unfortunately, up to the time of writing this paper there is no general result which either proves or disproves convergence of the node-heterogeneous HK model (although some partial results concerning stability of these dynamics are known [44, 43, 32]). In particular, in the recent work [35], the authors have used an algorithmic approach to show convergence of a special case of the node-heterogeneous HK model, namely 0-1 HK model, in which the confidence bound of each node is restricted to be either 0 or ϵ=1\epsilon=1 (i.e., ϵi{0,1},i[n]\epsilon_{i}\in\{0,1\},\forall i\in[n]). It is worth noting that due to heterogeneous confidence bounds, the communication network in the 0-1 HK model is no longer undirected (symmetric). While the proof in [35] is fairly long and algorithmic, here we provide a simple argument based on the BCD framework to establish Lyapunov stability of the 0-1 HK dynamics. An important advantage of our approach is that i) it provides an improved Lyapunov drift which can be useful towards convergence rate analysis, and ii) it provides a clear explanation of why the asymmetric 0-1 HK dynamics can still be treated as the symmetric homogeneous HK model.

Theorem 3

The 0-1 HK dynamics are Lyapunov stable.

Proof:

Let us consider the same function f(𝒚,𝝀)f(\boldsymbol{y},\boldsymbol{\lambda}) as in (12). We show that this function is nonincreasing over the trajectories of the 0-1 HK. Let S0S_{0} and S1=[n]S0S_{1}=[n]\setminus S_{0} denote the set of agents with confidence bounds 00 and ϵ=1\epsilon=1, respectively. As before, given a fixed state variable 𝒚:=𝒙(t)\boldsymbol{y}:=\boldsymbol{x}(t), minimizing ff with respect to 𝝀Λ:=[0,1]n×n\boldsymbol{\lambda}\in\Lambda:=[0,1]^{n\times n} we obtain (𝝀t)ij=(𝝀t)ji=1(\boldsymbol{\lambda}_{t})_{ij}=(\boldsymbol{\lambda}_{t})_{ji}=1 if |xi(t)xj(t)|ϵ|x_{i}(t)-x_{j}(t)|\leq\epsilon, and (𝝀t)ij=(𝝀t)ji=0(\boldsymbol{\lambda}_{t})_{ij}=(\boldsymbol{\lambda}_{t})_{ji}=0, otherwise. This correctly captures the communication links adjacent to the agents in S1S_{1} of the actual communication network in the 0-1 HK model at time tt. However, it is possible that 𝝀t\boldsymbol{\lambda}_{t} incorrectly sets (𝝀t)ij=1(\boldsymbol{\lambda}_{t})_{ij}=1 for an agent iS0i\in S_{0} so that 𝝀t\boldsymbol{\lambda}_{t} and the actual 0-1 HK communication network at time tt can only deviate from each other on edges {(i,j),iS0}\{(i,j),i\in S_{0}\}. Nevertheless, as far as it concerns the agents in S0S_{0}, this will not be an issue since the agents in S0S_{0} will never use their adjacent links to update their states (these agents are always fixed). Therefore, we are only left to show that for fixed undirected graph 𝝀t\boldsymbol{\lambda}_{t} with corresponding Laplacian t\mathcal{L}_{t}, Φ(𝒚,𝝀t)=𝒚Tt𝒚\Phi(\boldsymbol{y},\boldsymbol{\lambda}_{t})=\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y} still serves as a Lyapunov function for the 0-1 HK update rule.

Given an arbitrary state vector 𝒚\boldsymbol{y}, let us decompose it into 𝒚=(𝒚0,𝒚1)\boldsymbol{y}=(\boldsymbol{y}_{0},\boldsymbol{y}_{1}), where 𝒚0\boldsymbol{y}_{0} and 𝒚1\boldsymbol{y}_{1} are associated to the agents in S0S_{0} and S1S_{1}, respectively. Define R0:=𝝀t[S0]R_{0}:=\boldsymbol{\lambda}_{t}[S_{0}] and R1:=𝝀t[S1]R_{1}:=\boldsymbol{\lambda}_{t}[S_{1}] to be the adjacency matrices induced by 𝝀t\boldsymbol{\lambda}_{t} over the agents in S0S_{0} and S1S_{1}, respectively. Moreover, let M:=𝝀t[S1,S0]M:=\boldsymbol{\lambda}_{t}[S_{1},S_{0}] be the adjacency matrix of the bipartite graph induced by 𝝀t\boldsymbol{\lambda}_{t} between agents in S0S_{0} and S1S_{1}. Therefore, we can write

𝝀t=[R0MTMR1],t=[D0R0MTMD1R1],\displaystyle\boldsymbol{\lambda}_{t}=\left[\begin{array}[]{c|c}R_{0}&M^{T}\\ \hline\cr M&R_{1}\end{array}\right],\ \ \mathcal{L}_{t}=\left[\begin{array}[]{c|c}D_{0}-R_{0}&-M^{T}\\ \hline\cr-M&D_{1}-R_{1}\end{array}\right],

where D0=diag(R0𝟏)D_{0}=diag(R_{0}\boldsymbol{1}) and D1=diag(R1𝟏)D_{1}=diag(R_{1}\boldsymbol{1}). Moreover, the actual 0-1 HK update rule at time tt can be written as 𝒙(t+1)=At𝒙(t)\boldsymbol{x}(t+1)=A_{t}\boldsymbol{x}(t), where

At=[I𝟎D11MD11R1].\displaystyle A_{t}=\left[\begin{array}[]{c|c}I&\boldsymbol{0}\\ \hline\cr D_{1}^{-1}M&D_{1}^{-1}R_{1}\end{array}\right].

Since At𝒚=(𝒚0,D11M𝒚0+D11R1𝒚1)A_{t}\boldsymbol{y}=(\boldsymbol{y}_{0},D_{1}^{-1}M\boldsymbol{y}_{0}+D_{1}^{-1}R_{1}\boldsymbol{y}_{1}), we can write

𝒚Tt𝒚\displaystyle\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y} =𝒚0T(D0R0)𝒚0+𝒚1T(D1R1)𝒚12𝒚1TM𝒚0,\displaystyle=\boldsymbol{y}^{T}_{0}(D_{0}\!-\!R_{0})\boldsymbol{y}_{0}+\boldsymbol{y}^{T}_{1}(D_{1}\!-\!R_{1})\boldsymbol{y}_{1}\!-\!2\boldsymbol{y}^{T}_{1}M\boldsymbol{y}_{0},
𝒚TAtTtAt𝒚\displaystyle\boldsymbol{y}^{T}\!A_{t}^{T}\mathcal{L}_{t}A_{t}\boldsymbol{y}\! =𝒚0T(D0R0)𝒚02(D11M𝒚0+D11R𝒚1)TM𝒚0\displaystyle=\!\boldsymbol{y}^{T}_{0}(\!D_{0}\!-\!R_{0}\!)\boldsymbol{y}_{0}\!-\!2(\!D_{1}^{-1}\!M\boldsymbol{y}_{0}\!+\!D_{1}^{-1}\!R\boldsymbol{y}_{1}\!)^{T}\!M\boldsymbol{y}_{0} (30)
+(D11M𝒚0+D11R𝒚1)T(D1R1)(D11M𝒚0+D11R𝒚1).\displaystyle\qquad\!+\!(D_{1}^{-1}M\boldsymbol{y}_{0}\!+\!D_{1}^{-1}R\boldsymbol{y}_{1})^{T}(D_{1}\!-\!R_{1})(D_{1}^{-1}M\boldsymbol{y}_{0}\!+\!D_{1}^{-1}R\boldsymbol{y}_{1}). (31)

Subtracting these two expressions from each other, using the fact that R1T=R1R_{1}^{T}=R_{1}, and simplifying the terms we obtain

𝒚Tt𝒚𝒚TAtTtAt𝒚\displaystyle\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y}-\boldsymbol{y}^{T}A_{t}^{T}\mathcal{L}_{t}A_{t}\boldsymbol{y} =𝒚0T(2MTD11MMTD11(D1R1)D11M)𝒚0\displaystyle=\boldsymbol{y}^{T}_{0}\Big(2M^{T}D_{1}^{-1}M-M^{T}D_{1}^{-1}(D_{1}-R_{1})D_{1}^{-1}M\Big)\boldsymbol{y}_{0} (32)
+𝒚1T(D1R1R1D11(D1R1)D11R1)𝒚1\displaystyle\qquad+\boldsymbol{y}^{T}_{1}\Big(D_{1}-R_{1}-R_{1}D_{1}^{-1}(D_{1}-R_{1})D_{1}^{-1}R_{1}\Big)\boldsymbol{y}_{1} (33)
+2𝒚1T(R1D11R1D11MM)𝒚0.\displaystyle\qquad+2\boldsymbol{y}^{T}_{1}\Big(R_{1}D_{1}^{-1}R_{1}D_{1}^{-1}M-M\Big)\boldsymbol{y}_{0}. (34)

Now a straightforward calculation shows that the right-hand side of (32) can be factorized as PT(D1+R1)PP^{T}(D_{1}+R_{1})P, where P:=(ID11R1)𝒚1D11M𝒚0P:=(I-D_{1}^{-1}R_{1})\boldsymbol{y}_{1}-D_{1}^{-1}M\boldsymbol{y}_{0}. Finally, since 𝒚At𝒚=(𝟎,(ID11R1)𝒚1D11M𝒚0)\boldsymbol{y}-A_{t}\boldsymbol{y}=(\boldsymbol{0},(I-D_{1}^{-1}R_{1})\boldsymbol{y}_{1}-D_{1}^{-1}M\boldsymbol{y}_{0}), we can rewrite (32) as

𝒚Tt𝒚(At𝒚)Tt(At𝒚)=(𝒚At𝒚)T[I𝟎𝟎D1+R1](𝒚At𝒚).\displaystyle\boldsymbol{y}^{T}\!\mathcal{L}_{t}\boldsymbol{y}\!-\!(A_{t}\boldsymbol{y})^{T}\!\mathcal{L}_{t}(A_{t}\boldsymbol{y})\!=\!(\boldsymbol{y}\!-\!A_{t}\boldsymbol{y})^{T}\!\!\left[\!\!\begin{array}[]{cc}I&\!\!\!\boldsymbol{0}\\ \boldsymbol{0}&\!\!\!D_{1}\!+\!R_{1}\end{array}\!\!\!\!\right]\!\!(\boldsymbol{y}\!-\!A_{t}\boldsymbol{y}).

Therefore, if we define QQ to be the middle matrix in (III-C), QQ would be a positive definite matrix (as its diagonal elements are strictly positive and dominate the row-sums) and thus 𝒚Tt𝒚(At𝒚)Tt(At𝒚)=𝒚At𝒚Q20\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y}-(A_{t}\boldsymbol{y})^{T}\!\mathcal{L}_{t}(A_{t}\boldsymbol{y})=\|\boldsymbol{y}-A_{t}\boldsymbol{y}\|^{2}_{Q}\geq 0. Finally, choosing 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t) we get,

𝒙T(t)t𝒙(t)𝒙T(t+1)t𝒙(t+1)=𝒙1(t)𝒙1(t+1)Q2.\displaystyle\boldsymbol{x}^{T}\!(t)\mathcal{L}_{t}\boldsymbol{x}(t)\!-\!\boldsymbol{x}^{T}\!(t\!+\!1)\mathcal{L}_{t}\boldsymbol{x}(t\!+\!1)\!=\!\|\boldsymbol{x}_{1}(t)\!-\!\boldsymbol{x}_{1}(t\!+\!1)\|^{2}_{Q}.

Therefore, V(𝒚):=min𝝀[0,1]n2f(𝒚,𝝀)=i,j((yiyj)2ϵ2)V(\boldsymbol{y}):=\min_{\boldsymbol{\lambda}\in[0,1]^{n^{2}}}f(\boldsymbol{y},\boldsymbol{\lambda})=\sum_{i,j}\big((y_{i}-y_{j})^{2}-\epsilon^{2}\big)^{-} serves as a Lyapunov function for the 0-1 HK model so that V(𝒙(t))V(𝒙(t+1))𝒙1(t)𝒙1(t+1)Q2V(\boldsymbol{x}(t))-V(\boldsymbol{x}(t+1))\geq\|\boldsymbol{x}_{1}(t)\!-\!\boldsymbol{x}_{1}(t\!+\!1)\|^{2}_{Q}. Q.E.D.

Remark 3

One can view the update matrix AtA_{t} of the 0-1 HK model at a given time tt as the transition matrix of a lazy simple random walk with absorbing states S0S_{0} on the fixed actual communication graph at time tt. Therefore, in the second part of the proof of Theorem 3 we have shown that although the actual graph might have one-sided directed edges from S1S_{1} to the absorbing states S0S_{0}, 𝐲Tt𝐲\boldsymbol{y}^{T}\mathcal{L}_{t}\boldsymbol{y} still serves as a Lyapunov function for such absorbing random walks where t\mathcal{L}_{t} is the Laplacian of the symmetrized actual network 𝛌t\boldsymbol{\lambda}_{t} (i.e., viewing one-sided edges as undirected edges).

As we close this section, we would like to mention that unlike homogeneous HK model which is known to converge to an equilibrium point after finitely many steps [21], it may take arbitrarily long time until the 0-1 HK dynamics converge. In fact, it seems impossible to show that the drift of the above Lyapunov function is bounded below by a time-invariant quantity (as is the case for homogeneous HK model [43]). As an example, consider a set of n=3n=3 agents initially positioned at x1(0)=1+122mx_{1}(0)=-1+\frac{1}{2^{2^{m}}}, x2(0)=0x_{2}(0)=0, and x3(0)=1x_{3}(0)=1 where mm can be any arbitrary large integer. Also assume ϵ1=ϵ2=0\epsilon_{1}=\epsilon_{2}=0, and ϵ3=1\epsilon_{3}=1 so that agent 33 is the only moving agent. Then it takes 2m2^{m} steps until a switch in the communication network occurs so that agent 33 be able to observe agent 11. In particular, the drift of the above Lyapunov function during iteration t{1,,2m}t\in\{1,\ldots,2^{m}\} is 12t12t+1=12t+1\frac{1}{2^{t}}-\frac{1}{2^{t+1}}=\frac{1}{2^{t+1}} which can be arbitrarily small.

IV Nearest Neighbor Opinion Dynamics

In general, loosing symmetry in communication networks of multi-agent systems can substantially complicate their stability analysis which in turn requires novel techniques. In fact, unlike the symmetric case, existing results concerning stability of state-dependent networks of multi-agent systems with asymmetric communication typologies are quite limited. Nevertheless, this shall not eliminate the possibility of convergence of asymmetric dynamics to an equilibrium point, as it is shown in this section for a special class of nearest neighbor dynamics. More specifically, in this section we establish convergence of a class of nearest neighbor dynamics under both asynchronous and synchronous settings, where in the former at each time instance only one of the agents updates its opinion (state), while in the latter at each iteration all the agents simultaneously update their opinions.

IV-A Asynchronous Nearest Neighbor Dynamics

Consider a set of [n][n] agents where the opinion of agent i[n]i\in[n] at time t=0,1,2,t=0,1,2,\ldots is given by a vector xi(t)dx_{i}(t)\in\mathbb{R}^{d}. At each iteration tt one agent i[n]i\in[n] is selected based on some selection rule (e.g. uniformly at random) and updates its opinion at the next time step to xi(t+1)=μixi(t)+(1μi)xr(i)(t)x_{i}(t+1)=\mu_{i}x_{i}(t)+(1-\mu_{i})x_{r(i)}(t), where here r(i):=argminj[n]{i}xi(t)xj(t)r(i):=\arg\min_{j\in[n]\setminus\{i\}}\|x_{i}(t)-x_{j}(t)\| denotes the closest agent to ii with respect to profile 𝒙(t)n×d\boldsymbol{x}(t)\in\mathbb{R}^{n\times d}, and μi(0,1)\mu_{i}\in(0,1) is an agent-specific parameter. For all other agents jij\neq i, we set xj(t+1)=xj(t)x_{j}(t+1)=x_{j}(t).

The rationale behind introducing nearest neighbor dynamics is that often individuals get influenced by their closest friend/partner/leader so that depending on their stubbornness (captured by μi\mu_{i}) they are willing to compromise in order to get closer to their friend/partner/leader. It is important to note that the communication network in the nearest neighbor dynamics is asymmetric so that if r(i)r(i) is the closest agent to ii, it does not imply that ii is also the closest agent to r(i)r(i) (i.e., in general r(r(i))ir(r(i))\neq i).55 5 In fact, one can show that the communication network at each time instance is comprised of disjoint directed trees where the out-degree of each node is equal to 1. Moreover, the communication network which determines the “closest relationships” evolves as a function of agents’ opinions which can switch many times based on trajectory of the dynamics. Nevertheless, as we shall see in Theorem 4 such heterogeneous asymmetric state-dependent dynamics will converge to an equilibrium point as defined below:

Definition 2

An ϵ\epsilon-equilibrium for the nearest neighbor dynamics is an opinion profile where the maximum distance between every agent’s opinion and its closest neighbor is at most ϵ>0\epsilon>0. Moreover, given an initial opinion profile 𝐱(0)\boldsymbol{x}(0), we let tϵt_{\epsilon} be the first time instance when 𝐱(tϵ)\boldsymbol{x}(t_{\epsilon}) becomes an ϵ\epsilon-equilibrium.

It is worth noting that the definition of ϵ\epsilon-equilibrium implies that each agent will lie within a distance of at most ϵn\epsilon n from its limit point. This is because if the communication network in an ϵ\epsilon-equilibrium is connected, the maximum distance between every two nodes is at most ϵn\epsilon n so that the convex hull of all the opinions can have diameter at most ϵn\epsilon n. Since the nearest neighbor dynamics evolve inside of this convex hull for all the future iterations, the limit point (if it exists) will also lie in this convex hull. Similarly, if an ϵ\epsilon-equilibrium contains more than one connected component, then either the distance between their convex hulls is less than ϵ\epsilon, in which case a similar argument as above for the convex hull of the union of those components can be applied, or the distance between those components is more than ϵ\epsilon, in which case those components evolve separately from each other so that the limit points of each component lie within its own convex hull (and again the above argument applies).

Definition 3

Given two vectors 𝐮,𝐯n\boldsymbol{u},\boldsymbol{v}\in\mathbb{R}^{n}, we say 𝐮\boldsymbol{u} is lexicographically smaller than 𝐯\boldsymbol{v} (𝐮<Lex𝐯\boldsymbol{u}<_{Lex}\boldsymbol{v}), if there exists some knk\leq n for which u1=v1,,uk1=vk1u_{1}=v_{1},\ldots,u_{k-1}=v_{k-1}, and uk<vku_{k}<v_{k}. Note that there is no specific relation between components of 𝐮\boldsymbol{u} and 𝐯\boldsymbol{v} for indices larger than kk.

Theorem 4

The asynchronous nearest neighbor dynamics are Lyapunov stable and asymptotically converge to an equilibrium point. Moreover, if at each iteration an agent is selected uniformly at random to update its opinion, then the expected number of steps until the dynamics reach an ϵ\epsilon-equilibrium is bounded above by 𝔼[tϵ]n2nD0(1μmax)ϵ\mathbb{E}[t_{\epsilon}]\leq\frac{n2^{n}D_{0}}{(1-\mu_{\max})\epsilon}, where D0=maxi,jxi(0)xj(0)D_{0}=\max_{i,j}\|x_{i}(0)-x_{j}(0)\| and μmax=maxiμi\mu_{\max}=\max_{i}\mu_{i}.

Proof:

Given a vector 𝒗\boldsymbol{v}, let sort(𝒗)sort(\boldsymbol{v}) be a vector obtained by sorting all the components of 𝒗\boldsymbol{v} in a nondecreasing order. Let f:[0,1]n2×n×dnf:[0,1]^{n^{2}}\times\mathbb{R}^{n\times d}\to\mathbb{R}^{n} be the vector function f(𝝀,𝒚):=sort(j=1nλijyiyj,i[n])f(\boldsymbol{\lambda},\boldsymbol{y}):=sort\Big(\sum_{j=1}^{n}\lambda_{ij}\|y_{i}-y_{j}\|,i\in[n]\Big), and consider the lexicographical minimization problem, minLex{f(𝝀,𝒚):𝒚n×d,𝝀Λ}\min_{Lex}\{f(\boldsymbol{\lambda},\boldsymbol{y}):\boldsymbol{y}\in\mathbb{R}^{n\times d},\boldsymbol{\lambda}\in\Lambda\}, where Λ={(λij)[0,1]n2:j=1nλij=1,λii=0,i[n]}\Lambda=\{(\lambda_{ij})\in[0,1]^{n^{2}}:\sum_{j=1}^{n}\lambda_{ij}=1,\ \lambda_{ii}=0,\forall i\in[n]\}.66 6 In terms of Proposition 1 terminology, here we have (S,S)=(n,<Lex)(S,\leq_{S})=(\mathbb{R}^{n},<_{Lex}). Now given a fixed block variable 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t) at a generic time tt, minimizing f(𝝀,𝒙(t))f(\boldsymbol{\lambda},\boldsymbol{x}(t)) lexicographically with respect to 𝝀Λ\boldsymbol{\lambda}\in\Lambda gives us,

(𝝀t)ij={1if j=r(i),0else.\displaystyle(\boldsymbol{\lambda}_{t})_{ij}=\begin{cases}1&\mbox{if }j=r(i),\\ 0&\mbox{else}.\end{cases} (37)

This is because minimizing f(𝝀,𝒙(t))f(\boldsymbol{\lambda},\boldsymbol{x}(t)) over 𝝀Λ\boldsymbol{\lambda}\in\Lambda decomposes into minimizing f()f(\cdot) componentwise, and it is achieved by setting the coefficient λij\lambda_{ij} corresponding to the smallest term xi(t)xr(i)(t)\|x_{i}(t)-x_{r(i)}(t)\| of the iith component equal to 1, and to 0, otherwise. As a result, given a fixed state 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t), minimizing f(𝝀,𝒙(t))f(\boldsymbol{\lambda},\boldsymbol{x}(t)) with respect to 𝝀S\boldsymbol{\lambda}\in S accurately captures the directed communication network 𝝀t\boldsymbol{\lambda}_{t} of the nearest neighbor dynamics for the state 𝒙(t)\boldsymbol{x}(t). Next let us fix the communication network to 𝝀t\boldsymbol{\lambda}_{t} so that

f(𝝀t,𝒙(t))\displaystyle f(\boldsymbol{\lambda}_{t},\boldsymbol{x}(t)) =sort(xi(t)xr(i)(t),i[n])\displaystyle=sort\Big(\|x_{i}(t)-x_{r(i)}(t)\|,i\in[n]\Big)
=(x1(t)xr(1)(t),,xn(t)xr(n)(t)),\displaystyle=\Big(\|x_{1}(t)-x_{r(1)}(t)\|,\ldots,\|x_{n}(t)-x_{r(n)}(t)\|\Big), (38)

where in the second equality and without loss of generality (by relabeling the agents if necessary) we have assumed that x1(t)xr(1)(t)xn(t)xr(n)(t)\|x_{1}(t)-x_{r(1)}(t)\|\leq\ldots\leq\|x_{n}(t)-x_{r(n)}(t)\|.

To study the effect of state update on f(𝝀t,𝒙(t))f(\boldsymbol{\lambda}_{t},\boldsymbol{x}(t)), let us assume that at time tt agent [n]\ell\in[n] is selected to update its opinion. Then we obtain 𝒙(t+1)\boldsymbol{x}(t+1) for which x(t+1)=μx(t)+(1μ)xr()(t)x_{\ell}(t+1)=\mu_{\ell}x_{\ell}(t)+(1-\mu_{\ell})x_{r(\ell)}(t), and xj(t+1)=xj(t),jx_{j}(t+1)=x_{j}(t),\forall j\neq\ell. In particular,

x(t+1)xr()(t+1)\displaystyle\|x_{\ell}(t\!+\!1)-x_{r^{\prime}(\ell)}(t\!+\!1)\| x(t+1)xr()(t+1)\displaystyle\leq\|x_{\ell}(t\!+\!1)-x_{r(\ell)}(t\!+\!1)\| (39)
=x(t+1)xr()(t)=μx(t)xr()(t)\displaystyle=\|x_{\ell}(t\!+\!1)-x_{r(\ell)}(t)\|=\mu_{\ell}\|x_{\ell}(t)-x_{r(\ell)}(t)\| (40)
<x(t)xr()(t),\displaystyle<\|x_{\ell}(t)-x_{r(\ell)}(t)\|, (41)

where r()r^{\prime}(\ell) denotes the closest agent to \ell with respect to the opinion profile 𝒙(t+1)\boldsymbol{x}(t+1). Furthermore, for every i<i<\ell we have two possibilities: Case I) r(i)r(i)\neq\ell, in which case

xi(t+1)xr(i)(t+1)\displaystyle\|x_{i}(t\!+\!1)\!-\!x_{r^{\prime}(i)}(t\!+\!1)\| =min{xi(t)xr(i)(t),xi(t)x(t+1)}xi(t)xr(i)(t).\displaystyle=\min\{\|x_{i}(t)\!-\!x_{r(i)}(t)\|,\|x_{i}(t)\!-\!x_{\ell}(t\!+\!1)\|\}\leq\|x_{i}(t)\!-\!x_{r(i)}(t)\|.

Case II) r(i)=r(i)=\ell, in which case by definition of r()r(\ell) we must have x(t)xr()(t)x(t)xi(t)=xr(i)(t)xi(t)\|x_{\ell}(t)-x_{r(\ell)}(t)\|\leq\|x_{\ell}(t)-x_{i}(t)\|=\|x_{r(i)}(t)-x_{i}(t)\|. However, as i<i<\ell (and thus xr(i)(t)xi(t)x(t)xr()(t)\|x_{r(i)}(t)-x_{i}(t)\|\leq\|x_{\ell}(t)-x_{r(\ell)}(t)\|), one can see that Case II cannot happen unless x(t)xr()(t)=xi(t)xr(i)(t)\|x_{\ell}(t)-x_{r(\ell)}(t)\|=\|x_{i}(t)-x_{r(i)}(t)\|. As a result r()=ir(\ell)=i, and we can write

xi(t+1)xr(i)(t+1)\displaystyle\|x_{i}(t\!+\!1)-x_{r^{\prime}(i)}(t\!+\!1)\| =xi(t)x(t+1)\displaystyle=\|x_{i}(t)-x_{\ell}(t\!+\!1)\|
=xi(t)μx(t)(1μ)xr()(t)\displaystyle=\|x_{i}(t)-\mu_{\ell}x_{\ell}(t)-(1-\mu_{\ell})x_{r(\ell)}(t)\| (42)
=xi(t)μxr(i)(t)(1μ)xi(t)\displaystyle=\|x_{i}(t)-\mu_{\ell}x_{r(i)}(t)-(1-\mu_{\ell})x_{i}(t)\| (43)
=μxi(t)xr(i)(t)<xi(t)xr(i)(t).\displaystyle=\mu_{\ell}\|x_{i}(t)-x_{r(i)}(t)\|<\|x_{i}(t)-x_{r(i)}(t)\|. (44)

Therefore, we have shown that xi(t+1)xr(i)(t+1)xi(t)xr(i)(t),i<\|x_{i}(t\!+\!1)-x_{r^{\prime}(i)}(t\!+\!1)\|\leq\|x_{i}(t)-x_{r(i)}(t)\|,\forall i<\ell, and moreover x(t+1)xr()(t+1)<x(t)xr()(t)\|x_{\ell}(t\!+\!1)-x_{r^{\prime}(\ell)}(t\!+\!1)\|<\|x_{\ell}(t)-x_{r(\ell)}(t)\|. In other words, after agent \ell’s update, f(𝝀t,𝒙(t))f(\boldsymbol{\lambda}_{t},\boldsymbol{x}(t)) decreases lexicographically, i.e., f(𝝀t+1,𝒙(t+1))<Lexf(𝝀t,𝒙(t))f(\boldsymbol{\lambda}_{t+1},\boldsymbol{x}(t\!+\!1))\!<_{Lex}\!f(\boldsymbol{\lambda}_{t},\boldsymbol{x}(t)). In particular, V(𝒚):=min𝝀Λf(𝝀,𝒚)=sort(ykyr(k),k[n])V(\boldsymbol{y}):=\min_{\boldsymbol{\lambda}\in\Lambda}f(\boldsymbol{\lambda},\boldsymbol{y})=sort\Big(\|y_{k}-y_{r(k)}\|,k\in[n]\Big) serves as a Lyapunov function for the asynchronous nearest neighbor dynamics.

To show convergence of the asynchronous dynamics to an equilibrium point, let us convert V(𝒚)V(\boldsymbol{y}) into a scalar function V^(𝒚):=i=1nmini({ykyr(k),k[n]})2ni\hat{V}(\boldsymbol{y}):=\sum_{i=1}^{n}\min_{i}(\{\|y_{k}-y_{r(k)}\|,k\in[n]\})2^{n-i} by giving appropriate weights to its coordinates, where here mini(S)\min_{i}(S) denotes the iith smallest element of a finite set SS.77 7 This conversion encodes lexicographical decrease of V(𝒚)V(\boldsymbol{y}) into a scalar decrease in V^(𝒚)\hat{V}(\boldsymbol{y}). This will allow us to quantify a convergence rate for the asynchronous nearest neighbor dynamics. As before, let us assume that at time tt agent \ell is selected to update its opinion, and mini({xk(t)xr(k)(t),k[n]})=xi(t)xr(i)(t)\min_{i}(\{\|x_{k}(t)-x_{r(k)}(t)\|,k\in[n]\})=\|x_{i}(t)-x_{r(i)}(t)\|. For any ii\leq\ell, we have

mini({xk(t+1)xr(k)(t+1),k[n]})\displaystyle\min_{i}\Big(\{\|x_{k}(t+1)-x_{r^{\prime}(k)}(t+1)\|,k\in[n]\}\Big) (45)
max{xk(t+1)xr(k)(t+1),k[i]}\displaystyle\qquad\qquad\leq\max\{\|x_{k}(t+1)-x_{r^{\prime}(k)}(t+1)\|,k\in[i]\} (46)
max{xk(t)xr(k)(t),k[i]}=xi(t)xr(i)(t),\displaystyle\qquad\qquad\leq\max\{\|x_{k}(t)-x_{r^{\prime}(k)}(t)\|,k\in[i]\}=\|x_{i}(t)-x_{r(i)}(t)\|, (47)

where the last inequality is by xk(t+1)xr(k)(t+1)xk(t)xr(k)(t),k\|x_{k}(t+1)-x_{r^{\prime}(k)}(t+1)\|\leq\|x_{k}(t)-x_{r(k)}(t)\|,\forall k\leq\ell. Now we can write,

i(mini({xk(t)xr(k)(t),k[n]})mini({xk(t+1)xr(k)(t+1),i[n]}))2ni\displaystyle\sum_{i\leq\ell}\Big(\min_{i}\big(\{\|x_{k}(t)-x_{r(k)}(t)\|,k\in[n]\}\big)-\min_{i}\big(\{\|x_{k}(t+1)\!-\!x_{r^{\prime}(k)}(t+1)\|,i\in[n]\}\big)\Big)2^{n-i} (48)
=i(xi(t)xr(i)(t)mini({xk(t+1)xr(k)(t+1),k[n]}))2ni\displaystyle\qquad=\sum_{i\leq\ell}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\min_{i}\big(\{\|x_{k}(t+1)\!-\!x_{r^{\prime}(k)}(t+1)\|,k\in[n]\}\big)\Big)2^{n-i} (49)
2ni(xi(t)xr(i)(t)mini({xk(t+1)xr(k)(t+1),k[n]}))\displaystyle\qquad\geq 2^{n-\ell}\sum_{i\leq\ell}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\min_{i}(\{\|x_{k}(t+1)\!-\!x_{r^{\prime}(k)}(t+1)\|,k\in[n]\})\Big) (50)
=2n(ixi(t)xr(i)(t)imini({xk(t+1)xr(k)(t+1),k[n]}))\displaystyle\qquad=2^{n-\ell}\Big(\sum_{i\leq\ell}\|x_{i}(t)-x_{r(i)}(t)\|-\sum_{i\leq\ell}\min_{i}(\{\|x_{k}(t+1)\!-\!x_{r^{\prime}(k)}(t+1)\|,k\in[n]\})\Big) (51)
2n(ixi(t)xr(i)(t)ixi(t+1)xr(i)(t+1))\displaystyle\qquad\geq 2^{n-\ell}\Big(\sum_{i\leq\ell}\|x_{i}(t)-x_{r(i)}(t)\|-\sum_{i\leq\ell}\|x_{i}(t+1)\!-\!x_{r^{\prime}(i)}(t+1)\|\Big) (52)
=2ni(xi(t)xr(i)(t)xi(t+1)xr(i)(t+1))\displaystyle\qquad=2^{n-\ell}\sum_{i\leq\ell}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\|x_{i}(t+1)\!-\!x_{r^{\prime}(i)}(t+1)\|\Big) (53)
2n(1μ)x(t)xr()(t),\displaystyle\qquad\geq 2^{n-\ell}(1-\mu_{\ell})\|x_{\ell}(t)-x_{r(\ell)}(t)\|, (54)

where the first inequality is due (45), and the second inequality holds because sum of \ell smallest elements of a set is always smaller than sum of any \ell elements in that set. Finally, the last inequality is valid since the first 1\ell-1 summands are nonnegative and by (39) the \ellth summand is greater than or equal to (1μ)x(t)xr()(t)(1-\mu_{\ell})\|x_{\ell}(t)-x_{r(\ell)}(t)\|.

On the other hand, we note that for any agent kk, xk(t+1)xr(k)(t+1)xk(t)xr(k)(t)+(1μ)xk(t)xr(k)(t)\|x_{k}(t+1)\!-\!x_{r^{\prime}(k)}(t+1)\|\leq\|x_{k}(t)\!-\!x_{r^{\prime}(k)}(t)\|+(1-\mu_{\ell})\|x_{k}(t)\!-\!x_{r^{\prime}(k)}(t)\|. This is because the amount of movement of agent \ell at time tt is equal to x(t+1)x(t)=(1μ)x(t)xr()(t)\|x_{\ell}(t+1)-x_{\ell}(t)\|=(1-\mu_{\ell})\|x_{\ell}(t)-x_{r(\ell)}(t)\|. Therefore, by triangle inequality the distance of any agent kk to its nearest neighbor at the next time step can increase by at most (1μ)x(t)xr()(t)(1-\mu_{\ell})\|x_{\ell}(t)-x_{r(\ell)}(t)\|. Now similar as in (45) we can write,

mini({xk(t+1)xr(k)(t+1),k[n]})\displaystyle\min_{i}\Big(\{\|x_{k}(t+1)-x_{r^{\prime}(k)}(t+1)\|,k\in[n]\}\Big) (55)
max{xk(t+1)xr(k)(t+1),k[i]}\displaystyle\qquad\qquad\leq\max\{\|x_{k}(t+1)-x_{r^{\prime}(k)}(t+1)\|,k\in[i]\} (56)
max{xk(t)xr(k)(t)+(1μ)x(t)xr()(t),k[i]}\displaystyle\qquad\qquad\leq\max\{\|x_{k}(t)-x_{r^{\prime}(k)}(t)\|+(1-\mu_{\ell})\|x_{\ell}(t)-x_{r^{\prime}(\ell)}(t)\|,k\in[i]\} (57)
=xi(t)xr(i)(t)+(1μ)x(t)xr()(t).\displaystyle\qquad\qquad=\|x_{i}(t)-x_{r(i)}(t)\|+(1-\mu_{\ell})\|x_{\ell}(t)-x_{r^{\prime}(\ell)}(t)\|. (58)

As a result we obtain,

i>(mini({xk(t)xr(k)(t),k[n]})mini({xk(t+1)xr(k)(t+1),i[n]}))2ni\displaystyle\sum_{i>\ell}\Big(\min_{i}(\{\|x_{k}(t)-x_{r(k)}(t)\|,k\in[n]\})-\min_{i}(\{\|x_{k}(t+1)\!-\!x_{r^{\prime}(k)}(t+1)\|,i\in[n]\})\Big)2^{n-i} (59)
=i>(xi(t)xr(i)(t)mini({xk(t+1)xr(k)(t+1),k[n]}))2ni\displaystyle\qquad=\sum_{i>\ell}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\min_{i}(\{\|x_{k}(t\!+\!1)\!-\!x_{r^{\prime}(k)}(t+1)\|,k\in[n]\})\Big)2^{n-i} (60)
i>(1μ)x(t)xr()(t)2ni\displaystyle\qquad\geq-\sum_{i>\ell}(1-\mu_{\ell})\|x_{\ell}(t)-x_{r^{\prime}(\ell)}(t)\|2^{n-i} (61)
=(1μ)x(t)xr()(t)(2n1),\displaystyle\qquad=-(1-\mu_{\ell})\|x_{\ell}(t)-x_{r^{\prime}(\ell)}(t)\|(2^{n-\ell}-1), (62)

where the inequality is due to (55). Finally, summing (59) and (48), and using the definition of V^()\hat{V}(\cdot), we obtain V^(𝒙(t))V^(𝒙(t+1))(1μ)x(t)xr()(t)\hat{V}(\boldsymbol{x}(t))-\hat{V}(\boldsymbol{x}(t+1))\geq(1-\mu_{\ell})\|x_{\ell}(t)\!-\!x_{r^{\prime}(\ell)}(t)\|. As agent \ell is selected uniformly at random, this implies that as long as 𝒙(t)\boldsymbol{x}(t) is not an ϵ\epsilon-equilibrium, the expected decrease of V^()\hat{V}(\cdot) at time tt is at least,

𝔼[V^(𝒙(t))V^(𝒙(t+1))]1n=1n(1μ)x(t)xr()(t)(1μmax)ϵn.\displaystyle\mathbb{E}[\hat{V}(\boldsymbol{x}(t))-\hat{V}(\boldsymbol{x}(t+1))]\geq\frac{1}{n}\sum_{\ell=1}^{n}(1-\mu_{\ell})\|x_{\ell}(t)\!-\!x_{r^{\prime}(\ell)}(t)\|\geq\frac{(1-\mu_{\max})\epsilon}{n}. (63)

Finally, we note that V^()\hat{V}(\cdot) is a nonnegative function such that V^(𝒙(0))2nD0\hat{V}(\boldsymbol{x}(0))\leq 2^{n}D_{0}. This in view of (63) shows that the expected number of steps before the asynchronous dynamics reach an ϵ\epsilon-equilibrium is bounded above by 2nD0(1μmax)ϵn\frac{2^{n}D_{0}}{\frac{(1-\mu_{\max})\epsilon}{n}}. Q.E.D.

IV-B Synchronous Nearest Neighbor Dynamics

In this part we consider the synchronous nearest neighbor dynamics whose formal definition is as follows: Consider a set of [n][n] agents where the opinion of agent i[n]i\in[n] at time t=0,1,2,t=0,1,2,\ldots is given by xi(t)dx_{i}(t)\in\mathbb{R}^{d}. Given the current opinion profile 𝒙(t)\boldsymbol{x}(t), in the next iteration every agent i[n]i\in[n] updates its opinion to xi(t+1)=μixi(t)+(1μi)xr(i)(t)x_{i}(t+1)=\mu_{i}x_{i}(t)+(1-\mu_{i})x_{r(i)}(t), where as before r(i):=argminj[n]{i}xi(t)xj(t)r(i):=\arg\min_{j\in[n]\setminus\{i\}}\|x_{i}(t)-x_{j}(t)\| denotes the closest agent to ii with respect to the opinion profile 𝒙(t)\boldsymbol{x}(t), and μi(0,1),i[n]\mu_{i}\in(0,1),i\in[n] are mixing constants. As before, we note that the communication network of the synchronous dynamics is asymmetric whose evolution depends on the opinion profiles.

In the following we show that if all the agents have the same mixing parameter μi=μ,i[n]\mu_{i}=\mu,\forall i\in[n], then the synchronous nearest neighbor dynamics reach an ϵ\epsilon-equilibrium very fast.

Theorem 5

The synchronous nearest neighbor dynamics with μi:=μ(0,1)\mu_{i}:=\mu\in(0,1), i[n]\forall i\in[n] converge to an ϵ\epsilon-equilibrium point after at most tϵn(2D0ϵ+log|12μ|(ϵ2D0))t_{\epsilon}\leq n(\frac{2D_{0}}{\epsilon}+\log_{|1-2\mu|}(\frac{\epsilon}{2D_{0}})) iterations, where D0=maxi,jxi(0)xj(0)D_{0}=\max_{i,j}\|x_{i}(0)-x_{j}(0)\|.

Proof:

Consider BCD method for component-wise minimizing of the function:

min{f(𝒚,𝝀):\displaystyle\min\Big\{f(\boldsymbol{y},\boldsymbol{\lambda})\!: =(j=1nλ1jy1yj,,j=1nλnjynyj)|𝝀Λ,𝒚n×d},\displaystyle=\!\Big(\sum_{j=1}^{n}\lambda_{1j}\|y_{1}-y_{j}\|,\ldots,\sum_{j=1}^{n}\lambda_{nj}\|y_{n}-y_{j}\|\Big)|\ \boldsymbol{\lambda}\in\Lambda,\ \boldsymbol{y}\in\mathbb{R}^{n\times d}\Big\},

where Λ={(λij)[0,1]n2:j=1nλij=1,λii=0,i}\Lambda=\{(\lambda_{ij})\in[0,1]^{n^{2}}:\sum_{j=1}^{n}\lambda_{ij}=1,\ \lambda_{ii}=0,\forall i\}. As before, given a fixed opinion state 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t), minimizing f()f(\cdot) over 𝝀Λ\boldsymbol{\lambda}\in\Lambda gives us (𝝀t)ij=1(\boldsymbol{\lambda}_{t})_{ij}=1 if j=r(i)j=r(i), and (𝝀t)ij=0(\boldsymbol{\lambda}_{t})_{ij}=0 otherwise, where r()r(\cdot) is defined with respect to the opinion profile 𝒙(t)\boldsymbol{x}(t). Therefore, for a fixed state 𝒚=𝒙(t)\boldsymbol{y}=\boldsymbol{x}(t), the communication network is precisely captured by minimizing f()f(\cdot) over 𝝀Λ\boldsymbol{\lambda}\in\Lambda. Now by fixing the network to this minimizer 𝝀t\boldsymbol{\lambda}_{t}, we get f(𝒙(t),𝝀t)=(x1(t)xr(1)(t),,xn(t)xr(n)(t))f(\boldsymbol{x}(t),\boldsymbol{\lambda}_{t})=(\|x_{1}(t)-x_{r(1)}(t)\|,\ldots,\|x_{n}(t)-x_{r(n)}(t)\|), which we next evaluate the effect of opinion updates on it.

Let 𝒙(t+1)\boldsymbol{x}(t+1) be the updated opinion vector obtained from 𝒙(t)\boldsymbol{x}(t), given the fixed network 𝝀t\boldsymbol{\lambda}_{t}. By definition of synchronous dynamics xi(t+1)=μxi(t)+(1μ)xr(i)(t),ix_{i}(t+1)=\mu x_{i}(t)+(1-\mu)x_{r(i)}(t),\forall i,

xi(t+1)xr(i)(t+1)\displaystyle\|x_{i}(t\!+\!1)\!-\!x_{r(i)}(t\!+\!1)\| =μxi(t)+(μ)xr(i)(t)(μxr(i)(t)+(μ)xr(r(i))(t))\displaystyle=\|\mu x_{i}(t)\!+\!(1\!-\!\mu)x_{r(i)}(t)\!-\!\big(\mu x_{r(i)}(t)\!+\!(1\!-\!\mu)x_{r(r(i))}(t)\big)\|
=μ(xi(t)xr(i)(t))+(1μ)(xr(i)(t)xr(r(i))(t))\displaystyle=\|\mu(x_{i}(t)-x_{r(i)}(t))+(1-\mu)\big(x_{r(i)}(t)-x_{r(r(i))}(t)\big)\| (64)
μxi(t)xr(i)(t)+(1μ)xr(i)(t)xr(r(i))(t)\displaystyle\leq\mu\|x_{i}(t)-x_{r(i)}(t)\|+(1-\mu)\|x_{r(i)}(t)-x_{r(r(i))}(t)\| (65)
μxi(t)xr(i)(t)+(1μ)xi(t)xr(i)(t)\displaystyle\leq\mu\|x_{i}(t)-x_{r(i)}(t)\|+(1-\mu)\|x_{i}(t)-x_{r(i)}(t)\| (66)
=xi(t)xr(i)(t),\displaystyle=\|x_{i}(t)-x_{r(i)}(t)\|, (67)

where the first inequality is due to the triangle inequality, and the last equality holds since by definition of r()r(\cdot) we have xr(i)(t)xr(r(i))(t)xi(t)xr(i)(t)\|x_{r(i)}(t)-x_{r(r(i))}(t)\|\leq\|x_{i}(t)-x_{r(i)}(t)\|. Therefore,

f(𝒙(t+1),𝝀t)\displaystyle f(\boldsymbol{x}(t\!+\!1),\boldsymbol{\lambda}_{t}) =(xi(t+1)xr(i)(t+1),i[n])(xi(t)xr(i)(t),i[n])=f(𝒙(t),𝝀t).\displaystyle\!=\!(\|x_{i}(t\!+\!1)\!-\!x_{r(i)}(t\!+\!1)\|,i\in[n])\leq(\|x_{i}(t)\!-\!x_{r(i)}(t)\|,i\in[n])\!=\!f(\boldsymbol{x}(t),\boldsymbol{\lambda}_{t}).

where the above inequality is component-wise. As a result, we have

f(𝒙(t+1),𝝀t+1)\displaystyle f(\boldsymbol{x}(t\!+\!1),\boldsymbol{\lambda}_{t+1}) =min𝝀Λf(𝒙(t+1),𝝀)f(𝒙(t+1),𝝀t)f(𝒙(t),𝝀t),\displaystyle=\min_{\boldsymbol{\lambda}\in\Lambda}f(\boldsymbol{x}(t\!+\!1),\boldsymbol{\lambda})\leq f(\boldsymbol{x}(t\!+\!1),\boldsymbol{\lambda}_{t})\leq f(\boldsymbol{x}(t),\boldsymbol{\lambda}_{t}),

which shows that V(𝒚):=min𝝀Λf(𝝀,𝒚)=(yi(t)yr(i)(t),i[n])V(\boldsymbol{y}):=\min_{\boldsymbol{\lambda}\in\Lambda}f(\boldsymbol{\lambda},\boldsymbol{y})=(\|y_{i}(t)-y_{r(i)}(t)\|,i\in[n]) serves as a Lyapunov function for the synchronous nearest neighbor dynamics.

To evaluate the convergence speed of the dynamics to an ϵ\epsilon-equilibrium point, let us consider the scalar function V^(𝒙)=i=1nxixr(i)\hat{V}(\boldsymbol{x})=\sum_{i=1}^{n}\|x_{i}\!-\!x_{r(i)}\|. Let CtC_{t} be a connected component of the communication graph at time tt which contains the longest edge Dt:=maxixi(t)xr(i)(t)D_{t}:=\max_{i}\|x_{i}(t)-x_{r(i)}(t)\|, and define dtd_{t} to be the length of the shortest edge in CtC_{t}. A similar analysis as above shows that

V^(x(t))V^(x(tCLOSECLOSE\displaystyle\hat{V}(x(t))-\hat{V}(x(t OPENOPEN+1))i=1n(xi(t)xr(i)(t)xi(t+1)xr(i)(t+1))\displaystyle+1))\geq\sum_{i=1}^{n}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\|x_{i}(t+1)-x_{r(i)}(t+1)\|\Big) (68)
iCt(xi(t)xr(i)(t)xi(t+1)xr(i)(t+1))\displaystyle\geq\sum_{i\in C_{t}}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\|x_{i}(t+1)-x_{r(i)}(t+1)\|\Big) (69)
(1μ)iCt(xi(t)xr(i)(t)xr(i)(t)xr(r(i))(t))\displaystyle\geq(1-\mu)\sum_{i\in C_{t}}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\|x_{r(i)}(t)-x_{r(r(i))}(t)\|\Big) (70)
(1μ)iPt(xi(t)xr(i)(t)xr(i)(t)xr(r(i))(t))\displaystyle\geq(1-\mu)\sum_{i\in P_{t}}\Big(\|x_{i}(t)-x_{r(i)}(t)\|-\|x_{r(i)}(t)-x_{r(r(i))}(t)\|\Big) (71)
=(1μ)(Dtdt),\displaystyle=(1-\mu)(D_{t}-d_{t}), (72)

where as before r()r(\cdot) is associated with state x(t)x(t), and PtP_{t} is the unique directed path connecting the longest edge to the shortest edge in CtC_{t}. Note that the last equality in (68) holds due to the telescopic sum over the nodes of PtP_{t}.

On the other hand, by definition of tϵt_{\epsilon} we know that Dtϵ,t<tϵD_{t}\geq\epsilon,\forall t<t_{\epsilon}. We claim that for at most nlog|12μ|(ϵ2D0)n\log_{|1-2\mu|}(\frac{\epsilon}{2D_{0}}) time instances t{0,,tϵ}t\in\{0,\ldots,t_{\epsilon}\} we can have dtϵ2d_{t}\geq\frac{\epsilon}{2}. This is because whenever an agent ii is an endpoint of the shortest edge in CtC_{t}, we get r(r(i))=ir(r(i))=i, and thus:

xi(t+1)xrt+1(i)(t+1)\displaystyle\|x_{i}(t+1)-x_{r_{t+1}(i)}(t+1)\| xi(t+1)xr(i)(t+1)\displaystyle\leq\|x_{i}(t+1)-x_{r(i)}(t+1)\|
=μxi(t)+(1μ)xr(i)(t)(μxr(i)(t)(1μ)xi(t))\displaystyle=\|\mu x_{i}(t)+(1-\mu)x_{r(i)}(t)-(\mu x_{r(i)}(t)-(1-\mu)x_{i}(t))\| (73)
=|12μ|xi(t)xr(i)(t).\displaystyle=|1-2\mu|\|x_{i}(t)-x_{r(i)}(t)\|. (74)

As a result the distance between agent ii and its nearest neighbor in the next time step will reduce by a factor of |12μ|(0,1)|1-2\mu|\in(0,1), and as we saw earlier, for each i[n]i\in[n] this distance can only decrease in the future iterations of the dynamics. This implies that agent i[n]i\in[n] can be incident to the shortest edge in CtC_{t} for at most log|12μ|(ϵ2D0)\log_{|1-2\mu|}(\frac{\epsilon}{2D_{0}}) time instances before its distance to its closest neighbor shrinks below ϵ2\frac{\epsilon}{2}. As there are in total nn agents, the claim follows. Therefore, for at least tϵnlog|12μ|(ϵ2D0)t_{\epsilon}-n\log_{|1-2\mu|}(\frac{\epsilon}{2D_{0}}) time instances we have dt<ϵ2d_{t}<\frac{\epsilon}{2} and DtϵD_{t}\geq\epsilon. By (68), for such instances the function V^()\hat{V}(\cdot) must decrease by at least Dtdtϵ2D_{t}-d_{t}\geq\frac{\epsilon}{2}. Since V^()\hat{V}(\cdot) is always nonnegative and V^(𝒙(0))nD0\hat{V}(\boldsymbol{x}(0))\leq nD_{0}, we must have (tϵnlog|12μ|(ϵ2D0))ϵ2<nD0(t_{\epsilon}-n\log_{|1-2\mu|}(\frac{\epsilon}{2D_{0}}))\frac{\epsilon}{2}<nD_{0}, as desired. Q.E.D.

V Applications to Game Theory

In this section we provide a game-theoretic application of our successive optimization framework and show how it can be leveraged to establish convergence of natural best response dynamics (or its other variants) towards equilibrium points of the underlying game. For this purpose let us consider a Stackelberg game with one leader (network designer) and nn followers (players). At each stage the leader decides about the network structure and the followers best respond to the leader’s action. The leader’s action is to choose a matrix 𝝀:=(λij)Λ\boldsymbol{\lambda}:=(\lambda_{ij})\in\Lambda, where λij\lambda_{ij} indicates the amount by which the cost of player ii is influenced by the action of player jj. After that each follower i[n]i\in[n] incurs a cost of

ci(xi,𝒙i,𝝀)=j=1nλijϕi(xi,xj),\displaystyle c_{i}(x_{i},\boldsymbol{x}_{-i},\boldsymbol{\lambda})=\sum_{j=1}^{n}\lambda_{ij}\phi_{i}(x_{i},x_{j}),

where xiXix_{i}\in X_{i} denotes the action taken by player ii, 𝒙i\boldsymbol{x}_{-i} denotes the actions of all players other than ii, and ϕi(xi,xj)\phi_{i}(x_{i},x_{j}) is a player-specific function which captures the coupling cost incurred by player ii due to the action of player jj. For each player i[n]i\in[n], we assume that ϕi(xi,xj)\phi_{i}(x_{i},x_{j}) is a continuous and strictly convex function of its own action xix_{i}, given fixed actions of all others (including the leader). Finally, in this game we assume that the leader’s objective is to minimize the social cost defined by c(𝒙,𝝀)=i=1nci(𝒙,𝝀)c(\boldsymbol{x},\boldsymbol{\lambda})=\sum_{i=1}^{n}c_{i}(\boldsymbol{x},\boldsymbol{\lambda}). This fully determines a Stackelberg game between the leader and the followers where the leader first determines the network structure and the followers best respond to it by playing a noncooperative game among themselves.

Definition 4

Let XX be a compact set, and consider a continuous function f(x,y):X×Xf(x,y):X\times X\to\mathbb{R}, where for each xXx\in X the mapping Z(x):=argminyXf(x,y)Z(x):=\arg\min_{y\in X}f(x,y) is singular value. f()f(\cdot) is called min-symmetric, if f(Z(x),Z(x))f(x,Z(x)),xXf(Z(x),Z(x))\leq f(x,Z(x)),\forall x\in X.

Remark 4

Quadratic functions of the form f(𝐱,𝐲):=𝐱TQ1𝐱+𝐲TQ2𝐲2𝐲TR𝐱f(\boldsymbol{x},\boldsymbol{y}):=\boldsymbol{x}^{T}Q_{1}\boldsymbol{x}+\boldsymbol{y}^{T}Q_{2}\boldsymbol{y}-2\boldsymbol{y}^{T}R\boldsymbol{x}, where Q1,Q2Q_{1},Q_{2} are positive definite matrices, and R=RTR=R^{T} is a symmetric matrix, are a subclass of min-symmetric functions (see, e.g., [40, Lemma 1]).

Theorem 6

Let Λ[0,1]n×n\Lambda\subseteq[0,1]^{n\times n} be a polytope, and Xi,i[n]X_{i},i\in[n] be compact sets. Moreover, assume that for every 𝛌Λ\boldsymbol{\lambda}\in\Lambda, the function ρ(𝐱,𝐲,𝛌):=i=1nci(yi,𝐱i,𝛌)\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda}):=\sum_{i=1}^{n}c_{i}(y_{i},\boldsymbol{x}_{-i},\boldsymbol{\lambda}) is min-symmetric. Then every limit point of the best response dynamics generated by the leader and followers will be a Stackelberg equilibrium.

Proof:

Given a generic time instance tt, let us first fix leader’s action to 𝝀t\boldsymbol{\lambda}_{t} and analyze the game among followers. Note that for fixed 𝝀t\boldsymbol{\lambda}_{t}, the cost function of each player ii is continuous and strictly convex function of its own action xix_{i}. Therefore, the unique best-response profile of all the followers at time t+1t+1 is given by 𝒙(t+1)=argmin𝒚Xρ(𝒙(t),𝒚,𝝀t)\boldsymbol{x}(t+1)=\arg\min_{\boldsymbol{y}\in X}\rho(\boldsymbol{x}(t),\boldsymbol{y},\boldsymbol{\lambda}_{t}), where X:=X1××XnX\!:=X_{1}\!\times\cdots\times X_{n} is a compact set. By min-symmetric property of ρ(𝒙,𝒚,𝝀t)\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda}_{t}), we can write,

c(𝒙(t+1),𝝀t)\displaystyle c(\boldsymbol{x}(t+1),\boldsymbol{\lambda}_{t}) =ρ(𝒙(t+1),𝒙(t+1),𝝀t)\displaystyle=\rho(\boldsymbol{x}(t+1),\boldsymbol{x}(t+1),\boldsymbol{\lambda}_{t}) (75)
ρ(𝒙(t),𝒙(t+1),𝝀t)\displaystyle\leq\rho(\boldsymbol{x}(t),\boldsymbol{x}(t+1),\boldsymbol{\lambda}_{t}) (76)
=min𝒚Xρ(𝒙(t),𝒚,𝝀t)\displaystyle=\min_{\boldsymbol{y}\in X}\rho(\boldsymbol{x}(t),\boldsymbol{y},\boldsymbol{\lambda}_{t}) (77)
ρ(𝒙(t),𝒙(t),𝝀t)=c(𝒙(t),𝝀t).\displaystyle\leq\rho(\boldsymbol{x}(t),\boldsymbol{x}(t),\boldsymbol{\lambda}_{t})=c(\boldsymbol{x}(t),\boldsymbol{\lambda}_{t}). (78)

Finally, the leader’s best action to the followers’ actions 𝒙(t+1)\boldsymbol{x}(t+1) is given by 𝝀t+1=argmin𝝀Λc(𝒙(t+1),𝝀)\boldsymbol{\lambda}_{t+1}=\arg\min_{\boldsymbol{\lambda}\in\Lambda}c(\boldsymbol{x}(t+1),\boldsymbol{\lambda}), which in view of (75) shows that for any time tt we have,

c(𝒙(t+1),𝝀t+1)c(𝒙(t),𝝀t).\displaystyle c(\boldsymbol{x}(t+1),\boldsymbol{\lambda}_{t+1})\leq c(\boldsymbol{x}(t),\boldsymbol{\lambda}_{t}). (79)

Next we note that c(𝒙,𝝀)c(\boldsymbol{x},\boldsymbol{\lambda}) is a linear function of 𝝀\boldsymbol{\lambda} which at each iteration is minimized over the polytope Λ\Lambda. Therefore, the set of leader’s best actions {λt}t=0\{\lambda_{t}\}_{t=0}^{\infty} is a finite set comprised of at most all the extreme points of Λ\Lambda. Now let (𝒙,𝝀)(\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*}) be an arbitrary limit point of the sequence {(𝒙(t),λt)}t=0\{(\boldsymbol{x}(t),\lambda_{t})\}_{t=0}^{\infty} (which does exist since this sequence belongs to the compact set X×ΛX\times\Lambda). As {λt}t=0\{\lambda_{t}\}_{t=0}^{\infty} is a finite set, this implies that there exists a subsequence {𝒙(tk)}k=0\{\boldsymbol{x}(t_{k})\}_{k=0}^{\infty} which converges to 𝒙\boldsymbol{x}^{*} under the fixed network topology 𝝀tk=𝝀,k0\boldsymbol{\lambda}_{t_{k}}=\boldsymbol{\lambda}^{*},\forall k\geq 0. To derive a contradiction, assume that 𝝀\boldsymbol{\lambda}^{*} is not the best response of the leader to the followers’ actions 𝒙\boldsymbol{x}^{*}. This means that there exists a 𝝀~𝝀\tilde{\boldsymbol{\lambda}}\neq\boldsymbol{\lambda}^{*} such that c(𝒙,𝝀~)<c(𝒙,𝝀)c(\boldsymbol{x}^{*},\tilde{\boldsymbol{\lambda}})<c(\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*}). By continuity of c(𝒙,𝝀)c(\boldsymbol{x},\boldsymbol{\lambda}) we have

limkc(𝒙(tk),𝝀~)=c(𝒙,𝝀~)<c(𝒙,𝝀)=limkc(𝒙(tk),𝝀tk).\displaystyle\lim_{k\to\infty}c(\boldsymbol{x}(t_{k}),\tilde{\boldsymbol{\lambda}})=c(\boldsymbol{x}^{*},\tilde{\boldsymbol{\lambda}})<c(\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})=\lim_{k\to\infty}c(\boldsymbol{x}(t_{k}),\boldsymbol{\lambda}_{t_{k}}).

This means that there exists a sufficiently large integer r>0r>0 such that c(𝒙(tr),𝝀~)<c(𝒙(tr),𝝀tr)c(\boldsymbol{x}(t_{r}),\tilde{\boldsymbol{\lambda}})<c(\boldsymbol{x}(t_{r}),\boldsymbol{\lambda}_{t_{r}}). But we already know that c(𝒙(tr),𝝀tr)=min𝝀Λc(𝒙(tr),𝝀)c(𝒙tk,𝝀~)c(\boldsymbol{x}(t_{r}),\boldsymbol{\lambda}_{t_{r}})=\min_{\boldsymbol{\lambda}\in\Lambda}c(\boldsymbol{x}(t_{r}),\boldsymbol{\lambda})\leq c(\boldsymbol{x}_{t_{k}},\tilde{\boldsymbol{\lambda}}), which is in contrast with the former inequality. This shows that 𝝀\boldsymbol{\lambda}^{*} is the best response of the leader to followers’ actions 𝒙\boldsymbol{x}^{*}.

Similarly, given leader’s action 𝝀\boldsymbol{\lambda}^{*}, let us assume that xx^{*} is not a Nash equilibrium among the followers. Define Z(𝒙):=argmin𝒚Xρ(𝒙,𝒚,𝝀)Z(\boldsymbol{x}):=\arg\min_{\boldsymbol{y}\in X}\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda}^{*}) to be the best response function of the followers, given the fixed leader’s action 𝝀\boldsymbol{\lambda}^{*}. Note that Z(x):XXZ(x):X\to X is a continuous function due to uniform continuity of ρ(,,𝝀)\rho(\cdot,\cdot,\boldsymbol{\lambda}^{*}) over the compact set X×XX\times X. Since xx^{*} is not a Nash equilibrium, this means that ρ(𝒙,Z(𝒙),𝝀)<ρ(𝒙,𝒙,𝝀)\rho(\boldsymbol{x}^{*},Z(\boldsymbol{x}^{*}),\boldsymbol{\lambda}^{*})<\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*}). As before, let {𝒙(tk)}k=0\{\boldsymbol{x}(t_{k})\}_{k=0}^{\infty} be a subsequence converging to 𝒙\boldsymbol{x}^{*} under the fixed leader’s action 𝝀tk=𝝀,k0\boldsymbol{\lambda}_{t_{k}}=\boldsymbol{\lambda}^{*},\forall k\geq 0. By continuity of Z(x)Z(x) and ρ(,,𝝀)\rho(\cdot,\cdot,\boldsymbol{\lambda}^{*}), we get

limkρ(𝒙(tk),Z(𝒙(tk)),𝝀tk)\displaystyle\lim_{k\to\infty}\rho(\boldsymbol{x}(t_{k}),Z(\boldsymbol{x}(t_{k})),\boldsymbol{\lambda}_{t_{k}}) =limkρ(𝒙(tk),Z(𝒙(tk)),𝝀)\displaystyle=\lim_{k\to\infty}\rho(\boldsymbol{x}(t_{k}),Z(\boldsymbol{x}(t_{k})),\boldsymbol{\lambda}^{*}) (80)
=ρ(𝒙,Z(𝒙),𝝀)<ρ(𝒙,𝒙,𝝀).\displaystyle=\rho(\boldsymbol{x}^{*},Z(\boldsymbol{x}^{*}),\boldsymbol{\lambda}^{*})<\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*}). (81)

Let δ:=ρ(𝒙,𝒙,𝝀)ρ(𝒙,Z(𝒙),𝝀)>0\delta:=\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})-\rho(\boldsymbol{x}^{*},Z(\boldsymbol{x}^{*}),\boldsymbol{\lambda}^{*})>0. From (80) and for sufficiently large integer rr, we have ρ(𝒙(tr),Z(𝒙(tr)),𝝀tr)<ρ(𝒙,𝒙,𝝀)δ2\rho(\boldsymbol{x}(t_{r}),Z(\boldsymbol{x}(t_{r})),\boldsymbol{\lambda}_{t_{r}})<\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})-\frac{\delta}{2}. By definition we know that 𝒙(tr+1)=Z(𝒙(tr))\boldsymbol{x}(t_{r}+1)=Z(\boldsymbol{x}(t_{r})), thus by following the same argument as in (75) and (79),

c(𝒙(tr+1),𝝀tr+1)\displaystyle c(\boldsymbol{x}(t_{r}+1),\boldsymbol{\lambda}_{t_{r}+1}) =ρ(𝒙(tr+1),𝒙(tr+1),𝝀tr+1)\displaystyle=\rho(\boldsymbol{x}(t_{r}+1),\boldsymbol{x}(t_{r}+1),\boldsymbol{\lambda}_{t_{r}+1})
ρ(𝒙(tr+1),𝒙(tr+1),𝝀tr)\displaystyle\leq\rho(\boldsymbol{x}(t_{r}+1),\boldsymbol{x}(t_{r}+1),\boldsymbol{\lambda}_{t_{r}}) (82)
ρ(𝒙(tr),Z(𝒙(tr)),𝝀tr)\displaystyle\leq\rho(\boldsymbol{x}(t_{r}),Z(\boldsymbol{x}(t_{r})),\boldsymbol{\lambda}_{t_{r}}) (83)
<ρ(𝒙,𝒙,𝝀)δ2.\displaystyle<\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})-\frac{\delta}{2}. (84)

But from (79) we know that c(𝒙(t),𝝀t)c(\boldsymbol{x}(t),\boldsymbol{\lambda}_{t}) is a nonincreasing sequence so that for all k>rk>r, we must have c(𝒙(tk),𝝀tk)c(𝒙(tr+1),𝝀tr+1)<ρ(𝒙,𝒙,𝝀)δ2c(\boldsymbol{x}(t_{k}),\boldsymbol{\lambda}_{t_{k}})\leq c(\boldsymbol{x}(t_{r}+1),\boldsymbol{\lambda}_{t_{r}+1})<\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})-\frac{\delta}{2}. Thus,

ρ(𝒙,𝒙,𝝀)=c(𝒙,𝝀)=limkc(𝒙(tk),𝝀tk)ρ(𝒙,𝒙,𝝀)δ2.\displaystyle\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})=c(\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})=\lim_{k\to\infty}c(\boldsymbol{x}(t_{k}),\boldsymbol{\lambda}_{t_{k}})\leq\rho(\boldsymbol{x}^{*},\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*})-\frac{\delta}{2}.

This contradiction shows that 𝒙\boldsymbol{x}^{*} is a Nash equilibrium among the followers, given the leader’s action 𝝀\boldsymbol{\lambda}^{*}. Therefore, the limit point (𝒙,𝝀)(\boldsymbol{x}^{*},\boldsymbol{\lambda}^{*}) is a Stackelberg equilibrium. Q.E.D.

Example 1: Consider a special case where ϕi(x,y)=(xy)2ϵ2,i[n]\phi_{i}(x,y)=(x-y)^{2}-\epsilon^{2},\forall i\in[n], and the leader’s best response polytope is given by the symmetric matrices with self-loops, i.e., Λ:={𝝀[0,1]n×n:λij=λji,i,j,λii=1,i}\Lambda:=\{\boldsymbol{\lambda}\in[0,1]^{n\times n}:\ \lambda_{ij}=\lambda_{ji},\forall i,j,\ \lambda_{ii}=1,\forall i\}. In this case, we have

ρ(𝒙,𝒚,𝝀)=i=1nj=1nλij((yixj)2ϵ2)=𝒚TQ𝒚+𝒙TQ𝒙2𝒚T𝝀𝒚ϵ2(𝟏T𝝀𝟏),\displaystyle\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda})=\sum_{i=1}^{n}\sum_{j=1}^{n}\lambda_{ij}\big((y_{i}-x_{j})^{2}-\epsilon^{2}\big)=\boldsymbol{y}^{T}Q\boldsymbol{y}+\boldsymbol{x}^{T}Q\boldsymbol{x}-2\boldsymbol{y}^{T}\boldsymbol{\lambda}\boldsymbol{y}-\epsilon^{2}(\boldsymbol{1}^{T}\boldsymbol{\lambda}\boldsymbol{1}),

where again 𝝀\boldsymbol{\lambda} is a symmetric matrix, Q=diag(𝝀𝟏)Q=diag(\boldsymbol{\lambda}\boldsymbol{1}) is positive definite, and hence by Remark 4, ρ(𝒙,𝒚,𝝀)\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda}) is min-symmetric. Therefore, the best response dynamics of the leader-followers converge to a Stackelberg equilibrium. Interestingly, one can easily check that the best response of followers are exactly governed by the homogeneous HK update rule which shows that the steady-state of the homogeneous HK model is simply a Stackelberg equilibrium of the above game.

Example 2: Let us consider another special case where ϕi(x,y)=(xy)2,i[n]\phi_{i}(x,y)=(x-y)^{2},\forall i\in[n]. Assume that at each instance the leader’s objective is to keep the communication network among the followers (agents) connected, so that the followers can communicate with each other and eventually rendezvous at a consensus point. Perhaps one can think of agents as robots, and the leader as a bandwidth provider who aims to bring all the agents together with minimum bandwidth cost. In this case the leader’s action set can be represent by the polytope

Λ:={𝝀[0,1]n×n:iS,jSλij1,S[n],λij=λjii,j}.\displaystyle\Lambda:=\{\boldsymbol{\lambda}\in[0,1]^{n\times n}:\ \sum_{i\in S,j\notin S}\lambda_{ij}\geq 1,\forall S\subset[n],\ \lambda_{ij}=\lambda_{ji}\forall i,j\}.

Note that the constraints iS,jSλij1,S[n]\sum_{i\in S,j\notin S}\lambda_{ij}\geq 1,\forall S\subset[n] guarantee that there is a communication path between every pair of agents. It is worth noting that the best response of the leader over the polytope Λ\Lambda can be solved efficiently in polynomial time despite the fact that Λ\Lambda contains exponentially many linear constraints. This can be done using Ellipsoid algorithm where the separation oracle can be implemented by solving at most n2n^{2} network-flow problems. On the other hand, we have

ρ(𝒙,𝒚,𝝀)=i=1nj=1nλij(yixj)2=𝒚TQ𝒚+𝒙TQ𝒙2𝒚T𝝀𝒚,\displaystyle\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda})=\sum_{i=1}^{n}\sum_{j=1}^{n}\lambda_{ij}(y_{i}-x_{j})^{2}=\boldsymbol{y}^{T}Q\boldsymbol{y}+\boldsymbol{x}^{T}Q\boldsymbol{x}-2\boldsymbol{y}^{T}\boldsymbol{\lambda}\boldsymbol{y},

where 𝝀\boldsymbol{\lambda} is a symmetric matrix, and Q=diag(𝝀𝟏)Q=diag(\boldsymbol{\lambda}\boldsymbol{1}) is a diagonal matrix with diagonal elements Qii=jλij1,i[n]Q_{ii}=\sum_{j}\lambda_{ij}\geq 1,\forall i\in[n]. Again using Remark 4 we conclude that ρ(𝒙,𝒚,𝝀)\rho(\boldsymbol{x},\boldsymbol{y},\boldsymbol{\lambda}) is a min-symmetric function, which in view of Theorem 6 implies that the best response dynamics of the leader-followers converge to a Stackelberg equilibrium. Now let us find an explicit form for the best response dynamics. Given that at iteration tt the followers are at locations 𝒙(t)\boldsymbol{x}(t), and the current min-cost network designed by the leader is 𝝀t\boldsymbol{\lambda}_{t}, at the next time instance the followers move to the average point of their neighbors given by 𝒙(t+1)=argmin𝒚ρ(𝒙(t),𝒚,𝝀t)=(Qt1𝝀t)𝒙(t)\boldsymbol{x}(t+1)=\arg\min_{\boldsymbol{y}}\rho(\boldsymbol{x}(t),\boldsymbol{y},\boldsymbol{\lambda}_{t})=(Q_{t}^{-1}\boldsymbol{\lambda}_{t})\boldsymbol{x}(t), or rewritten componentwise xi(t+1)=j=1n(𝝀t)ij(𝝀t𝟏)ixj(t),i[n]x_{i}(t+1)=\sum_{j=1}^{n}\frac{(\boldsymbol{\lambda}_{t})_{ij}}{(\boldsymbol{\lambda}_{t}\boldsymbol{1})_{i}}x_{j}(t),\forall i\in[n]. Consequently, the leader looks at the current positions of the followers 𝒙(t+1)\boldsymbol{x}(t+1), and assuming that making a communication link between two followers at positions xi(t+1)x_{i}(t+1) and xj(t+1)x_{j}(t+1) is proportional to their Euclidean distance (xi(t+1)xj(t+1))2(x_{i}(t+1)-x_{j}(t+1))^{2}, the leader decides what links to make (i.e., selects 𝝀t+1Λ\boldsymbol{\lambda}_{t+1}\in\Lambda) so that the network remains connected while the network cost c(𝒙(t+1),𝝀t+1)c(\boldsymbol{x}(t+1),\boldsymbol{\lambda}_{t+1}) is minimized. In other words, the leader solves the optimization problem c(𝒙(t+1),𝝀t+1)=min𝝀Λc(𝒙(t+1),𝝀)c(\boldsymbol{x}(t+1),\boldsymbol{\lambda}_{t+1})=\min_{\boldsymbol{\lambda}\in\Lambda}c(\boldsymbol{x}(t+1),\boldsymbol{\lambda}). Clearly as the network remains connected at all the time instances while the followers take average over it, the convergent Stackelberg equilibrium must be a consensus point in which all the followers rendezvous at the same point.

As we close this section, we would like to note that in general, one can define different action polytopes Λ\Lambda for the leader and different influence functions ϕi()\phi_{i}(\cdot) for the followers. This allows us to recover various state-dependent network dynamics as the best response updates of leader-followers in a Stackelberg game, which by Theorem 6 are guaranteed to converge to a Stackelberg equilibrium, assuming the min-symmetric property. For instance, settings Λ:={𝝀[0,1]n×n:jiλij=1,i[n],λij=λjii,j}\Lambda:=\{\boldsymbol{\lambda}\in[0,1]^{n\times n}:\ \sum_{j\neq i}\lambda_{ij}=1,\forall i\in[n],\ \lambda_{ij}=\lambda_{ji}\forall i,j\} and quadratic influence functions ϕi(x,y)=(xy)2\phi_{i}(x,y)=(x-y)^{2} one can recover dynamics of a different variant of the nearest neighbor model considered in Section IV, and establish their convergence to an equilibrium point. In conclusion, game theory can provide alternative shortcuts for the analysis of dynamic networks of multi-agent systems. An interesting fact about this approach is that many of the existing results developed in the game-theory literature such as [45] do not require stringent homogeneity or symmetric assumptions among the players. This provides ample room for analysis of heterogeneous multi-agent network dynamics by bridging them to game-theoretic settings.

VI Conclusions and Further Discusssions

In this paper, we have studied Lyapunov stability and convergence of multi-agent network systems with state-dependent switching dynamics. By incorporating the network structure into our framework as a new variable, we have shown that how the evolution of multi-agent network dynamics can be viewed as trajectories of successive optimization problems. Leveraging this framework, we were able to establish Lyapunov stability and convergence of several well-known models from social science, and extended our results to scenarios with asymmetric communication structures. In particular, we showed how these results can be viewed from a game-theoretic perspective, so that convergence of multi-agent network systems can be interpreted as selfish behavior of players in a well-designed Stackelberg game.

We believe that this work opens many new directions for stability analysis of multi-agent network systems with rich state-dependent switching dynamics. In the following we have listed a few of such directions:

1) Often in successive optimization framework one adds smooth proximal functions to the original objective function (or approximates it with a smooth upper bound) in order to facilitates the minimization sub-problems [46]. Incorporating this idea into our framework will give us perturbed multi-agent network dynamics whose update rules are affected by the proximal term (while their convergence to a stationary point is still guaranteed). For instance, one can add a proximal term to the objective function f(𝒚,𝝀)f(\boldsymbol{y},\boldsymbol{\lambda}) of the homogeneous HK model in terms of Bregman distance of a strongly convex function and recover noisy versions of HK dynamics.

2) In Section III, for a fixed network 𝝀\boldsymbol{\lambda}^{*} with associate Laplacian \mathcal{L}^{*}, we only used the quadratic Lyapunov function Φ(𝒙,𝝀)=𝒙T𝒙\Phi(\boldsymbol{x},\boldsymbol{\lambda}^{*})=\boldsymbol{x}^{T}\mathcal{L}^{*}\boldsymbol{x} into our BCD framework. While this Lyapunov function seems quite suitable when the network structure is symmetric (undirected), it cannot serve as a Lyapunov function for general asymmetric (directed) networks. However, for a general fixed directed network with adjacency matrix 𝝀\boldsymbol{\lambda}^{*}, one can choose an alternative Lyapunov function given by Φ(𝒙,𝝅)=𝒙T(diag(𝝅)𝝅(𝝅)T)𝒙=i,jπiπj(xixj)2\Phi(\boldsymbol{x},\boldsymbol{\pi}^{*})=\boldsymbol{x}^{T}(diag(\boldsymbol{\pi}^{*})-\boldsymbol{\pi}^{*}(\boldsymbol{\pi}^{*})^{T})\boldsymbol{x}=\sum_{i,j}\pi^{*}_{i}\pi^{*}_{j}(x_{i}-x_{j})^{2}, where 𝝅\boldsymbol{\pi}^{*} is the Perron-left eigenvector of the transition matrix of a simple random walk over 𝝀\boldsymbol{\lambda}^{*}, i.e., the Peron-left eigenvector of A=(diag(𝝀𝟏))1𝝀A=(diag(\boldsymbol{\lambda}^{*}\boldsymbol{1}))^{-1}\boldsymbol{\lambda}^{*} so that (𝝅)TA=(𝝅)T(\boldsymbol{\pi}^{*})^{T}A=(\boldsymbol{\pi}^{*})^{T}. In this case, one can replace the network variable matrix 𝝀\boldsymbol{\lambda}^{*} by its Perron-left eigenvector 𝝅\boldsymbol{\pi}^{*} in the BCD framework. Therefore, as far as minimizing 𝝅\boldsymbol{\pi} with respect to its constraint set matches the left-Perron eigenvector of the actual network dynamics, one can use Proposition 1 with Φ(𝒙,𝝅)=𝒙T(diag(𝝅)𝝅(𝝅)T)𝒙\Phi(\boldsymbol{x},\boldsymbol{\pi}^{*})=\boldsymbol{x}^{T}(diag(\boldsymbol{\pi}^{*})-\boldsymbol{\pi}^{*}(\boldsymbol{\pi}^{*})^{T})\boldsymbol{x} to show Lyapunov stability of joint state-network dynamics under more general asymmetric (directed) environment. Therefore, it is interesting to study stability of state-dependent multi-agent networks under highly asymmetric setting using 𝝅\boldsymbol{\pi} variables, rather than simply the edge variables 𝝀=(λij)\boldsymbol{\lambda}=(\lambda_{ij}).

3) In Section V, we used a Stackelberg game where the followers’ cost functions are convex (i.e. the noncooperative game among followers is a convex game [45]) and the leader’s objective function is to minimize the social cost. These conditions can be generalized to other settings. For instance, one can consider a Stackelberg network game where the game among followers is a potential game [47] whose potential function can serve as the Lyapunov function Φ(𝒙,𝝀)\Phi(\boldsymbol{x},\boldsymbol{\lambda}^{*}) into our BCD framework, given a fixed leader’s action 𝝀\boldsymbol{\lambda}^{*}.

Finally, in this paper we mainly focused on the stability and convergence analysis of the underlying dynamics. Therefore, an interesting direction here is to characterize the structure of the equilibrium points in more detail.

References

  • [1] M. H. DeGroot, “Reaching a consensus,” Journal of the American Statistical Association, vol. 69, no. 345, pp. 118–121, 1974.
  • [2] N. E. Friedkin and E. C. Johnsen, “Social positions in influence networks,” Social Networks, vol. 19, no. 3, pp. 209–222, 1997.
  • [3] A. Nedić, A. Olshevsky, and M. G. Rabbat, “Network topology and communication-computation tradeoffs in decentralized optimization,” Proceedings of the IEEE, vol. 106, no. 5, pp. 953–976, 2018.
  • [4] A. V. Proskurnikov and R. Tempo, “A tutorial on modeling and analysis of dynamic social networks. Part I,” Annual Reviews in Control, vol. 43, pp. 65–79, 2017.
  • [5] A. Olshevsky and J. N. Tsitsiklis, “Convergence speed in distributed consensus and averaging,” SIAM Journal on Control and Optimization, vol. 48, no. 1, pp. 33–55, 2009.
  • [6] A. Nedic, A. Olshevsky, A. Ozdaglar, and J. N. Tsitsiklis, “On distributed averaging algorithms and quantization effects,” IEEE Transactions on Automatic Control, vol. 54, no. 11, pp. 2506–2517, 2009.
  • [7] A. Nedić and A. Olshevsky, “Distributed optimization over time-varying directed graphs,” IEEE Transactions on Automatic Control, vol. 60, no. 3, pp. 601–615, 2015.
  • [8] T. Başar, S. R. Etesami, and A. Olshevsky, “Convergence time of quantized Metropolis consensus over time-varying networks,” IEEE Transactions on Automatic Control, vol. 61, no. 12, pp. 4048–4054, 2016.
  • [9] M. Zhu and S. Martínez, “On the convergence time of asynchronous distributed quantized averaging algorithms,” IEEE Transactions on Automatic Control, vol. 56, no. 2, pp. 386–390, 2011.
  • [10] T. Tatarenko and B. Touri, “Non-convex distributed optimization,” IEEE Transactions on Automatic Control, vol. 62, no. 8, pp. 3744–3757, 2017.
  • [11] R. Olfati-Saber and R. M. Murray, “Consensus problems in networks of agents with switching topology and time-delays,” IEEE Transactions on Automatic Control, vol. 49, no. 9, pp. 1520–1533, 2004.
  • [12] J. M. Hendrickx and J. N. Tsitsiklis, “Convergence of type-symmetric and cut-balanced consensus seeking systems,” IEEE Transactions on Automatic Control, vol. 58, no. 1, pp. 214–218, 2013.
  • [13] A. Jadbabaie, J. Lin, and A. S. Morse, “Coordination of groups of mobile autonomous agents using nearest neighbor rules,” IEEE Transactions on automatic control, vol. 48, no. 6, pp. 988–1001, 2003.
  • [14] B. Touri, Product of random stochastic matrices and distributed averaging. Springer Science & Business Media, 2012.
  • [15] I. M. Sonin et al., “The decomposition-separation theorem for finite nonhomogeneous Markov chains and related problems,” in Markov Processes and Related Topics: A Festschrift for Thomas G. Kurtz. Institute of Mathematical Statistics, 2008, pp. 1–15.
  • [16] B. Touri and A. Nedić, “On existence of a quadratic comparison function for random weighted averaging dynamics and its implications,” in 50th IEEE Conference on Decision and Control and European Control Conference (CDC-ECC). IEEE, 2011, pp. 3806–3811.
  • [17] S. R. Etesami, T. Başar, A. Nedić, and B. Touri, “Termination time of multidimensional Hegselmann-Krause opinion dynamics,” in American Control Conference (ACC). IEEE, 2013, pp. 1255–1260.
  • [18] M. Razaviyayn, M. Hong, and Z.-Q. Luo, “A unified convergence analysis of block successive minimization methods for nonsmooth optimization,” SIAM Journal on Optimization, vol. 23, no. 2, pp. 1126–1153, 2013.
  • [19] H. Howson and N. Sancho, “A new algorithm for the solution of multi-state dynamic programming problems,” Mathematical Programming, vol. 8, no. 1, pp. 104–116, 1975.
  • [20] J. Hartigan and M. Wong, “A k-means clustering algorithm,” Journal of the Royal Statistical Society. Series C (Applied Statistics), vol. 28, no. 1, pp. 100–108, 1979.
  • [21] R. Hegselmann and U. Krause, “Opinion dynamics and bounded confidence models, analysis and simulation,” Journal of Artificial Societies and Social Simulation, vol. 5, no. 3, pp. 1–33, 2002.
  • [22] N. E. Friedkin and E. C. Johnsen, “Social influence networks and opinion change,” Advances in Group Processes, vol. 16, no. 1, pp. 1–29, 1999.
  • [23] ——, Social influence network theory: A sociological examination of small group dynamics. Cambridge University Press, 2011, vol. 33.
  • [24] J. Lorenz, “Continuous opinion dynamics under bounded confidence: A survey,” International Journal of Modern Physics C, vol. 18, no. 12, pp. 1819–1838, 2007.
  • [25] A. V. Proskurnikov and R. Tempo, “A tutorial on modeling and analysis of dynamic social networks. Part II,” Annual Reviews in Control, vol. 45, pp. 166–190, 2018.
  • [26] F. Bullo, J. Cortes, and S. Martinez, Distributed control of robotic networks: A mathematical approach to motion coordination algorithms. Princeton University Press, 2009, vol. 27.
  • [27] B. Chazelle, “The total s-energy of a multiagent system,” SIAM Journal on Control and Optimization, vol. 49, no. 4, pp. 1680–1706, 2011.
  • [28] Y. Dong, X. Chen, H. Liang, and C.-C. Li, “Dynamics of linguistic opinion formation in bounded confidence model,” Information Fusion, vol. 32, pp. 52–61, 2016.
  • [29] M. B. Ye, J. Liu, B. D. Anderson, C. B. Yu, and T. Basar, “Evolution of social power in social networks with dynamic topology,” IEEE Transactions on Automatic Control, vol. 63, pp. 3793 – 3808, 2018.
  • [30] H.-T. Wai, A. Scaglione, and A. Leshem, “Identifying trust in social networks with stubborn agents, with application to market decisions,” in 53rd Annual Allerton Conference on Communication, Control, and Computing. IEEE, 2015, pp. 747–754.
  • [31] J. Lorenz, “Repeated averaging and bounded-confidence, modeling, analysis and simulation of continuous opinion dynamics,” Ph.D. dissertation, University of Bremen, 2007.
  • [32] A. Mirtabatabaei and F. Bullo, “Opinion dynamics in heterogeneous networks: Convergence conjectures and theorems,” SIAM Journal on Control and Optimization, vol. 50, no. 5, pp. 2763–2785, 2012.
  • [33] J. Lorenz, “Heterogeneous bounds of confidence: Meet, discuss and find consensus!” Complexity, vol. 15, no. 4, pp. 43–52, 2010.
  • [34] J. M. Hendrickx, “Graphs and networks for the analysis of autonomous agent systems,” Ph.D. Thesis, Universite Catholique de Louvain, 2011.
  • [35] B. Chazelle and C. Wang, “Inertial Hegselmann-Krause systems,” IEEE Transactions on Automatic Control, vol. 62, no. 8, pp. 3905–3913, 2017.
  • [36] M. Pineda, R. Toral, and E. Hernández-García, “The noisy Hegselmann-Krause model for opinion dynamics,” The European Physical Journal B, vol. 86, no. 12, pp. 1–10, 2013.
  • [37] V. D. Blondel, J. M. Hendrickx, and J. N. Tsitsiklis, “On the 2R2R conjecture for multi-agent systems,” in European Control Conference (ECC). IEEE, 2007, pp. 874–881.
  • [38] J. M. Hendrickx and A. Olshevsky, “On symmetric continuum opinion dynamics,” SIAM Journal on Control and Optimization, vol. 54, no. 5, pp. 2893–2918, 2016.
  • [39] A. Bhattacharyya, M. Braverman, B. Chazelle, and H. L. Nguyen, “On the convergence of the Hegselmann-Krause system,” in Proceedings of the 4th Conference on Innovations in Theoretical Computer Science. ACM, 2013, pp. 61–66.
  • [40] M. Roozbehani, A. Megretski, and E. Frazzoli, “Lyapunov analysis of quadratically symmetric neighborhood consensus algorithms,” in Proc. 47th IEEE Conference on Decision and Control (CDC). IEEE, 2008, pp. 2252–2257.
  • [41] H. Zhang, F. L. Lewis, and Z. Qu, “Lyapunov, adaptive, and optimal design techniques for cooperative systems on directed communication graphs,” IEEE Transactions on Industrial Electronics, vol. 59, no. 7, pp. 3026–3041, 2012.
  • [42] A. Olshevsky and J. Tsitsiklis, “On the nonexistence of quadratic Lyapunov functions for consensus algorithms,” IEEE Transactions on Automatic Control, vol. 53, pp. 2642–2645, 2008.
  • [43] S. R. Etesami and T. Başar, “Game-theoretic analysis of the Hegselmann-Krause model for opinion dynamics in finite dimensions,” IEEE Transactions on Automatic Control, vol. 60, no. 7, pp. 1886–1897, 2015.
  • [44] W. Su, Y. Gu, S. Wang, and Y. Yu, “Partial convergence of heterogeneous Hegselmann-Krause opinion dynamics,” Science China Technological Sciences, vol. 60, no. 9, pp. 1433–1438, 2017.
  • [45] J. B. Rosen, “Existence and uniqueness of equilibrium points for concave nn-person games,” Econometrica: Journal of the Econometric Society, pp. 520–534, 1965.
  • [46] H. Attouch, J. Bolte, P. Redont, and A. Soubeyran, “Proximal alternating minimization and projection methods for nonconvex problems: An approach based on the kurdyka-Lojasiewicz inequality,” Mathematics of Operations Research, vol. 35, no. 2, pp. 438–457, 2010.
  • [47] D. Monderer and L. S. Shapley, “Potential games,” Games and Economic Behavior, vol. 14, no. 1, pp. 124–143, 1996.

VII Appendix I

Lemma 1

Consider the restricted HK model and choose δ<1n2\delta<\frac{1}{n^{2}}. Let tδt_{\delta} be a time instance for which τ=tδi=1n(xi(τ+1)xi(τ))2<δ2\sum_{\tau=t_{\delta}}^{\infty}\sum_{i=1}^{n}(x_{i}(\tau+1)-x_{i}(\tau))^{2}<\delta^{2}. Then for any ttδt\geq t_{\delta} the communication network 𝛌t\boldsymbol{\lambda}_{t} remains unchanged.

Proof:

By contradiction, let us assume that the lemma does not hold and ttδt\geq t_{\delta} be a time instance for which the communication network changes, i.e., 𝝀t𝝀t+1\boldsymbol{\lambda}_{t}\neq\boldsymbol{\lambda}_{t+1}. We construct an undirected graph \mathcal{B}, so-called balanced graph, as follows:

  • Nodes of \mathcal{B} are the points {x1(t+1),x2(t+1),,xn(t+1)}\{x_{1}(t+1),x_{2}(t+1),\ldots,x_{n}(t+1)\} which are positioned on the real axis. For simplicity, we refer to these nodes by their indicies such that ii represents the node positioned at xi(t+1)x_{i}(t+1). Note that in \mathcal{B} both labels and geometric positions of the vertices are important.

  • The edges of \mathcal{B} are partitioned into two groups: solid and dashed. There is a solid edge between two nodes ii and jj if and only if {i,j},|xi(t)xj(t)|>1\{i,j\}\in\mathcal{E},|x_{i}(t)-x_{j}(t)|>1, and |xi(t+1)xj(t+1)|1|x_{i}(t+1)-x_{j}(t+1)|\leq 1. Thus a solid edge between nodes ii and jj in \mathcal{B} indicates that agents ii and jj were not each others’ neighbors at time tt (i.e., (𝝀t)ij=0(\boldsymbol{\lambda}_{t})_{ij}=0), while they become neighbor at time t+1t+1 (i.e., (𝝀t+1)ij=1(\boldsymbol{\lambda}_{t+1})_{ij}=1). There is a dashed edge between ii and jj if and only if {i,j},|xi(t)xj(t)|1\{i,j\}\in\mathcal{E},|x_{i}(t)-x_{j}(t)|\leq 1, and |xi(t+1)xj(t+1)|>1|x_{i}(t+1)-x_{j}(t+1)|>1. Thus a dashed edge in \mathcal{B} shows that agents ii and jj were each others’ neighbors at time tt, but they separate at time t+1t+1.

Next we show that if ttδt\geq t_{\delta} is a time instance for which the communication network changes, then the following three facts must hold: Fact 1) For every two nodes ii and jj which are connected by an edge in \mathcal{B}, we have 12δ<|xi(t+1)xj(t+1)|<1+2δ1-2\delta<|x_{i}(t+1)-x_{j}(t+1)|<1+2\delta. This is an immediate consequence of the edge definition given above together with the fact that no agent can move by a distance more than δ\delta from time tt to t+1t+1. Fact 2) The geometric distance between any two nodes which are connected by a dashed edge is strictly greater than the distance between any other two nodes which are connected by a solid edge. This is because the former is strictly greater than 1, while the later is at most 1. Fact 3) Let sL(i)s_{L}(i) and dL(i)d_{L}(i) denote, respectively, the number of solid and dashed edges which are incident to node ii from left hand side. Similarly, let sR(i)s_{R}(i) and dR(i)d_{R}(i) be the number of solid/dashed edges which are incident to node ii from right hand side. Then a straightforward calculation as in [35, Equation 2] shows that for sufficiently small δ1n2\delta\leq\frac{1}{n^{2}}, we must have sL(i)dL(i)=sR(i)dR(i),is_{L}(i)-d_{L}(i)=s_{R}(i)-d_{R}(i),\forall i. In other words, every node in \mathcal{B} must be balanced with respect to the ‘effective’ amount of change in its neighbors from time tt to t+1t+1. Otherwise moving from time t+1t+1 to t+2t+2, an unbalanced agent will move by a distance of at least δ\delta, contradicting the fact that ttδt\geq t_{\delta}.88 8 Intuitively, sL(i)dL(i)s_{L}(i)-d_{L}(i) measures the ‘effective’ change due to change of left-neighbors of agent ii from time tt to t+1t+1. This must be equal to sR(i)dR(i)s_{R}(i)-d_{R}(i) which is the effective change in the right-neighbors of ii. Otherwise, agent ii will move by a large distance at the next time instance t+2t+2.

To derive a contradiction, consider the time ttδt\geq t_{\delta} for which a switch in the communication network occurs. This means that the associated balance graph \mathcal{B} has at least one edge. Henceforth, and by some abuse of notation, we denote a nontrivial connected component of this balanced graph by \mathcal{B} (which is a connected graph with at least one edge and no isolated vertex). Since Fact 1 holds, the vertices of \mathcal{B} can be covered by the union of some disjoint intervals {I=(a,b):=1,,k}\{I_{\ell}=(a_{\ell},b_{\ell}):\ell=1,\ldots,k\}, where a1<b1<a2<b2,<bka_{1}<b_{1}<a_{2}<b_{2},\ldots<b_{k}. Moreover, each of these intervals has a very small length (at most ba<4nδb_{\ell}-a_{\ell}<4n\delta), and any edge of \mathcal{B} has its end points in two consecutive intervals I,I+1I_{\ell},I_{\ell+1} (see Figure 2 for an illustration).99 9 Note that this also implies that every two consecutive intervals are apart from each other by a large distance of at least 12δ8nδ1-2\delta-8n\delta.

Refer to caption
Fig. 2: An illustration of the balanced graph \mathcal{B} with solid/dashed edges and covering intervals I1,I2,I3I_{1},I_{2},I_{3}, and I4I_{4}. Here a maximal walk W=(e1,e2,e3,e4,e5,e6)W=(e_{1},e_{2},e_{3},e_{4},e_{5},e_{6}) with initial node i1i_{1} and endpoint ii is represented by thick edges. Note that for node i1i_{1} we have sL(i1)=dL(i1)=0s_{L}(i_{1})=d_{L}(i_{1})=0 and sR(i1)=dR(i1)=1s_{R}(i_{1})=d_{R}(i_{1})=1. Therefore, sL(i1)dL(i1)=sR(i1)dR(i1)s_{L}(i_{1})-d_{L}(i_{1})=s_{R}(i_{1})-d_{R}(i_{1}), which means that node i1i_{1} is balanced. However for node ii we have sL(i)=dL(i)=sR(i)=0s_{L}(i)=d_{L}(i)=s_{R}(i)=0, and dR(i)=1d_{R}(i)=1. Thus sL(i)dL(i)=01=sR(i)dR(i)s_{L}(i)-d_{L}(i)=0\neq-1=s_{R}(i)-d_{R}(i), which means that node ii is not balanced.

Next let us consider the most left agent in the connected balance graph \mathcal{B} and call it i1i_{1} (breaking ties arbitrarily). By Fact 3 every node in \mathcal{B} (and in particular i1i_{1}) is balanced. Since i1i_{1} does not have any edge incident to it from left hand side, it must have at least one solid edge incident to it from the right hand side. Now starting from node i1i_{1}, let us consider a maximal alternating walk WW which sequentially traverses over the edges of \mathcal{B} from left to right using solid edges, and from right to left using dashed edges (Figure 2). As we argued above, WW is nonempty and contains at least one solid edge. We claim that WW must be a path, meaning that no vertex can be visited more than once by WW. Otherwise, consider the first time when WW visits a vertex for the second time, which results to an alternating cycle CC. As each vertex of CC lies in one of the disjoint intervals II_{\ell}, and each edge of CC has its endpoints in two consecutive intervals, this implies that CC must have the same number of solid and dashed edges. This in view of Fact 2 shows that the sum of lengths of solid edges in CC is strictly less than the sum of lengths of dashed edges in CC. However, we know that for any alternating cycle the sum of lengths of solid edges (the total movement from left to right) must be equal to the sum of lengths of dashed edges (the total movement from right to left). This contradiction shows that WW must be a path.

Finally, let us denote the other endpoint of path WW by ii. Either node ii is reached through the path WW using a dashed edge (and hence from right to left), in which case dR(i)1d_{R}(i)\geq 1 and sR(i)=dL(i)=0s_{R}(i)=d_{L}(i)=0 (otherwise WW can be extended and is not maximal). Or node ii is reached using a solid edge (and hence from left to right), in which case sL(i)1s_{L}(i)\geq 1 and sR(i)=dL(i)=0s_{R}(i)=d_{L}(i)=0. Now an easy calculation shows that in either case sL(i)dL(i)sR(i)dR(i)s_{L}(i)-d_{L}(i)\neq s_{R}(i)-d_{R}(i), meaning that node ii is not a balanced node. This is in contradiction with Fact 3, which completes the proof. Q.E.D.