Source author record

Marianne Akian

Marianne Akian appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

20works
12topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

20 published item(s)

preprint2016arXiv

A probabilistic max-plus numerical method for solving stochastic control problems

We consider fully nonlinear Hamilton-Jacobi-Bellman equations associated to diffusion control problems involving a finite set-valued (or switching) control and possibly a continuum-valued control. We construct a lower complexity probabilistic numerical algorithm by combining the idempotent expansion properties obtained by McEneaney, Kaise and Han (2011) for solving such problems with a numerical probabilistic method such as the one proposed by Fahim, Touzi and Warin (2011) for solving some fully nonlinear parabolic partial differential equations. Numerical tests on a small example of pricing and hedging an option are presented.

preprint2016arXiv

Log-majorization of the moduli of the eigenvalues of a matrix polynomial by tropical roots

We show that the sequence of moduli of the eigenvalues of a matrix polynomial is log-majorized, up to universal constants, by a sequence of "tropical roots" depending only on the norms of the matrix coefficients. These tropical roots are the non-differentiability points of an auxiliary tropical polynomial, or equivalently, the opposites of the slopes of its Newton polygon. This extends to the case of matrix polynomials some bounds obtained by Hadamard, Ostrowski and Pólya for the roots of scalar polynomials. We also obtain new bounds in the scalar case, which are accurate for "fewnomials" or when the tropical roots are well separated.

preprint2015arXiv

Ergodicity conditions for zero-sum games

A basic question for zero-sum repeated games consists in determining whether the mean payoff per time unit is independent of the initial state. In the special case of "zero-player" games, i.e., of Markov chains equipped with additive functionals, the answer is provided by the mean ergodic theorem. We generalize this result to repeated games. We show that the mean payoff is independent of the initial state for all state-dependent perturbations of the rewards if and only if an ergodicity condition is verified. The latter is characterized by the uniqueness modulo constants of nonlinear harmonic functions (fixed points of the recession function associated to the Shapley operator), or, in the special case of stochastic games with finite action spaces and perfect information, by a reachability condition involving conjugate subsets of states in directed hypergraphs. We show that the ergodicity condition for games only depends on the support of the transition probability, and that it can be checked in polynomial time when the number of states is fixed.

preprint2015arXiv

Hypergraph conditions for the solvability of the ergodic equation for zero-sum games

The ergodic equation is a basic tool in the study of mean-payoff stochastic games. Its solvability entails that the mean payoff is independent of the initial state. Moreover, optimal stationary strategies are readily obtained from its solution. In this paper, we give a general sufficient condition for the solvability of the ergodic equation, for a game with finite state space but arbitrary action spaces. This condition involves a pair of directed hypergraphs depending only on the ``growth at infinity'' of the Shapley operator of the game. This refines a recent result of the authors which only applied to games with bounded payments, as well as earlier nonlinear fixed point results for order preserving maps, involving graph conditions.

preprint2014arXiv

A Collatz-Wielandt characterization of the spectral radius of order-preserving homogeneous maps on cones

Several notions of spectral radius arise in the study of nonlinear order-preserving positively homogeneous self-maps of cones in Banach spaces. We give conditions that guarantee that all these notions lead to the same value. In particular, we give a Collatz-Wielandt type formula, which characterizes the growth rate of the orbits in terms of eigenvectors in the closed cone or super-eigenvectors in the interior of the cone. This characterization holds when the cone is normal and when a quasi-compactness condition, involving an essential spectral radius defined in terms of $k$-set-contractions, is satisfied. Some fixed point theorems for non-linear maps on cones are derived as intermediate results. We finally apply these results to show that non-linear spectral radii commute with respect to suprema and infima of families of order preserving maps satisfying selection properties.

preprint2014arXiv

Generic uniqueness of the bias vector of mean payoff zero-sum games

Zero-sum mean payoff games can be studied by means of a nonlinear spectral problem. When the state space is finite, the latter consists in finding an eigenpair $(u,λ)$ solution of $T(u)=λ\mathbf{1} + u$ where $T:\mathbb{R}^n \to \mathbb{R}^n$ is the Shapley (dynamic programming) operator, $λ$ is a scalar, $\mathbf{1}$ is the unit vector, and $u \in \mathbb{R}^n$. The scalar $λ$ yields the mean payoff per time unit, and the vector $u$, called the bias, allows one to determine optimal stationary strategies. The existence of the eigenpair $(u,λ)$ is generally related to ergodicity conditions. A basic issue is to understand for which classes of games the bias vector is unique (up to an additive constant). In this paper, we consider perfect information zero-sum stochastic games with finite state and action spaces, thinking of the transition payments as variable parameters, transition probabilities being fixed. We identify structural conditions on the support of the transition probabilities which guarantee that the spectral problem is solvable for all values of the transition payments. Then, we show that the bias vector, thought of as a function of the transition payments, is generically unique (up to an additive constant). The proof uses techniques of max-plus (tropical) algebra and nonlinear Perron-Frobenius theory.

preprint2014arXiv

Uniqueness of the fixed point of nonexpansive semidifferentiable maps

We consider semidifferentiable (possibly nonsmooth) maps, acting on a subset of a Banach space, that are nonexpansive either in the norm of the space or in the Hilbert's or Thompson's metric inherited from a convex cone. We show that the global uniqueness of the fixed point of the map, as well as the geometric convergence of every orbit to this fixed point, can be inferred from the semidifferential of the map at this point. In particular, we show that the geometric convergence rate of the orbits to the fixed point can be bounded in terms of Bonsall's nonlinear spectral radius of the semidifferential. We derive similar results concerning the uniqueness of the eigenline and the geometric convergence of the orbits to it, in the case of positively homogeneous maps acting on the interior of a cone, or of additively homogeneous maps acting on an AM-space with unit. This is motivated in particular by the analysis of dynamic programming operators (Shapley operators) of zero-sum stochastic games.

preprint2013arXiv

Policy iteration for perfect information stochastic mean payoff games with bounded first return times is strongly polynomial

Recent results of Ye and Hansen, Miltersen and Zwick show that policy iteration for one or two player (perfect information) zero-sum stochastic games, restricted to instances with a fixed discount rate, is strongly polynomial. We show that policy iteration for mean-payoff zero-sum stochastic games is also strongly polynomial when restricted to instances with bounded first mean return time to a given state. The proof is based on methods of nonlinear Perron-Frobenius theory, allowing us to reduce the mean-payoff problem to a discounted problem with state dependent discount rate. Our analysis also shows that policy iteration remains strongly polynomial for discounted problems in which the discount rate can be state dependent (and even negative) at certain states, provided that the spectral radii of the nonnegative matrices associated to all strategies are bounded from above by a fixed constant strictly less than 1.

preprint2013arXiv

Tropical bounds for eigenvalues of matrices

We show that for all k = 1,...,n the absolute value of the product of the k largest eigenvalues of an n-by-n matrix A is bounded from above by the product of the k largest tropical eigenvalues of the matrix |A| (entrywise absolute value), up to a combinatorial constant depending only on k and on the pattern of the matrix. This generalizes an inequality by Friedland (1986), corresponding to the special case k = 1.

preprint2013arXiv

Tropical Cramer Determinants Revisited

We prove general Cramer type theorems for linear systems over various extensions of the tropical semiring, in which tropical numbers are enriched with an information of multiplicity, sign, or argument. We obtain existence or uniqueness results, which extend or refine earlier results of Gondran and Minoux (1978), Plus (1990), Gaubert (1992), Richter-Gebert, Sturmfels and Theobald (2005) and Izhakian and Rowen (2009). Computational issues are also discussed; in particular, some of our proofs lead to Jacobi and Gauss-Seidel type algorithms to solve linear systems in suitably extended tropical semirings.

preprint2012arXiv

Policy iteration algorithm for zero-sum multichain stochastic games with mean payoff and perfect information

We consider zero-sum stochastic games with finite state and action spaces, perfect information, mean payoff criteria, without any irreducibility assumption on the Markov chains associated to strategies (multichain games). The value of such a game can be characterized by a system of nonlinear equations, involving the mean payoff vector and an auxiliary vector (relative value or bias). We develop here a policy iteration algorithm for zero-sum stochastic games with mean payoff, following an idea of two of the authors (Cochet-Terrasson and Gaubert, C. R. Math. Acad. Sci. Paris, 2006). The algorithm relies on a notion of nonlinear spectral projection (Akian and Gaubert, Nonlinear Analysis TMA, 2003), which is analogous to the notion of reduction of super-harmonic functions in linear potential theory. To avoid cycling, at each degenerate iteration (in which the mean payoff vector is not improved), the new relative value is obtained by reducing the earlier one. We show that the sequence of values and relative values satisfies a lexicographical monotonicity property, which implies that the algorithm does terminate. We illustrate the algorithm by a mean-payoff version of Richman games (stochastic tug-of-war or discrete infinity Laplacian type equation), in which degenerate iterations are frequent. We report numerical experiments on large scale instances, arising from the latter games, as well as from monotone discretizations of a mean-payoff pursuit-evasion deterministic differential game.

preprint2011arXiv

Ergodic Control and Polyhedral approaches to PageRank Optimization

We study a general class of PageRank optimization problems which consist in finding an optimal outlink strategy for a web site subject to design constraints. We consider both a continuous problem, in which one can choose the intensity of a link, and a discrete one, in which in each page, there are obligatory links, facultative links and forbidden links. We show that the continuous problem, as well as its discrete variant when there are no constraints coupling different pages, can both be modeled by constrained Markov decision processes with ergodic reward, in which the webmaster determines the transition probabilities of websurfers. Although the number of actions turns out to be exponential, we show that an associated polytope of transition measures has a concise representation, from which we deduce that the continuous problem is solvable in polynomial time, and that the same is true for the discrete problem when there are no coupling constraints. We also provide efficient algorithms, adapted to very large networks. Then, we investigate the qualitative features of optimal outlink strategies, and identify in particular assumptions under which there exists a "master" page to which all controlled pages should point. We report numerical results on fragments of the real web graph.

preprint2011arXiv

Multigrid methods for two-player zero-sum stochastic games

We present a fast numerical algorithm for large scale zero-sum stochastic games with perfect information, which combines policy iteration and algebraic multigrid methods. This algorithm can be applied either to a true finite state space zero-sum two player game or to the discretization of an Isaacs equation. We present numerical tests on discretizations of Isaacs equations or variational inequalities. We also present a full multi-level policy iteration, similar to FMG, which allows to improve substantially the computation time for solving some variational inequalities.

preprint2011arXiv

Tropical polyhedra are equivalent to mean payoff games

We show that several decision problems originating from max-plus or tropical convexity are equivalent to zero-sum two player game problems. In particular, we set up an equivalence between the external representation of tropical convex sets and zero-sum stochastic games, in which tropical polyhedra correspond to deterministic games with finite action spaces. Then, we show that the winning initial positions can be determined from the associated tropical polyhedron. We obtain as a corollary a game theoretical proof of the fact that the tropical rank of a matrix, defined as the maximal size of a submatrix for which the optimal assignment problem has a unique solution, coincides with the maximal number of rows (or columns) of the matrix which are linearly independent in the tropical sense. Our proofs rely on techniques from non-linear Perron-Frobenius theory.

preprint2010arXiv

Best approximation in max-plus semimodules

We establish new results concerning projectors on max-plus spaces, as well as separating half-spaces, and derive an explicit formula for the distance in Hilbert's projective metric between a point and a half-space over the max-plus semiring, as well as explicit descriptions of the set of minimizers. As a consequence, we obtain a cyclic projection type algorithm to solve systems of max-plus linear inequalities.

preprint2010arXiv

Stability and convergence in discrete convex monotone dynamical systems

We study the stable behaviour of discrete dynamical systems where the map is convex and monotone with respect to the standard positive cone. The notion of tangential stability for fixed points and periodic points is introduced, which is weaker than Lyapunov stability. Among others we show that the set of tangentially stable fixed points is isomorphic to a convex inf-semilattice, and a criterion is given for the existence of a unique tangentially stable fixed point. We also show that periods of tangentially stable periodic points are orders of permutations on $n$ letters, where $n$ is the dimension of the underlying space, and a sufficient condition for global convergence to periodic orbits is presented.

preprint2008arXiv

The optimal assignment problem for a countable state space

Given a square matrix B=(b_{ij}) with real entries, the optimal assignment problem is to find a bijection s between the rows and the columns maximising the sum of the b_{is(i)}. In discrete optimal control and in the theory of discrete event systems, one often encounters the problem of solving the equation Bf=g for a given vector g, where the same symbol B denotes the corresponding max-plus linear operator, (Bf)_i:=max_j (b_{ij}+f_j). The matrix B is said to be strongly regular when there exists a vector g such that the equation Bf=g has a unique solution f. A result of Butkovic and Hevery shows that B is strongly regular if and only if the associated optimal assignment problem has a unique solution. We establish here an extension of this result which applies to max-plus linear operators over a countable state space. The proofs use the theory developed in a previous work in which we characterised the unique solvability of equations involving Moreau conjugacies over an infinite state space, in terms of the minimality of certain coverings of the state space by generalised subdifferentials.

preprint2005arXiv

Solutions of max-plus linear equations and large deviations

We generalise the Gartner-Ellis theorem of large deviations theory. Our results allow us to derive large deviation type results in stochastic optimal control from the convergence of generalised logarithmic moment generating functions. They rely on the characterisation of the uniqueness of the solutions of max-plus linear equations. We give an illustration for a simple investment model, in which logarithmic moment generating functions represent risk-sensitive values.

preprint2005arXiv

The max-plus finite element method for optimal control problems: further approximation results

We develop the max-plus finite element method to solve finite horizon deterministic optimal control problems. This method, that we introduced in a previous work, relies on a max-plus variational formulation, and exploits the properties of projectors on max-plus semimodules. We prove here a convergence result, in arbitrary dimension, showing that for a subclass of problems, the error estimate is of order $δ+Δx(δ)^{-1}$, where $δ$ and $Δx$ are the time and space steps respectively. We also show how the max-plus analogues of the mass and stiffness matrices can be computed by convex optimization, even when the global problem is non convex. We illustrate the method by numerical examples in dimension 2.