Source author record

Salar Rahili

Salar Rahili appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
1topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2020arXiv

Distributed Adaptive Reinforcement Learning: A Method for Optimal Routing

In this paper, a learning-based optimal transportation algorithm for autonomous taxis and ridesharing vehicles is presented. The goal is to design a mechanism to solve the routing problem for multiple autonomous vehicles and multiple customers in order to maximize the transportation company's profit. As a result, each vehicle selects the customer whose request maximizes the company's profit in the long run. To solve this problem, the system is modeled as a Markov Decision Process (MDP) using past customers data. By solving the defined MDP, a centralized high-level planning recommendation is obtained, where this offline solution is used as an initial value for the real-time learning. Then, a distributed SARSA reinforcement learning algorithm is proposed to capture the model errors and the environment changes, such as variations in customer distributions in each area, traffic, and fares, thereby providing optimal routing policies in real-time. Vehicles, or agents, use only their local information and interaction, such as current passenger requests and estimates of neighbors' tasks and their optimal actions, to obtain the optimal policies in a distributed fashion. An optimal adaptive rate is introduced to make the distributed SARSA algorithm capable of adapting to changes in the environment and tracking the time-varying optimal policies. Furthermore, a game-theory-based task assignment algorithm is proposed, where each agent uses the optimal policies and their values from distributed SARSA to select its customer from the set of local available requests in a distributed manner. Finally, the customers data provided by the city of Chicago is used to validate the proposed algorithms.

preprint2016arXiv

Distributed Average Tracking for Second-order Agents with Nonlinear Dynamics

This paper addresses distributed average tracking of physical second-order agents with nonlinear dynamics, where the interaction among the agents is described by an undirected graph. In both agents' and reference inputs' dynamics, there is a nonlinear term that satisfying the Lipschitz-type condition. To achieve the distributed average tracking problem in the presence of nonlinear term, a non-smooth filter and a control input are designed for each agent. The idea is that each filter outputs converge to the average of the reference inputs and the reference velocities asymptotically and in parallel each agent's position and velocity are driven to track its filter outputs. To overcome the nonlinear term unboundedness effect, novel state-dependent time varying gains are employed in each agent's filter and control input. In the proposed algorithm, each agent needs its neighbors' filters outputs besides its own filter outputs, absolute position and absolute velocity and its neighbors' reference inputs and reference velocities. Finally, the algorithm is simplified to achieve the distributed average tracking of physical second-order agents in the presence of an unknown bounded term in both agents' and reference inputs' dynamics.

preprint2016arXiv

Distributed Convex Optimization for Continuous-Time Dynamics with Time-Varying Cost Function

In this paper, a time-varying distributed convex optimization problem is studied for continuous-time multi-agent systems. Control algorithms are designed for the cases of single-integrator and double-integrator dynamics. Two discontinuous algorithms based on the signum function are proposed to solve the problem in each case. Then in the case of double-integrator dynamics, two continuous algorithms based on, respectively, a time-varying and a fixed boundary layer are proposed as continuous approximations of the signum function. Also, to account for inter-agent collision for physical agents, a distributed convex optimization problem with swarm tracking behavior is introduced for both single-integrator and double-integrator dynamics.

preprint2016arXiv

Heterogeneous Distributed Average Tracking

This paper addresses distributed average tracking for a group of heterogeneous physical agents consisting of single-integrator, double-integrator and Euler-Lagrange dynamics. Here, the goal is that each agent uses local information and local interaction to calculate the average of individual time-varying reference inputs, one per agent. Two nonsmooth algorithms are proposed to achieve the distributed average tracking goal. In our first proposed algorithm, each agent tracks the average of the reference inputs, where each agent is required to have access to only its own position and the relative positions between itself and its neighbors. To relax the restrictive assumption on admissible reference inputs, we propose the second algorithm. A filter is introduced for each agent to generate an estimation of the average of the reference inputs. Then, each agent tracks its own generated signal to achieve the average tracking goal in a distributed manner. Finally, numerical example is included for illustration.

preprint2015arXiv

Distributed Convex Optimization of Time-Varying Cost Functions with Swarm Tracking Behavior for Continuous-time Dynamics

In this paper, a distributed convex optimization problem with swarm tracking behavior is studied for continuous-time multi-agent systems. The agents' task is to drive their center to track an optimal trajectory which minimizes the sum of local time-varying cost functions through local interaction, while maintaining connectivity and avoiding inter-agent collision. Each local cost function is only known to an individual agent and the team's optimal solution is time-varying. Here two cases are considered, single-integrator dynamics and double-integrator dynamics. For each case, a distributed convex optimization algorithm with swarm tracking behavior is proposed where each agent relies only on its own position and the relative positions (and velocities in the double-integrator case) between itself and its neighbors. It is shown that the center of the agents tracks the optimal trajectory, the the connectivity of the agents will be maintained and inter-agent collision is avoided. Finally, numerical examples are included for illustration.