Source author record

Prabir Barooah

Prabir Barooah appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

21works
11topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

21 published item(s)

preprint2022arXiv

Reinforcement Learning for Optimal Control of a District Cooling Energy Plant

District cooling energy plants (DCEPs) consisting of chillers, cooling towers, and thermal energy storage (TES) systems consume a considerable amount of electricity. Optimizing the scheduling of the TES and chillers to take advantage of time-varying electricity price is a challenging optimal control problem. The classical method, model predictive control (MPC), requires solving a high dimensional mixed-integer nonlinear program (MINLP) because of the on/off actuation of the chillers and charging/discharging of TES, which are computationally challenging. RL is an attractive alternative to MPC: the real time control computation is a low-dimensional optimization problem that can be easily solved. However, the performance of an RL controller depends on many design choices. In this paper, we propose a Q-learning based reinforcement learning (RL) controller for this problem. Numerical simulation results show that the proposed RL controller is able to reduce energy cost over a rule-based baseline controller by approximately 8%, comparable to savings reported in the literature with MPC for similar DCEPs. We describe the design choices in the RL controller, including basis functions, reward function shaping, and learning algorithm parameters. Compared to existing work on RL for DCEPs, the proposed controller is designed for continuous state and actions spaces.

preprint2021arXiv

An adaptive MPC scheme for energy-efficient control of building HVAC systems

An autonomous adaptive MPC architecture is presented for control of heating, ventilation and air condition (HVAC) systems to maintain indoor temperature while reducing energy use. Although equipment use and occupant changes with time, existing MPC methods are not capable of automatically relearning models and computing control decisions reliably for extended periods without intervention from a human expert. We seek to address this weakness. Two major features are embedded in the proposed architecture to enable autonomy: (i) a system identification algorithm from our prior work that periodically re-learns building dynamics and unmeasured internal heat loads from data without requiring re-tuning by experts. The estimated model is guaranteed to be stable and has desirable physical properties irrespective of the data; (ii) an MPC planner with a convex approximation of the original nonconvex problem. The planner uses a descent and convergent method, with the underlying optimization problem being feasible and convex. A year long simulation with a realistic plant shows that both of the features of the proposed architecture - periodic model and disturbance update and convexification of the planning problem - are essential to get the performance improvement over a commonly used baseline controller. Without these features, though MPC can outperform the baseline controller in certain situations, the benefits may not be substantial enough to warrant the investment in MPC.

preprint2021arXiv

Increasing Energy Resiliency to Hurricanes with Battery and Rooftop Solar Through Intelligent Control

Rooftop solar photovoltaic (PV) panels together with batteries can provide resiliency to blackouts during natural disasters such as hurricanes. Without intelligent and automated decision making that can trade off conflicting requirements, a large PV system and a large battery is needed to provide meaningful resiliency. By utilizing the flexibility of various household demands, an intelligent system can ensure that critical loads are serviced longer than a non-intelligent system. As a result a smaller (and thus lower cost) system can provide the same energy resilience that a much larger system will be needed otherwise. In this paper we propose such an intelligent control system that uses a model predictive control (MPC) architecture. The optimization problem is formulated as a MILP (mixed integer linear program) due to the on/off decisions for the loads. Performance is compared with two rule based controllers, a simple all-or-none controller that mimics what is available now commercially, and a Rule-Based controller that uses the same information that the MPC controller uses. The controllers are tested through simulation on a PV-battery system chosen carefully for a small single family house in Florida. Simulations are conducted for a one week period during hurricane Irma in 2017. Simulations show that the size of the PV+battery system to provide a certain resiliency performance can be halved by the proposed control system.

preprint2021arXiv

MPC-Based Hierarchical Control of a Multi-Zone Commercial HVAC System

This paper presents a novel architecture for model predictive control (MPC) based indoor climate control of multi-zone buildings to provide energy efficiency. Unlike prior works we do not assume the availability of a high-resolution multi-zone building model, which is challenging to obtain. Instead, the architecture uses a low-resolution model of the building which is divided into a small number of "meta-zones" that can be easily identified using existing data-driven modeling techniques. The proposed architecture is hierarchical. At the higher level, an MPC controller uses the low-resolution model to make decisions for the air handling unit (AHU) and the meta-zones. Since the meta-zones are fictitious, a lower level controller converts the high-level MPC decisions into commands for the individual zones by solving a projection problem that strikes a trade-off between two potentially conflicting goals: the AHU-level decisions made by the MPC are respected while the climate of the individual zones is maintained within the comfort bounds. The performance of the proposed controller is assessed via simulations in a high-fidelity simulation testbed and compared to that of a rule-based controller that is used in practice. Simulations in multiple weather conditions show the effectiveness of the proposed controller in terms of energy savings, climate control, and computational tractability.

preprint2020arXiv

Aggregation and Data Driven Identification of Building Thermal Dynamic Model and Unmeasured Disturbance

An aggregate model is a single-zone equivalent of a multi-zone building, and is useful for many purposes, including model based control of large heating, ventilation and air conditioning (HVAC) equipment. This paper deals with the problem of simultaneously identifying an aggregate thermal dynamic model and unknown disturbances from input-output data. The unknown disturbance is a key challenge since it is not measurable but non-negligible. We first present a principled method to aggregate a multi-zone building model into a single zone model, and show the aggregation is not as trivial as it has been assumed in the prior art. We then provide a method to identify the parameters of the model and the unknown disturbance for this aggregate (single-zone) model. Finally, we test our proposed identification algorithm to data collected from a multi-zone building testbed in Oak Ridge National Laboratory. A key insight provided by the aggregation method allows us to recognize under what conditions the estimation of the disturbance signal will be necessarily poor and uncertain, even in the case of a specially designed test in which the disturbances affecting each zone are known (as the case of our experimental testbed). This insight is used to provide a heuristic that can be used to assess when the identification results are likely to have high or low accuracy.

preprint2020arXiv

Characterizing capacity of flexible loads for providing grid support

Flexible loads are a resource for the Balancing Authority (BA) of the future to aid in the balance of supply and demand in the power grid. Consequently, it is of interest for a BA to know how much flexibility a collection of loads has, so to successfully incorporate flexible loads into grid level resource allocation. Loads' flexibility is limited by all their Quality of Service (QoS) requirements. In this work we present a characterization of capacity for a collection of flexible loads. This characterization is in terms of the Power Spectral Density (PSD) of the reference signal. Two advantages of our characterization are: (i) it easily allows for a BA to use the characterization for resource allocation of flexible loads and (ii) it allows for precise definitions of the power and energy capacity for a collection of flexible loads.

preprint2020arXiv

Predictive resource allocation for flexible loads with local QoS

Loads that can vary their power consumption without violating their Quality of service (QoS), that is flexible loads, are an invaluable resource for grid operators. Utilizing flexible loads as a resource requires the grid operator to incorporate them into a resource allocation problem. Since flexible loads are often consumers, for concerns of privacy it is desirable for this problem to have a distributed implementation. Technically, this distributed implementation manifests itself as a time varying convex optimization problem constrained by the QoS of each load. In the literature, a time invariant form of this problem without all of the necessary QoS metrics for the flexible loads is often considered. Moving to a more realistic setup introduces additional technical challenges, due to the problems' time-varying nature. In this work, we develop an algorithm to account for the challenges introduced when considering a time varying setup with appropriate QoS metrics.

preprint2020arXiv

Simultaneous identification of linear building dynamic model and disturbance using sparsity-promoting optimization

We propose a method that simultaneously identifies a linear time-invariant model of a building's temperature dynamics and a transformed version of the unmeasured disturbance affecting the building. Our method uses l1-regularization to encourage the identified disturbance to be approximately sparse, which is motivated by the slowly-varying nature of occupancy that determines the disturbance. The proposed method involves solving a convex optimization problem that guarantees the identified black-box model possesses known properties of the plant, especially input-output stability and positive DC gains. These features enable one to use the method as part of a self-learning control system in which the model of the building is updated periodically without requiring human intervention. Results from the application of the method on data from a simulated and real building are provided.

preprint2020arXiv

Smart Home Energy Management System for Power System Resiliency

The need for resiliency of electricity supply is increasing due to increasing frequency of natural disasters---such as hurricanes---that disrupt supply from the power grid. Rooftop solar photovoltaic (PV) panels together with batteries can provide resiliency in many scenarios. Without intelligent and automated decision making that can trade off conflicting requirements, a large PV system and a large battery is needed to provide meaningful resiliency. By using forecast of solar generation and household demand, an intelligent decision maker can operate the equipment (battery and critical loads) to ensure that the critical loads are serviced to the maximum duration possible. With the aid of such an intelligent control system, a smaller (and thus lower cost) system can service the primary loads for the same duration that a much larger system will be needed to service otherwise. In this paper we propose such an intelligent control system. A model predictive control (MPC) architecture is used that uses available measurements and forecasts to make optimal decisions for batteries and critical loads in real time. The optimization problem is formulated as a MILP (mixed integer linear program) due to the on/off decisions for the loads. Performance is compared with a non-intelligent baseline controller, for a PV-battery system chosen carefully for a single family house in Florida. Simulations are conducted for a one week period during hurricane Irma in 2017. Simulations show that the cost of the PV+battery system to provide a certain resiliency performance, duration the primary load can be serviced successfully, can be halved by the proposed control system.

preprint2020arXiv

The COVID-19 pandemic's impact on U.S. electricity demand and supply: an early view from the data

After the onset of the recent COVID-19 pandemic, a number of studies reported on possible changes in electricity consumption trends. The overall theme of these reports was that ``electricity use has decreased during the pandemic, but the power grid is still reliable''---mostly due to reduced economic activity. In this paper we analyze electricity data upto end of May 2020, examining both electricity demand and variables that can indicate stress on the power grid, such as peak demand and demand ramp-rate. We limit this study to three states in the USA: New York, California, and Florida. The results indicate that the effect of the pandemic on electricity demand is not a simple reduction from comparable time frames, and there are noticeable differences among regions. The variables that can indicate stress on the grid also conveyed mixed messages: some indicate an increase in stress, some indicate a decrease, and some do not indicate any clear difference. A positive message is that some of the changes that were observed around the time stay-at-home orders were issued appeared to revert back by May 2020. A key challenge in ascribing any observed change to the pandemic is correcting for weather. We provide a weather-correction method, apply it to a small city-wide area, and discuss the implications of the estimated changes in demand. The weather correction exercise underscored that weather-correction is as challenging as it is important.

preprint2014arXiv

Accurate Distributed Time Synchronization in Mobile Wireless Sensor Networks from Noisy Difference Measurements

We propose a distributed algorithm for time synchronization in mobile wireless sensor networks. Each node can employ the algorithm to estimate the global time based on its local clock time. The problem of time synchronization is formulated as nodes estimating their skews and offsets from noisy difference measurements of offsets and logarithm of skews; the measurements acquired by time-stamped message exchanges between neighbors. A distributed stochastic approximation based algorithm is proposed to ensure that the estimation error is mean square convergent (variance converging to 0) under certain conditions. A sequence of scheduled update instants is used to meet the requirement of decreasing time-varying gains that need to be synchronized across nodes with unsynchronized clocks. Moreover, a modification on the algorithm is also presented to improve the initial convergence speed. Simulations indicate that highly accurate global time estimates can be achieved with the proposed algorithm for long time durations, while the errors in competing algorithms increase over time.

preprint2014arXiv

Ancillary Service to the Grid Using Intelligent Deferrable Loads

Renewable energy sources such as wind and solar power have a high degree of unpredictability and time-variation, which makes balancing demand and supply challenging. One possible way to address this challenge is to harness the inherent flexibility in demand of many types of loads. Introduced in this paper is a technique for decentralized control for automated demand response that can be used by grid operators as ancillary service for maintaining demand-supply balance. A Markovian Decision Process (MDP) model is introduced for an individual load. A randomized control architecture is proposed, motivated by the need for decentralized decision making, and the need to avoid synchronization that can lead to large and detrimental spikes in demand. An aggregate model for a large number of loads is then developed by examining the mean field limit. A key innovation is an LTI-system approximation of the aggregate nonlinear model, with a scalar signal as the input and a measure of the aggregate demand as the output. This makes the approximation particularly convenient for control design at the grid level. The second half of the paper contains a detailed application of these results to a network of residential pools. Simulations are provided to illustrate the accuracy of the approximations and effectiveness of the proposed control approach.

preprint2013arXiv

Estimation from Relative Measurements in Mobile Networks with Markovian Switching Topology: Clock Skew and Offset Estimation for Time Synchronization

We analyze a distributed algorithm for estimation of scalar parameters belonging to nodes in a mobile network from noisy relative measurements. The motivation comes from the problem of clock skew and offset estimation for the purpose of time synchronization. The time variation of the network was modeled as a Markov chain. The estimates are shown to be mean square convergent under fairly weak assumptions on the Markov chain, as long as the union of the graphs is connected. Expressions for the asymptotic mean and correlation are also provided. The Markovian switching topology model of mobile networks is justified for certain node mobility models through empirically estimated conditional entropy measures.

preprint2012arXiv

Improving Convergence Rate of Distributed Consensus Through Asymmetric Weights

We propose a weight design method to increase the convergence rate of distributed consensus. Prior work has focused on symmetric weight design due to computational tractability. We show that with proper choice of asymmetric weights, the convergence rate can be improved significantly over even the symmetric optimal design. In particular, we prove that the convergence rate in a lattice graph can be made independent of the size of the graph with asymmetric weights. We then use a Sturm-Liouville operator to approximate the graph Laplacian of more general graphs. A general weight design method is proposed based on this continuum approximation. Numerical computations show that the resulting convergence rate with asymmetric weight design is improved considerably over that with symmetric optimal weights and Metropolis-Hastings weights.

preprint2012arXiv

On achieving size-independent stability margin of vehicular lattice formations with distributed control

We study the stability margin of a vehicular formation with distributed control, in which the control at each vehicle only depends on the information from its neighbors in an information graph. We consider a D-dimensional lattice as information graph, of which the 1-D platoon is a special case. The stability margin is measured by the real part of the least stable eigenvalue of the closed-loop state matrix, which quantifies the rate of decay of initial errors. In [1], it was shown that with symmetric control, in which two neighbors put equal weight on information received from each other, the stability margin of a 1-D vehicular platoon decays to 0 as O(1/N^2), where N is the number of vehicles. Moreover, a perturbation analysis was used to show that with vanishingly small amount of asymmetry in the control gains, the stability margin scaling can be improved to O(1/N). In this paper, we show that, with judicious choice of non-vanishing asymmetry in control, the stability margin of the closed loop can be bounded away from zero uniformly in N. Asymmetry in control gains thus makes the control architecture highly scalable. The results are also generalized to D-dimensional lattice information graphs that were studied in [2], and the correspondingly stronger conclusions than those derived in [2] are obtained. In addition, we show that the size-independent stability margin can be achieved with relative position and relative velocity (RPRV) feedback as well as relative position and absolute velocity (RPAV) feedback, while the analysis in [1], [2] was only for the RPAV case.

preprint2011arXiv

Control of large 1D networks of double integrator agents: role of heterogeneity and asymmetry on stability margin

We consider the distributed control of a network of heterogeneous agents with double integrator dynamics to maintain a rigid formation in 1D Euclidean space. The control signal at a vehicle is allowed to use relative position and velocity with its two nearest neighbors. Most of the work on this problem, though extensive, has been limited to homogeneous networks, in which agents have identical masses and control gains, and symmetric control, in which information from front and back neighbors are weighted equally. We examine the effect of heterogeneity and asymmetry on the closed loop stability margin, which is measured by the real part of the least stable pole of the closed-loop system. By using a PDE (partial differential equation) approximation in the limit of large number of vehicles, we show that heterogeneity has little effect while asymmetry has a significant effect on the stability margin. When control is symmetric, the stability margin decays to 0 as $O(1/N^2)$, where $N$ is the number of agents, even when the agents are heterogeneous in their masses and control gains. In contrast, we show that arbitrarily small amount of asymmetry in the velocity feedback gains can improve the decay of the stability margin to $O(1/N)$. Poor design of such asymmetry makes the closed loop unstable for sufficiently large $N$. With equal amount of asymmetry in both position and velocity feedback gains, the closed loop is stable for arbitrary $N$ and the stability margin scaling trend can also be improved to $O(1/N)$, but the sensitivity to disturbance becomes worse. Effect of asymmetry in position feedback gains alone and unequal amount of asymmetry in position and velocity feedback are open problems. Numerical computations are provided to corroborate the analysis.

preprint2011arXiv

Decentralized control of large vehicular formations: stability margin and sensitivity to external disturbances

We study the stability and robustness of large-scale vehicular formations, in which each vehicle is modeled as a double-integrator. Two types of information graphs are considered: directed trees and undirected graphs. We prove stability of the formation with arbitrary number of vehicles for linear as well as a class of nonlinear controllers. In the case of linear control, we provide quantitative scaling laws of the stability margin and sensitivity to external disturbances (H-infinity norm) with respect to the number of vehicles $N$ in the formation. It is shown that the formation with directed tree graph achieves size-independent stability margin but suffers from high algebraic growth of initial errors. The stability margin in case of the undirected graph decays to 0 as at least $O(1/N)$. In addition, we show that the sensitivity to external disturbances in directed tree graphs is geometric in $k$ where $k <= N$ is the number of generations of the directed tree, while that of the undirected graph is only quadratic in $N$. In particular, for 1-D vehicular platoons, we obtain precise formulae for the H-infinity norm of the transfer function from the disturbances to the position errors. It is shown that the H-infinity norm scales as $O(α^N) (α>1)$ for predecessor-following architecture, but only as $O(N^3)$ for symmetric bidirectional architectures. For a class of nonlinear controllers, numerical simulations show that the transient response due to initial errors and sensitivity to external disturbances are improved considerably for the formation with directed tree graphs. However, by using the nonlinear controller considered, little improvement can be made for that with undirected graphs.

preprint2011arXiv

Detecting Separation in Robotic and Sensor Networks

In this paper we consider the problem of monitoring detecting separation of agents from a base station in robotic and sensor networks. Such separation can be caused by mobility and/or failure of the agents. While separation/cut detection may be performed by passing messages between a node and the base in static networks, such a solution is impractical for networks with high mobility, since routes are constantly changing. We propose a distributed algorithm to detect separation from the base station. The algorithm consists of an averaging scheme in which every node updates a scalar state by communicating with its current neighbors. We prove that if a node is permanently disconnected from the base station, its state converges to $0$. If a node is connected to the base station in an average sense, even if not connected in any instant, then we show that the expected value of its state converges to a positive number. Therefore, a node can detect if it has been separated from the base station by monitoring its state. The effectiveness of the proposed algorithm is demonstrated through simulations, a real system implementation and experiments involving both static as well as mobile networks.

preprint2010arXiv

Stability Margin Scaling Laws for Distributed Formation Control as a Function of Network Structure

We consider the problem of distributed formation control of a large number of vehicles. An individual vehicle in the formation is assumed to be a fully actuated point mass. A distributed control law is examined: the control action on an individual vehicle depends on (i) its own velocity and (ii) the relative position measurements with a small subset of vehicles (neighbors) in the formation. The neighbors are defined according to an information graph. In this paper we describe a methodology for modeling, analysis, and distributed control design of such vehicular formations whose information graph is a D-dimensional lattice. The modeling relies on an approximation based on a partial differential equation (PDE) that describes the spatio-temporal evolution of position errors in the formation. The analysis and control design is based on the PDE model. We deduce asymptotic formulae for the closed-loop stability margin (absolute value of the real part of the least stable eigenvalue) of the controlled formation. The stability margin is shown to approach 0 as the number of vehicles N goes to infinity. The exponent on the scaling law for the stability margin is influenced by the dimension and the structure of the information graph. We show that the scaling law can be improved by employing a higher dimensional information graph. Apart from analysis, the PDE model is used for a mistuning-based design of control gains to maximize the stability margin. Mistuning here refers to small perturbation of control gains from their nominal symmetric values. We show that the mistuned design can have a significantly better stability margin even with a small amount of perturbation. The results of the analysis with the PDE model are corroborated with numerical computation of eigenvalues with the state-space model of the formation.

preprint2009arXiv

Error Scaling Laws for Linear Optimal Estimation from Relative Measurements

We study the problem of estimating vector-valued variables from noisy "relative" measurements. This problem arises in several sensor network applications. The measurement model can be expressed in terms of a graph, whose nodes correspond to the variables and edges to noisy measurements of the difference between two variables. We take an arbitrary variable as the reference and consider the optimal (minimum variance) linear unbiased estimate of the remaining variables. We investigate how the error in the optimal linear unbiased estimate of a node variable grows with the distance of the node to the reference node. We establish a classification of graphs, namely, dense or sparse in Rd,1<= d <=3, that determines how the linear unbiased optimal estimation error of a node grows with its distance from the reference node. In particular, if a graph is dense in 1,2, or 3D, then a node variable's estimation error is upper bounded by a linear, logarithmic, or bounded function of distance from the reference, respectively. Corresponding lower bounds are obtained if the graph is sparse in 1, 2 and 3D. Our results also show that naive measures of graph density, such as node degree, are inadequate predictors of the estimation error. Being true for the optimal linear unbiased estimate, these scaling laws determine algorithm-independent limits on the estimation accuracy achievable in large graphs.

preprint2008arXiv

Mistuning-based Control Design to Improve Closed-Loop Stability of Vehicular Platoons

We consider a decentralized bidirectional control of a platoon of N identical vehicles moving in a straight line. The control objective is for each vehicle to maintain a constant velocity and inter-vehicular separation using only the local information from itself and its two nearest neighbors. Each vehicle is modeled as a double integrator. To aid the analysis, we use continuous approximation to derive a partial differential equation (PDE) approximation of the discrete platoon dynamics. The PDE model is used to explain the progressive loss of closed-loop stability with increasing number of vehicles, and to devise ways to combat this loss of stability. If every vehicle uses the same controller, we show that the least stable closed-loop eigenvalue approaches zero as O(1/N^2) in the limit of a large number (N) of vehicles. We then show how to ameliorate this loss of stability by small amounts of "mistuning", i.e., changing the controller gains from their nominal values. We prove that with arbitrary small amounts of mistuning, the asymptotic behavior of the least stable closed loop eigenvalue can be improved to O(1/N) All the conclusions drawn from analysis of the PDE model are corroborated via numerical calculations of the state-space platoon model.