Source author record

Peng Hao

Peng Hao appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

6works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

6 published item(s)

preprint2022arXiv

Hybrid Reinforcement Learning-Based Eco-Driving Strategy for Connected and Automated Vehicles at Signalized Intersections

Taking advantage of both vehicle-to-everything (V2X) communication and automated driving technology, connected and automated vehicles are quickly becoming one of the transformative solutions to many transportation problems. However, in a mixed traffic environment at signalized intersections, it is still a challenging task to improve overall throughput and energy efficiency considering the complexity and uncertainty in the traffic system. In this study, we proposed a hybrid reinforcement learning (HRL) framework which combines the rule-based strategy and the deep reinforcement learning (deep RL) to support connected eco-driving at signalized intersections in mixed traffic. Vision-perceptive methods are integrated with vehicle-to-infrastructure (V2I) communications to achieve higher mobility and energy efficiency in mixed connected traffic. The HRL framework has three components: a rule-based driving manager that operates the collaboration between the rule-based policies and the RL policy; a multi-stream neural network that extracts the hidden features of vision and V2I information; and a deep RL-based policy network that generate both longitudinal and lateral eco-driving actions. In order to evaluate our approach, we developed a Unity-based simulator and designed a mixed-traffic intersection scenario. Moreover, several baselines were implemented to compare with our new design, and numerical experiments were conducted to test the performance of the HRL model. The experiments show that our HRL method can reduce energy consumption by 12.70% and save 11.75% travel time when compared with a state-of-the-art model-based Eco-Driving approach.

preprint2022arXiv

Spatiotemporal Transformer Attention Network for 3D Voxel Level Joint Segmentation and Motion Prediction in Point Cloud

Environment perception including detection, classification, tracking, and motion prediction are key enablers for automated driving systems and intelligent transportation applications. Fueled by the advances in sensing technologies and machine learning techniques, LiDAR-based sensing systems have become a promising solution. The current challenges of this solution are how to effectively combine different perception tasks into a single backbone and how to efficiently learn the spatiotemporal features directly from point cloud sequences. In this research, we propose a novel spatiotemporal attention network based on a transformer self-attention mechanism for joint semantic segmentation and motion prediction within a point cloud at the voxel level. The network is trained to simultaneously outputs the voxel level class and predicted motion by learning directly from a sequence of point cloud datasets. The proposed backbone includes both a temporal attention module (TAM) and a spatial attention module (SAM) to learn and extract the complex spatiotemporal features. This approach has been evaluated with the nuScenes dataset, and promising performance has been achieved.

preprint2020arXiv

End-to-End Vision-Based Adaptive Cruise Control (ACC) Using Deep Reinforcement Learning

This paper presented a deep reinforcement learning method named Double Deep Q-networks to design an end-to-end vision-based adaptive cruise control (ACC) system. A simulation environment of a highway scene was set up in Unity, which is a game engine that provided both physical models of vehicles and feature data for training and testing. Well-designed reward functions associated with the following distance and throttle/brake force were implemented in the reinforcement learning model for both internal combustion engine (ICE) vehicles and electric vehicles (EV) to perform adaptive cruise control. The gap statistics and total energy consumption are evaluated for different vehicle types to explore the relationship between reward functions and powertrain characteristics. Compared with the traditional radar-based ACC systems or human-in-the-loop simulation, the proposed vision-based ACC system can generate either a better gap regulated trajectory or a smoother speed trajectory depending on the preset reward function. The proposed system can be well adaptive to different speed trajectories of the preceding vehicle and operated in real-time.

preprint2014arXiv

Extending cascading gravity model to lower dimensions

The cascading gravity model was proposed to eliminate instabilities of the original DGP model by embedding our 4D universe into a 5D brane, which is itself embedded in a 6D bulk. Thus gravity cascades from 6D down to 4D as we decrease the length scales. We show that it is possible to extend this setup to lower dimensions as well, i.e. there is a self-consistent embedding of a 3D brane into a 4D brane, which is itself embedded in a 5D bulk and so on. This extension fits well into the "vanishing dimensions" framework in which dimensions open up as we increase the length scales.

preprint2011arXiv

Time Evolution of Temperature and Entropy of a Gravitationally Collapsing Cylinder

We investigate the time evolution of the temperature and entropy of a gravitationally collapsing cylinder, represented by an infinitely thin domain wall, as seen by an asymptotic observer. Previous work has shown that the entropy of a spherically symmetric collapsing domain approaches a constant, and we follow this procedure using a (3+1) BTZ metric to see if a different topology will yield different results. We do this by coupling a scalar field to the background of the domain wall and analyzing the spectrum of radiation as a function of time. We find that the spectrum is quasi-thermal, with the degree of thermality increasing as the domain wall approaches the horizon. The thermal distribution allows for the determination of the temperature as a function of time, and we find that the late time temperature is very close to the Hawking temperature and that it also exhibits the proper scaling with the mass. From the temperature we find the entropy. Since the collapsing domain wall is what forms a black hole, we can compare the results to those of the standard entropy-area relation. We find that the entropy does in fact approach a constant that is close to the Hawking entropy. However, the time dependence of the entropy shows that the entropy decreases with time, indicating that a (3+1) BTZ domain wall will not collapse spontaneously.

preprint2010arXiv

Classical and Quantum Equations of Motion for a BTZ Black String in AdS Space

We investigate gravitational collapse of a $(3+1)$-dimensional BTZ black string in AdS space in the context of both classical and quantum mechanics. This is done by first deriving the conserved mass per unit length of the cylindrically symmetric domain wall, which is taken as the classical Hamiltonian of the black string. In the quantum mechanical context, we take primary interest in the behavior of the collapse near the horizon and near the origin (classical singularity) from the point of view of an infalling observer. In the absence of radiation, quantum effects near the horizon do not change the classical conclusions for an infalling observer, meaning that the horizon is not an obstacle for him/her. The most interesting quantum mechanical effect comes in when investigating near the origin. First, quantum effects are able to remove the classical singularity at the origin, since the wave function is non-singular at the origin. Second, the Schrödinger equation describing the behavior near the origin displays non-local effects, which depend on the energy density of the domain wall. This is manifest in that derivatives of the wavefunction at one point are related to the value of the wavefunction at some other distant point.