Source author record

Yi Hou

Yi Hou appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

9works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

9 published item(s)

preprint2022arXiv

BBA-net: A bi-branch attention network for crowd counting

In the field of crowd counting, the current mainstream CNN-based regression methods simply extract the density information of pedestrians without finding the position of each person. This makes the output of the network often found to contain incorrect responses, which may erroneously estimate the total number and not conducive to the interpretation of the algorithm. To this end, we propose a Bi-Branch Attention Network (BBA-NET) for crowd counting, which has three innovation points. i) A two-branch architecture is used to estimate the density information and location information separately. ii) Attention mechanism is used to facilitate feature extraction, which can reduce false responses. iii) A new density map generation method combining geometric adaptation and Voronoi split is introduced. Our method can integrate the pedestrian's head and body information to enhance the feature expression ability of the density map. Extensive experiments performed on two public datasets show that our method achieves a lower crowd counting error compared to other state-of-the-art methods.

preprint2022arXiv

Enhancing and Dissecting Crowd Counting By Synthetic Data

In this article, we propose a simulated crowd counting dataset CrowdX, which has a large scale, accurate labeling, parameterized realization, and high fidelity. The experimental results of using this dataset as data enhancement show that the performance of the proposed streamlined and efficient benchmark network ESA-Net can be improved by 8.4\%. The other two classic heterogeneous architectures MCNN and CSRNet pre-trained on CrowdX also show significant performance improvements. Considering many influencing factors determine performance, such as background, camera angle, human density, and resolution. Although these factors are important, there is still a lack of research on how they affect crowd counting. Thanks to the CrowdX dataset with rich annotation information, we conduct a large number of data-driven comparative experiments to analyze these factors. Our research provides a reference for a deeper understanding of the crowd counting problem and puts forward some useful suggestions in the actual deployment of the algorithm.

preprint2021arXiv

A Modular and Transferable Reinforcement Learning Framework for the Fleet Rebalancing Problem

Mobility on demand (MoD) systems show great promise in realizing flexible and efficient urban transportation. However, significant technical challenges arise from operational decision making associated with MoD vehicle dispatch and fleet rebalancing. For this reason, operators tend to employ simplified algorithms that have been demonstrated to work well in a particular setting. To help bridge the gap between novel and existing methods, we propose a modular framework for fleet rebalancing based on model-free reinforcement learning (RL) that can leverage an existing dispatch method to minimize system cost. In particular, by treating dispatch as part of the environment dynamics, a centralized agent can learn to intermittently direct the dispatcher to reposition free vehicles and mitigate against fleet imbalance. We formulate RL state and action spaces as distributions over a grid partitioning of the operating area, making the framework scalable and avoiding the complexities associated with multiagent RL. Numerical experiments, using real-world trip and network data, demonstrate that this approach has several distinct advantages over baseline methods including: improved system cost; high degree of adaptability to the selected dispatch method; and the ability to perform scale-invariant transfer learning between problem instances with similar vehicle and request distributions.

preprint2020arXiv

High Yield Growth and Doping of Black Phosphorus with Tunable Electronic Properties

Black phosphorus (BP) has recently attracted significant interest due to its unique electronic and optical properties. Doping is an effective strategy to tune a material's electronic structures, however, the direct and controllable growth of BP with a high yield and its doping remain a great challenge. Here we report an efficient short-distance transport (SDT) growth approach and achieve the controlled growth of high quality BP with the highest yield so far, where 98% of the red phosphorus is converted to BP. The doping of BP by As, Sb, Bi, Se and Te are also achieved by this SDT growth approach. Spectroscopic results show that doping systematically changes its electronic structures including band gap, work function, and energy band position. As a result, we have found that the air-stability of doped BP samples (Sb and Te-doped BP) improves compared with pristine BP, due to the downshift of the conduction band minimum with doping. This work develops a new method to grow BP and doped BP with tunable electronic structures and improved stability, and should extend the uses of these class of materials in various areas.

preprint2016arXiv

A Gb/s Parallel Block-based Viterbi Decoder for Convolutional Codes on GPU

In this paper, we propose a parallel block-based Viterbi decoder (PBVD) on the graphic processing unit (GPU) platform for the decoding of convolutional codes. The decoding procedure is simplified and parallelized, and the characteristic of the trellis is exploited to reduce the metric computation. Based on the compute unified device architecture (CUDA), two kernels with different parallelism are designed to map two decoding phases. Moreover, the optimal design of data structures for several kinds of intermediate information are presented, to improve the efficiency of internal memory transactions. Experimental results demonstrate that the proposed decoder achieves high throughput of 598Mbps on NVIDIA GTX580 and 1802Mbps on GTX980 for the 64-state convolutional code, which are 1.5 times speedup compared to the existing fastest works on GPUs.

preprint2015arXiv

Convolutional Neural Network-Based Image Representation for Visual Loop Closure Detection

Deep convolutional neural networks (CNN) have recently been shown in many computer vision and pattern recog- nition applications to outperform by a significant margin state- of-the-art solutions that use traditional hand-crafted features. However, this impressive performance is yet to be fully exploited in robotics. In this paper, we focus one specific problem that can benefit from the recent development of the CNN technology, i.e., we focus on using a pre-trained CNN model as a method of generating an image representation appropriate for visual loop closure detection in SLAM (simultaneous localization and mapping). We perform a comprehensive evaluation of the outputs at the intermediate layers of a CNN as image descriptors, in comparison with state-of-the-art image descriptors, in terms of their ability to match images for detecting loop closures. The main conclusions of our study include: (a) CNN-based image representations perform comparably to state-of-the-art hand- crafted competitors in environments without significant lighting change, (b) they outperform state-of-the-art competitors when lighting changes significantly, and (c) they are also significantly faster to extract than the state-of-the-art hand-crafted features even on a conventional CPU and are two orders of magnitude faster on an entry-level GPU.

preprint2015arXiv

Time-domain simulation of ultrasound propagation in a tissue-like medium based on the resolution of the nonlinear acoustic constitutive relations

A time-domain numerical code based on the constitutive relations of nonlinear acoustics for simulating ultrasound propagation is presented. To model frequency power law attenuation, such as observed in biological media, multiple relaxation processes are included and relaxation parameters are fitted to both exact frequency power law attenuation and empirically measured attenuation of a variety of tissues that does not fit an exact power law. A computational technique based on artificial relaxation is included to correct the non-negligible numerical dispersion of the numerical method and to improve stability when shock waves are present. This technique avoids the use of high order finite difference schemes, leading to fast calculations. The numerical code is especially suitable to study high intensity and focused axisymmetric acoustic beams in tissue-like medium, as it is based on the full constitutive relations that overcomes the limitations of the parabolic approximations, while some specific effects not contemplated by the Westervelt equation can be also studied. The accuracy of the method is discussed by comparing the proposed simulation solutions to one-dimensional analytical ones, to $k$-space numerical solutions and also to experimental data from a focused beam propagating in a frequency power law attenuation media.

preprint2014arXiv

Joint Successive Cancellation Decoding of Polar Codes over Intersymbol Interference Channels

Polar codes are a class of capacity-achieving codes for the binary-input discrete memoryless channels (B-DMCs). However, when applied in channels with intersymbol interference (ISI), the codes may perform poorly with BCJR equalization and conventional decoding methods. To deal with the ISI problem, in this paper a new joint successive cancellation (SC) decoding algorithm is proposed for polar codes in ISI channels, which combines the equalization and conventional decoding. The initialization information of the decoding method is the likelihood functions of ISI codeword symbols rather than the codeword symbols. The decoding adopts recursion formulas like conventional SC decoding and is without iterations. This is in contrast to the conventional iterative algorithm which performs iterations between the equalizer and decoder. In addition, the proposed SC trellis decoding can be easily extended to list decoding which can further improve the performance. Simulation shows that the proposed scheme significantly outperforms the conventional decoding schemes in ISI channels.

preprint2014arXiv

Nonlinear Acoustics FDTD method including Frequency Power Law Attenuation for Soft Tissue Modeling

This paper describes a model for nonlinear acoustic wave propagation through absorbing and weakly dispersive media, and its numerical solution by means of finite differences in time domain method (FDTD). The attenuation is based on multiple relaxation processes, and provides frequency dependent absorption and dispersion without using computational expensive convolutional operators. In this way, by using an optimization algorithm the coefficients for the relaxation processes can be obtained in order to fit a frequency power law that agrees the experimentally measured attenuation data for heterogeneous media over the typical frequency range for ultrasound medical applications. Our results show that two relaxation processes are enough to fit attenuation data for most soft tissues in this frequency range including the fundamental and the first ten harmonics. Furthermore, this model can fit experimental attenuation data that do not follow exactly a frequency power law over the frequency range of interest. The main advantage of the proposed method is that only one auxiliary field per relaxation process is needed, which implies less computational resources compared with time-domain fractional derivatives solvers based on convolutional operators.