Source author record

Yang Jun

Yang Jun appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

2works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

2 published item(s)

preprint2021arXiv

Modeling the Interaction between Agents in Cooperative Multi-Agent Reinforcement Learning

Value-based methods of multi-agent reinforcement learning (MARL), especially the value decomposition methods, have been demonstrated on a range of challenging cooperative tasks. However, current methods pay little attention to the interaction between agents, which is essential to teamwork in games or real life. This limits the efficiency of value-based MARL algorithms in the two aspects: collaborative exploration and value function estimation. In this paper, we propose a novel cooperative MARL algorithm named as interactive actor-critic~(IAC), which models the interaction of agents from the perspectives of policy and value function. On the policy side, a multi-agent joint stochastic policy is introduced by adopting a collaborative exploration module, which is trained by maximizing the entropy-regularized expected return. On the value side, we use the shared attention mechanism to estimate the value function of each agent, which takes the impact of the teammates into consideration. At the implementation level, we extend the value decomposition methods to continuous control tasks and evaluate IAC on benchmark tasks including classic control and multi-agent particle environments. Experimental results indicate that our method outperforms the state-of-the-art approaches and achieves better performance in terms of cooperation.

preprint2013arXiv

The Digital discrimination of neutron and γ ray using organic scintillation detector based on wavelet transform modulus maximum

A novel algorithm for the discrimination of neutron and γ-ray with wavelet transform modulus maximum (WTMM) in an organic scintillation has been investigated. Voltage pulses arising from a BC501A organic liquid scintillation detector in a mixed radiation field have been recorded with a fast digital sampling oscilloscope. The performances of most pulse shape discrimination methods in scintillation detection systems using time-domain features of the pulses are affected intensively by noise. However, the WTMM method using frequency-domain features exhibits a strong insensitivity to noise and can be used to discriminate neutron and γ-ray events based on their different asymptotic decay trend between the positive modulus maximum curve and the negative modulus maximum curve in the scale-space plane. This technique has been verified by the corresponding mixed-field data assessed by the time-of-flight (TOF) method and the frequency gradient analysis (FGA) method. It is shown that the characterization of neutron and gamma achieved by the discrimination method based on WTMM is consistent with that afforded by TOF and better than FGA. Moreover, because the WTMM method is it self presented to eliminate the noise, there is no need to make any pretreatment for the pulses.