Source author record

Zongye Zhang

Zongye Zhang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2025arXiv

SkeletonX: Data-Efficient Skeleton-based Action Recognition via Cross-sample Feature Aggregation

While current skeleton action recognition models demonstrate impressive performance on large-scale datasets, their adaptation to new application scenarios remains challenging. These challenges are particularly pronounced when facing new action categories, diverse performers, and varied skeleton layouts, leading to significant performance degeneration. Additionally, the high cost and difficulty of collecting skeleton data make large-scale data collection impractical. This paper studies one-shot and limited-scale learning settings to enable efficient adaptation with minimal data. Existing approaches often overlook the rich mutual information between labeled samples, resulting in sub-optimal performance in low-data scenarios. To boost the utility of labeled data, we identify the variability among performers and the commonality within each action as two key attributes. We present SkeletonX, a lightweight training pipeline that integrates seamlessly with existing GCN-based skeleton action recognizers, promoting effective training under limited labeled data. First, we propose a tailored sample pair construction strategy on two key attributes to form and aggregate sample pairs. Next, we develop a concise and effective feature aggregation module to process these pairs. Extensive experiments are conducted on NTU RGB+D, NTU RGB+D 120, and PKU-MMD with various GCN backbones, demonstrating that the pipeline effectively improves performance when trained from scratch with limited data. Moreover, it surpasses previous state-of-the-art methods in the one-shot setting, with only 1/10 of the parameters and much fewer FLOPs. The code and data are available at: https://github.com/zzysteve/SkeletonX

preprint2025arXiv

Towards Robust and Controllable Text-to-Motion via Masked Autoregressive Diffusion

Generating 3D human motion from text descriptions remains challenging due to the diverse and complex nature of human motion. While existing methods excel within the training distribution, they often struggle with out-of-distribution motions, limiting their applicability in real-world scenarios. Existing VQVAE-based methods often fail to represent novel motions faithfully using discrete tokens, which hampers their ability to generalize beyond seen data. Meanwhile, diffusion-based methods operating on continuous representations often lack fine-grained control over individual frames. To address these challenges, we propose a robust motion generation framework MoMADiff, which combines masked modeling with diffusion processes to generate motion using frame-level continuous representations. Our model supports flexible user-provided keyframe specification, enabling precise control over both spatial and temporal aspects of motion synthesis. MoMADiff demonstrates strong generalization capability on novel text-to-motion datasets with sparse keyframes as motion prompts. Extensive experiments on two held-out datasets and two standard benchmarks show that our method consistently outperforms state-of-the-art models in motion quality, instruction fidelity, and keyframe adherence. The code is available at: https://github.com/zzysteve/MoMADiff

preprint2020arXiv

Selected strong decays of pentaquark State $P_c(4312)$ in a chiral constituent quark model

The newly confirmed pentaquark state $P_c(4312)$ has been treated as a weakly bound $(Σ_c\bar{D})$ state by a well-established chiral constituent quark model and by a dynamical calculation on quark degrees of freedom where the quark exchange effect is accounted for. The obtained mass $4308$ MeV agrees with data. In this work, the selected strong decays of the $P_c(4312)$ state are studied with the obtained wave function. It is shown that the width of the $Λ_c\bar{D}^*$ decay is overwhelmed and the branching ratios of the $p\,η_c$ and $p\,J/ψ$ decays are both less than 1 percentage.

preprint2016arXiv

Decay width of $d^*(2380)\to NN ππ$ processes

The decay widths of four-body double-pion decays $\ds\to pn π^0π^0$, $\ds\to pn π^+π^-$, and iso-scalar parts of $\ds\to pp π^0π^-$ and $\ds\to nn π^+π^0$ are explicitly calculated with the help of the $d^*$ wave function obtained in a chiral SU(3) quark model calculation. The effect of the dynamical structure on $\ds$'s width is analyzed both in the single $ΔΔ$ channel and coupled $ΔΔ$ and $CC$ channel approximations. It is found that in the coupled-channel approximation, the obtained partial decay widths of $\ds\to pn π^0π^0$, $\ds\to pn π^+π^-$, and those of $d^*$ to the iso-scalar parts of $pp π^0π^-$ and $nn π^+π^0$ are about $7.4$MeV, $16.4$MeV, $3.5$MeV and $3.5$MeV, respectively As a consequence, the total width is about $64.5$MeV. These widths are consistent with those estimated by using the corresponding cross section data in our previous investigation and also the observed data. But in the single $ΔΔ$ channel approximation, the widths are still almost 2-times larger than the measured values. Apparently, the explicitly calculated width together with the evaluated mass of $d^*$ in the coupled $ΔΔ$ and $CC$ channel approximation can well explain the observed data, which again supports our assertion that the $\ds$ resonance is a six-quark dominated exotic state.

preprint2015arXiv

A study of $d^*(2380)\to d ππ$ decay width

The decay widths of the $\ds\to d π^0π^0$ and $\ds\to d π^+π^-$ processes are explicitly calculated in terms of our chiral quark model. By using the experimental ratios of cross sections between various decay channels, the partial widths of the $\ds\to pn π^0π^0$, $\ds\to pn π^+π^-$, $\ds\to pp π^0π^-$, and $\ds\to nn π^+π^0$ channels are also extracted. Further including the estimated partial width for the $\ds\to pn $ process, the total width of the $\ds$ resonance is obtained. In the first step of the practical calculation, the effect of the dynamical structure on the width of $\ds$ is studied in the single $ΔΔ$ channel approximation. It is found that the width is reduced by few tens of MeV, in comparison with the one obtained by considering the effect of the kinematics only. This presents the importance of such effect from the dynamical structure. However, the obtained width with the single $ΔΔ$ channel wave function is still too large to explain the data. It implies that the $\ds$ resonance will not consist of the $ΔΔ$ structure only, and instead there should be enough room for other structure such as the hidden-color (CC) component. Thus, in the second step, the width of $\ds$ is further evaluated by using a wave function obtained in the coupled $ΔΔ$ and CC channel calculation in the framework of the Resonating Group Method (RGM). It is shown that the resultant total width for $\ds$ is about 69 MeV, which is compatible with the experimental observation of about 75 MeV and justifies our assertion that the $\ds$ resonance is a hexaquark-dominated exotic state.