Source author record

Xiumei Wang

Xiumei Wang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

7works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

7 published item(s)

preprint2026arXiv

Inference-Time Scaling for Visual AutoRegressive modeling by Searching Representative Samples

While inference-time scaling has significantly enhanced generative quality in large language and diffusion models, its application to vector-quantized (VQ) visual autoregressive modeling (VAR) remains unexplored. We introduce VAR-Scaling, the first general framework for inference-time scaling in VAR, addressing the critical challenge of discrete latent spaces that prohibit continuous path search. We find that VAR scales exhibit two distinct pattern types: general patterns and specific patterns, where later-stage specific patterns conditionally optimize early-stage general patterns. To overcome the discrete latent space barrier in VQ models, we map sampling spaces to quasi-continuous feature spaces via kernel density estimation (KDE), where high-density samples approximate stable, high-quality solutions. This transformation enables effective navigation of sampling distributions. We propose a density-adaptive hybrid sampling strategy: Top-k sampling focuses on high-density regions to preserve quality near distribution modes, while Random-k sampling explores low-density areas to maintain diversity and prevent premature convergence. Consequently, VAR-Scaling optimizes sample fidelity at critical scales to enhance output quality. Experiments in class-conditional and text-to-image evaluations demonstrate significant improvements in inference process. The code is available at https://github.com/WD7ang/VAR-Scaling.

preprint2022arXiv

Seeking Subjectivity in Visual Emotion Distribution Learning

Visual Emotion Analysis (VEA), which aims to predict people's emotions towards different visual stimuli, has become an attractive research topic recently. Rather than a single label classification task, it is more rational to regard VEA as a Label Distribution Learning (LDL) problem by voting from different individuals. Existing methods often predict visual emotion distribution in a unified network, neglecting the inherent subjectivity in its crowd voting process. In psychology, the \textit{Object-Appraisal-Emotion} model has demonstrated that each individual's emotion is affected by his/her subjective appraisal, which is further formed by the affective memory. Inspired by this, we propose a novel \textit{Subjectivity Appraise-and-Match Network (SAMNet)} to investigate the subjectivity in visual emotion distribution. To depict the diversity in crowd voting process, we first propose the \textit{Subjectivity Appraising} with multiple branches, where each branch simulates the emotion evocation process of a specific individual. Specifically, we construct the affective memory with an attention-based mechanism to preserve each individual's unique emotional experience. A subjectivity loss is further proposed to guarantee the divergence between different individuals. Moreover, we propose the \textit{Subjectivity Matching} with a matching loss, aiming at assigning unordered emotion labels to ordered individual predictions in a one-to-one correspondence with the Hungarian algorithm. Extensive experiments and comparisons are conducted on public visual emotion distribution datasets, and the results demonstrate that the proposed SAMNet consistently outperforms the state-of-the-art methods. Ablation study verifies the effectiveness of our method and visualization proves its interpretability.

preprint2020arXiv

Progressive Perception-Oriented Network for Single Image Super-Resolution

Recently, it has been demonstrated that deep neural networks can significantly improve the performance of single image super-resolution (SISR). Numerous studies have concentrated on raising the quantitative quality of super-resolved (SR) images. However, these methods that target PSNR maximization usually produce blurred images at large upscaling factor. The introduction of generative adversarial networks (GANs) can mitigate this issue and show impressive results with synthetic high-frequency textures. Nevertheless, these GAN-based approaches always have a tendency to add fake textures and even artifacts to make the SR image of visually higher-resolution. In this paper, we propose a novel perceptual image super-resolution method that progressively generates visually high-quality results by constructing a stage-wise network. Specifically, the first phase concentrates on minimizing pixel-wise error, and the second stage utilizes the features extracted by the previous stage to pursue results with better structural retention. The final stage employs fine structure features distilled by the second phase to produce more realistic results. In this way, we can maintain the pixel, and structural level information in the perceptual image as much as possible. It is useful to note that the proposed method can build three types of images in a feed-forward process. Also, we explore a new generator that adopts multi-scale hierarchical features fusion. Extensive experiments on benchmark datasets show that our approach is superior to the state-of-the-art methods. Code is available at https://github.com/Zheng222/PPON.

preprint2016arXiv

A note on Matching Cover Algorithm

A $k$-matching cover of a graph $G$ is a union of $k$ matchings of $G$ which covers $V(G)$. A matching cover of $G$ is optimal if it consists of the fewest matchings of $G$. In this paper, we present an algorithm for finding an optimal matching cover of a graph on $n$ vertices and $m$ edges in $O(nm)$ time. This algorithm corrects an error of Matching Cover Algorithm in (Xiumei Wang, Xiaoxin Song, Jinjiang Yuan, On matching cover of graphs, Math. Program. Ser. A (2014)147: 499-518).

preprint2016arXiv

Improvement in medium-long term frequency stability of integrating sphere cold atom clock

The medium-long term frequency stability of the integrating sphere cold atom clock was improved.During the clock operation, Rb atoms were cooled and manipulated using cooling light diffusely reflected by the inner surface of a microwave cavity in the clock. This light heated the cavity and caused a frequency drift from the resonant frequency of the cavity. Power fluctuations of the cooling light led to atomic density variations in the cavity's central area, which increased the clock frequency instability through a cavity pulling effect. We overcame these limitations with appropriate solutions. A frequency stability of 3.5E-15 was achieved when the integrating time ? increased to 2E4 s.

preprint2016arXiv

Observation of magneto-optical rotation effects in cold $^{87}$Rb atoms in an integrating sphere

We present a modified scheme for detection of the magneto-optical rotation (MOR) effect, where a linearly polarized laser field is interacting with cold $^{87}$Rb atoms in an integrating sphere. The rotation angle of the probe beam's polarization plane is detected in the experiment. The results indicate that the biased magnetic field, the probe light intensity and detuning, and the cold atoms' temperature are key parameters for the MOR effect. This scheme may improve the contrast of the rotation signal and provide an useful approach for high contrast cold atom clocks and magnetometers.

preprint2015arXiv

A new scheme of compact cold atom clock based on diffuse laser cooling in a cylindrical cavity

We present a new scheme of compact Rubidium cold-atom clock which performs the diffuse light cooling, the microwave interrogation and the detection of the clock signal in a cylindrical microwave cavity. The diffuse light is produced by the reflection of the laser light at the inner surface of the microwave cavity. The pattern of injected laser beams is specially designed to make most of the cold atoms accumulate in the center of the microwave cavity. The microwave interrogation of cold atoms in the cavity leads to Ramsey fringes whose line-width is 24.5 Hz and the contrast of 95.6% when the free evolution time is 20 ms. The frequency stability of $7.3\times10^{-13}τ^{-1/2}$ has been achieved recently. The scheme of this physical package can largely reduce the complexity of the cold atom clock, and increase the performance of the clock.