Source author record

Qian Cao

Qian Cao appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2026arXiv

DPWriter: Reinforcement Learning with Diverse Planning Branching for Creative Writing

Reinforcement learning (RL)-based enhancement of large language models (LLMs) often leads to reduced output diversity, undermining their utility in open-ended tasks like creative writing. Current methods lack explicit mechanisms for guiding diverse exploration and instead prioritize optimization efficiency and performance over diversity. This paper proposes an RL framework structured around a semi-structured long Chain-of-Thought (CoT), in which the generation process is decomposed into explicitly planned intermediate steps. We introduce a Diverse Planning Branching method that strategically introduces divergence at the planning phase based on diversity variation, alongside a group-aware diversity reward to encourage distinct trajectories. Experimental results on creative writing benchmarks demonstrate that our approach significantly improves output diversity without compromising generation quality, consistently outperforming existing baselines.

preprint2026arXiv

Qwen-Scope: Turning Sparse Features into Development Tools for Large Language Models

Large language models have achieved remarkable capabilities across diverse tasks, yet their internal decision-making processes remain largely opaque, limiting our ability to inspect, control, and systematically improve them. This opacity motivates a growing body of research in mechanistic interpretability, with sparse autoencoders (SAEs) emerging as one of the most promising tools for decomposing model activations into sparse, interpretable feature representations. We introduce Qwen-Scope, an open-source suite of SAEs built on the Qwen model family, comprising 14 groups of SAEs across 7 model variants from the Qwen3 and Qwen3.5 series, covering both dense and mixture-of-expert architectures. Built on top of these SAEs, we show that SAEs can go beyond post-hoc analysis to serve as practical interfaces for model development along four directions: (i) inference-time steering, where SAE feature directions control language, concepts, and preferences without modifying model weights; (ii) evaluation analysis, where activated SAE features provide a representation-level proxy for benchmark redundancy and capability coverage; (iii) data-centric workflows, where SAE features support multilingual toxicity classification and safety-oriented data synthesis; and (iv) post-training optimization, where SAE-derived signals are incorporated into supervised fine-tuning and reinforcement learning objectives to mitigate undesirable behaviors such as code-switching and repetition. Together, these results demonstrate that SAEs can serve not only as post-hoc analysis tools, but also as reusable representation-level interfaces for diagnosing, controlling, evaluating, and improving large language models. By open-sourcing Qwen-Scope, we aim to support mechanistic research and accelerate practical workflows that connect model internals to downstream behavior.

preprint2022arXiv

Superfluid-Mott insulator quantum phase transition in a cavity optomagnonic system

The emerging hybrid cavity optomagnonic system is a very promising quantum information processing platform for its strong or ultrastrong photon-magnon interaction on the scale of micrometers in the experiment. In this paper, the superfluid-Mott insulator quantum phase transition in a two-dimensional cavity optomagnonic array system has been studied based on this characteristic. The analytical solution of the critical hopping rate is obtained by the mean field approach, second perturbation theory and Landau second order phase transition theory. The numerical results show that the increasing coupling strength and the positive detunings of the photon and the magnon favor the coherence and then the stable areas of Mott lobes are compressed correspondingly. Moreover, the analytical results agree with the numerical ones when the total excitation number is lower. Finally, an effective repulsive potential is constructed to exhibit the corresponding mechanism. The results obtained here provide an experimentally feasible scheme for characterizing the quantum phase transitions in a cavity optomagnonic array system, which will offer valuable insight for quantum simulations.

preprint2021arXiv

Photonic toroidal vortex

Toroidal vortices are whirling disturbances rotating about a ring-shaped core while advancing in the direction normal to the ring orifice. Toroidal vortices are commonly found in nature and being studied in a wide range of disciplines. Here we report the experimental observation of photonic toroidal vortex as a new solution to Maxwell's equations with the use of conformal mapping. The helical phase twists around a closed loop leading to an azimuthal local orbital angular momentum density. The preparation of such intriguing light field may offer insights of extending toroidal vortex to other disciplines and find important applications in light-matter interaction, optical manipulation, photonic symmetry and topology, and quantum information.

preprint2012arXiv

Power Allocation and Pricing in Multi-User Relay Networks Using Stackelberg and Bargaining Games

This paper considers a multi-user single-relay wireless network, where the relay gets paid for helping the users forward signals, and the users pay to receive the relay service. We study the relay power allocation and pricing problems, and model the interaction between the users and the relay as a two-level Stackelberg game. In this game, the relay, modeled as the service provider and the leader of the game, sets the relay price to maximize its revenue; while the users are modeled as customers and the followers who buy power from the relay for higher transmission rates. We use a bargaining game to model the negotiation among users to achieve a fair allocation of the relay power. Based on the proposed fair relay power allocation rule, the optimal relay power price that maximizes the relay revenue is derived analytically. Simulation shows that the proposed power allocation scheme achieves higher network sum-rate and relay revenue than the even power allocation. Furthermore, compared with the sum-rate-optimal solution, simulation shows that the proposed scheme achieves better fairness with comparable network sum-rate for a wide range of network scenarios. The proposed pricing and power allocation solutions are also shown to be consistent with the laws of supply and demand.