Source author record

Dong Qiu

Dong Qiu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
3topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2025arXiv

Group Deliberation Oriented Multi-Agent Conversational Model for Complex Reasoning

This paper proposes a group deliberation oriented multi-agent conversational model to address the limitations of single large language models in complex reasoning tasks. The model adopts a three-level role division architecture consisting of generation, verification, and integration. An opinion generation agent produces diverse reasoning perspectives, an evidence verification agent retrieves external knowledge and quantifies factual support, and a consistency arbitration agent integrates logically coherent conclusions. A self-game mechanism is introduced to expand multi-path reasoning trajectories, while a retrieval enhancement module dynamically supplements external knowledge. A composite reward function combining factual consistency and logical coherence is designed, and an improved proximal policy optimization strategy is applied for collaborative training. Experimental results show that the proposed model improves multi-hop reasoning accuracy by 16.8 percent on HotpotQA, 14.3 percent on 2WikiMultihopQA, and 19.2 percent on MeetingBank, while improving consistency by 21.5 percent. The model achieves higher reasoning efficiency than mainstream multi-agent approaches, providing an effective and stable solution for complex reasoning tasks.

preprint2025arXiv

Reinforcement Learning-Augmented LLM Agents for Collaborative Decision Making and Performance Optimization

Large Language Models (LLMs) perform well in language tasks but often lack collaborative awareness and struggle to optimize global performance in multi-agent settings. We present a reinforcement learning-augmented LLM agent framework that formulates cooperation as a decentralized partially observable Markov decision process (Dec-POMDP) and adopts centralized training with decentralized execution (CTDE). We introduce Group Relative Policy Optimization (GRPO) to jointly optimize agent policies with access to global signals during training, together with a simplified joint reward that balances task quality, speed, and coordination cost. On collaborative writing and coding benchmarks, our framework delivers a 3x increase in task processing speed over single-agent baselines, 98.7% structural/style consistency in writing, and a 74.6% test pass rate in coding. The approach consistently outperforms strong multi-agent LLM baselines and provides a practical path toward reliable collaboration in complex workflows.

preprint2019arXiv

Supercontinuum generation without residual pump peak through multiple coherent pump seeds

Residual pump peak in fiber-based supercontinuum, as a general phenomenon, limits its practical application. We report a novel supercontinuum generation (SCG) in a conventional highly nonlinear fiber (HNLF) through multiple coherent pump technique, which eliminates the residual pump peak existed in conventional SCG. The multiple coherent pump technique is realized by double bound-state solitons achieved from a homemade modelocked fiber laser. We further compare the SCGs pumped by conventional bound-state soliton and single soliton. It confirms that the effective elimination of the residual pump peak in supercontinuum owes to higher transferring efficiency of the pump energy to new generated frequencies in the multiple coherent pump scheme. The use of multiple coherent pump scheme, i.e., double bound-state solitons, provides a new, simple and promising method to obtain flat supercontinuum source.

preprint2016arXiv

Electron Transmission through Modified Benzene

The renormalization method is applied to investigate the electron transmission properties of a circuit containing a benzene molecule, in which one of the carbon atoms has been modified so as to simulate displacement in position or replacement by another atom. Consideration of the different possible attachments of the leads, and the relative location of the modified atom, results in 9 distinct configurations to examine. For each configuration, the number and locations of anti-resonances, and whether they shift upon variation of the parameters, is seen to be the key to determining the shape of the electron-transmission curve. In particular, those configurations, in which the perturbed atom is not directly attached to a lead, are seen to have the most variation in their structure, compared to pure benzene.