Source author record

Wenxuan Lu

Wenxuan Lu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

4works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

4 published item(s)

preprint2026arXiv

PRISMA: Reinforcement Learning Guided Two-Stage Policy Optimization in Multi-Agent Architecture for Open-Domain Multi-Hop Question Answering

Answering real-world open-domain multi-hop questions over massive corpora is a critical challenge in Retrieval-Augmented Generation (RAG) systems. Recent research employs reinforcement learning (RL) to end-to-end optimize the retrieval-augmented reasoning process, directly enhancing its capacity to resolve complex queries. However, reliable deployment is hindered by two obstacles. 1) Retrieval Collapse: iterative retrieval over large corpora fails to locate intermediate evidence containing bridge answers without reasoning-guided planning, causing downstream reasoning to collapse. 2) Learning Instability: end-to-end trajectory training suffers from weak credit assignment across reasoning chains and poor error localization across modules, causing overfitting to benchmark-specific heuristics that limit transferability and stability. To address these problems, we propose PRISMA, a decoupled RL-guided framework featuring a Plan-Retrieve-Inspect-Solve-Memoize architecture. PRISMA's strength lies in reasoning-guided collaboration: the Inspector provides reasoning-based feedback to refine the Planner's decomposition and fine-grained retrieval, while enforcing evidence-grounded reasoning in the Solver. We optimize individual agent capabilities via Two-Stage Group Relative Policy Optimization (GRPO). Stage I calibrates the Planner and Solver as specialized experts in planning and reasoning, while Stage II utilizes Observation-Aware Residual Policy Optimization (OARPO) to enhance the Inspector's ability to verify context and trigger targeted recovery. Experiments show that PRISMA achieves state-of-the-art performance on ten benchmarks and can be deployed efficiently in real-world scenarios.

preprint2012arXiv

Stability Conditions and Mirror Symmetry of K3 Surfaces in Attractor Backgrounds

We study the space of stability conditions on $K3$ surfaces from the perspective of mirror symmetry. It is done in the so called attractor backgrounds (moduli) which can be far from the conventional large complex limits and are selected by the attractor mechanism for certain black holes. We find certain highly non-generic behaviors of stability walls (a key notion in the study of wall crossings) in the space of stability conditions. They correspond via mirror symmetry to some non-generic behaviors of special Lagrangians in an attractor background. The main results can be understood as a mirror correspondence in a synthesis of homological mirror conjecture and SYZ mirror conjecture.

preprint2012arXiv

SYZ Mirror Symmetry of Hitchin's Moduli Spaces Near Singular Fibers I

We study hyperkahler metrics and hyperholomorphic connections of Hitchin's moduli spaces after Gaiotto, Moore and Neitzke. Their construction via the twistor technique produces intricate wall crossing behaviors. For certain four dimensional Hitchin's moduli spaces local models and degeneration to local models near singular fibers of the Hitchin's fibration are understood.

preprint2011arXiv

Instanton Correction, Wall Crossing And Mirror Symmetry Of Hitchin's Moduli Spaces

We study two instanton correction problems of Hitchin's moduli spaces along with their wall crossing formulas. The hyperkahler metric of a Hitchin's moduli space can be put into an instanton-corrected form according to physicists Gaiotto, Moore and Neitzke. The problem boils down to the construction of a set of special coordinates which can be constructed as Fock-Goncharov coordinates associated with foliations of quadratic differentials on a Riemann surface. A wall crossing formula of Kontsevich and Soibelman arises both as a crucial consistency condition and an effective computational tool. On the other hand Gross and Siebert have succeeded in determining instanton corrections of complex structures of Calabi-Yau varieties in the context of mirror symmetry from a singular affine structure with additional data. We will show that the two instanton correction problems are equivalent in an appropriate sense via the identification of the wall crossing formulas in the metric problem with consistency conditions in the complex structure problem. This result provides examples of Calabi-Yau varieties where the instanton correction (in the sense of mirror symmetry) of metrics and complex structures can be determined.