Source author record

Li Gong

Li Gong appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

5works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2026arXiv

MindWatcher: Toward Smarter Multimodal Tool-Integrated Reasoning

Traditional workflow-based agents exhibit limited intelligence when addressing real-world problems requiring tool invocation. Tool-integrated reasoning (TIR) agents capable of autonomous reasoning and tool invocation are rapidly emerging as a powerful approach for complex decision-making tasks involving multi-step interactions with external environments. In this work, we introduce MindWatcher, a TIR agent integrating interleaved thinking and multimodal chain-of-thought (CoT) reasoning. MindWatcher can autonomously decide whether and how to invoke diverse tools and coordinate their use, without relying on human prompts or workflows. The interleaved thinking paradigm enables the model to switch between thinking and tool calling at any intermediate stage, while its multimodal CoT capability allows manipulation of images during reasoning to yield more precise search results. We implement automated data auditing and evaluation pipelines, complemented by manually curated high-quality datasets for training, and we construct a benchmark, called MindWatcher-Evaluate Bench (MWE-Bench), to evaluate its performance. MindWatcher is equipped with a comprehensive suite of auxiliary reasoning tools, enabling it to address broad-domain multimodal problems. A large-scale, high-quality local image retrieval database, covering eight categories including cars, animals, and plants, endows model with robust object recognition despite its small size. Finally, we design a more efficient training infrastructure for MindWatcher, enhancing training speed and hardware utilization. Experiments not only demonstrate that MindWatcher matches or exceeds the performance of larger or more recent models through superior tool invocation, but also uncover critical insights for agent training, such as the genetic inheritance phenomenon in agentic RL.

preprint2023arXiv

Sensitivity of $CP$ Violation of $Λ$ decay in $J/ψ\to Λ\barΛ$ at STCF

The process of $J/ψ\to Λ\barΛ$ is studied using $1.0\times10^{12}$ $J/ψ$ Monte Carlo (MC) events at $\sqrt{s}$=3.097 GeV with a fast simulation software at future Super Tau Charm Facility (STCF). The statistical sensitivity for $CP$ violation is determined to be the order of $\mathcal{O} (10^{-4})$ by measuring the asymmetric parameters of the $Λ$ decay. Furthermore, the decay of $J/ψ\to Λ\barΛ$ also serves as a benchmark process to optimize the detector responses using the interface provided by the fast simulation software.

preprint2017arXiv

Coherent anti-Stokes Raman Scattering Lidar Using Slow Light: A Theoretical Study

We theoretically investigate a scheme in which backward coherent anti-Stokes Raman scattering (CARS) is significantly enhanced by using slow light. Specifically, we reduce the group velocity of the Stokes excitation pulse by introducing a coupling laser that causes electromagnetically induced transparency (EIT). When the Stokes pulse has a spatial length shorter than the CARS wavelength, the backward CARS emission is significantly enhanced. We also investigated the possibility of applying this scheme as a CARS lidar with O2 or N2 as the EIT medium. We found that if nanosecond laser with large pulse energy (>1 J) and a telescope with large aperture (~10 m) are equipped in the lidar system, a CARS lidar could become much more sensitive than a spontaneous Raman lidar.

preprint2016arXiv

Revealing travel patterns and city structure with taxi trip data

Detecting regional spatial structures based on spatial interactions is crucial in applications ranging from urban planning to traffic control. In the big data era, various movement trajectories are available for studying spatial structures. This research uses large scale Shanghai taxi trip data extracted from GPS-enabled taxi trajectories to reveal traffic flow patterns and urban structure of the city. Using the network science methods, 15 temporally stable regions reflecting the scope of people's daily travels are found using community detection method on the network built from short trips, which represent residents' daily intra-urban travels and exhibit a clear pattern. In each region, taxi traffic flows are dominated by a few 'hubs' and 'hubs' in suburbs impact more trips than 'hubs' in urban areas. Land use conditions in urban regions are different from those in suburban areas. Additionally, 'hubs' in urban area associate with office buildings and commercial areas more, whereas residential land use is more common in suburban hubs. The taxi flow structures and land uses reveal the polycentric and layered concentric structure of Shanghai. Finally, according to the temporal variations of taxi flows and the diversity levels of taxi trip lengths, we explore the total taxi traffic properties of each region and proved the city structure we find. External trips across regions also take large proportion of the total traffic in each region, especially in suburbs. The results could help transportation policy making and shed light on the way to reveal urban structures with big data.

preprint2012arXiv

Near-Infrared Super Resolution Imaging with Metallic Nanoshell Particle Chain Array

We propose a near-infrared super resolution imaging system without a lens or a mirror but with an array of metallic nanoshell particle chain. The imaging array can plasmonically transfer the near-field components of dipole sources in the incoherent and coherent manners and the super resolution images can be reconstructed in the output plane. By tunning the parameters of the metallic nanoshell particle, the plasmon resonance band of the isolate nanoshell particle red-shifts to the near-infrared region. The near-infrared super resolution images are obtained subsequently. We calculate the field intensity distribution at the different planes of imaging process using the finite element method and find that the array has super resolution imaging capability at near-infrared wavelengths. We also show that the image formation highly depends on the coherence of the dipole sources and the image-array distance.