Source author record

Xin Xie

Xin Xie appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

cond-mat.mes-hall physics.optics Machine Learning Computer Vision Artificial Intelligence Computation and Language physics.atom-ph cond-mat.quant-gas Cryptography and Security Distributed, Parallel, and Cluster Computing Genomics quant-ph

Catalog footprint

What is connected

17works

12topics

4close collaborators

Actions

Connect this record

Open graph Browse works

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

preprint2026arXiv

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding

Dynamic spatial reasoning from monocular video is essential for bridging visual intelligence and the physical world, yet remains challenging for vision-language models (VLMs). Prior approaches either verbalize spatial-temporal reasoning entirely as text, which is inherently verbose and imprecise for complex dynamics, or rely on external geometric modules that increase inference complexity without fostering intrinsic model capability. In this paper, we present 4DThinker, the first framework that enables VLMs to "think with 4D" through dynamic latent mental imagery, i.e., internally simulating how scenes evolve within the continuous hidden space. Specifically, we first introduce a scalable, annotation-free data generation pipeline that synthesizes 4D reasoning data from raw videos. We then propose Dynamic-Imagery Fine-Tuning (DIFT), which jointly supervises textual tokens and 4D latents to ground the model in dynamic visual semantics. Building on this, 4D Reinforcement Learning (4DRL) further tackles complex reasoning tasks via outcome-based rewards, restricting policy gradients to text tokens to ensure stable optimization. Extensive experiments across multiple dynamic spatial reasoning benchmarks demonstrate that 4DThinker consistently outperforms strong baselines and offers a new perspective toward 4D reasoning in VLMs. Our code is available at https://github.com/zhangquanchen/4DThinker.

preprint2026arXiv

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

General reasoning represents a long-standing and formidable challenge in artificial intelligence. Recent breakthroughs, exemplified by large language models (LLMs) and chain-of-thought prompting, have achieved considerable success on foundational reasoning tasks. However, this success is heavily contingent upon extensive human-annotated demonstrations, and models' capabilities are still insufficient for more complex problems. Here we show that the reasoning abilities of LLMs can be incentivized through pure reinforcement learning (RL), obviating the need for human-labeled reasoning trajectories. The proposed RL framework facilitates the emergent development of advanced reasoning patterns, such as self-reflection, verification, and dynamic strategy adaptation. Consequently, the trained model achieves superior performance on verifiable tasks such as mathematics, coding competitions, and STEM fields, surpassing its counterparts trained via conventional supervised learning on human demonstrations. Moreover, the emergent reasoning patterns exhibited by these large-scale models can be systematically harnessed to guide and enhance the reasoning capabilities of smaller models.

preprint2024arXiv

DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

The rapid development of open-source large language models (LLMs) has been truly remarkable. However, the scaling law described in previous literature presents varying conclusions, which casts a dark cloud over scaling LLMs. We delve into the study of scaling laws and present our distinctive findings that facilitate scaling of large scale models in two commonly used open-source configurations, 7B and 67B. Guided by the scaling laws, we introduce DeepSeek LLM, a project dedicated to advancing open-source language models with a long-term perspective. To support the pre-training phase, we have developed a dataset that currently consists of 2 trillion tokens and is continuously expanding. We further conduct supervised fine-tuning (SFT) and Direct Preference Optimization (DPO) on DeepSeek LLM Base models, resulting in the creation of DeepSeek Chat models. Our evaluation results demonstrate that DeepSeek LLM 67B surpasses LLaMA-2 70B on various benchmarks, particularly in the domains of code, mathematics, and reasoning. Furthermore, open-ended evaluations reveal that DeepSeek LLM 67B Chat exhibits superior performance compared to GPT-3.5.

preprint2022arXiv

A Survey on Gradient Inversion: Attacks, Defenses and Future Directions

Recent studies have shown that the training samples can be recovered from gradients, which are called Gradient Inversion (GradInv) attacks. However, there remains a lack of extensive surveys covering recent advances and thorough analysis of this issue. In this paper, we present a comprehensive survey on GradInv, aiming to summarize the cutting-edge research and broaden the horizons for different domains. Firstly, we propose a taxonomy of GradInv attacks by characterizing existing attacks into two paradigms: iteration- and recursion-based attacks. In particular, we dig out some critical ingredients from the iteration-based attacks, including data initialization, model training and gradient matching. Second, we summarize emerging defense strategies against GradInv attacks. We find these approaches focus on three perspectives covering data obscuration, model improvement and gradient protection. Finally, we discuss some promising directions and open problems for further research.

preprint2022arXiv

Federated Unlearning via Class-Discriminative Pruning

We explore the problem of selectively forgetting categories from trained CNN classification models in the federated learning (FL). Given that the data used for training cannot be accessed globally in FL, our insights probe deep into the internal influence of each channel. Through the visualization of feature maps activated by different channels, we observe that different channels have a varying contribution to different categories in image classification. Inspired by this, we propose a method for scrubbing the model clean of information about particular categories. The method does not require retraining from scratch, nor global access to the data used for training. Instead, we introduce the concept of Term Frequency Inverse Document Frequency (TF-IDF) to quantize the class discrimination of channels. Channels with high TF-IDF scores have more discrimination on the target categories and thus need to be pruned to unlearn. The channel pruning is followed by a fine-tuning process to recover the performance of the pruned model. Evaluated on CIFAR10 dataset, our method accelerates the speed of unlearning by 8.9x for the ResNet model, and 7.9x for the VGG model under no degradation in accuracy, compared to retraining from scratch. For CIFAR100 dataset, the speedups are 9.9x and 8.4x, respectively. We envision this work as a complementary block for FL towards compliance with legal and ethical criteria.

preprint2022arXiv

Single charge control of localized excitons in heterostructures with ferroelectric thin films and two-dimensional transition metal dichalcogenides

Single charge control of localized excitons (LXs) in two-dimensional transition metal dichalcogenides (TMDCs) is crucial for potential applications in quantum information processing and storage. However, traditional electrostatic doping method with applying metallic gates onto TMDCs may cause the inhomogeneous charge distribution, optical quench, and energy loss. Here, by locally controlling the ferroelectric polarization of the ferroelectric thin film BiFeO3 (BFO) with a scanning probe, we can deterministically manipulate the doping type of monolayer WSe2 to achieve the p-type and n-type doping. This nonvolatile approach can maintain the doping type and hold the localized excitonic charges for a long time without applied voltage. Our work demonstrated that ferroelectric polarization of BFO can control the charges of LXs effectively. Neutral and charged LXs have been observed in different ferroelectric polarization regions, confirmed by magnetic optical measurement. Highly circular polarization degree about 90 % of the photon emission from these quantum emitters have been achieved in high magnetic fields. Controlling single charge of LXs in a non-volatile way shows a great potential for deterministic photon emission with desired charge states for photonic long-term memory.

preprint2021arXiv

Coherent control of spin tunneling in a spin-orbit coupled bosonic triple well

We study the coherent control of spin tunneling for a spin-orbit (SO) coupled boson held in a driven triple well. Under high-frequency approximation, we analytically obtain the quasienergies of the SO-coupled bosonic triple-well system and fine energy band structure is displayed. By adjusting the driving parameters, we reveal that the directed selective spin-flipping or spin-conserving tunneling of a SO-coupled boson occurs along different pathways and in different directions. The analytical results are numerically confirmed and perfect agreements are found. Further, a scheme of quantum spin switch with or without spin-flipping is presented. These results may be useful for quantum information processing and the design of spintronic devices.

preprint2021arXiv

Position-dependent chiral coupling between single quantum dots and cross waveguides

Chiral light-matter interaction between photonic nanostructures with quantum emitters shows great potential to implement spin-photon interfaces for quantum information processing. Position-dependent spin momentum locking of the quantum emitter is important for these chiral coupled nanostructures. Here, we report the position-dependent chiral coupling between quantum dots (QDs) and cross waveguides both numerically and experimentally. Four quantum dots distributed at different positions in the cross section are selected to characterize the chiral properties of the device. Directional emission is achieved in a single waveguide as well as in both two waveguides simultaneously. In addition, the QD position can be determined with the chiral contrasts from four outputs. Therefore, the cross waveguide can function as a one-way unidirectional waveguide and a circularly polarized beam splitter by placing the QD in a rational position, which has potential applications in spin-to-path encoding for complex quantum optical networks at the single-photon level.

preprint2021arXiv

Strong Triplet-Exciton-LO-Phonon Coupling in Two-Dimensional Layered Organic-Inorganic Hybrid Perovskite Single Crystal Microflakes

Two-dimensional (2D) layered hybrid perovskites provide an ideal platform for studying the properties of excitons. Here, we report on a strong triplet-exciton and longitudinal-optical (LO) phonon coupling in 2D (C6H5CH2CH2NH3, PEA)2PbBr4 perovskites. The triplet excitons exhibit strong photoluminescence (PL) in thick perovskite microflakes, and the PL is not detectable for monolayer microflakes. The coupling strength of the triplet exciton-LO phonon is approximately two to three times greater than that of the singlet exciton-LO phonon with a LO phonon energy of about 21 meV. This difference might due to the different locations of singlet excitons located in the well and triplet excitons located in the barrier in the 2D layered perovskite. Revealing the strong coupling of triplet exciton-LO phonon provides a fundamental understanding of many-body interaction in hybrid perovskites, which is useful to develop and optimize the optoelectronic devices based on 2D perovskites in the future.

preprint2020arXiv

Cavity Quantum Electrodynamics with Second-Order Topological Corner State

Topological photonics provides a new paradigm in studying cavity quantum electrodynamics with robustness to disorder. In this work, we demonstrate the coupling between single quantum dots and the second-order topological corner state. Based on the second-order topological corner state, a topological photonic crystal cavity is designed and fabricated into GaAs slabs with quantum dots embedded. The coexistence of corner state and edge state with high quality factor close to 2000 is observed. The enhancement of photoluminescence intensity and emission rate are both observed when the quantum dot is on resonance with the corner state. This result enables the application of topology into cavity quantum electrodynamics, offering an approach to topological devices for quantum information processing.

preprint2020arXiv

Diabolical Points in Coupled Active Cavities with Quantum Emitters

In single microdisks, embedded active emitters intrinsically affect the cavity mode of microdisks, which results in a trivial symmetric backscattering and a low controllability. Here we propose a macroscopical control of the backscattering direction by optimizing the cavity size. The signature of positive and negative backscattering directions in each single microdisk is confirmed with two strongly coupled microdisks. Furthermore, the diabolical points are achieved at the resonance of two microdisks, which agrees well with the theoretical calculations considering backscattering directions. The diabolical points in active optical structures pave a way to implement quantum information processing with geometric phase in quantum photonic networks.

preprint2020arXiv

Electron and hole g tensors of neutral and charged excitons in single quantum dots by high-resolution photocurrent spectroscopy

We report a high-resolution photocurrent (PC) spectroscopy of a single self-assembled InAs/GaAs quantum dot (QD) embedded in an n-i-Schottky device with an applied vector magnetic field. The PC spectra of positively charged exciton (X$^+$) and neutral exciton (X$^0$) are obtained by two-color resonant excitation. With an applied magnetic field in Voigt geometry, the double $Λ$ energy level structure of X$^+$ and the dark states of X$^0$ are observed in PC spectra clearly. In Faraday geometry, the PC amplitude of X$^+$ decreases and then quenches with the increasing of the magnetic field, which provides a new way to determine the relative sign of the electron and the hole g-factors. With an applied vector magnetic field, the electron and the hole g-factor tensors of X$^+$ and X$^0$ are obtained. The anisotropy of the hole g-factors of both X$^+$ and X$^0$ is larger than that of the electron.

preprint2020arXiv

Identifying defect-related quantum emitters in monolayer WSe$_2$

Monolayer transition metal dichalcogenides have recently attracted great interests because the quantum dots embedded in monolayer can serve as optically active single photon emitters. Here, we provide an interpretation of the recombination mechanisms of these quantum emitters through polarization-resolved and magneto-optical spectroscopy at low temperature. Three types of defect-related quantum emitters in monolayer tungsten diselenide (WSe$_2$) are observed, with different exciton g factors of 2.02, 9.36 and unobservable Zeeman shift, respectively. The various magnetic response of the spatially localized excitons strongly indicate that the radiative recombination stems from the different transitions between defect-induced energy levels, valance and conduction bands. Furthermore, the different g factors and zero-field splittings of the three types of emitters strongly show that quantum dots embedded in monolayer have various types of confining potentials for localized excitons, resulting in electron-hole exchange interaction with a range of values in the presence of anisotropy. Our work further sheds light on the recombination mechanisms of defect-related quantum emitters and paves a way toward understanding the role of defects in single photon emitters in atomically thin semiconductors.

preprint2020arXiv

Large photoluminescence enhancement by an out-of-plane magnetic field in exfoliated WS$_2$ flakes

We report an out-of-plane magnetic field induced large photoluminescence enhancement in WS${}_2$ flakes at $4$ K, in contrast to the photoluminescence enhancement provided by in-plane field in general. Two mechanisms for the enhancement are proposed. One is a larger overlap of electron and hole caused by the magnetic field induced confinement. The other is that the energy difference between $Λ$ and K valleys is reduced by magnetic field, and thus enhancing the corresponding indirect-transition trions. Meanwhile, the Landé g factor of the trion is measured as $-0.8$, whose absolute value is much smaller than normal exciton, which is around $|-4|$. A model for the trion g factor is presented, confirming that the smaller absolute value of Landé g factor is a behavior of this $Λ$-K trion. By extending the valley space, we believe this work provides a further understanding of the valleytronics in monolayer transition metal dichalcogenides.

preprint2020arXiv

Low-threshold topological nanolasers based on second-order corner state

The topological lasers, which are immune to imperfections and disorders, have been recently demonstrated based on many kinds of robust edge states, being mostly at microscale. The realization of 2D on-chip topological nanolasers, having the small footprint, low threshold and high energy efficiency, is still to be explored. Here, we report on the first experimental demonstration of the topological nanolaser with high performance in 2D photonic crystal slab. Based on the generalized 2D Su-Schrieffer-Heeger model, a topological nanocavity is formed with the help of the Wannier-type 0D corner state. Laser behaviors with low threshold about 1 $μW$ and high spontaneous emission coupling factor of 0.25 are observed with quantum dots as the active material. Such performance is much better than that of topological edge lasers and comparable to conventional photonic crystal nanolasers. Our experimental demonstration of the low-threshold topological nanolaser will be of great significance to the development of topological nanophotonic circuitry for manipulation of photons in classical and quantum regimes.

preprint2020arXiv

Observation of Efimov Universality across a Non-Universal Feshbach Resonance in \textsuperscript{39}K

We study three-atom inelastic scattering in ultracold \textsuperscript{39}K near a Feshbach resonance of intermediate coupling strength. The non-universal character of such resonance leads to an abnormally large Efimov absolute length scale and a relatively small effective range $r_e$, allowing the features of the \textsuperscript{39}K Efimov spectrum to be better isolated from the short-range physics. Meticulous characterization of and correction for finite temperature effects ensure high accuracy on the measurements of these features at large-magnitude scattering lengths. For a single Feshbach resonance, we unambiguously locate four distinct features in the Efimov structure. Three of these features form ratios that obey the Efimov universal scaling to within 10\%, while the fourth feature, occurring at a value of scattering length closest to $r_e$, instead deviates from the universal value.

preprint2010arXiv

Towards automated high-throughput screening of C. elegans on agar

High-throughput screening (HTS) using model organisms is a promising method to identify a small number of genes or drugs potentially relevant to human biology or disease. In HTS experiments, robots and computers do a significant portion of the experimental work. However, one remaining major bottleneck is the manual analysis of experimental results, which is commonly in the form of microscopy images. This manual inspection is labor intensive, slow and subjective. Here we report our progress towards applying computer vision and machine learning methods to analyze HTS experiments that use Caenorhabditis elegans (C. elegans) worms grown on agar. Our main contribution is a robust segmentation algorithm for separating the worms from the background using brightfield images. We also show that by combining the output of this segmentation algorithm with an algorithm to detect the fluorescent dye, Nile Red, we can reliably distinguish different fluorescence-based phenotypes even though the visual differences are subtle. The accuracy of our method is similar to that of expert human analysts. This new capability is a significant step towards fully automated HTS experiments using C. elegans.

Xin Xie

What is connected

Connect this record

See the researcher in context

Building this map preview

17 published item(s)

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

A Survey on Gradient Inversion: Attacks, Defenses and Future Directions

Federated Unlearning via Class-Discriminative Pruning

Single charge control of localized excitons in heterostructures with ferroelectric thin films and two-dimensional transition metal dichalcogenides

Coherent control of spin tunneling in a spin-orbit coupled bosonic triple well

Position-dependent chiral coupling between single quantum dots and cross waveguides

Strong Triplet-Exciton-LO-Phonon Coupling in Two-Dimensional Layered Organic-Inorganic Hybrid Perovskite Single Crystal Microflakes

Cavity Quantum Electrodynamics with Second-Order Topological Corner State

Diabolical Points in Coupled Active Cavities with Quantum Emitters

Electron and hole g tensors of neutral and charged excitons in single quantum dots by high-resolution photocurrent spectroscopy

Identifying defect-related quantum emitters in monolayer WSe$_2$

Large photoluminescence enhancement by an out-of-plane magnetic field in exfoliated WS$_2$ flakes

Low-threshold topological nanolasers based on second-order corner state

Observation of Efimov Universality across a Non-Universal Feshbach Resonance in \textsuperscript{39}K

Towards automated high-throughput screening of C. elegans on agar