Source author record

Rui Wu

Rui Wu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

25works
19topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

25 published item(s)

preprint2026arXiv

Chinese Labor Law Large Language Model Benchmark

Recent advances in large language models (LLMs) have led to substantial progress in domain-specific applications, particularly within the legal domain. However, general-purpose models such as GPT-4 often struggle with specialized subdomains that require precise legal knowledge, complex reasoning, and contextual sensitivity. To address these limitations, we present LabourLawLLM, a legal large language model tailored to Chinese labor law. We also introduce LabourLawBench, a comprehensive benchmark covering diverse labor-law tasks, including legal provision citation, knowledge-based question answering, case classification, compensation computation, named entity recognition, and legal case analysis. Our evaluation framework combines objective metrics (e.g., ROUGE-L, accuracy, F1, and soft-F1) with subjective assessment based on GPT-4 scoring. Experiments show that LabourLawLLM consistently outperforms general-purpose and existing legal-specific LLMs across task categories. Beyond labor law, our methodology provides a scalable approach for building specialized LLMs in other legal subfields, improving accuracy, reliability, and societal value of legal AI applications.

preprint2026arXiv

Cryogenic interface-state filling and tunneling mechanisms in strained Ge/SiGe heterostructures

Traps at the semiconductor-oxide interface are considered as a major source of instability in strained Ge/SiGe quantum devices, yet the quantified study of their cryogenic behavior remains limited. In this work, we investigate interface-state trapping using Hall-bar field-effect transistors fabricated on strained Ge/SiGe heterostructures. Combining transport measurements with long-term stabilization and Schrödinger-Poisson modelling, we reconstruct the gradual filling process of interface states at cryogenic condition. Using the calculated valence band profiles, we further evaluate the tunneling current density between the quantum well and the semiconductor-oxide interface. Our calculation demonstrates that the total tunneling current is consistent with a crossover from trap-assisted-tunneling-dominated transport to Fowler-Nordheim-tunneling-dominated transport under different gate bias regimes. These results refine the conventional Fowler-Nordheim-based picture of interface trapping in strained Ge/SiGe heterostructures and provide guidelines for improving Ge-based quantum device performance by improving barrier crystalline qualities and reducing dislocation-related trap densities.

preprint2026arXiv

Reasoning over Precedents Alongside Statutes: Case-Augmented Deliberative Alignment for LLM Safety

Ensuring that Large Language Models (LLMs) adhere to safety principles without refusing benign requests remains a significant challenge. While OpenAI introduces deliberative alignment (DA) to enhance the safety of its o-series models through reasoning over detailed ``code-like'' safety rules, the effectiveness of this approach in open-source LLMs, which typically lack advanced reasoning capabilities, is understudied. In this work, we systematically evaluate the impact of explicitly specifying extensive safety codes versus demonstrating them through illustrative cases. We find that referencing explicit codes inconsistently improves harmlessness and systematically degrades helpfulness, whereas training on case-augmented simple codes yields more robust and generalized safety behaviors. By guiding LLMs with case-augmented reasoning instead of extensive code-like safety rules, we avoid rigid adherence to narrowly enumerated rules and enable broader adaptability. Building on these insights, we propose CADA, a case-augmented deliberative alignment method for LLMs utilizing reinforcement learning on self-generated safety reasoning chains. CADA effectively enhances harmlessness, improves robustness against attacks, and reduces over-refusal while preserving utility across diverse benchmarks, offering a practical alternative to rule-only DA for improving safety while maintaining helpfulness.

preprint2023arXiv

Continuous odor profile monitoring to study olfactory navigation in small animals

Olfactory navigation is observed across species and plays a crucial role in locating resources for survival. In the laboratory, understanding the behavioral strategies and neural circuits underlying odor-taxis requires a detailed understanding of the animal's sensory environment. For small model organisms like C. elegans and larval D. melanogaster, controlling and measuring the odor environment experienced by the animal can be challenging, especially for airborne odors, which are subject to subtle effects from airflow, temperature variation, and from the odor's adhesion, adsorption or reemission. Here we present a method to flexibly control and precisely measure airborne odor concentration in an arena with agar while imaging animal behavior. Crucially and unlike previous methods, our method allows continuous monitoring of the odor profile during behavior. We construct stationary chemical landscapes in an odor flow chamber through spatially patterned odorized air. The odor concentration is measured with a spatially distributed array of digital gas sensors. Careful placement of the sensors allows the odor concentration across the arena to be accurately inferred and continuously monitored at all points in time. We use this approach to measure the precise odor concentration that each animal experiences as it undergoes chemotaxis behavior and report chemotaxis strategies for C. elegans and D. melanogaster larvae populations under different spatial odor landscapes.

preprint2022arXiv

Commentary: analyzing binary data using MCPMod when zero counts are expected

Bretz et al (2005) proposed multiple Comparison Procedure and Modeling (MCPMod) method to design and analyze dose-finding study. Pinheiro (2014) then generalized it to various types of endpoint, including but not limited to binary endpoint, survival endpoint, count data, and longitudinal data. Pinheiro (2013) recommended to use the estimated covariance matrix from the observed data to recalculate the optimal contrast and the critical value of the test For many phase II studies it is common to have small sample sizes per arm with low placebo response rates jointly. Under such circumstances, it cannot be excluded to have a zero count observed. For example, when the placebo response rate is 10%, there is about 4% chance to observe zero responders in the placebo group, or other dose group(s), which has a similar response rate as placebo. In this manuscript, we would like to illustrate the potential problem of Pinheiro (2013) using a case study and simulations. An alternative method using Firth's logistic regression was evaluated to get a stable estimate of response for each dose group. In addition, we evaluated two options to address the issue with problematic contrast coefficients.

preprint2022arXiv

Conditional Generation Net for Medication Recommendation

Medication recommendation targets to provide a proper set of medicines according to patients' diagnoses, which is a critical task in clinics. Currently, the recommendation is manually conducted by doctors. However, for complicated cases, like patients with multiple diseases at the same time, it's difficult to propose a considerate recommendation even for experienced doctors. This urges the emergence of automatic medication recommendation which can help treat the diagnosed diseases without causing harmful drug-drug interactions.Due to the clinical value, medication recommendation has attracted growing research interests.Existing works mainly formulate medication recommendation as a multi-label classification task to predict the set of medicines. In this paper, we propose the Conditional Generation Net (COGNet) which introduces a novel copy-or-predict mechanism to generate the set of medicines. Given a patient, the proposed model first retrieves his or her historical diagnoses and medication recommendations and mines their relationship with current diagnoses. Then in predicting each medicine, the proposed model decides whether to copy a medicine from previous recommendations or to predict a new one. This process is quite similar to the decision process of human doctors. We validate the proposed model on the public MIMIC data set, and the experimental results show that the proposed model can outperform state-of-the-art approaches.

preprint2022arXiv

Many-body slow quench dynamics and nonadiabatic characterization of topological phases

Previous studies have shown that the bulk topology of single-particle systems can be captured by the band inversion surface or by the spin inversion surface emerged on the time-averaged spin polarization. Most of the studies, however, are based on the single-particle picture even though the systems are fermionic and of multi-bands. Here, we study the many-body quench dynamics of topological systems with all the valence bands fully occupied, and show that the concepts of band inversion surface and spin inversion surface are still valid. More importantly, the many-body quench dynamics is shown to be reduced to a nontrivial three-level Landau-Zener model, which can be solved exactly. Based on the analytical results, the topological spin texture revealed by the time-averaged spin polarization can be applied to characterize the bulk topology and thus provides a direct comparison for future experiments.

preprint2020arXiv

Data Efficient Training for Reinforcement Learning with Adaptive Behavior Policy Sharing

Deep Reinforcement Learning (RL) is proven powerful for decision making in simulated environments. However, training deep RL model is challenging in real world applications such as production-scale health-care or recommender systems because of the expensiveness of interaction and limitation of budget at deployment. One aspect of the data inefficiency comes from the expensive hyper-parameter tuning when optimizing deep neural networks. We propose Adaptive Behavior Policy Sharing (ABPS), a data-efficient training algorithm that allows sharing of experience collected by behavior policy that is adaptively selected from a pool of agents trained with an ensemble of hyper-parameters. We further extend ABPS to evolve hyper-parameters during training by hybridizing ABPS with an adapted version of Population Based Training (ABPS-PBT). We conduct experiments with multiple Atari games with up to 16 hyper-parameter/architecture setups. ABPS achieves superior overall performance, reduced variance on top 25% agents, and equivalent performance on the best agent compared to conventional hyper-parameter tuning with independent training, even though ABPS only requires the same number of environmental interactions as training a single agent. We also show that ABPS-PBT further improves the convergence speed and reduces the variance.

preprint2020arXiv

Drying of porous media by concurrent drainage and evaporation: A pore network modeling study

Drainage and evaporation can occur simultaneously during the drying of porous media, but the interactions between these processes and their effects on drying are rarely studied. In this work, we develop a pore network model that considers drainage, evaporation, and rarefied multi-component gas transport in porous media with nanoscale pores. Using this model, we investigate the drying of a liquid solvent-saturated porous medium enabled by the flow of purge gas through it. Simulations show that drying progresses in three stages, and the solvent removal by drainage effects (evaporation effects) becomes increasingly weak (strong) as drying progresses through these stages. Interestingly, drainage can contribute considerably to solvent removal even after evaporation effects become very strong, especially when the applied pressure difference across the porous medium is low. We show that these phenomena are the results of the coupling between the drainage and evaporation effects and this coupling depends on the operating conditions and the stage of drying.

preprint2020arXiv

Experimental observations indicating the topological nature of the edge states on HfTe5

The topological edge states of two-dimensional topological insulators with large energy gap furnish ideal conduction channels for dissipationless current transport. Transition metal tellurides XTe5 (X=Zr, Hf) are theoretically predicted to be large-gap two-dimensional topological insulators and the experimental observations of their bulk insulating gap and in-gap edge states have been reported, but the topological nature of these edge states still remains to be further elucidated. Here, we report our low temperature scanning tunneling microscopy/spectroscopy study on single crystals of HfTe5. We demonstrate a full energy gap of ~80 meV near the Fermi level on the surface monolayer of HfTe5 and that such insulating energy gap gets filled with finite energy states when measured at the monolayer step edges. Remarkably, such states are absent at the edges of a narrow monolayer strip of one-unit-cell in width but persist at both step edges of a unit-cell wide monolayer groove. These experimental observations strongly indicate that the edge states of HfTe5 monolayers are not trivially caused by translational symmetry breaking, instead they are topological in nature protected by the 2D nontrivial bulk properties.

preprint2020arXiv

Influence of Atomic Roughness at The Uncompensated Fe/CoO (111) Interface on Exchange Bias Effect

The effect of interface roughness of ferromagnetic and antiferromagnetic layers on exchange bias is still not well understood. In this report we have investigated the effect of surface roughness in (111)-oriented antiferromagnetic CoO films on exchange bias with ferromagnetic Fe grown on top. The surface roughness is controlled at the atomic scale, over a range below ~ 0.35 nm, by varying layer thickness of the CoO films. It is observed that both exchange bias field ($H_{E}$) and coercivity ($H_{C}$) extensively depend on the atomic scale roughness of the CoO (111) at the interface with Fe film. An opposite dependence of $H_{E}$ and $H_{C}$ on interface roughness was found, which was ascribed to partially compensated spin states induced by the atomic roughness at the fully uncompensated CoO (111) surfaces and was corroborated using the Monte Carlo simulations. Moreover, the onset temperature for $H_{C}$ is found to be up to ~ 80 K below the blocking temperature ($T_{B}$) and the temperature dependence of $H_{C}$ follows the power law with a critical exponent equal to one, which indicates that, in this system, $H_{C}$ is more of an interface-related property than $H_{E}$.

preprint2020arXiv

MCEN: Bridging Cross-Modal Gap between Cooking Recipes and Dish Images with Latent Variable Model

Nowadays, driven by the increasing concern on diet and health, food computing has attracted enormous attention from both industry and research community. One of the most popular research topics in this domain is Food Retrieval, due to its profound influence on health-oriented applications. In this paper, we focus on the task of cross-modal retrieval between food images and cooking recipes. We present Modality-Consistent Embedding Network (MCEN) that learns modality-invariant representations by projecting images and texts to the same embedding space. To capture the latent alignments between modalities, we incorporate stochastic latent variables to explicitly exploit the interactions between textual and visual features. Importantly, our method learns the cross-modal alignments during training but computes embeddings of different modalities independently at inference time for the sake of efficiency. Extensive experimental results clearly demonstrate that the proposed MCEN outperforms all existing approaches on the benchmark Recipe1M dataset and requires less computational cost.

preprint2019arXiv

Bubble formation due to capillary instability during evaporation of a porous medium

We show that during evaporation of a pore network, liquid can refill the gas occupied pores, snapping off a gas bubble, which then moves to a stable configuration. This phenomenon is induced by the capillary instability due to the wettability heterogeneity of the pore network and has a much smaller time scale as compared to the evaporation process. The capillary instability induced liquid refilling and bubble movement are explained in detail based on the analysis of the images obtained from the visualization experiment. The capillary valve effect, which hinders the movement of the gas-liquid interface and is induced by the sudden geometrical expansion between small and large pores, can be suppressed by the residual liquid in the large pore. For better understanding of the capillary instability induced gas-liquid two-phase transport during evaporation, a novel pore network model is developed, which considers not only the capillary and viscous forces but also the inertial forces that are seldom taken into account in the previous models. The pore network modeling results are in good agreement with the experimental data, demonstrating the effectiveness of the developed pore network model, which opens up a new route for better understanding of the role of inertial forces in two-phase transport in porous media.

preprint2019arXiv

Pore network model of evaporation in porous media with continuous and discontinuous corner films

During evaporation in porous media, two types of corner films are distinguished. A continuous corner film is connected to the bulk liquid, while a discontinuous one is not. To disclose their effects on evaporation in porous media, a pore network model with both continuous and discontinuous corner films is developed, which considers the capillary and viscous forces as well as the effects of corner films on the threshold pressures of pores. The capillary valve effect induced by the sudden geometrical expansion between the small and large pores is also taken into account in the model. The developed pore network model agrees well with the evaporation experiment with a quasi 2D micro model porous medium, in terms of not only the variation of the liquid saturation in each pore but also the variation of the total evaporation rate. The pore network models that neglect the corner films or the liquid viscosity are also compared with the experiment so as to shed light on the roles of the corner films. The continuous corner films, which contribute to sustain the high evaporation rate, can be interrupted to be the discontinuous ones not only by the gas invasion into pores but also by the capillary scissors effect due to the local convex topology of the solid matrix.

preprint2016arXiv

Cooper Pairing and Phase Coherence in Iron Superconductor Fe1+x(Te,Se)

The Cooper pairing and phase coherence are two fundamental aspects of superconductivity. Due to breaking time reversal symmetry, magnetic impurities are detrimental to superconductivity, yet microscopically how they affect the pairing strength and phase coherence in a real material is less understood. Recently we observed a robust zero-energy bound state at an interstitial Fe impurity (IFI) in superconducting Fe1+x(Te,Se), signifying intense impurity scattering. Here we report a comprehensive study, using scanning tunnelling microscopy/spectroscopy (STM/S) technique, of the global effects of IFIs on the ground state of Fe1+x(Te,Se) over a wide range of IFI concentration x. Our high resolution tunnelling spectroscopy and quasi-particle interference data at very low temperature demonstrate that IFIs hardly affect the electron pairing strength, while they cause significant decoherence of Cooper pairs in precedence of the Coulomb correlation, eventually driving the ground state of the system from strong-coupling-superconductor to diffusive-metal with incoherent electron pairs.

preprint2015arXiv

Clustering and Inference From Pairwise Comparisons

Given a set of pairwise comparisons, the classical ranking problem computes a single ranking that best represents the preferences of all users. In this paper, we study the problem of inferring individual preferences, arising in the context of making personalized recommendations. In particular, we assume that there are $n$ users of $r$ types; users of the same type provide similar pairwise comparisons for $m$ items according to the Bradley-Terry model. We propose an efficient algorithm that accurately estimates the individual preferences for almost all users, if there are $r \max \{m, n\}\log m \log^2 n$ pairwise comparisons per type, which is near optimal in sample complexity when $r$ only grows logarithmically with $m$ or $n$. Our algorithm has three steps: first, for each user, compute the \emph{net-win} vector which is a projection of its $\binom{m}{2}$-dimensional vector of pairwise comparisons onto an $m$-dimensional linear subspace; second, cluster the users based on the net-win vectors; third, estimate a single preference for each cluster separately. The net-win vectors are much less noisy than the high dimensional vectors of pairwise comparisons and clustering is more accurate after the projection as confirmed by numerical experiments. Moreover, we show that, when a cluster is only approximately correct, the maximum likelihood estimation for the Bradley-Terry model is still close to the true preference.

preprint2015arXiv

In situ resonant photoemission and X-ray absorption study of the BiFeO3 thin film

Multiferroic bismuth ferrite (BiFeO3) thin films were prepared by pulsed laser deposition (PLD) technique. Electronic structures of the film have been studied by in situ photoemission spectroscopy (PES) and x-ray absorption spectroscopy (XAS). Both the Fe 2p PES and XAS spectra show that Fe ion is formally in +3 valence state. The Fe 2p and O K edge XAS spectra indicate that the oxygen octahedral crystal ligand field splits the unoccupied Fe 3d state to t2g and eg states. Valence band Fe 2p-3d resonant photoemission results indicate that hybridization between Fe 3d and O 2p plays important role in the multiferroic BiFeO3 thin films.

preprint2014arXiv

Collaborative Filtering with Information-Rich and Information-Sparse Entities

In this paper, we consider a popular model for collaborative filtering in recommender systems where some users of a website rate some items, such as movies, and the goal is to recover the ratings of some or all of the unrated items of each user. In particular, we consider both the clustering model, where only users (or items) are clustered, and the co-clustering model, where both users and items are clustered, and further, we assume that some users rate many items (information-rich users) and some users rate only a few items (information-sparse users). When users (or items) are clustered, our algorithm can recover the rating matrix with $ω(MK \log M)$ noisy entries while $MK$ entries are necessary, where $K$ is the number of clusters and $M$ is the number of items. In the case of co-clustering, we prove that $K^2$ entries are necessary for recovering the rating matrix, and our algorithm achieves this lower bound within a logarithmic factor when $K$ is sufficiently large. We compare our algorithms with a well-known algorithms called alternating minimization (AM), and a similarity score-based algorithm known as the popularity-among-friends (PAF) algorithm by applying all three to the MovieLens and Netflix data sets. Our co-clustering algorithm and AM have similar overall error rates when recovering the rating matrix, both of which are lower than the error rate under PAF. But more importantly, the error rate of our co-clustering algorithm is significantly lower than AM and PAF in the scenarios of interest in recommender systems: when recommending a few items to each user or when recommending items to users who only rated a few items (these users are the majority of the total user population). The performance difference increases even more when noise is added to the datasets.

preprint2014arXiv

Interference Channels with Half-Duplex Source Cooperation

The performance gain by allowing half-duplex source cooperation is studied for Gaussian interference channels. The source cooperation is {\em in-band}, meaning that each source can listen to the other source's transmission, but there is no independent (or orthogonal) channel between the sources. The half-duplex constraint supposes that at each time instant the sources can either transmit or listen, but not do both. Our main result is a characterization of the sum capacity when the cooperation is bidirectional and the channel gains are symmetric. With unidirectional cooperation, we essentially have a cognitive radio channel. By requiring the primary to achieve a rate close to its link capacity, the best possible rate for the secondary is characterized within a constant. Novel inner and outer bounds are derived as part of these characterizations.

preprint2014arXiv

Jointly Clustering Rows and Columns of Binary Matrices: Algorithms and Trade-offs

In standard clustering problems, data points are represented by vectors, and by stacking them together, one forms a data matrix with row or column cluster structure. In this paper, we consider a class of binary matrices, arising in many applications, which exhibit both row and column cluster structure, and our goal is to exactly recover the underlying row and column clusters by observing only a small fraction of noisy entries. We first derive a lower bound on the minimum number of observations needed for exact cluster recovery. Then, we propose three algorithms with different running time and compare the number of observations needed by them for successful cluster recovery. Our analytical results show smooth time-data trade-offs: one can gradually reduce the computational complexity when increasingly more observations are available.

preprint2014arXiv

Learning Loosely Connected Markov Random Fields

We consider the structure learning problem for graphical models that we call loosely connected Markov random fields, in which the number of short paths between any pair of nodes is small, and present a new conditional independence test based algorithm for learning the underlying graph structure. The novel maximization step in our algorithm ensures that the true edges are detected correctly even when there are short cycles in the graph. The number of samples required by our algorithm is C*log p, where p is the size of the graph and the constant C depends on the parameters of the model. We show that several previously studied models are examples of loosely connected Markov random fields, and our algorithm achieves the same or lower computational complexity than the previously designed algorithms for individual cases. We also get new results for more general graphical models, in particular, our algorithm learns general Ising models on the Erdos-Renyi random graph G(p, c/p) correctly with running time O(np^5).

preprint2014arXiv

Rapidly reconfigurable radio-frequency arbitrary waveforms synthesized on a CMOS photonic chip

Photonic methods of radio-frequency waveform generation and processing provide performance and flexibility over electronic methods due to the ultrawide bandwidth offered by the optical carriers. However, they suffer from lack of integration and slow reconfiguration speed. Here we propose an architecture of integrated photonic RF waveform generation and processing, and implement it on a silicon chip fabricated in a semiconductor manufacturing foundry. Our device can generate programmable RF bursts or continuous waveforms with only the light source, electrical drives/controls and detectors being off chip. It turns on and off an individual pulse in the RF burst within 4 nanoseconds, achieving a reconfiguration speed three orders of magnitude faster than thermal tuning. The on-chip optical delay elements offers an integrated approach to accurately manipulate individual RF waveform features without constrains set by the speed and timing jitter of electronics, and should find broad applications ranging from high-speed wireless to defense electronics.

preprint2012arXiv

Comb-Based Radio-Frequency Photonic Filters with Rapid Tunability and High Selectivity

Photonic technologies have received considerable attention for enhancement of radio-frequency (RF) electrical systems, including high-frequency analog signal transmission, control of phased arrays, analog-to-digital conversion, and signal processing. Although the potential of radio-frequency photonics for implementation of tunable electrical filters over broad RF bandwidths has been much discussed, realization of programmable filters with highly selective filter lineshapes and rapid reconfigurability has faced significant challenges. A new approach for RF photonic filters based on frequency combs offers a potential route to simultaneous high stopband attenuation, fast tunability, and bandwidth reconfiguration. In one configuration tuning of the RF passband frequency is demonstrated with unprecedented (~40 ns) speed by controlling the optical delay between combs. In a second, fixed filter configuration, cascaded four-wave mixing simultaneously broadens and smoothes comb spectra, resulting in Gaussian RF filter lineshapes exhibiting extremely high (>60 dB) main lobe to sidelobe suppression ratio and (>70 dB) stopband attenuation.

preprint2010arXiv

Generation of very flat optical frequency combs from continuous-wave lasers using cascaded intensity and phase modulators driven by tailored radio frequency waveforms

We demonstrate a scheme, based on a cascade of lithium niobate intensity and phase modulators driven by specially tailored radio frequency waveforms to generate an optical frequency comb with very high spectral flatness. In this work we demonstrate a 10 GHz comb with ~40 lines with spectral power variation below 1-dB and ~60 lines in total. The number of lines that can be generated is limited by the power handling capability of the phase modulator, and this can be scaled without compromising the spectral flatness. Furthermore, the spectral phase of the generated combs in our scheme is almost purely quadratic which, as we will demonstrate, allows for very high quality pulse compression using only single mode fiber.