Source author record

Weichen Wang

Weichen Wang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

16works
15topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

16 published item(s)

preprint2026arXiv

A monolithic fabrication platform for intrinsically stretchable polymer transistors and complementary circuits

Soft, stretchable organic field-effect transistors (OFETs) can provide powerful on-skin signal conditioning, but current fabrication methods are often material-specific: each new polymer semiconductor (PSC) requires a tailored process. The challenge is even greater for complementary OFET circuits, where two PSCs must be patterned sequentially, which often leads to device degradation. Here, we introduce a universal, monolithic photolithography process that enables high-yield, high-resolution stretchable complementary OFETs and circuits. This approach is enabled by a process-design framework that includes (i) a direct, photopatternable, solvent-resistant, crosslinked dielectric/semiconductor interface, (ii) broadly applicable crosslinked PSC blends that preserve high mobility, and (iii) a patterning strategy that provides simultaneous etch masking and encapsulation. Using this platform, we achieve record integration density for stretchable OTFTs (55,000 cm^-2), channel lengths down to 2 um, and low-voltage operation at 5 V. We demonstrate photopatterning across multiple PSC types and realize complementary circuits, including 3 kHz stretchable ring oscillators, the first to exceed 1 kHz and representing more than a 60-fold increase in stage switching speed over the state of the art. Finally, we demonstrate the first stretchable complementary OTFT neuron circuit, where the output frequency is modulated by the input current to mimic neuronal signal processing. This scalable approach can be readily extended to diverse high-performance stretchable materials, accelerating the development and manufacturing of skin-like electronics.

preprint2023arXiv

Inference on Time Series Nonparametric Conditional Moment Restrictions Using General Sieves

General nonlinear sieve learnings are classes of nonlinear sieves that can approximate nonlinear functions of high dimensional variables much more flexibly than various linear sieves (or series). This paper considers general nonlinear sieve quasi-likelihood ratio (GN-QLR) based inference on expectation functionals of time series data, where the functionals of interest are based on some nonparametric function that satisfy conditional moment restrictions and are learned using multilayer neural networks. While the asymptotic normality of the estimated functionals depends on some unknown Riesz representer of the functional space, we show that the optimally weighted GN-QLR statistic is asymptotically Chi-square distributed, regardless whether the expectation functional is regular (root-$n$ estimable) or not. This holds when the data are weakly dependent beta-mixing condition. We apply our method to the off-policy evaluation in reinforcement learning, by formulating the Bellman equation into the conditional moment restriction framework, so that we can make inference about the state-specific value functional using the proposed GN-QLR method with time series data. In addition, estimating the averaged partial means and averaged partial derivatives of nonparametric instrumental variables and quantile IV models are also presented as leading examples. Finally, a Monte Carlo study shows the finite sample performance of the procedure

preprint2023arXiv

Ranking Inferences Based on the Top Choice of Multiway Comparisons

This paper considers ranking inference of $n$ items based on the observed data on the top choice among $M$ randomly selected items at each trial. This is a useful modification of the Plackett-Luce model for $M$-way ranking with only the top choice observed and is an extension of the celebrated Bradley-Terry-Luce model that corresponds to $M=2$. Under a uniform sampling scheme in which any $M$ distinguished items are selected for comparisons with probability $p$ and the selected $M$ items are compared $L$ times with multinomial outcomes, we establish the statistical rates of convergence for underlying $n$ preference scores using both $\ell_2$-norm and $\ell_\infty$-norm, with the minimum sampling complexity. In addition, we establish the asymptotic normality of the maximum likelihood estimator that allows us to construct confidence intervals for the underlying scores. Furthermore, we propose a novel inference framework for ranking items through a sophisticated maximum pairwise difference statistic whose distribution is estimated via a valid Gaussian multiplier bootstrap. The estimated distribution is then used to construct simultaneous confidence intervals for the differences in the preference scores and the ranks of individual items. They also enable us to address various inference questions on the ranks of these items. Extensive simulation studies lend further support to our theoretical results. A real data application illustrates the usefulness of the proposed methods convincingly.

preprint2022arXiv

Some statistics on generalized Motzkin paths with vertical steps

Recently, several authors have considered lattice paths with various steps, including vertical steps permitted. In this paper, we consider a kind of generalized Motzkin paths, called {\it G-Motzkin paths} for short, that is lattice paths from $(0, 0)$ to $(n, 0)$ in the first quadrant of the $XOY$-plane that consist of up steps $\mathbf{u}=(1, 1)$, down steps $\mathbf{d}=(1, -1)$, horizontal steps $\mathbf{h}=(1, 0)$ and vertical steps $\mathbf{v}=(0, -1)$. We mainly count the number of G-Motzkin paths of length $n$ with given number of $\mathbf{z}$-steps for $\mathbf{z}\in \{\mathbf{u}, \mathbf{h}, \mathbf{v}, \mathbf{d}\}$, and enumerate the statistics "number of $\mathbf{z}$-steps" at given level in G-Motzkin paths for $\mathbf{z}\in \{\mathbf{u}, \mathbf{h}, \mathbf{v}, \mathbf{d}\}$, some explicit formulas and combinatorial identities are given by bijective and algebraic methods, some enumerative results are linked with Riordan arrays according to the structure decompositions of G-Motzkin paths. We also discuss the statistics "number of $\mathbf{z}_1\mathbf{z}_2$-steps" in G-Motzkin paths for $\mathbf{z}_1, \mathbf{z}_2\in \{\mathbf{u}, \mathbf{h}, \mathbf{v}, \mathbf{d}\}$, the exact counting formulas except for $\mathbf{z}_1\mathbf{z}_2=\mathbf{dd}$ are obtained by the Lagrange inversion formula and their generating functions.

preprint2022arXiv

The $\mathbf{uvu}$-avoiding $(a,b,c)$-Generalized Motzkin paths with vertical steps: bijections and statistic enumerations

A generalized Motzkin path, called G-Motzkin path for short, of length $n$ is a lattice path from $(0, 0)$ to $(n, 0)$ in the first quadrant of the XOY-plane that consists of up steps $\mathbf{u}=(1, 1)$, down steps $\mathbf{d}=(1, -1)$, horizontal steps $\mathbf{h}=(1, 0)$ and vertical steps $\mathbf{v}=(0, -1)$. An $(a,b,c)$-G-Motzkin path is a weighted G-Motzkin path such that the $\mathbf{u}$-steps, $\mathbf{h}$-steps, $\mathbf{v}$-steps and $\mathbf{d}$-steps are weighted respectively by $1, a, b$ and $c$. In this paper, we first give bijections between the set of $\mathbf{uvu}$-avoiding $(a,b,b^2)$-G-Motzkin paths of length $n$ and the set of $(a,b)$-Schröder paths as well as the set of $(a+b,b)$-Dyck paths of length $2n$, between the set of $\{\mathbf{uvu, uu}\}$-avoiding $(a,b,b^2)$-G-Motzkin paths of length $n$ and the set of $(a+b,ab)$-Motzkin paths of length $n$, between the set of $\{\mathbf{uvu,uu}\}$-avoiding $(a,b,b^2)$-G-Motzkin paths of length $n+1$ beginning with an $\mathbf{h}$-step weighted by $a$ and the set of $(a,b)$-Dyck paths of length $2n+2$. In the last section, we focus on the enumeration of statistics "number of $\mathbf{z}$-steps" for $\mathbf{z}\in \{\mathbf{u}, \mathbf{h}, \mathbf{v}, \mathbf{d}\}$ and "number of points" at given level in $\mathbf{uvu}$-avoiding G-Motzkin paths. These counting results are linked with Riordan arrays.

preprint2022arXiv

The Baltimore Oriole's Nest: Cool Winds from the Inner and Outer Parts of a Star-Forming Galaxy at z=1.3

Strong galactic winds are ubiquitous at $z\gtrsim 1$. However, it is not well known where inside galaxies these winds are launched from. We study the cool winds ($\sim 10^4$\,K) in two spatial regions of a massive galaxy at $z=1.3$, which we nickname the "Baltimore Oriole's Nest." The galaxy has a stellar mass of $10^{10.3\pm 0.3} M_\odot$, is located on the star-forming main sequence, and has a morphology indicative of a recent merger. Gas kinematics indicate a dynamically complex system with velocity gradients ranging from 0 to 60 $\mathrm{km}\cdot\mathrm{s}^{-1}$. The two regions studied are: a dust-reddened center (Central region), and a blue arc at 7 kpc from the center (Arc region). We measure the \ion{Fe}{2} and \ion{Mg}{2} absorption line profiles from deep Keck/DEIMOS spectra. Blueshifted wings up to 450 km$\cdot$s$^{-1}$ are found for both regions. The \ion{Fe}{2} column densities of winds are $10^{14.7\pm 0.2}\,\mathrm{cm}^{-2}$ and $10^{14.6\pm 0.2}\,\mathrm{cm}^{-2}$ toward the Central and Arc regions, respectively. Our measurements suggest that the winds are most likely launched from both regions. The winds may be driven by the spatially extended star formation, the surface density of which is around 0.2 $M_\odot\,\mathrm{yr}^{-1}\cdot \mathrm{kpc}^{-2}$ in both regions. The mass outflow rates are estimated to be $4\,M_\odot\,\mathrm{yr}^{-1}$ and $3\,M_\odot\,\mathrm{yr}^{-1}$ for the Central and Arc regions, with uncertainties of one order-of-magnitude or more. Findings of this work and a few previous studies suggest that the cool galactic winds at $z\gtrsim 1$ might be commonly launched from the entire spatial extents of their host galaxies due to extended galaxy star formation.

preprint2022arXiv

The Dwarf Galaxy Population at $z\sim 0.7$: A Catalog of Emission Lines and Redshifts from Deep Keck Observations

We present a catalog of spectroscopically measured redshifts over $0 < z < 2$ and emission line fluxes for 1440 galaxies. The majority ($\sim$65\%) of the galaxies come from the HALO7D survey, with the remainder from the DEEPwinds program. This catalog includes redshifts for 646 dwarf galaxies with $\log(M_{\star}/M_{\odot}) < 9.5$. 810 catalog galaxies did not have previously published spectroscopic redshifts, including 454 dwarf galaxies. HALO7D used the DEIMOS spectrograph on the Keck II telescope to take very deep (up to 32 hours exposure, with a median of $\sim$7 hours) optical spectroscopy in the COSMOS, EGS, GOODS-North, and GOODS-South CANDELS fields, and in some areas outside CANDELS. We compare our redshift results to existing spectroscopic and photometric redshifts in these fields, finding only a 1\% rate of discrepancy with other spectroscopic redshifts. We measure a small increase in median photometric redshift error (from 1.0\% to 1.3\%) and catastrophic outlier rate (from 3.5\% to 8\%) with decreasing stellar mass. We obtained successful redshift fits for 75\% of massive galaxies, and demonstrate a similar 70-75\% successful redshift measurement rate in $8.5 < \log(M_{\star}/M_{\odot}) < 9.5$ galaxies, suggesting similar survey sensitivity in this low-mass range. We describe the redshift, mass, and color-magnitude distributions of the catalog galaxies, finding HALO7D galaxies representative of CANDELS galaxies up to \textit{i}-band magnitudes of 25. The catalogs presented will enable studies of star formation (SF), the mass-metallicity relation, SF-morphology relations, and other properties of the $z\sim0.7$ dwarf galaxy population.

preprint2020arXiv

Estimation of the Laser Frequency Nosie Spectrum by Continuous Dynamical Decoupling

Decoherence induced by the laser frequency noise is one of the most important obstacles in the quantum information processing. In order to suppress this decoherence, the noise power spectral density needs to be accurately characterized. In particular, the noise spectrum measurement based on the coherence characteristics of qubits would be a meaningful and still challenging method. Here, we theoretically analyze and experimentally obtain the spectrum of laser frequency noise based on the continuous dynamical decoupling technique. We first estimate the mixture-noise (including laser and magnetic noises) spectrum up to $(2π)$530 kHz by monitoring the transverse relaxation from an initial state $+X$, followed by a gradient descent data process protocol. Then the contribution from the laser noise is extracted by enconding the qubits on different Zeeman sublevels. We also investigate two sufficiently strong noise components by making an analogy between these noises and driving lasers whose linewidth assumed to be negligible. This method is verified experimentally and finally helps to characterize the noise.

preprint2020arXiv

Global Convergence of Policy Gradient for Linear-Quadratic Mean-Field Control/Game in Continuous Time

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem. Therefore, it has motivated new research directions for mean-field control (MFC) and mean-field game (MFG). In this paper, we study the policy gradient method for the linear-quadratic mean-field control and game, where we assume each agent has identical linear state transitions and quadratic cost functions. While most of the recent works on policy gradient for MFC and MFG are based on discrete-time models, we focus on the continuous-time models where some analyzing techniques can be interesting to the readers. For both MFC and MFG, we provide policy gradient update and show that it converges to the optimal solution at a linear rate, which is verified by a synthetic simulation. For MFG, we also provide sufficient conditions for the existence and uniqueness of the Nash equilibrium.

preprint2016arXiv

Heterogeneity Adjustment with Applications to Graphical Model Inference

Heterogeneity is an unwanted variation when analyzing aggregated datasets from multiple sources. Though different methods have been proposed for heterogeneity adjustment, no systematic theory exists to justify these methods. In this work, we propose a generic framework named ALPHA (short for Adaptive Low-rank Principal Heterogeneity Adjustment) to model, estimate, and adjust heterogeneity from the original data. Once the heterogeneity is adjusted, we are able to remove the biases of batch effects and to enhance the inferential power by aggregating the homogeneous residuals from multiple sources. Under a pervasive assumption that the latent heterogeneity factors simultaneously affect a large fraction of observed variables, we provide a rigorous theory to justify the proposed framework. Our framework also allows the incorporation of informative covariates and appeals to the "Bless of Dimensionality". As an illustrative application of this generic framework, we consider a problem of estimating high-dimensional precision matrix for graphical model inference based on multiple datasets. We also provide thorough numerical studies on both synthetic datasets and a brain imaging dataset to demonstrate the efficacy of the developed theory and methods.

preprint2016arXiv

Projected principal component analysis in factor models

This paper introduces a Projected Principal Component Analysis (Projected-PCA), which employs principal component analysis to the projected (smoothed) data matrix onto a given linear space spanned by covariates. When it applies to high-dimensional factor analysis, the projection removes noise components. We show that the unobserved latent factors can be more accurately estimated than the conventional PCA if the projection is genuine, or more precisely, when the factor loading matrices are related to the projected linear space. When the dimensionality is large, the factors can be estimated accurately even when the sample size is finite. We propose a flexible semiparametric factor model, which decomposes the factor loading matrix into the component that can be explained by subject-specific covariates and the orthogonal residual component. The covariates' effects on the factor loadings are further modeled by the additive model via sieve approximations. By using the newly proposed Projected-PCA, the rates of convergence of the smooth factor loading matrices are obtained, which are much faster than those of the conventional factor analysis. The convergence is achieved even when the sample size is finite and is particularly appealing in the high-dimension-low-sample-size situation. This leads us to developing nonparametric tests on whether observed covariates have explaining powers on the loadings and whether they fully explain the loadings. The proposed method is illustrated by both simulated data and the returns of the components of the S&P 500 index.

preprint2016arXiv

Robust Covariance Estimation for Approximate Factor Models

In this paper, we study robust covariance estimation under the approximate factor model with observed factors. We propose a novel framework to first estimate the initial joint covariance matrix of the observed data and the factors, and then use it to recover the covariance matrix of the observed data. We prove that once the initial matrix estimator is good enough to maintain the element-wise optimal rate, the whole procedure will generate an estimated covariance with desired properties. For data with only bounded fourth moments, we propose to use Huber loss minimization to give the initial joint covariance estimation. This approach is applicable to a much wider range of distributions, including sub-Gaussian and elliptical distributions. We also present an asymptotic result for Huber's M-estimator with a diverging parameter. The conclusions are demonstrated by extensive simulations and real data analysis.

preprint2016arXiv

The UV-optical Color Gradients in Star-Forming Galaxies at 0.5<z<1.5: Origins and Link to Galaxy Assembly

The rest-frame UV-optical (i.e., NUV-B) color index is sensitive to the low-level recent star formation and dust extinction, but it is insensitive to the metallicity. In this Letter, we have measured the rest-frame NUV-B color gradients in ~1400 large ($\rm r_e>0.18^{\prime\prime}$), nearly face-on (b/a>0.5) main-sequence star-forming galaxies (SFGs) between redshift 0.5 and 1.5 in the CANDELS/GOODS-S and UDS fields. With this sample, we study the origin of UV-optical color gradients in the SFGs at z~1 and discuss their link with the buildup of stellar mass. We find that the more massive, centrally compact, and more dust extinguished SFGs tend to have statistically more negative raw color gradients (redder centers) than the less massive, centrally diffuse, and less dusty SFGs. After correcting for dust reddening based on optical-SED fitting, the color gradients in the low-mass ($M_{\ast} <10^{10}M_{\odot}$) SFGs generally become quite flat, while most of the high-mass ($M_{\ast} > 10^{10.5}M_{\odot}$) SFGs still retain shallow negative color gradients. These findings imply that dust reddening is likely the principal cause of negative color gradients in the low-mass SFGs, while both increased central dust reddening and buildup of compact old bulges are likely the origins of negative color gradients in the high-mass SFGs. These findings also imply that at these redshifts the low-mass SFGs buildup their stellar masses in a self-similar way, while the high-mass SFGs grow inside out.

preprint2015arXiv

Asymptotics of Empirical Eigen-structure for Ultra-high Dimensional Spiked Covariance Model

We derive the asymptotic distributions of the spiked eigenvalues and eigenvectors under a generalized and unified asymptotic regime, which takes into account the spike magnitude of leading eigenvalues, sample size, and dimensionality. This new regime allows high dimensionality and diverging eigenvalue spikes and provides new insights into the roles the leading eigenvalues, sample size, and dimensionality play in principal component analysis. The results are proven by a technical device, which swaps the role of rows and columns and converts the high-dimensional problems into low-dimensional ones. Our results are a natural extension of those in Paul (2007) to more general setting with new insights and solve the rates of convergence problems in Shen et al. (2013). They also reveal the biases of the estimation of leading eigenvalues and eigenvectors by using principal component analysis, and lead to a new covariance estimator for the approximate factor model, called shrinkage principal orthogonal complement thresholding (S-POET), that corrects the biases. Our results are successfully applied to outstanding problems in estimation of risks of large portfolios and false discovery proportions for dependent test statistics and are illustrated by simulation studies.

preprint2015arXiv

Estimation of functionals of sparse covariance matrices

High-dimensional statistical tests often ignore correlations to gain simplicity and stability leading to null distributions that depend on functionals of correlation matrices such as their Frobenius norm and other $\ell_r$ norms. Motivated by the computation of critical values of such tests, we investigate the difficulty of estimation the functionals of sparse correlation matrices. Specifically, we show that simple plug-in procedures based on thresholded estimators of correlation matrices are sparsity-adaptive and minimax optimal over a large class of correlation matrices. Akin to previous results on functional estimation, the minimax rates exhibit an elbow phenomenon. Our results are further illustrated in simulated data as well as an empirical study of data arising in financial econometrics.

preprint2015arXiv

Large Covariance Estimation through Elliptical Factor Models

We proposed a general Principal Orthogonal complEment Thresholding (POET) framework for large-scale covariance matrix estimation based on an approximate factor model. A set of high level sufficient conditions for the procedure to achieve optimal rates of convergence under different matrix norms were brought up to better understand how POET works. Such a framework allows us to recover the results for sub-Gaussian in a more transparent way that only depends on the concentration properties of the sample covariance matrix. As a new theoretical contribution, for the first time, such a framework allows us to exploit conditional sparsity covariance structure for the heavy-tailed data. In particular, for the elliptical data, we proposed a robust estimator based on marginal and multivariate Kendall's tau to satisfy these conditions. In addition, conditional graphical model was also studied under the same framework. The technical tools developed in this paper are of general interest to high dimensional principal component analysis. Thorough numerical results were also provided to back up the developed theory.