Source author record

Jiahui Zhu

Jiahui Zhu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

7works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

7 published item(s)

preprint2026arXiv

DGAE: Diffusion-Guided Autoencoder for Efficient Latent Representation Learning

Autoencoders empower state-of-the-art image and video generative models by compressing pixels into a latent space through visual tokenization. Although recent advances have alleviated the performance degradation of autoencoders under high compression ratios, addressing the training instability caused by GAN remains an open challenge. While improving spatial compression, we also aim to minimize the latent space dimensionality, enabling more efficient and compact representations. To tackle these challenges, we focus on improving the decoder's expressiveness. Concretely, we propose DGAE, which employs a diffusion model to guide the decoder in recovering informative signals that are not fully decoded from the latent representation. With this design, DGAE effectively mitigates the performance degradation under high spatial compression rates. At the same time, DGAE achieves state-of-the-art performance with a 2x smaller latent space. When integrated with Diffusion Models, DGAE demonstrates competitive performance on image generation for ImageNet-1K and shows that this compact latent representation facilitates faster convergence of the diffusion model.

preprint2022arXiv

SVMAC: Unsupervised 3D Human Pose Estimation from a Single Image with Single-view-multi-angle Consistency

Recovering 3D human pose from 2D joints is still a challenging problem, especially without any 3D annotation, video information, or multi-view information. In this paper, we present an unsupervised GAN-based model consisting of multiple weight-sharing generators to estimate a 3D human pose from a single image without 3D annotations. In our model, we introduce single-view-multi-angle consistency (SVMAC) to significantly improve the estimation performance. With 2D joint locations as input, our model estimates a 3D pose and a camera simultaneously. During training, the estimated 3D pose is rotated by random angles and the estimated camera projects the rotated 3D poses back to 2D. The 2D reprojections will be fed into weight-sharing generators to estimate the corresponding 3D poses and cameras, which are then mixed to impose SVMAC constraints to self-supervise the training process. The experimental results show that our method outperforms the state-of-the-art unsupervised methods on Human 3.6M and MPI-INF-3DHP. Moreover, qualitative results on MPII and LSP show that our method can generalize well to unknown data.

preprint2015arXiv

Maximal inequality of Stochastic convolution driven by compensated Poisson random measures in Banach spaces

Let $(E, \| \cdot\|)$ be a Banach space such that, for some $q\geq 2$, the function $x\mapsto \|x\|^q$ is of $C^2$ class and its first and second Fréchet derivatives are bounded by some constant multiples of $(q-1)$-th power of the norm and $(q-2)$-th power of the norm and let $S$ be a $C_0$-semigroup of contraction type on $(E, \| \cdot\|)$. We consider the following stochastic convolution process \begin{align*} u(t)=\int_0^t\int_ZS(t-s)ξ(s,z)\,\tilde{N}(\mathrm{d} s,\mathrm{d} z), \;\;\; t\geq 0, \end{align*} where $\tilde{N}$ is a compensated Poisson random measure on a measurable space $(Z,\mathcal{Z})$ and $ξ:[0,\infty)\timesΩ\times Z\rightarrow E$ is an $\mathbb{F}\otimes \mathcal{Z}$-predictable function. We prove that there exists a càdlàg modification a $\tilde{u}$ of the process $u$ which satisfies the following maximal inequality \begin{align*} \mathbb{E} \sup_{0\leq s\leq t} \|\tilde{u}(s)\|^{q^\prime}\leq C\ \mathbb{E} \left(\int_0^t\int_Z \|ξ(s,z) \|^{p}\,N(\mathrm{d} s,\mathrm{d} z)\right)^{\frac{q^\prime}{p}}, \end{align*} for all $ q^\prime \geq q$ and $1<p\leq 2$ with $C=C(q,p)$.

preprint2013arXiv

Strong solutions for SPDE with locally monotone coefficients driven by Lévy noise

Motivated by applications to a manifold of semilinear and quasilinear stochastic partial differential equations (SPDEs) we establish the existence and uniqueness of strong solutions to coercive and locally monotone SPDEs driven by Lévy processes. We illustrate the main result of our paper by showing how it can be applied to various types of SPDEs such as stochastic reaction-diffusion equations, stochastic Burgers type equations, stochastic 2D hydrodynamical systems and stochastic equations of non-Newtonian fluids, which generalize many existing results in the literature.

preprint2011arXiv

A maximal inequality for stochastic convolutions in 2-smooth Banach spaces

Let (e^{tA})_{t \geq 0} be a C_0-contraction semigroup on a 2-smooth Banach space E, let (W_t)_{t \geq 0} be a cylindrical Brownian motion in a Hilbert space H, and let (g_t)_{t \geq 0} be a progressively measurable process with values in the space γ(H,E) of all γ-radonifying operators from H to E. We prove that for all 0<p<\infty there exists a constant C, depending only on p and E, such that for all T \geq 0 we have \E \sup_{0\le t\le T} || \int_0^t e^{(t-s)A} g_s dW_s \ ||^p \leq C \mathbb{E} (\int_0^T || g_t ||_{γ(H,E)}^2 dt)^\frac{p}{2}. For p \geq 2 the proof is based on the observation that ψ(x) = || x ||^p is Fréchet differentiable and its derivative satisfies the Lipschitz estimate || ψ'(x) - ψ'(y)|| \leq C(|| x || + || y ||)^{p-2} || x-y ||; the extension to 0<p<2 proceeds via Lenglart's inequality.