Source author record

Yanhui Wang

Yanhui Wang appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

3works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

3 published item(s)

preprint2023arXiv

MicroCinema: A Divide-and-Conquer Approach for Text-to-Video Generation

We present MicroCinema, a straightforward yet effective framework for high-quality and coherent text-to-video generation. Unlike existing approaches that align text prompts with video directly, MicroCinema introduces a Divide-and-Conquer strategy which divides the text-to-video into a two-stage process: text-to-image generation and image\&text-to-video generation. This strategy offers two significant advantages. a) It allows us to take full advantage of the recent advances in text-to-image models, such as Stable Diffusion, Midjourney, and DALLE, to generate photorealistic and highly detailed images. b) Leveraging the generated image, the model can allocate less focus to fine-grained appearance details, prioritizing the efficient learning of motion dynamics. To implement this strategy effectively, we introduce two core designs. First, we propose the Appearance Injection Network, enhancing the preservation of the appearance of the given image. Second, we introduce the Appearance Noise Prior, a novel mechanism aimed at maintaining the capabilities of pre-trained 2D diffusion models. These design elements empower MicroCinema to generate high-quality videos with precise motion, guided by the provided text prompts. Extensive experiments demonstrate the superiority of the proposed framework. Concretely, MicroCinema achieves SOTA zero-shot FVD of 342.86 on UCF-101 and 377.40 on MSR-VTT. See https://wangyanhui666.github.io/MicroCinema.github.io/ for video samples.

preprint2022arXiv

Quasi-semilattices on networks

This paper introduces the tensor representation of a network, here tensors are the primitive structures of the network. In view of tensor chains, two binary operations on tensor sets are defined: chain addition and reducing. Based on the reducing operation, the tensor chain representation of subnetworks of a network is given, and it is proved that all connected subnetworks of a network (here refers to the tensor chain generated by primitive structures) form a quasi-semilattice with respect to reducing, namely {\it network quasi-semilattices}. Here, quasi-semilattices refer to algebraic systems that are idempotent commutative and do not satisfy the association law. Then, we discuss the subalgebra structures of the network quasi-semilattice in terms of two equivalent relations $σ$ and $δ$. $δ$ is a congruence. Each $δ$-class forms a semilattice with respect to reducing, that is, an idempotent commutative semigroup, and also each $δ$-class has an order structure with the maximum element and minimum elements. Here, the minimum elements correspond to the spanning tree in graph theory. Finally, we discuss how three path algebras: graph inverse semigroups, Leavitt path algebra and Cuntz-Krieger graph $C^*$-algebra are constructed in terms of tensors with respect to chain-addition.

preprint2015arXiv

Universality for products of random matrices I: Ginibre and truncated unitary cases

Recently, the joint probability density functions of complex eigenvalues for products of independent complex Ginibre matrices have been explicitly derived as determinantal point processes. We express truncated series coming from the correlation kernels as multivariate integrals with singularity and investigate saddle point method for such a type of integrals. As an application, we prove that the eigenvalue correlation functions have the same scaling limits as those of the single complex Ginibre ensemble, both in the bulk and at the edge of the spectrum. We also prove that the similar results hold true for products of independent truncated unitary matrices.