Source author record

Kuai Yu

Kuai Yu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Artificial Intelligence Computation and Language Machine Learning math.CO physics.app-ph physics.optics

Catalog footprint

What is connected

4works

6topics

4close collaborators

Actions

Connect this record

Open graph Browse works

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

preprint2026arXiv

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

General reasoning represents a long-standing and formidable challenge in artificial intelligence. Recent breakthroughs, exemplified by large language models (LLMs) and chain-of-thought prompting, have achieved considerable success on foundational reasoning tasks. However, this success is heavily contingent upon extensive human-annotated demonstrations, and models' capabilities are still insufficient for more complex problems. Here we show that the reasoning abilities of LLMs can be incentivized through pure reinforcement learning (RL), obviating the need for human-labeled reasoning trajectories. The proposed RL framework facilitates the emergent development of advanced reasoning patterns, such as self-reflection, verification, and dynamic strategy adaptation. Consequently, the trained model achieves superior performance on verifiable tasks such as mathematics, coding competitions, and STEM fields, surpassing its counterparts trained via conventional supervised learning on human demonstrations. Moreover, the emergent reasoning patterns exhibited by these large-scale models can be systematically harnessed to guide and enhance the reasoning capabilities of smaller models.

preprint2026arXiv

mHC: Manifold-Constrained Hyper-Connections

Recently, studies exemplified by Hyper-Connections (HC) have extended the ubiquitous residual connection paradigm established over the past decade by expanding the residual stream width and diversifying connectivity patterns. While yielding substantial performance gains, this diversification fundamentally compromises the identity mapping property intrinsic to the residual connection, which causes severe training instability and restricted scalability, and additionally incurs notable memory access overhead. To address these challenges, we propose Manifold-Constrained Hyper-Connections (mHC), a general framework that projects the residual connection space of HC onto a specific manifold to restore the identity mapping property, while incorporating rigorous infrastructure optimization to ensure efficiency. Empirical experiments demonstrate that mHC is effective for training at scale, offering tangible performance improvements and superior scalability. We anticipate that mHC, as a flexible and practical extension of HC, will contribute to a deeper understanding of topological architecture design and suggest promising directions for the evolution of foundational models.

preprint2022arXiv

Tunable bilayer dielectric metasurface via stacking magnetic mirrors

Functional tunability, environmental adaptability, and easy fabrication are highly desired properties in metasurfaces. Here we provide a tunable bilayer metasurface composed of two stacked identical dielectric magnetic mirrors, which are excited by the dominant electric dipole and other magnetic multipoles, exhibiting nonlocal electric field enhancement near the interface and high reflection. Differ from the tunability through the direct superposition of two structures with different functionalities, we achieve the reversible conversion between high reflection and high transmission by manipulating the interlayer coupling near the interface between the two magnetic mirrors. The magnetic mirror effect boosts the interlayer coupling when the interlayer spacing is small. Decreasing the interlayer spacing of the bilayer metasurface leads to stronger interlayer coupling and scattering suppression of the meta-atom, which results in high transmission. On the contrary, increasing the spacing leads to weaker interlayer coupling and scattering enhancement, which results in high reflection. The high transmission of the bilayer metasurface has good robustness due to that the meta-atom with interlayer coupling can maintain the scattering suppression against adjacent meta-atom movement and disordered position perturbation. This work provides a straightforward method (i.e. stacking magnetic mirrors) to design tunable metasurface, shed new light on high-performance optical switches applied in communication and sensing.

preprint2014arXiv

G-parking functions and tree inversions

A depth-first search version of Dhar's burning algorithm is used to give a bijection between the parking functions of a graph and labeled spanning trees, relating the degree of the parking function with the number of inversions of the spanning tree. Specializing to the complete graph answers a problem posed by R. Stanley.