Source author record

Ziyang Meng

Ziyang Meng appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

10works
5topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

10 published item(s)

preprint2026arXiv

An Efficient and Multi-Modal Navigation System with One-Step World Model

Navigation is a fundamental capability for mobile robots. While the current trend is to use learning-based approaches to replace traditional geometry-based methods, existing end-to-end learning-based policies often struggle with 3D spatial reasoning and lack a comprehensive understanding of physical world dynamics. Integrating world models-which predict future observations conditioned on given actions-with iterative optimization planning offers a promising solution due to their capacity for imagination and flexibility. However, current navigation world models, typically built on pure transformer architectures, often rely on multi-step diffusion processes and autoregressive frame-by-frame generation. These mechanisms result in prohibitive computational latency, rendering real-time deployment impossible. To address this bottleneck, we propose a lightweight navigation world model that adopts a one-step generation paradigm and a 3D U-Net backbone equipped with efficient spatial-temporal attention. This design drastically reduces inference latency, enabling high-frequency control while achieving superior predictive performance. We also integrate this model into an optimization-based planning framework utilizing anchor-based initialization to handle multi-modal goal navigation tasks. Extensive closed-loop experiments in both simulation and real-world environments demonstrate our system's superior efficiency and robustness compared to state-of-the-art baselines.

preprint2026arXiv

Learning Diverse Skills for Behavior Models with Mixture of Experts

Imitation learning has demonstrated strong performance in robotic manipulation by learning from large-scale human demonstrations. While existing models excel at single-task learning, it is observed in practical applications that their performance degrades in the multi-task setting, where interference across tasks leads to an averaging effect. To address this issue, we propose to learn diverse skills for behavior models with Mixture of Experts, referred to as Di-BM. Di-BM associates each expert with a distinct observation distribution, enabling experts to specialize in sub-regions of the observation space. Specifically, we employ energy-based models to represent expert-specific observation distributions and jointly train them alongside the corresponding action models. Our approach is plug-and-play and can be seamlessly integrated into standard imitation learning methods. Extensive experiments on multiple real-world robotic manipulation tasks demonstrate that Di-BM significantly outperforms state-of-the-art baselines. Moreover, fine-tuning the pretrained Di-BM on novel tasks exhibits superior data efficiency and the reusable of expert-learned knowledge. Code is available at https://github.com/robotnav-bot/Di-BM.

preprint2026arXiv

STEP3-VL-10B Technical Report

We present STEP3-VL-10B, a lightweight open-source foundation model designed to redefine the trade-off between compact efficiency and frontier-level multimodal intelligence. STEP3-VL-10B is realized through two strategic shifts: first, a unified, fully unfrozen pre-training strategy on 1.2T multimodal tokens that integrates a language-aligned Perception Encoder with a Qwen3-8B decoder to establish intrinsic vision-language synergy; and second, a scaled post-training pipeline featuring over 1k iterations of reinforcement learning. Crucially, we implement Parallel Coordinated Reasoning (PaCoRe) to scale test-time compute, allocating resources to scalable perceptual reasoning that explores and synthesizes diverse visual hypotheses. Consequently, despite its compact 10B footprint, STEP3-VL-10B rivals or surpasses models 10$\times$-20$\times$ larger (e.g., GLM-4.6V-106B, Qwen3-VL-235B) and top-tier proprietary flagships like Gemini 2.5 Pro and Seed-1.5-VL. Delivering best-in-class performance, it records 92.2% on MMBench and 80.11% on MMMU, while excelling in complex reasoning with 94.43% on AIME2025 and 75.95% on MathVision. We release the full model suite to provide the community with a powerful, efficient, and reproducible baseline.

preprint2016arXiv

Modulus Consensus over Networks with Antagonistic Interactions and Switching Topologies

In this paper, we study the discrete-time consensus problem over networks with antagonistic and cooperative interactions. Following the work by Altafini [IEEE Trans. Automatic Control, 58 (2013), pp. 935--946], by an antagonistic interaction between a pair of nodes updating their scalar states we mean one node receives the opposite of the state of the other and naturally by an cooperative interaction we mean the former receives the true state of the latter. Here the pairwise communication can be either unidirectional or bidirectional and the overall network topology graph may change with time. The concept of modulus consensus is introduced to characterize the scenario that the moduli of the node states reach a consensus. It is proved that modulus consensus is achieved if the switching interaction graph is uniformly jointly strongly connected for unidirectional communications, or infinitely jointly connected for bidirectional communications. We construct a counterexample to underscore the rather surprising fact that quasi-strong connectivity of the interaction graph, i.e., the graph contains a directed spanning tree, is not sufficient to guarantee modulus consensus even under fixed topologies. Finally, simulation results using a discrete-time Kuramoto model are given to illustrate the convergence results showing that the proposed framework is applicable to a class of networks with general nonlinear node dynamics.

preprint2015arXiv

Multi-agent Systems with Compasses

This paper investigates agreement protocols over cooperative and cooperative--antagonistic multi-agent networks with coupled continuous-time nonlinear dynamics. To guarantee convergence for such systems, it is common in the literature to assume that the vector field of each agent is pointing inside the convex hull formed by the states of the agent and its neighbors, given that the relative states between each agent and its neighbors are available. This convexity condition is relaxed in this paper, as we show that it is enough that the vector field belongs to a strict tangent cone based on a local supporting hyperrectangle. The new condition has the natural physical interpretation of requiring shared reference directions in addition to the available local relative states. Such shared reference directions can be further interpreted as if each agent holds a magnetic compass indicating the orientations of a global frame. It is proven that the cooperative multi-agent system achieves exponential state agreement if and only if the time-varying interaction graph is uniformly jointly quasi-strongly connected. Cooperative--antagonistic multi-agent systems are also considered. For these systems, the relation has a negative sign for arcs corresponding to antagonistic interactions. State agreement may not be achieved, but instead it is shown that all the agents' states asymptotically converge, and their limits agree componentwise in absolute values if and in general only if the time-varying interaction graph is uniformly jointly strongly connected.

preprint2015arXiv

Network Synchronization with Nonlinear Dynamics and Switching Interactions

This paper considers the synchronization problem for networks of coupled nonlinear dynamical systems under switching communication topologies. Two types of nonlinear agent dynamics are considered. The first one is non-expansive dynamics (stable dynamics with a convex Lyapunov function $φ(\cdot)$) and the second one is dynamics that satisfies a global Lipschitz condition. For the non-expansive case, we show that various forms of joint connectivity for communication graphs are sufficient for networks to achieve global asymptotic $φ$-synchronization. We also show that $φ$-synchronization leads to state synchronization provided that certain additional conditions are satisfied. For the globally Lipschitz case, unlike the non-expansive case, joint connectivity alone is not sufficient for achieving synchronization. A sufficient condition for reaching global exponential synchronization is established in terms of the relationship between the global Lipschitz constant and the network parameters. We also extend the results to leader-follower networks.

preprint2015arXiv

Sampled-Data Consensus over Random Networks

This paper considers the consensus problem for a network of nodes with random interactions and sampled-data control actions. We first show that consensus in expectation, in mean square, and almost surely are equivalent for a general random network model when the inter-sampling interval and network size satisfy a simple relation. The three types of consensus are shown to be simultaneously achieved over an independent or a Markovian random network defined on an underlying graph with a directed spanning tree. For both independent and Markovian random network models, necessary and sufficient conditions for mean-square consensus are derived in terms of the spectral radius of the corresponding state transition matrix. These conditions are then interpreted as the existence of critical value on the inter-sampling interval, below which global mean-square consensus is achieved and above which the system diverges in mean-square sense for some initial states. Finally, we establish an upper bound on the inter-sampling interval below which almost sure consensus is reached, and a lower bound on the inter-sampling interval above which almost sure divergence is reached. Some numerical simulations are given to validate the theoretical results and some discussions on the critical value of the inter-sampling intervals for the mean-square consensus are provided.

preprint2014arXiv

Cooperative Set Aggregation for Multiple Lagrangian Systems

In this paper, we study the cooperative set tracking problem for a group of Lagrangian systems. Each system observes a convex set as its local target. The intersection of these local sets is the group aggregation target. We first propose a control law based on each system's own target sensing and information exchange with neighbors. With necessary connectivity for both cases of fixed and switching communication graphs, multiple Lagrangian systems are shown to achieve rendezvous on the intersection of all the local target sets while the vectors of generalized coordinate derivatives are driven to zero. Then, we introduce the collision avoidance control term into set aggregation control to ensure group dispersion. By defining an ultimate bound on the final generalized coordinate between each system and the intersection of all the local target sets, we show that multiple Lagrangian systems approach a bounded region near the intersection of all the local target sets while the collision avoidance is guaranteed during the movement. In addition, the vectors of generalized coordinate derivatives of all the mechanical systems are shown to be driven to zero. Simulation results are given to validate the theoretical results.

preprint2014arXiv

Coordinated Output Regulation of Heterogeneous Linear Systems under Switching Topologies

This paper constructs a framework to describe and study the coordinated output regulation problem for multiple heterogeneous linear systems. Each agent is modeled as a general linear multiple-input multiple-output system with an autonomous exosystem which represents the individual offset from the group reference for the agent. The multi-agent system as a whole has a group exogenous state which represents the tracking reference for the whole group. Under the constraints that the group exogenous output is only locally available to each agent and that the agents have only access to their neighbors' information, we propose observer-based feedback controllers to solve the coordinated output regulation problem using output feedback information. A high-gain approach is used and the information interactions are allowed to be switched over a finite set of fixed networks containing both graphs that have a directed spanning tree and graphs that do not. The fundamental relationship between the information interactions, the dwell time, the non-identical dynamics of different agents, and the high-gain parameters is given. Simulations are shown to validate the theoretical results.

preprint2014arXiv

Periodic Behaviors in Constrained Multi-agent Systems

In this paper, we provide two discrete-time multi-agent models which generate periodic behaviors. The first one is a multi-agent system of identical double integrators with input saturation constraints, while the other one is a multi-agent system of identical neutrally stable system with input saturation constraints. In each case, we show that if the feedback gain parameters of the local controller satisfy a certain condition, the multi-agent system exhibits a periodic solution.