Source author record

Mingxi Chen

Mingxi Chen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

2works
3topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

2 published item(s)

preprint2025arXiv

Enhanced TM-Mode 3D Coupled Wave Theory for Photonic Crystal Surface-Emitting Terahertz Quantum Cascade Lasers

In this study, we propose and develop an enhanced three-dimensional coupled wave theory (3D CWT) to investigate the optical field behavior in photonic crystal surface-emitting terahertz quantum cascade lasers (THz-QCLs). By incorporating an effective permittivity enhancement (EP) model and a self-consistent iteration (SCI) method, we successfully address the numerical dispersion issues encountered in analytical methods when dealing with metallic waveguide structures. The results demonstrate that the EP and SCI-enhanced 3D TM mode CWT achieves computational accuracy comparable to traditional numerical simulation methods such as finite-difference time-domain (FDTD), while significantly reducing the required computational resources, including time and memory, to just tens of minutes. Moreover, this method provides a clear physical insight, revealing the reasons behind the current low extraction efficiency in surface-emitting THz-QCLs. Our study showcases the potential of the EP and SCI-enhanced 3D CWT as a powerful simulation tool in the research of photonic crystal surface-emitting lasers, offering a new theoretical foundation and optimization direction for future laser designs.

preprint2025arXiv

RSAgent: Learning to Reason and Act for Text-Guided Segmentation via Multi-Turn Tool Invocations

Text-guided object segmentation requires both cross-modal reasoning and pixel grounding abilities. Most recent methods treat text-guided segmentation as one-shot grounding, where the model predicts pixel prompts in a single forward pass to drive an external segmentor, which limits verification, refocusing and refinement when initial localization is wrong. To address this limitation, we propose RSAgent, an agentic Multimodal Large Language Model (MLLM) which interleaves reasoning and action for segmentation via multi-turn tool invocations. RSAgent queries a segmentation toolbox, observes visual feedback, and revises its spatial hypothesis using historical observations to re-localize targets and iteratively refine masks. We further build a data pipeline to synthesize multi-turn reasoning segmentation trajectories, and train RSAgent with a two-stage framework: cold-start supervised fine-tuning followed by agentic reinforcement learning with fine-grained, task-specific rewards. Extensive experiments show that RSAgent achieves a zero-shot performance of 66.5% gIoU on ReasonSeg test, improving over Seg-Zero-7B by 9%, and reaches 81.5% cIoU on RefCOCOg, demonstrating state-of-the-art performance on both in-domain and out-of-domain benchmarks.