Source author record

Donghyeong Kim

Donghyeong Kim appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

2works
4topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

2 published item(s)

preprint2026arXiv

OTT-Vid: Optimal Transport Temporal Token Compression for Video Large Language Models

As Video Large Language Models (Video-LLMs) scale to longer and more complex videos, their inference cost grows rapidly due to the large volume of visual tokens accumulated across frames. Training-free token compression has emerged as a practical solution to this bottleneck. However, existing temporal compression methods rely primarily on cross-frame token similarity or segmentation heuristics, overlooking each token's semantic role within its frame and failing to adapt compression strength to the compressibility of each frame pair. In this work, we propose OTT-Vid, a transport-derived allocation framework for temporal token compression. Our approach consists of two stages: spatial pruning identifies representative content within each frame, and optimal transport (OT) is then solved between neighboring frames to estimate temporal compressibility. We formulate this OT with non-uniform token mass, which protects semantically important tokens from aggressive compression, and a locality-aware cost that captures both feature and spatial disparities. The resulting transport plan jointly balances token importance and matching cost, while its total cost defines the transport difficulty of each frame pair, which we use to allocate compression budgets dynamically. Experiments on six benchmarks spanning video question answering and temporal grounding show that OTT-Vid preserves 95.8% of VQA and 73.9% of VTG performance while retaining only 10% of tokens, consistently outperforming existing state-of-the-art training-free compression methods.

preprint2022arXiv

Customising radiative decay dynamics of two-dimensional excitons via position- and polarisation-dependent vacuum-field interference

Embodying bosonic and electrically interactive characteristics in two-dimensional space, excitons in transition-metal dichalcogenides (TMDCs) have garnered considerable attention. The realisation and application of strong-correlation effects, long-range transport, and valley-dependent optoelectronic properties require customising exciton decay dynamics. Strains, defects, and electrostatic doping effectively control the decay dynamics but significantly disturb the intrinsic properties of TMDCs, such as electron band structure and exciton binding energy. Meanwhile, vacuum-field manipulation provides an optical alternative for engineering radiative decay dynamics. Planar mirrors and cavities have been employed to manage the light-matter interactions of two-dimensional excitons. However, the conventional flat platforms cannot customise the radiative decay landscape in the horizontal TMDC plane or independently control vacuum field interference at different pumping and emission frequencies. Here, we present a meta-mirror resolving the issues with more optical freedom. For neutral excitons of the monolayer MoSe2, the meta-mirror manipulated the radiative decay rate by two orders of magnitude, depending on its geometry. Moreover, we experimentally identified the correlation between emission intensity and spectral linewidth. The anisotropic meta-mirror demonstrated polarisation-dependent radiative decay control. We expect that the meta-mirror platform will be promising to tailor the two-dimensional distributions of lifetime, density, and diffusion of TMDC excitons in advanced opto-excitonic applications.