Source author record

Kai Dong

Kai Dong appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Artificial Intelligence Computation and Language Machine Learning Applications eess.SP eess.SY gr-qc math-ph math.AP math.DG math.MP Systems and Control

Catalog footprint

What is connected

5works

12topics

4close collaborators

Actions

Connect this record

Open graph Browse works

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

preprint2026arXiv

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

General reasoning represents a long-standing and formidable challenge in artificial intelligence. Recent breakthroughs, exemplified by large language models (LLMs) and chain-of-thought prompting, have achieved considerable success on foundational reasoning tasks. However, this success is heavily contingent upon extensive human-annotated demonstrations, and models' capabilities are still insufficient for more complex problems. Here we show that the reasoning abilities of LLMs can be incentivized through pure reinforcement learning (RL), obviating the need for human-labeled reasoning trajectories. The proposed RL framework facilitates the emergent development of advanced reasoning patterns, such as self-reflection, verification, and dynamic strategy adaptation. Consequently, the trained model achieves superior performance on verifiable tasks such as mathematics, coding competitions, and STEM fields, surpassing its counterparts trained via conventional supervised learning on human demonstrations. Moreover, the emergent reasoning patterns exhibited by these large-scale models can be systematically harnessed to guide and enhance the reasoning capabilities of smaller models.

preprint2026arXiv

The past stability of Kasner singularities for the $(3+1)$-dimensional Einstein vacuum spacetime under polarized $U(1)$-symmetry

In this paper, we give a new proof to a past stability result established in Fournodavlos-Rodnianski-Speck (arXiv:2012.05888), for Kasner solutions of the $(3+1)$-dimensional Einstein vacuum equations under polarized $U(1)$-symmetry. Our method, inspired by Beyer-Oliynyk-Olvera-Santamar{\'ı}a-Zheng (arXiv:1907.04071, arXiv:2502.09210), relies on a newly developed $(2+1)$ orthonormal-frame decomposition and a careful symmetrization argument, after which the Fuchsian techniques can be applied. We show that the perturbed solutions are asymptotically pointwise Kasner, geodesically incomplete and crushing at the Big Bang singularity. They are achieved by reducing the $(3+1)$ Einstein vacuum equations to a Fuchsian system coupled with several constraint equations, with the symmetry assumption playing an important role in the reduction. Using Fuchsian theory together with finite speed of constraints propagation, we obtain global existence and precise asymptotics of the solutions up to the singularities.

preprint2024arXiv

DeepSeek LLM: Scaling Open-Source Language Models with Longtermism

The rapid development of open-source large language models (LLMs) has been truly remarkable. However, the scaling law described in previous literature presents varying conclusions, which casts a dark cloud over scaling LLMs. We delve into the study of scaling laws and present our distinctive findings that facilitate scaling of large scale models in two commonly used open-source configurations, 7B and 67B. Guided by the scaling laws, we introduce DeepSeek LLM, a project dedicated to advancing open-source language models with a long-term perspective. To support the pre-training phase, we have developed a dataset that currently consists of 2 trillion tokens and is continuously expanding. We further conduct supervised fine-tuning (SFT) and Direct Preference Optimization (DPO) on DeepSeek LLM Base models, resulting in the creation of DeepSeek Chat models. Our evaluation results demonstrate that DeepSeek LLM 67B surpasses LLaMA-2 70B on various benchmarks, particularly in the domains of code, mathematics, and reasoning. Furthermore, open-ended evaluations reveal that DeepSeek LLM 67B Chat exhibits superior performance compared to GPT-3.5.

preprint2022arXiv

Conformal Metasurfaces: a Novel Solution for Vehicular Communications

In future 6G millimeter wave (mmWave)/sub-THz vehicle-to-everything (V2X) communication systems, vehicles are expected to be equipped with massive antenna arrays to realize beam-based links capable of compensating for the severe path loss. However, vehicle-to-vehicle (V2V) direct links are prone to be blocked by surrounding vehicles. Emerging metasurface technologies enable the control of the electromagnetic wave reflection towards the desired direction, enriching the channel scattering to boost communication performance. Reconfigurable intelligent surfaces (RIS), and mostly the pre-configured counterpart intelligent reflecting surfaces (IRS), are a promising low-cost relaying system for 6G. This paper proposes using conformal metasurfaces (either C-RIS or C-IRS) deployed on vehicles' body to mitigate the blockage impact in a highway multi-lane scenario. In particular, conformal metasurfaces create artificial reflections to mitigate blockage by compensating for the non-flat shape of the vehicle's body, such as the lateral doors, with proper phase patterns. We analytically derive the phase pattern to apply to a cylindrical C-RIS/C-IRS approximating the shape of the car body, as a function of both incidence and reflection angles, considering cylindrical RIS/IRS as a generalization of conventional planar ones. We propose a novel design for optimally pre-configured C-IRS to mimic the behavior of an EM flat surface on car doors, proving the benefits of C-RIS and C-IRS in a multi-lane V2V highway scenario. The results show a consistent reduction of blockage probability when exploiting C-RIS/C-IRS, 20% for pre-configured C-IRS, and 70% for C-RIS, as well as a remarkable improvement in terms of average signal-to-noise ratio, respectively 10-20 dB for C-IRS and 30-40 dB for C-RIS.

preprint2015arXiv

NBLDA: Negative Binomial Linear Discriminant Analysis for RNA-Seq Data

RNA-sequencing (RNA-Seq) has become a powerful technology to characterize gene expression profiles because it is more accurate and comprehensive than microarrays. Although statistical methods that have been developed for microarray data can be applied to RNA-Seq data, they are not ideal due to the discrete nature of RNA-Seq data. The Poisson distribution and negative binomial distribution are commonly used to model count data. Recently, Witten (2011) proposed a Poisson linear discriminant analysis for RNA-Seq data. The Poisson assumption may not be as appropriate as negative binomial distribution when biological replicates are available and in the presence of overdispersion (i.e., when the variance is larger than the mean). However, it is more complicated to model negative binomial variables because they involve a dispersion parameter that needs to be estimated. In this paper, we propose a negative binomial linear discriminant analysis for RNA-Seq data. By Bayes' rule, we construct the classifier by fitting a negative binomial model, and propose some plug-in rules to estimate the unknown parameters in the classifier. The relationship between the negative binomial classifier and the Poisson classifier is explored, with a numerical investigation of the impact of dispersion on the discriminant score. Simulation results show the superiority of our proposed method. We also analyze four real RNA-Seq data sets to demonstrate the advantage of our method in real-world applications.