Researcher profile

Mingming Li

Mingming Li contributes to research discovery and scholarly infrastructure.

ResearcherAffiliation not importedOpen to collaborate

Trust snapshot

Quick read

Trust 15 - UnverifiedVerification L1Unclaimed author
3works
0followers
6topics
4close collaborators

Actions

Decide how to stay connected

Follow researcher0

Identity and collaboration

How to connect with this researcher

Claiming links this public author record to a researcher profile and unlocks direct collaboration workflows.

Log in to claim

Direct collaboration

Open a focused conversation when the fit is right

Claim this author entity first to unlock direct invitations.

Research graph

See the researcher in context

Open full explorer

Inspect adjacent work, topics, institutions and collaborators without jumping out to a separate graph page.

Building this graph slice

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

3 published item(s)

preprint2026arXiv

DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

General reasoning represents a long-standing and formidable challenge in artificial intelligence. Recent breakthroughs, exemplified by large language models (LLMs) and chain-of-thought prompting, have achieved considerable success on foundational reasoning tasks. However, this success is heavily contingent upon extensive human-annotated demonstrations, and models' capabilities are still insufficient for more complex problems. Here we show that the reasoning abilities of LLMs can be incentivized through pure reinforcement learning (RL), obviating the need for human-labeled reasoning trajectories. The proposed RL framework facilitates the emergent development of advanced reasoning patterns, such as self-reflection, verification, and dynamic strategy adaptation. Consequently, the trained model achieves superior performance on verifiable tasks such as mathematics, coding competitions, and STEM fields, surpassing its counterparts trained via conventional supervised learning on human demonstrations. Moreover, the emergent reasoning patterns exhibited by these large-scale models can be systematically harnessed to guide and enhance the reasoning capabilities of smaller models.

preprint2022arXiv

Optical Properties of C$-$rich ($^{12}$C, SiC and FeC) Dust Layered Structure of Massive Stars

The composition and structure of interstellar dust are important and complex for the study of the evolution of stars and the \textbf{interstellar medium} (ISM). However, there is a lack of corresponding experimental data and model theories. By theoretical calculations based on ab-initio method, we have predicted and geometry optimized the structures of Carbon-rich (C-rich) dusts, carbon ($^{12}$C), iron carbide (FeC), silicon carbide (SiC), even silicon ($^{28}$Si), iron ($^{56}$Fe), and investigated the optical absorption coefficients and emission coefficients of these materials in 0D (zero$-$dimensional), 1D, and 2D nanostructures. Comparing the \textbf{nebular spectra} of the supernovae (SN) with the coefficient of dust, we find that the optical absorption coefficient of the 2D $^{12}$C, $^{28}$Si, $^{56}$Fe, SiC and FeC structure corresponds to the absorption peak displayed in the infrared band (5$-$8) $μ$$m$ of the spectrum at 7554 days after the SN1987A explosion. And it also corresponds to the spectrum of 535 days after the explosion of SN2018bsz, when the wavelength in the range of (0.2$-$0.8) and (3$-$10) $μ$$m$. Nevertheless, 2D SiC and FeC corresponds to the spectrum of 844 days after the explosion of SN2010jl, when the wavelength is within (0.08$-$10) $μ$$m$. Therefore, FeC and SiC may be the second type of dust in SN1987A corresponding to infrared band (5$-$8) $μ$$m$ of dust and may be in the ejecta of SN2010jl and SN2018bsz.

preprint2019arXiv

Fisher-Rao Geometry and Jeffreys Prior for Pareto Distribution

In this paper, we investigate the Fisher-Rao geometry of the two-parameter family of Pareto distribution. We prove that its geometrical structure is isometric to the Poincaré upper half-plane model, and then study the corresponding geometrical features by presenting explicit expressions for connection, curvature and geodesics. It is then applied to Bayesian inference by considering the Jeffreys prior determined by the volume form. In addition, the posterior distribution from the prior is computed, providing a systematic method to the Bayesian inference for Pareto distribution.