Researcher profile

Sheng Gong

Sheng Gong contributes to research discovery and scholarly infrastructure.

ResearcherAffiliation not importedOpen to collaborate

Trust snapshot

Quick read

Trust 19 - UnverifiedVerification L1Unclaimed author
5works
0followers
5topics
4close collaborators

Actions

Decide how to stay connected

Follow researcher0

Identity and collaboration

How to connect with this researcher

Claiming links this public author record to a researcher profile and unlocks direct collaboration workflows.

Log in to claim

Direct collaboration

Open a focused conversation when the fit is right

Claim this author entity first to unlock direct invitations.

Research graph

See the researcher in context

Open full explorer

Inspect adjacent work, topics, institutions and collaborators without jumping out to a separate graph page.

Building this graph slice

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

5 published item(s)

preprint2026arXiv

Bridging Quantum Mechanics to Organic Liquid Properties via a Universal Force Field

Molecular dynamics (MD) simulations are essential tools for unraveling atomistic insights into the structure and dynamics of condensed-phase systems. However, the universal and accurate prediction of macroscopic properties from ab initio calculations remains a significant challenge, often hindered by the trade-off between computational cost and simulation accuracy. Here, we present ByteFF-Pol, a graph neural network (GNN)-parameterized polarizable force field, trained exclusively on high-level quantum mechanics (QM) data. Leveraging physically-motivated force field forms and training strategies, ByteFF-Pol exhibits exceptional performance in predicting thermodynamic and transport properties for a wide range of small-molecule liquids and electrolytes, outperforming state-of-the-art (SOTA) classical and machine learning force fields. The zero-shot prediction capability of ByteFF-Pol bridges the gap between microscopic QM calculations and macroscopic liquid properties, enabling the exploration of previously intractable chemical spaces. This advancement holds transformative potential for applications such as electrolyte design and custom-tailored solvent, representing a pivotal step toward data-driven materials discovery.

preprint2022arXiv

A cloud platform for automating and sharing analysis of raw simulation data from high throughput polymer molecular dynamics simulations

Open material databases storing hundreds of thousands of material structures and their corresponding properties have become the cornerstone of modern computational materials science. Yet, the raw outputs of the simulations, such as the trajectories from molecular dynamics simulations and charge densities from density functional theory calculations, are generally not shared due to their huge size. In this work, we describe a cloud-based platform to facilitate the sharing of raw data and enable the fast post-processing in the cloud to extract new properties defined by the user. As an initial demonstration, our database currently includes 6286 molecular dynamics trajectories for amorphous polymer electrolytes and 5.7 terabytes of data. We create a public analysis library at https://github.com/TRI-AMDD/htp_md to extract multiple properties from the raw data, using both expert designed functions and machine learning models. The analysis is run automatically with computation in the cloud, and results then populate a database that can be accessed publicly. Our platform encourages users to contribute both new trajectory data and analysis functions via public interfaces. Newly analyzed properties will be incorporated into the database. Finally, we create a front-end user interface at https://www.htpmd.matr.io for browsing and visualization of our data. We envision the platform to be a new way of sharing raw data and new insights for the computational materials science community.

preprint2022arXiv

Calibrating DFT formation enthalpy calculations by multi-fidelity machine learning

Machine learning materials properties measured by experiments is valuable yet difficult due to the limited amount of experimental data. In this work, we use a multi-fidelity random forest model to learn the experimental formation enthalpy of materials with prediction accuracy higher than the empirically corrected PBE functional (PBEfe) and meta-GGA functional (SCAN), and it outperforms the hotly studied deep neural-network based representation learning and transfer learning. We then use the model to calibrate the DFT formation enthalpy in the Materials Project database, and discover materials with underestimated stability. The multi-fidelity model is also used as a data-mining approach to find how DFT deviates from experiments by the explaining the model output.

preprint2020arXiv

Charting Lattice Thermal Conductivity of Inorganic Crystals

Thermal conductivity is a fundamental material property but challenging to predict, with less than 5% out of about $10^5$ synthesized inorganic materials being documented. In this work, we extract the structural chemistry that governs lattice thermal conductivity, by combining graph neural networks and random forest approaches. We show that both mean and variation of unit-cell configurational properties, such as atomic volume and bond length, are the most important features, followed by mass and elemental electronegativity. We chart the structural chemistry of lattice thermal conductivity into extended van-Arkel triangles, and predict the thermal conductivity of all known inorganic materials in the Inorganic Crystal Structure Database. For the latter, we develop a transfer learning framework extendable for other applications.

preprint2020arXiv

Screening and understanding Li adsorption on 2-dimensional metallic materials by learning physics

Two-dimensional (2D) materials have received considerable attention as possible electrodes in Li-ion batteries (LIBs), although a deeper understanding of the Li adsorption behavior as well as broad screening of the materials space is still needed. In this work, we build a high-throughput screening scheme that incorporates a learned interaction. First, density functional theory and graph convolution networks are utilized to calculate minimum Li adsorption energies for a small set of 2D metallic materials. The data is then used to find a dependence of the minimum Li adsorption energies on the sum of ionization potential, work function of the 2D metal, and coupling energy between Li+ and substrate. Our results show that variances of elemental properties and density are the most correlated features with coupling. To illustrate the applicability of this approach, the model is employed to show that some fluorides and chromium oxides are potential high-voltage materials with adsorption energies < -7 eV, and the found physics is used as the design principle to enhance the Li adsorption ability of graphene. This physics-driven approach shows higher accuracy and transferability compared with purely data-driven models.