Source author record

Pu Tian

Pu Tian appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2025arXiv

LeafTutor: An AI Agent for Programming Assignment Tutoring

High enrollment in STEM-related degree programs has created increasing demand for scalable tutoring support, as universities experience a shortage of qualified instructors and teaching assistants (TAs). To address this challenge, LeafTutor, an AI tutoring agent powered by large language models (LLMs), was developed to provide step-by-step guidance for students. LeafTutor was evaluated through real programming assignments. The results indicate that the system can deliver step-by-step programming guidance comparable to human tutors. This work demonstrates the potential of LLM-driven tutoring solutions to enhance and personalize learning in STEM education. If any reader is interested in collaboration with our team to improve or test LeafTutor, please contact Pu Tian (pu.tian@stockton.edu) or Yalong Wu (wuy@uhcl.edu).

preprint2016arXiv

Configurational space continuity and free energy calculations

Free energy is arguably the most importance function(al) for understanding of molecular systems. A number of rigorous and approximate free energy calculation/estimation methods have been developed over many decades. One important issue, the continuity of an interested macrostate (or path) in configurational space, has not been well articulated, however. As a matter of fact, some important special cases have been intensively discussed. In this perspective, I discuss the relevance of configurational space continuity in development of more efficient and reliable next generation free energy methodologies.

preprint2016arXiv

Shot Range and High Order Correlations in Proteins

The main chain dihedral angles play an important role to decide the protein conformation. The native states of a protein can be regard as an ensemble of a lot of similar conformations, in different conformations the main chain dihedral angles vary in a certain range. Each dihedral angle value can be described as a distribution, but only using the distribution can't describe the real conformation space. The reason is that the dihedral angle has correlation with others, especially the neighbor dihedral angles in primary sequence. In our study we analysis extensive molecular dynamics (MD) simulation trajectories of eleven proteins with different sizes and folds, we found that in stable second structure the correlations only exist between the dihedrals near to each other in primary sequence, long range correlations are rare. But in unstable structures (loop) long range correlations exist. Further we observed some characteristics of the short range correlations in different second structures (α-helix, β-sheet) and we found that we can approximate good high order dihedral angle distribution good only use three order distribution in stable second structure which illustrates that high order correlations (over three order) is small in stable second structure.

preprint2016arXiv

Utility of potential energy span as an approximate free energy proxy

Free energy calculation is critical in predictive tasks such as protein folding, docking and design. However, rigorous calculation of free energy change is prohibitively expensive in these practical applications. The minimum potential energy is therefore widely utilized to approximate free energy. In this study, based on analysis of extensive molecular dynamics (MD) simulation trajectories of a few native globular proteins, we found that change of minimum and corresponding maximum potential energy terms exhibit similar level of correlation with change of free energy. More importantly, we demonstrated that change of span (maximum - minimum) of potential energy terms, which engender negligible additional computational cost, exhibit considerably stronger correlations with change of free energy than the corresponding change of minimum and maximum potential energy terms. Therefore, potential energy span may serve as an alternative efficient approximate free energy proxy.

preprint2015arXiv

Configurational space discretization and free energy calculation in complex molecular systems

Trajectories provide dynamical information that is discarded in free energy calculations, for which we sought to design a scheme with the hope of saving cost for generating dynamical information. We first demonstrated that snapshots in a converged trajectory set are associated with implicit conformers that have invariant statistical weight distribution (ISWD). Based on the thought that infinite number of sets of implicit conformers with ISWD may be created through independent converged trajectory sets, we hypothesized that explicit conformers with ISWD may be constructed for complex molecular systems through systematic increase of conformer fineness, and tested the hypothesis in lipid molecule palmitoyloleoylphosphatidylcholine (POPC). Furthermore, when explicit conformers with ISWD were utilized as basic states to define conformational entropy, change of which between two given macrostates was found to be equivalent to change of free energy except a mere difference of a negative temperature factor, and change of enthalpy essentially cancels corresponding change of average intra-conformer entropy. These findings suggest that entropy enthalpy compensation is inherently a local phenomenon in configurational space. By implicitly taking advantage of entropy enthalpy compensation and forgoing all dynamical information, constructing explicit conformers with ISWD and counting thermally accessible number of which for interested end macrostates is likely to be an efficient and reliable alternative end point free energy calculation strategy.

preprint2015arXiv

Ideal gas behavior of rotamerically defined conformers in native globular proteins

Protein conformational transitions, which are essential for function, may be driven either by entropy or enthalpy when molecular systems comprising solute and solvent molecules are the focus. Revealing thermodynamic origin of a given molecular process is an important but difficult task, and general principles governing protein conformational distributions remain elusive. Here we demonstrate that when protein molecules are taken as thermodynamic systems and solvents being treated as the environment, conformational entropy is an excellent proxy for free energy and is sufficient to explain protein conformational distributions. Specifically, by defining each unique combination of side chain torsional state as a conformer, the population distribution (or free energy) on an arbitrarily given order parameter is approximately a linear function of conformational entropy. Additionally, span of various microscopic potential energy terms is observed to be highly correlated with both conformational entropy and free energy. Presently widely utilized free energy proxies, including minimum potential energy, average potential energy terms by themselves or in combination with vibrational entropy\cite, are found to correlate with free energy rather poorly. Therefore, our findings provide a fundamentally new theoretical base for development of significantly more reliable and efficient next generation computational tools, where the number of available conformers,rather than poential energy of microscopic configurations, is the central focus. We anticipate that many related research fields, including structure based drug design and discovery, protein design, docking and prediction of general intermolecular interactions involving proteins, are expected to benefit greatly.

preprint2015arXiv

Pathway-based feature selection algorithms identify genes discriminating patients with multiple sclerosis apart from controls

Introduction The focus of analyzing data from microarray experiments and extracting biological insight from such data has experienced a shift from identification of individual genes in association with a phenotype to that of biological pathways or gene sets. Meanwhile, feature selection algorithm becomes imperative to cope with the high dimensional nature of many modeling tasks in bioinformatics. Many feature selection algorithms use information contained within a gene set as a biological priori, and select relevant features by incorporating such information. Thus, an integration of gene set analysis with feature selection is highly desired. Significance analysis of microarray to gene-set reduction analysis (SAM-GSR) algorithm is a novel direction of gene set analysis, aiming at further reduction of gene set into a core subset. Here, we explore the feature selection trait possessed by SAM-GSR and then modify SAM-GSR specifically to better fulfill this role. Results and Conclusions Training on a multiple sclerosis (MS) microarray data using both SAM-GSR and our modification of SAM-GSR, excellent discriminative performance on an independent test set was achieved. To conclude, absorbing biological information from a gene set may be helpful for classification and feature selection. Discussion Given the fact the complete pathway information is far from completeness, a statistical method capable of constructing biologically meaningful gene networks is in demand. The basic requirement is that interplay among genes must be taken into account.

preprint2012arXiv

Entropically Dominant State of Proteins

Configurational entropy is an important factor in the free energy change of many macromolecular recognition and binding processes, and has been intensively studied. Despite great progresses that have been made, the global sampling remains to be a grand challenge in computational analysis of relevant processes. Here we propose and demonstrate an entropy estimation method that is based on physical partition of configurational space and can be readily combined with currently available methodologies. Tests with two globular proteins suggest that for flexible macromolecules with large and complex configurational space, accurate configurational entropy estimation may be achieved simply by considering the entropically most important subspace. This conclusion effectively converts an exhaustive sampling problem into a local sampling one, and defines entropically dominant state for proteins and other complex macromolecules. The conceptional breakthrough is likely to positively impact future theoretical analysis, computational algorithm development and experimental design of diverse chemical and biological molecular systems.