Source author record

Maho Nakata

Maho Nakata appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

7works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

7 published item(s)

preprint2026arXiv

Tensor Decomposition Networks for Fast Machine Learning Interatomic Potential Computations

$\rm{SO}(3)$-equivariant networks are the dominant models for machine learning interatomic potentials (MLIPs). The key operation of such networks is the Clebsch-Gordan (CG) tensor product, which is computationally expensive. To accelerate the computation, we develop tensor decomposition networks (TDNs) as a class of approximately equivariant networks in which CG tensor products are replaced by low-rank tensor decompositions, such as the CANDECOMP/PARAFAC (CP) decomposition. With the CP decomposition, we prove (i) a uniform bound on the induced error of $\rm{SO}(3)$-equivariance, and (ii) the universality of approximating any equivariant bilinear map. To further reduce the number of parameters, we propose path-weight sharing that ties all multiplicity-space weights across the $\mathcal{O}(L^3)$ CG paths into a single shared parameter set without compromising equivariance, where $L$ is the maximum angular degree. The resulting layer acts as a plug-and-play replacement for tensor products in existing networks, and the computational complexity of tensor products is reduced from $\mathcal{O}(L^6)$ to $\mathcal{O}(L^4)$. We evaluate TDNs on PubChemQCR, a newly curated molecular relaxation dataset containing 105 million DFT-calculated snapshots. We also use existing datasets, including OC20, and OC22. Results show that TDNs achieve competitive performance with dramatic speedup in computations. Our code is publicly available as part of the AIRS library (\href{https://github.com/divelab/AIRS/tree/main/OpenMol/TDN}{https://github.com/divelab/AIRS/}).

preprint2022arXiv

MPLAPACK version 2.0.1 user manual

The MPLAPACK (formerly MPACK) is a multiple-precision version of LAPACK (https://www.netlib.org/lapack/). MPLAPACK version 2.0.1 is based on LAPACK version 3.9.1 and translated from Fortran 90 to C++ using FABLE, a Fortran to C++ source-to-source conversion tool (https://github.com/cctbx/cctbx_project/tree/master/fable/). MPLAPACK version 2.0.1 provides the real and complex version of MPBLAS, and the real and complex versions of MPLAPACK support all LAPACK features: solvers for systems of simultaneous linear equations, least-squares solutions of linear systems of equations, eigenvalue problems, and singular value problems, and related matrix factorizations except for mixed-precision routines. The MPLAPACK defines an API for numerical linear algebra, similar to LAPACK. It is easy to port legacy C/C++ numerical codes using MPLAPACK. MPLAPACK supports binary64, binary128, FP80 (extended double), MPFR, GMP, and QD libraries (double-double and quad-double). Users can choose MPFR or GMP for arbitrary accurate calculations, double-double or quad-double for fast 32 or 64-decimal calculations. We can consider the binary64 version as the C++ version of LAPACK. Moreover, it comes with an OpenMP accelerated version of MPBLAS for some routines and CUDA (A100 and V100 support) for double-double versions of Rgemm and Rsyrk. The peak performances of the OpenMP version are almost proportional to the number of cores, and the performances of the CUDA version are impressive, and approximately 400-600 GFlops. MPLAPACK is available at GitHub (https://github.com/nakatamaho/mplapack/) under the 2-clause BSD license.

preprint2020arXiv

White Paper from Workshop on Large-scale Parallel Numerical Computing Technology (LSPANC 2020): HPC and Computer Arithmetic toward Minimal-Precision Computing

In numerical computations, precision of floating-point computations is a key factor to determine the performance (speed and energy-efficiency) as well as the reliability (accuracy and reproducibility). However, precision generally plays a contrary role for both. Therefore, the ultimate concept for maximizing both at the same time is the minimal-precision computing through precision-tuning, which adjusts the optimal precision for each operation and data. Several studies have been already conducted for it so far (e.g. Precimoniuos and Verrou), but the scope of those studies is limited to the precision-tuning alone. Hence, we aim to propose a broader concept of the minimal-precision computing system with precision-tuning, involving both hardware and software stack. In 2019, we have started the Minimal-Precision Computing project to propose a more broad concept of the minimal-precision computing system with precision-tuning, involving both hardware and software stack. Specifically, our system combines (1) a precision-tuning method based on Discrete Stochastic Arithmetic (DSA), (2) arbitrary-precision arithmetic libraries, (3) fast and accurate numerical libraries, and (4) Field-Programmable Gate Array (FPGA) with High-Level Synthesis (HLS). In this white paper, we aim to provide an overview of various technologies related to minimal- and mixed-precision, to outline the future direction of the project, as well as to discuss current challenges together with our project members and guest speakers at the LSPANC 2020 workshop; https://www.r-ccs.riken.jp/labs/lpnctrt/lspanc2020jan/.

preprint2015arXiv

The PubChemQC Project: a large chemical database from the first principle calculations

In this research, we have been constructing a large database of molecules by {\it ab initio} calculations. Currently, we have over 1.53 million entries of 6-31G* B3LYP optimized geometries and ten excited states by 6-31+G* TDDFT calculations. To calculate molecules, we only refer the InChI (International Chemical Identifier) representation of chemical formula by the International Union of Pure and Applied Chemistry (IUPAC), thus, no reference to experimental data. These results are open to public at http://pubchemqc.riken.jp/. The molecular data have been taken from the PubChem Project (http://pubchem.ncbi.nlm.nih.gov/) which is one of the largest in the world (approximately 63 million molecules are listed) and free (public domain) database. Our final goal is, using these data, to develop a molecular search engine or molecular expert system to find molecules which have desired properties.

preprint2012arXiv

The second-order reduced density matrix method and the two-dimensional Hubbard model

The second-order reduced density matrix method (the RDM method) has performed well in determining energies and properties of atomic and molecular systems, achieving coupled-cluster singles and doubles with perturbative triples (CC SD(T)) accuracy without using the wave-function. One question that arises is how well does the RDM method perform with the same conditions that result in CCSD(T) accuracy in the strong correlation limit. The simplest and a theoretically important model for strongly correlated electronic systems is the Hubbard model. In this paper, we establish the utility of the RDM method when employing the $P$, $Q$, $G$, $T1$ and $T2^\prime$ conditions in the two-dimension al Hubbard model case and we conduct a thorough study applying the $4\times 4$ Hubbard model employing a coefficients. Within the Hubbard Hamilt onian we found that even in the intermediate setting, where $U/t$ is between 4 and 10, the $P$, $Q$, $G$, $T1$ and $T2^\prime$ conditions re produced good ground state energies.

preprint2011arXiv

On the size-consistency of the reduced-density-matrix method and the unitary invariant diagonal $N$-representability conditions

Variational calculation of the ground state energy and its properties using the second-order reduced density matrix (2-RDM) is a promising approach for quantum chemistry. A major obstacle with this approach is that the $N$-representability conditions are too difficult in general. Therefore, we usually employ some approximations such as the $P$, $Q$, $G$, $T1$ and $T2^\prime$ conditions, for realistic calculations. The results of using these approximations and conditions in 2-RDM are comparable to those of CCSD(T). However, these conditions do not incorporate an important property; size-consistency. Size-consistency requires that energies $E(A)$, $E(B)$ and $E(A...B)$ for two infinitely separated systems $A$, $B$, and their respective combined system $A...B$, to satisfy $E(A...B) = E(A) + E(B)$. In this study, we show that the size-consistency can be satisfied if 2-RDM satisfies the following conditions: (i) 2-RDM is unitary invariant diagonal $N$-representable; (ii) 2-RDM corresponding to each subsystem is the eigenstate of the number of corresponding electrons; and (iii) 2-RDM satisfies at least one of the $ P$, $Q$, $G$, $T1$ and $T2^\prime$ conditions.

preprint2011arXiv

Variational approach for the electronic structure calculation on the second-order reduced density matrices and the $N$-representability problem

The reduced-density-matrix method is an promising candidate for the next generation electronic structure calculation method; it is equivalent to solve the Schrödinger equation for the ground state. The number of variables is the same as a four electron system and constant regardless of the electrons in the system. Thus many researchers have been dreaming of a much simpler method for quantum mechanics. In this chapter, we give a overview of the reduced-density matrix method; details of the theories, methods, history, and some new computational results. Typically, the results are comparable to the CCSD(T) which is a sophisticated traditional approach in quantum chemistry.