Source author record

Edward F. Valeev

Edward F. Valeev appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
9topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2026arXiv

Tensor Algebra Processing Primitives (TAPP): Towards a Standard for Tensor Operations

To address the absence of a universal standard interface for tensor operations, we introduce the Tensor Algebra Processing Primitives (TAPP), a C-based interface designed to decouple the application layer from hardware-specific implementations. We provide a mathematical formulation of tensor contractions and a reference implementation to ensure correctness and facilitate the validation of optimized kernels. Developed through community consensus involving academic and industrial stakeholders, TAPP aims to enable performance portability and resolving dependency challenges. The viability of the standard is demonstrated through successful integrations with the TBLIS and cuTENSOR libraries, as well as the DIRAC quantum chemistry package.

preprint2022arXiv

Accurate quantum simulation of molecular ground and excited states with a transcorrelated Hamiltonian

NISQ era devices suffer from a number of challenges like limited qubit connectivity, short coherence times and sizable gate error rates. Thus, quantum algorithms are desired that require shallow circuit depths and low qubit counts to take advantage of these devices. We attempt to realize this with the help of classical quantum chemical theories of canonical transformation and explicit correlation. In this work, compact ab initio Hamiltonians are generated classically through an approximate similarity transformation of the Hamiltonian with a) an explicitly correlated two-body unitary operator with generalized pair excitations that remove the Coulombic electron-electron singularities from the Hamiltonian and b) a unitary one-body operator to efficiently capture the orbital relaxation effects required for accurate description of the excited states. The resulting transcorelated Hamiltonians are able to describe both ground and excited states of molecular systems in a balanced manner. Using the fermionic-ADAPT-VQE method based on the unitary coupled cluster with singles and doubles (UCCSD) ansatz and only a minimal basis set (ANO-RCC-MB), we demonstrate that the transcorrelated Hamiltonians can produce ground state energies comparable to the much larger cc-pVTZ basis. This leads to a potential reduction in the number of required CNOT gates by more than three orders of magnitude for the chemical species studied in this work. Furthermore, using the qEOM formalism in conjunction with the transcorrelated Hamiltonian, we reduce the errors in excitation energies by an order of magnitude. The transcorrelated Hamiltonians developed here are Hermitian and contain only one- and two-body interaction terms and thus can be easily combined with any quantum algorithm for accurate electronic structure simulations.

preprint2021arXiv

Quantum simulation of electronic structure with a transcorrelated Hamiltonian: improved accuracy with a smaller footprint on the quantum computer

Quantum simulations of electronic structure with a transformed Hamiltonian that includes some electron correlation effects are demonstrated. The transcorrelated Hamiltonian used in this work is efficiently constructed classically, at polynomial cost, by an approximate similarity transformation with an explicitly correlated two-body unitary operator. This Hamiltonian is Hermitian, includes no more than two-particle interactions, and is free of electron-electron singularities. We investigate the effect of such a transformed Hamiltonian on the accuracy and computational cost of quantum simulations by focusing on a widely used solver for the Schrodinger equation, namely the variational quantum eigensolver method, based on the unitary coupled cluster with singles and doubles (q-UCCSD) Ansatz. Nevertheless, the formalism presented here translates straightforwardly to other quantum algorithms for chemistry. Our results demonstrate that a transcorrelated Hamiltonian, paired with extremely compact bases, produces explicitly correlated energies comparable to those from much larger bases. For the chemical species studied here, explicitly correlated energies based on an underlying 6-31G basis had cc-pVTZ quality. The use of the very compact transcorrelated Hamiltonian reduces the number of CNOT gates required to achieve cc-pVTZ quality by up to two orders of magnitude, and the number of qubits by a factor of three.

preprint2018arXiv

Exploration of Reduced Scaling Formulation of Equation of Motion Coupled-Cluster Singles and Doubles Based on State-Averaged Pair Natural Orbitals

A reduced-complexity variant of equation-of-motion coupled-cluster singles and doubles (EOM-CCSD) method is formulated in terms of state-averaged excited state pair natural orbitals (PNO) designed to describe manifolds of excited states. State-averaged excited state PNOs for the {\em target} manifold are determined by averaging CIS(D) pair densities over the computational manifold. To assess the performance of PNO-EOM-CCSD approach on extended systems the new massively parallel canonical EOM-CCSD program has been developed in the Massively Parallel Quantum Chemistry program that allows treatment of systems with 50+ atoms using realistic basis sets with 1000+ functions. The use of state-averaged PNOs offers several potential advantages relative to the recently proposed state-specific PNOs: our approach is robust with respect to root flipping and state degeneracies, it is more economical when computing large manifolds of states, and it simplifies evaluation of transition-specific observables such as dipole moments. With the PNO truncation threshold of $10^{-7}$, the errors in excitation energies are on average below 0.02 eV for the first six singlet states of 28 organic molecules included in the standard test set of Thiel and co-workers (J. Chem. Phys. 2008, 128, 134110) with 50-70 state-averaged PNOs per pair.

preprint2018arXiv

Optimized pair natural orbitals for the coupled cluster methods

We present the coupled-cluster singles and doubles method formulated in terms of truncated pair-natural orbitals (PNO) that are optimized to minimize the effect of truncation. Compared to the standard ground-state PNO coupled-cluster approaches, in which truncated PNOs derived from first-order Møller-Plesset (MP1) amplitudes are used to compress the CC wave operator, the iteratively-optimized PNOs ("iPNOs") offer moderate improvement for small PNO ranks but rapidly increase their effectiveness for large PNO ranks. The error introduced by PNO truncation in the CCSD energy is reduced by orders of magnitude in the asymptotic regime, with an insignificant increase in PNO ranks. The effect of PNO optimization is particularly effective when combined with Neese's perturbative correction for the PNO incompleteness of the CCSD energy. The use of the perturbative correction in combination with the PNO optimization procedure seems to produce the most precise approximation to the canonical CCSD energies for small and large PNO ranks. For the standard benchmark set of noncovalent binding energies remarkable improvements with respect to standard PNO approach range from a factor of 3 with PNO truncation threshold $τ_\text{PNO}=10^{-6}$ (with the maximum PNO truncation error in the binding energy of only 0.1 kcal/mol) to more than 2 orders of magnitude with $τ_\text{PNO}=10^{-9}$.

preprint2015arXiv

MADNESS: A Multiresolution, Adaptive Numerical Environment for Scientific Simulation

MADNESS (multiresolution adaptive numerical environment for scientific simulation) is a high-level software environment for solving integral and differential equations in many dimensions that uses adaptive and fast harmonic analysis methods with guaranteed precision based on multiresolution analysis and separated representations. Underpinning the numerical capabilities is a powerful petascale parallel programming environment that aims to increase both programmer productivity and code scalability. This paper describes the features and capabilities of MADNESS and briefly discusses some current applications in chemistry and several areas of physics.

preprint2015arXiv

Scalable Task-Based Algorithm for Multiplication of Block-Rank-Sparse Matrices

A task-based formulation of Scalable Universal Matrix Multiplication Algorithm (SUMMA), a popular algorithm for matrix multiplication (MM), is applied to the multiplication of hierarchy-free, rank-structured matrices that appear in the domain of quantum chemistry (QC). The novel features of our formulation are: (1) concurrent scheduling of multiple SUMMA iterations, and (2) fine-grained task-based composition. These features make it tolerant of the load imbalance due to the irregular matrix structure and eliminate all artifactual sources of global synchronization.Scalability of iterative computation of square-root inverse of block-rank-sparse QC matrices is demonstrated; for full-rank (dense) matrices the performance of our SUMMA formulation usually exceeds that of the state-of-the-art dense MM implementations (ScaLAPACK and Cyclops Tensor Framework).

preprint2015arXiv

Task-Based Algorithm for Matrix Multiplication: A Step Towards Block-Sparse Tensor Computing

Distributed-memory matrix multiplication (MM) is a key element of algorithms in many domains (machine learning, quantum physics). Conventional algorithms for dense MM rely on regular/uniform data decomposition to ensure load balance. These traits conflict with the irregular structure (block-sparse or rank-sparse within blocks) that is increasingly relevant for fast methods in quantum physics. To deal with such irregular data we present a new MM algorithm based on Scalable Universal Matrix Multiplication Algorithm (SUMMA). The novel features are: (1) multiple-issue scheduling of SUMMA iterations, and (2) fine-grained task-based formulation. The latter eliminates the need for explicit internodal synchronization; with multiple-iteration scheduling this allows load imbalance due to nonuniform matrix structure. For square MM with uniform and nonuniform block sizes (the latter simulates matrices with general irregular structure) we found excellent performance in weak and strong-scaling regimes, on commodity and high-end hardware.

preprint2014arXiv

A tight distance-dependent estimator for screening three-center Coulomb integrals over Gaussian basis functions

A new estimator for three-center two-particle Coulomb integrals is presented. Our estimator is exact for some classes of integrals and is much more efficient than the standard Schwartz counterpart due to the proper account of distance decay. Although it is not a rigorous upper bound, the maximum degree of underestimation can be controlled by two adjustable parameters. We also give numerical evidence of the excellent tightness of the estimator. The use of the estimator will lead to increased efficiency in reduced-scaling one- and many-body electronic structure theories.

preprint2014arXiv

Fast construction of the exchange operator in an atom-centered basis with concentric atomic density fitting

A linear-scaling algorithm is presented for computing the Hartree-Fock (HF) exchange matrix using concentric atomic density fitting. The algorithm utilizes the stronger distance dependence of the three-center electron repulsion integrals along with the rapid decay of the density matrix to accelerate the construction of the exchange matrix. The new algorithm is tested with computations on systems with up to 1536 atoms and 15585 basis functions, the latter of which represents, to our knowledge, the largest quadruple-zeta HF computation ever performed. Our method handles screening of high angular momentum contributions in a particularly efficient manner, allowing the use of larger basis sets for large molecules without a prohibitive increase in cost.