Source author record

Thomas C. Schulthess

Thomas C. Schulthess appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2022arXiv

Performance optimizations for porting the openQ$^\star$D package to GPUs

OpenQ$^\star$D code has been used by the RC$^\star$ collaboration for the generation of fully dynamical QCD+QED gauge configurations with C$^\star$ boundary conditions. In this talk, optimization of solvers provided with the openQ$^\star$D package relevant for porting the code on GPU-accelerated supercomputing platforms is discussed. We present the analysis of the current implementations of the GCR solver preconditioned with Schwarz alternating procedure for ill-conditioned Dirac-operators. With the goal of enabling support for GPUs from various vendors, a novel method of adaptive CPU/GPU-hybrid implementation is proposed.

preprint2020arXiv

Continuous momentum dependence in the dynamical cluster approximation

The dynamical cluster approximation (DCA) is a quantum cluster extension to the single-site dynamical mean-field theory that incorporates spatially nonlocal dynamic correlations systematically and nonperturbatively. The DCA$^+$ algorithm addresses the cluster shape dependence of the DCA and improves the convergence with cluster size by introducing a lattice self-energy with continuous momentum dependence. However, we show that the DCA$^+$ algorithm is plagued by a fundamental problem when its self-consistency equations are formulated using the bare Green's function of the cluster. This problem is most severe in the strongly correlated regime at low doping, where the DCA$^+$ self-energy becomes overly metallic and local, and persists to cluster sizes where the standard DCA has long converged. In view of the failure of the DCA$^+$ algorithm, we propose to complement DCA simulations with a post-interpolation procedure for single-particle and two-particle correlation functions to preserve continuous momentum dependence and the associated benefits in the DCA. We demonstrate the effectiveness of this practical approach with results for the half-filled and hole-doped two-dimensional Hubbard model.

preprint2016arXiv

All-electron self-consistent GW in the Matsubara-time domain: implementation and benchmarks of semiconductors and insulators

The GW approximation is a well-known method to improve electronic structure predictions calculated within density functional theory. In this work, we have implemented a computationally efficient GW approach that calculates central properties within the Matsubara-time domain using the modified version of Elk, the full-potential linearized augmented plane wave (FP-LAPW) package. Continuous-pole expansion (CPE), a recently proposed analytic continuation method, has been incorporated and compared to the widely used Pade approximation. Full crystal symmetry has been employed for computational speedup. We have applied our approach to 18 well-studied semiconductors/insulators that cover a wide range of band gaps computed at the levels of single-shot G0W0, partially self-consistent GW0, and fully self-consistent GW (scGW). Our calculations show that G0W0 leads to band gaps that agree well with experiment for the case of simple s-p electron systems, whereas scGW is required for improving the band gaps in 3-d electron systems. In addition, GW0 almost always predicts larger band gap values compared to scGW, likely due to the substantial underestimation of screening effects. Both the CPE method and Pade approximation lead to similar band gaps for most systems except strontium titantate, suggesting further investigation into the latter approximation is necessary for strongly correlated systems. Our computed band gaps serve as important benchmarks for the accuracy of the Matsubara-time GW approach.

preprint2014arXiv

All-Electron GW Quasiparticle Band Structures of Group 14 Nitride Compounds

We have investigated the group 14 nitrides (M$_3$N$_4$) in the spinel phase ($γ$-M$_3$N$_4$ with M= C, Si, Ge and Sn) and $β$ phase ($β$-M$_3$N$_4$ with M= Si, Ge and Sn) using density functional theory with the local density approximation and the GW approximation. The Kohn-Sham energies of these systems have been first calculated within the framework of full-potential linearized augmented plane waves and then corrected using single-shot G$_0$W$_0$ calculations, which we have implemented in the modified version of the Elk full-potential LAPW code. Direct band gaps at the $Γ$ point have been found for spinel-type nitrides $γ$-M$_3$N$_4$ with M= Si, Ge and Sn. The corresponding GW-corrected band gaps agree with experiment. We have also found that the GW calculations with and without the plasmon-pole approximation give very similar results, even when the system contains semi-core $d$ electrons. These spinel-type nitrides are novel materials for potential optoelectronic applications because of their direct and tunable band gaps.

preprint2014arXiv

First Experiences With Validating and Using the Cray Power Management Database Tool

In October 2013 CSCS installed the first hybrid Cray XC-30 system, dubbed Piz Daint. This system features the power management database (PMDB), that was recently introduced by Cray to collect detailed power consumption information in a non-intrusive manner. Power measurements are taken on each node, with additional measurements for the Aries network and blowers, and recorded in a database. This enables fine-grained reporting of power consumption that is not possible with external power meters, and is useful to both application developers and facility operators. This paper will show how benchmarks of representative applications at CSCS were used to validate the PMDB on Piz Daint. Furthermore we will elaborate, with the well-known HPL benchmark serving as prototypical application, on how the PMDB streamlines the tuning for optimal power efficiency in production, which lead to Piz Daint being recognised as the most energy efficient petascale supercomputer presently in operation.

preprint2013arXiv

DCA$^+$: Dynamical Cluster Approximation with continuous lattice self-energy

The dynamical cluster approximation (DCA) is a systematic extension beyond the single site approximation in dynamical mean field theory (DMFT), to include spatially non-local correlations in quantum many-body simulations of strongly correlated systems. We extend the DCA with a continuous lattice self-energy in oder to achieve better convergence with cluster size. The new method, which we call DCA$^+$, cures the cluster shape dependence problems of the DCA, without suffering from causality violations of previous attempts to interpolate the cluster self-energy. A practical approach based on standard inference techniques is given to deduce the continuous lattice self-energy from an interpolated cluster self-energy. We study the pseudogap region of a hole-doped two-dimensional Hubbard model and find that in the DCA$^+$ algorithm, the self-energy and pseudo-gap temperature $T^*$ converge monotonously with cluster size. Introduction of a continuous lattice self-energy eliminates artificial long-rage correlations and thus significantly reduces the sign problem of the quantum Monte Carlo cluster solver in the DCA$^+$ algorithm compared to the normal DCA. Simulations with much larger cluster sizes thus become feasible, which, along with the improved convergence in cluster size, raises hope that precise extrapolations to the exact infinite cluster size limit can be reached for other physical quantities as well.

preprint2012arXiv

A hybrid Hermitian general eigenvalue solver

The adoption of hybrid GPU-CPU nodes in traditional supercomputing platforms opens acceleration opportunities for electronic structure calculations in materials science and chemistry applications, where medium sized Hermitian generalized eigenvalue problems must be solved many times. The small size of the problems limits the scalability on a distributed memory system, hence they can benefit from the massive computational performance concentrated on a single node, hybrid GPU-CPU system. However, new algorithms that efficiently exploit heterogeneity and massive parallelism of not just GPUs, but of multi/many-core CPUs as well are required. Addressing these demands, we implemented a novel Hermitian general eigensolver algorithm. This algorithm is based on a standard eigenvalue solver, and existing algorithms can be used. The resulting eigensolvers are state-of-the-art in HPC, significantly outperforming existing libraries. We analyze their performance impact on applications of interest, when different fractions of eigenvectors are needed by the host electronic structure code.

preprint2004arXiv

Phase transitions in ferro-antiferromagnetic bilayers with a stepped interface

We have studied magnetic ordering in ferro/antiferromagnetic (F/AF) bilayers using Monte Carlo simulations of classical Heisenberg spins. For both flat and stepped interfaces we observed order in the AF above the Neel temperature, with the AF spins aligning collinearly with the F moments. In the case of the stepped interface there is a transition from collinear to perpendicular alignment of the F and AF spins at a lower temperature.