Source author record

Markus Rampp

Markus Rampp appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
12topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2026arXiv

FROST-CLUSTERS -- III. Metallicity-dependent intermediate mass black hole formation by runaway collisions in dense star clusters

We explore the formation of intermediate mass black holes (IMBHs), potential seeds for supermassive black holes (SMBHs), via runaway stellar collisions for a wide range of star cluster (surface) densities ($4\times10^3 M_\odot$ pc$^{-2} \lesssim Σ_\mathrm{h}$ $\lesssim 4\times10^6 M_\odot$ pc$^{-2}$) and metallicities ($0.01 Z_\odot \lesssim Z \lesssim 1.0 Z_\odot)$. Our sample of isolated $(>1400)$ and hierarchical ($30$) simulations of young, massive star clusters with up to $N=1.8\times10^6$ stars includes collisional stellar dynamics, stellar evolution, and post-Newtonian equations of motion for black holes using the BIFROST code. High stellar wind rates suppress IMBH formation at high metallicities $(Z \gtrsim 0.2 Z_\odot)$ and low collision rates prevent their formation at low densities ($Σ_\mathrm{h} \lesssim 3\times10^4 M_\odot$ pc$^{-2}$). The assumptions about stellar wind loss rates strongly affect the maximum final IMBH masses $(M_\bullet \sim 6000 M_\odot$ vs. $25000 M_\odot$). The total stellar mass loss from collisions and collisionally boosted winds before $t=3$ Myr can together reach up to 5-10% of the final cluster mass. We present fitting formulae for IMBH masses as a function of host star cluster $Σ_\mathrm{h}$ and Z, and formulate a model for the cosmic IMBH formation rate density. Depending on the cluster birth densities, the IMBH formation rates peak at $z\sim2$-$4$ at up to $\sim10^{-7}$ yr$^{-1}$cMpc$^{-3}$. As more than 50% form below $z\lesssim1.5$-$3$, the model challenges a view in which all local IMBHs are failed early Universe SMBH seeds.

preprint2022arXiv

An Efficient Particle Tracking Algorithm for Large-Scale Parallel Pseudo-Spectral Simulations of Turbulence

Particle tracking in large-scale numerical simulations of turbulent flows presents one of the major bottlenecks in parallel performance and scaling efficiency. Here, we describe a particle tracking algorithm for large-scale parallel pseudo-spectral simulations of turbulence which scales well up to billions of tracer particles on modern high-performance computing architectures. We summarize the standard parallel methods used to solve the fluid equations in our hybrid MPI/OpenMP implementation. As the main focus, we describe the implementation of the particle tracking algorithm and document its computational performance. To address the extensive inter-process communication required by particle tracking, we introduce a task-based approach to overlap point-to-point communications with computations, thereby enabling improved resource utilization. We characterize the computational cost as a function of the number of particles tracked and compare it with the flow field computation, showing that the cost of particle tracking is very small for typical applications.

preprint2022arXiv

Importance of accurate consideration of the electron inertia in hybrid-kinetic simulations of collisionless plasma turbulence: 1. The 2D limit

The dissipation mechanism of the magnetic energy in turbulent collisionless space and astrophysical plasmas is still not well understood. Its investigation requires efficient kinetic simulations of the energy transfer in collisionless plasma turbulence. In this respect, hybrid-kinetic simulations, in which ions are treated as particles and electrons as an inertial fluid, have begun to attract a significant interest recently. Hybrid-kinetic models describe both ion- and electron scale processes by ignoring electron kinetic effects so that they are computationally much less demanding compared to fully kinetic plasma models. Hybrid-kinetic codes solve either the Vlasov equation for the ions (Eulerian Vlasov-hybrid codes) or the equations of motion of the ions as macro-particles (Lagrangian Particle-in-Cell (PIC)-hybrid codes). They consider the inertia of the electron fluid using different approximations. We check the validity of these approximations by employing our recently massively parallelized three-dimensional PIC-hybrid code CHIEF which considers the electron inertia without any of the common approximations. In particular we report the results of simulations of two-dimensional collisionless plasma turbulence. We conclude that the simulation results obtained using hybrid-kinetic codes which use approximations to describe the electron inertia need to be interpreted with caution. We also discuss the parallel scalability of CHIEF, to the best of our knowledge, the first PIC-hybrid code which without approximations describes the inertial electron fluid.

preprint2021arXiv

All-electron periodic $G_0W_0$ implementation with numerical atomic orbital basis functions: algorithm and benchmarks

We present an all-electron, periodic {\GnWn} implementation within the numerical atomic orbital (NAO) basis framework. A localized variant of the resolution-of-the-identity (RI) approximation is employed to significantly reduce the computational cost of evaluating and storing the two-electron Coulomb repulsion integrals. We demonstrate that the error arising from localized RI approximation can be reduced to an insignificant level by enhancing the set of auxiliary basis functions, used to expand the products of two single-particle NAOs. An efficient algorithm is introduced to deal with the Coulomb singularity in the Brillouin zone sampling that is suitable for the NAO framework. We perform systematic convergence tests and identify a set of computational parameters, which can serve as the default choice for most practical purposes. Benchmark calculations are carried out for a set of prototypical semiconductors and insulators, and compared to independent reference values obtained from an independent $G_0W_0$ implementation based on linearized augmented plane waves (LAPW) plus high-energy localized orbitals (HLOs) basis set, as well as experimental results. With a moderate (FHI-aims \textit{tier} 2) NAO basis set, our $G_0W_0$ calculations produce band gaps that typically lie in between the standard LAPW and the LAPW+HLO results. Complementing \textit{tier} 2 with highly localized Slater-type orbitals (STOs), we find that the obtained band gaps show an overall convergence towards the LAPW+HLO results. The algorithms and techniques developed in this work pave the way for efficient implementations of correlated methods within the NAO framework.

preprint2021arXiv

Convolutional neural network-assisted recognition of nanoscale L12 ordered structures in face-centred cubic alloys

Nanoscale L12-type ordered structures are widely used in face-centred cubic (FCC) alloys to exploit their hardening capacity and thereby improve mechanical properties. These fine-scale particles are typically fully coherent with matrix with the same atomic configuration disregarding chemical species, which makes them challenging to be characterized. Spatial distribution maps (SDMs) are used to probe local order by interrogating the three-dimensional (3D) distribution of atoms within reconstructed atom probe tomography (APT) data. However, it is almost impossible to manually analyse the complete point cloud ($>10$ million) in search for the partial crystallographic information retained within the data. Here, we proposed an intelligent L12-ordered structure recognition method based on convolutional neural networks (CNNs). The SDMs of a simulated L12-ordered structure and the FCC matrix were firstly generated. These simulated images combined with a small amount of experimental data were used to train a CNN-based L12-ordered structure recognition model. Finally, the approach was successfully applied to reveal the 3D distribution of L12-type $δ^\prime$-Al3(LiMg) nanoparticles with an average radius of 2.54 nm in a FCC Al-Li-Mg system. The minimum radius of detectable nanodomain is even down to 5 Å. The proposed CNN-APT method is promising to be extended to recognize other nanoscale ordered structures and even more-challenging short-range ordered phenomena in the near future.

preprint2020arXiv

NECI: N-Electron Configuration Interaction with emphasis on state-of-the-art stochastic methods

We present NECI, a state-of-the-art implementation of the Full Configuration Interaction Quantum Monte Carlo algorithm, a method based on a stochastic application of the Hamiltonian matrix on a sparse sampling of the wave function. The program utilizes a very powerful parallelization and scales efficiently to more than 24000 CPU cores. In this paper, we describe the core functionalities of NECI and recent developments. This includes the capabilities to calculate ground and excited state energies, properties via the one- and two-body reduced density matrices, as well as spectral and Green's functions for ab initio and model systems. A number of enhancements of the bare FCIQMC algorithm are available within NECI, allowing to use a partially deterministic formulation of the algorithm, working in a spin-adapted basis or supporting transcorrelated Hamiltonians. NECI supports the FCIDUMP file format for integrals, supplying a convenient interface to numerous quantum chemistry programs and it is licensed under GPL-3.0.

preprint2020arXiv

nsCouette -- A high-performance code for direct numerical simulations of turbulent Taylor-Couette flow

We present nsCouette, a highly scalable software tool to solve the Navier-Stokes equations for incompressible fluid flow between differentially heated and independently rotating, concentric cylinders. It is based on a pseudospectral spatial discretization and dynamic time-stepping. It is implemented in modern Fortran with a hybrid MPI-OpenMP parallelization scheme and thus designed to compute turbulent flows at high Reynolds and Rayleigh numbers. An additional GPU implementation (C-CUDA) for intermediate problem sizes and a basic version for turbulent pipe flow (nsPipe) are also provided.

preprint2019arXiv

Octopus, a computational framework for exploring light-driven phenomena and quantum dynamics in extended and finite systems

Over the last years extraordinary advances in experimental and theoretical tools have allowed us to monitor and control matter at short time and atomic scales with a high-degree of precision. An appealing and challenging route towards engineering materials with tailored properties is to find ways to design or selectively manipulate materials, especially at the quantum level. To this end, having a state-of-the-art ab initio computer simulation tool that enables a reliable and accurate simulation of light-induced changes in the physical and chemical properties of complex systems is of utmost importance. The first principles real-space-based Octopus project was born with that idea in mind, providing an unique framework allowing to describe non-equilibrium phenomena in molecular complexes, low dimensional materials, and extended systems by accounting for electronic, ionic, and photon quantum mechanical effects within a generalized time-dependent density functional theory framework. The present article aims to present the new features that have been implemented over the last few years, including technical developments related to performance and massive parallelism. We also describe the major theoretical developments to address ultrafast light-driven processes, like the new theoretical framework of quantum electrodynamics density-functional formalism (QEDFT) for the description of novel light-matter hybrid states. Those advances, and other being released soon as part of the Octopus package, will enable the scientific community to simulate and characterize spatial and time-resolved spectroscopies, ultrafast phenomena in molecules and materials, and new emergent states of matter (QED-materials).

preprint2016arXiv

GPEC, a real-time capable Tokamak equilibrium code

A new parallel equilibrium reconstruction code for tokamak plasmas is presented. GPEC allows to compute equilibrium flux distributions sufficiently accurate to derive parameters for plasma control within 1 ms of runtime which enables real-time applications at the ASDEX Upgrade experiment (AUG) and other machines with a control cycle of at least this size. The underlying algorithms are based on the well-established offline-analysis code CLISTE, following the classical concept of iteratively solving the Grad-Shafranov equation and feeding in diagnostic signals from the experiment. The new code adopts a hybrid parallelization scheme for computing the equilibrium flux distribution and extends the fast, shared-memory-parallel Poisson solver which we have described previously by a distributed computation of the individual Poisson problems corresponding to different basis functions. The code is based entirely on open-source software components and runs on standard server hardware and software environments. The real-time capability of GPEC is demonstrated by performing an offline-computation of a sequence of 1000 flux distributions which are taken from one second of operation of a typical AUG discharge and deriving the relevant control parameters with a time resolution of a millisecond. On current server hardware the new code allows employing a grid size of 32x64 zones for the spatial discretization and up to 15 basis functions. It takes into account about 90 diagnostic signals while using up to 4 equilibrium iterations and computing more than 20 plasma-control parameters, including the computationally expensive safety-factor q on at least 4 different levels of the normalized flux.

preprint2014arXiv

A Hybrid MPI-OpenMP Parallel Implementation for pseudospectral simulations with application to Taylor-Couette Flow

A hybrid-parallel direct-numerical-simulation method with application to turbulent Taylor-Couette flow is presented. The Navier-Stokes equations are discretized in cylindrical coordinates with the spectral Fourier-Galerkin method in the axial and azimuthal directions, and high-order finite differences in the radial direction. Time is advanced by a second-order, semi-implicit projection scheme, which requires the solution of five Helmholtz/Poisson equations, avoids staggered grids and renders very small slip velocities. Nonlinear terms are computed with the pseudospectral method. The code is parallelized using a hybrid MPI-OpenMP strategy, which is simpler to implement, reduces inter-node communications and is more efficient compared to a flat MPI parallelization. A strong scaling study shows that the hybrid code maintains very good scalability up to more than 20000 processor cores and thus allows to perform simulations at higher resolutions than previously feasible, and opens up the possibility to simulate turbulent Taylor-Couette flows at Reynolds numbers up to $\mathcal{O}(10^5)$. This enables to probe hydrodynamic turbulence in Keplerian flows in experimentally relevant regimes.

preprint2014arXiv

Towards Petaflops Capability of the VERTEX Supernova Code

The VERTEX code is employed for multi-dimensional neutrino-radiation hydrodynamics simulations of core-collapse supernova explosions from first principles. The code is considered state-of-the-art in supernova research and it has been used for modeling for more than a decade, resulting in numerous scientific publications. The computational performance of the code, which is currently deployed on several high-performance computing (HPC) systems up to the Tier-0 class (e.g. in the framework of the European PRACE initiative and the German GAUSS program), however, has so far not been extensively documented. This paper presents a high-level overview of the relevant algorithms and parallelization strategies and outlines the technical challenges and achievements encountered along the evolution of the code from the gigaflops scale with the first, serial simulations in 2000, up to almost petaflops capabilities, as demonstrated lately on the SuperMUC system of the Leibniz Supercomputing Centre (LRZ). In particular, we shall document the parallel scalability and computational efficiency of VERTEX at the large scale and on the major, contemporary HPC platforms. We will outline upcoming scientific requirements and discuss the resulting challenges for the future development and operation of the code.

preprint2013arXiv

Porting Large HPC Applications to GPU Clusters: The Codes GENE and VERTEX

We have developed GPU versions for two major high-performance-computing (HPC) applications originating from two different scientific domains. GENE is a plasma microturbulence code which is employed for simulations of nuclear fusion plasmas. VERTEX is a neutrino-radiation hydrodynamics code for "first principles"-simulations of core-collapse supernova explosions. The codes are considered state of the art in their respective scientific domains, both concerning their scientific scope and functionality as well as the achievable compute performance, in particular parallel scalability on all relevant HPC platforms. GENE and VERTEX were ported by us to HPC cluster architectures with two NVidia Kepler GPUs mounted in each node in addition to two Intel Xeon CPUs of the Sandy Bridge family. On such platforms we achieve up to twofold gains in the overall application performance in the sense of a reduction of the time to solution for a given setup with respect to a pure CPU cluster. The paper describes our basic porting strategies and benchmarking methodology, and details the main algorithmic and technical challenges we faced on the new, heterogeneous architecture.