Source author record

Olivier Coulaud

Olivier Coulaud appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

8works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

8 published item(s)

preprint2014arXiv

Asymptotic profiles for the second grade fluids equations in R^2

In the present paper, we study the long time behaviour of the solutions of the second grade fluids equations in dimension 3. Using scaling variables and energy estimates in weighted Sobolev spaces, we describe the first order asymptotic profiles of these solutions. In particular, we show that the solutions of the second grade fluids equations converge to self-similar solutions of the heat equations, which are explicit and depend on the initial data. Since this phenomenon occurs also for the Navier-Stokes equations, it shows that the fluids of second grade behave asymptotically like Newtonian fluids.

preprint2014arXiv

Asymptotic profiles for the third grade fluids equations

We study the long time behaviour of the solutions of the third grade fluids equations in dimension 2. Introducing scaled variables and performing several energy estimates in weighted Sobolev spaces, we describe the first order of an asymptotic expansion of these solutions. It shows in particular that, under smallness assumptions on the data, the solutions of the third grade fluids equations converge to self-similar solutions of the heat equations, which can be computed explicitly from the data.

preprint2013arXiv

Deflation and augmentation techniques in Krylov subspace methods for the solution of linear systems

In this paper we present deflation and augmentation techniques that have been designed to accelerate the convergence of Krylov subspace methods for the solution of linear systems of equations. We review numerical approaches both for linear systems with a non-Hermitian coefficient matrix, mainly within the Arnoldi framework, and for Hermitian positive definite problems with the conjugate gradient method.

preprint2013arXiv

Extensions of the siesta dft code for simulation of molecules

We describe extensions to the siesta density functional theory (dft) code [30], for the simulation of isolated molecules and their absorption spectra. The extensions allow for: - Use of a multi-grid solver for the Poisson equation on a finite dft mesh. Non-periodic, Dirichlet boundary conditions are computed by expansion of the electric multipoles over spherical harmonics. - Truncation of a molecular system by the method of design atom pseudo- potentials of Xiao and Zhang[32]. - Electrostatic potential fitting to determine effective atomic charges. - Derivation of electronic absorption transition energies and oscillator stren- gths from the raw spectra produced by a recently described, order O(N3), time-dependent dft code[21]. The code is furthermore integrated within siesta as a post-processing option.

preprint2012arXiv

Optimized M2L Kernels for the Chebyshev Interpolation based Fast Multipole Method

A fast multipole method (FMM) for asymptotically smooth kernel functions (1/r, 1/r^4, Gauss and Stokes kernels, radial basis functions, etc.) based on a Chebyshev interpolation scheme has been introduced in [Fong et al., 2009]. The method has been extended to oscillatory kernels (e.g., Helmholtz kernel) in [Messner et al., 2012]. Beside its generality this FMM turns out to be favorable due to its easy implementation and its high performance based on intensive use of highly optimized BLAS libraries. However, one of its bottlenecks is the precomputation of the multiple-to-local (M2L) operator, and its higher number of floating point operations (flops) compared to other FMM formulations. Here, we present several optimizations for that operator, which is known to be the costliest FMM operator. The most efficient ones do not only reduce the precomputation time by a factor up to 340 but they also speed up the matrix-vector product. We conclude with comparisons and numerical validations of all presented optimizations.

preprint2012arXiv

Pipelining the Fast Multipole Method over a Runtime System

Fast Multipole Methods (FMM) are a fundamental operation for the simulation of many physical problems. The high performance design of such methods usually requires to carefully tune the algorithm for both the targeted physics and the hardware. In this paper, we propose a new approach that achieves high performance across architectures. Our method consists of expressing the FMM algorithm as a task flow and employing a state-of-the-art runtime system, StarPU, in order to process the tasks on the different processing units. We carefully design the task flow, the mathematical operators, their Central Processing Unit (CPU) and Graphics Processing Unit (GPU) implementations, as well as scheduling schemes. We compute potentials and forces of 200 million particles in 48.7 seconds on a homogeneous 160 cores SGI Altix UV 100 and of 38 million particles in 13.34 seconds on a heterogeneous 12 cores Intel Nehalem processor enhanced with 3 Nvidia M2090 Fermi GPUs.

preprint2010arXiv

A Parallel Iterative Method for Computing Molecular Absorption Spectra

We describe a fast parallel iterative method for computing molecular absorption spectra within TDDFT linear response and using the LCAO method. We use a local basis of "dominant products" to parametrize the space of orbital products that occur in the LCAO approach. In this basis, the dynamical polarizability is computed iteratively within an appropriate Krylov subspace. The iterative procedure uses a a matrix-free GMRES method to determine the (interacting) density response. The resulting code is about one order of magnitude faster than our previous full-matrix method. This acceleration makes the speed of our TDDFT code comparable with codes based on Casida's equation. The implementation of our method uses hybrid MPI and OpenMP parallelization in which load balancing and memory access are optimized. To validate our approach and to establish benchmarks, we compute spectra of large molecules on various types of parallel machines. The methods developed here are fairly general and we believe they will find useful applications in molecular physics/chemistry, even for problems that are beyond TDDFT, such as organic semiconductors, particularly in photovoltaics.

preprint2010arXiv

Fast construction of the Kohn--Sham response function for molecules

The use of the LCAO (Linear Combination of Atomic Orbitals) method for excited states involves products of orbitals that are known to be linearly dependent. We identify a basis in the space of orbital products that is local for orbitals of finite support and with a residual error that vanishes exponentially with its dimension. As an application of our previously reported technique we compute the Kohn--Sham density response function $χ_{0}$ for a molecule consisting of $N$ atoms in $N^{2}N_ω$ operations, with $N_ω$ the number of frequency points. We test our construction of $χ_{0}$ by computing molecular spectra directly from the equations of Petersilka--Gossmann--Gross in $N^{2}N_ω$ operations rather than from Casida's equations which takes $N^{3}$ operations. We consider the good agreement with previously calculated molecular spectra as a validation of our construction of $χ_{0}$. Ongoing work indicates that our method is well suited for the computation of the GW self-energy $Σ=\mathrm{i}GW$ and we expect it to be useful in the analysis of exitonic effects in molecules.