The present work attempts to integrate the independent efforts in the fast N-body commu-nity to create the fastest N-body library for many-core and heterogenous architectures. Focus is placed on low accuracy optimizations, in response to the recent interest to use FMM as a preconditioner for sparse linear solvers. A direct comparison with other state-of-the-art fast N-body codes demonstrates that orders of magnitude increase in performance can be achieved by careful selection of the optimal algorithm and low-level optimization of the code. The cur-rent N-body solver uses a fast multipole method with an efficient strategy for finding the list of cell-cell interactions by a dual tree traversal. A task-based threading model is used to maximize...
Abstract—N-body simulations are computation-intensive ap-plications that calculate the motion of a l...
We are witnessing a dramatic change in computer architecture due to the multicore paradigm shift, as...
General N-body problems are a set of problems in which an update to a single element in the system d...
We present new analysis, algorithmic techniques, and implementations of the Fast Multipole Method (F...
The fast multipole method is an algorithm first developed to approximately solve the N-body problem ...
N-body problems, such as simulating the motion of stars in a galaxy and evaluating the spatial stati...
This work presents the first extensive study of single- node performance optimization, tuning, and a...
In the last two decades, physical constraints in chip design have spawned a paradigm shift in comput...
The integration of the equations of motion of N interacting particles, represents a classical proble...
The N-body problem appears in many computational physics simulations. At each time step the computat...
This thesis presents a top to bottom analysis on designing and implementing fast algorithms for curr...
We present parallel versions of a representative N-body application that uses Greengard and Rokhlin&...
O(N) algorithms for N-body simulations enable the simulation of particle systems with up to 100 mill...
We describe the design of several portable and efficient parallel implementations of adaptive N-body...
The Fast Multipole Method (FMM) is well known to possess a bottleneck arising from decreasing worklo...
Abstract—N-body simulations are computation-intensive ap-plications that calculate the motion of a l...
We are witnessing a dramatic change in computer architecture due to the multicore paradigm shift, as...
General N-body problems are a set of problems in which an update to a single element in the system d...
We present new analysis, algorithmic techniques, and implementations of the Fast Multipole Method (F...
The fast multipole method is an algorithm first developed to approximately solve the N-body problem ...
N-body problems, such as simulating the motion of stars in a galaxy and evaluating the spatial stati...
This work presents the first extensive study of single- node performance optimization, tuning, and a...
In the last two decades, physical constraints in chip design have spawned a paradigm shift in comput...
The integration of the equations of motion of N interacting particles, represents a classical proble...
The N-body problem appears in many computational physics simulations. At each time step the computat...
This thesis presents a top to bottom analysis on designing and implementing fast algorithms for curr...
We present parallel versions of a representative N-body application that uses Greengard and Rokhlin&...
O(N) algorithms for N-body simulations enable the simulation of particle systems with up to 100 mill...
We describe the design of several portable and efficient parallel implementations of adaptive N-body...
The Fast Multipole Method (FMM) is well known to possess a bottleneck arising from decreasing worklo...
Abstract—N-body simulations are computation-intensive ap-plications that calculate the motion of a l...
We are witnessing a dramatic change in computer architecture due to the multicore paradigm shift, as...
General N-body problems are a set of problems in which an update to a single element in the system d...