Showing posts with label Publications. Show all posts
Showing posts with label Publications. Show all posts

Wednesday, February 1, 2012

Fully reviewed journal publications

  • J. Habich, C. Feichtinger, H. Köstler, G. Hager, G. Wellein: Performance engineering for the Lattice Boltzmann method on GPGPUs: Architectural requirements and performance results, Computers & Fluids, Available online 27 February 2012, ISSN 0045-7930, 10.1016/j.compfluid.2012.02.013.
    DOI LINK

  • C. Feichtinger, J. Habich, H. Köstler , G. Hager , U. Rüde, G. Wellein: A flexible Patch-based lattice Boltzmann parallelization approach for heterogeneous GPU–CPU clusters, Parallel Computing, Volume 37, Issue 9, September 2011, Pages 536-549, ISSN 0167-8191, 10.1016/j.parco.2011.03.005.
    DOI LINK

  • J. Habich, T. Zeiser, G. Hager, G. Wellein, Performance analysis and optimization strategies for a D3Q19 Lattice Boltzmann Kernel on nVIDIA GPUs using CUDA, In Special issue of "Advances in Engineering Software and Computers & Structures".
    DOI LINK

  • J. Habich, T. Zeiser, G. Hager, and G. Wellein: Speeding up a Lattice Boltzmann Kernel on nVIDIA GPUs. In Proceedings of PARENG09-S01, the First International Conference on Parallel, Distributed and Grid Computing for Engineering, Pecs, Hungary, April 2009.


Wednesday, June 24, 2009

Lectures

G.Wellein G. Hager, J. Habich: Lecture on Programming Techniques for Supercomputers PTfS, Summer Term 2011.

G. Hager, J. Habich: Tutorials for the lecture on Programming Techniques for Supercomputers PTfS, Summer Term 2011.

G.Wellein G. Hager, J. Habich: Lecture on Programming Techniques for Supercomputers PTfS, Summer Term 2010.

G. Hager, J. Habich: Tutorials for the lecture on Programming Techniques for Supercomputers PTfS, Summer Term 2010.

G. Hager, J. Habich: Tutorials for the lecture on Programming Techniques for Supercomputers PTfS, Summer Term 2009.

Monday, October 13, 2008

Theses


  • Johannes Habich: Performance Evaluation of Numeric Compute Kernels on NVIDIA GPUs, Master's Thesis , RRZE-Erlangen, LSS-Erlangen, 2008.

  • Johannes Habich: Improving computational efficiency of Lattice Boltzmann methods on complex geometries , Bachelor's Thesis , RRZE-Erlangen, LSS-Erlangen, 2006.

Other publications (not fully reviewed)

  • G. Hager, J. Treibig, J. Habich, and G. Wellein: Exploring performance and power properties of modern multicore chips via simple machine models. Submitted. Preprint: arXiv:1208.2908

  • J. Habich, C. Feichtinger, G. Wellein: GPGPU implementation of the LBM: Architectural Requirements and Performance Result,
    Parallel CFD Conference 2011, BSC, Barcelona, Spain, May 2011.

  • G. Wellein, J. Habich, G. Hager, T. Zeiser: Node-level performance of the lattice Boltzmann method on recent multicore CPUs,
    Parallel CFD Conference 2011, BSC, Barcelona, Spain, May 2011.

  • C. Feichtinger, J. Habich, H. Köstler, U. Rüde,  G. Wellein: WaLBerla: Heterogeneous Simulation of Particulate Flows on GPU Clusters,
    Parallel CFD Conference 2011, BSC, Barcelona, Spain, May 2011.

  • J. Habich, C. Feichtinger, G. Hager, G. Wellein: Poster: Parallelizing Lattice Boltzmann Simulations on Heterogeneous GPU&CPU Clusters. 2010 ACM/IEEE International Conference for High Performance Computing, Networking, Storage and Analysis (Supercomputing '10, New Orleans, 13.11. -- 19.11.2010) , 2010.

  • J. Habich, T. Zeiser, G. Hager, G. Wellein: Enabling temporal blocking for a lattice Boltzmann flow solver through multicore aware wavefront parallelization. Parallel CFD Conference 2009, NASA AMES, Moffet Field (CA, USA), Mai, 2009.

  • S. Donath, T. Zeiser, G. Hager, J. Habich, G. Wellein: Optimizing performance of the lattice Boltzmann method for complex geometries on cache-based architectures, (In: F. Hülsemann, M. Kowarschik, U. Rüde (editors), Frontiers in Simulation -- Simulationstechnique, 18th Symposium in Erlangen, September 2005 (ASIM)), SCS Publishing, Fortschritte in der Simulationstechnik, ISBN 3-936150-41-9, (2005) 728-735.

Given or co-authored talks and presentations (see also section on lectures below)


  • J. Habich, C. Feichtinger, G. Wellein, waLBerla: MPI parallele Implementierung eines LBM Lösers auf dem Tsubame 2.0 GPU Cluster, Seminar Talk, Leibniz Rechenzentrum, München, Germany, Feb. 29th 2012.

  • J. Habich, C. Feichtinger, G. Wellein, Hochskalierbarer Lattice Boltzmann Löser für GPGPU Cluster , High Performance Computing Workshop , Leogang, Austria, Feb. 27th 2012.

  • G. Wellein, J.Habich, G. Hager, T. Zeiser, Node-level performance of the lattice Boltzmann method on recent multicore CPUs I,
    Parallel CFD Conference 2011, Barcelona, Spain, May 2011.

  • G. Wellein, J.Habich, G. Hager, T. Zeiser, Node-level performance of the lattice Boltzmann method on recent multicore CPUs II,
    Parallel CFD Conference 2011, Barcelona, Spain, May 2011.

  • J.Habich, C. Feichtinger, G. Wellein, GPGPU implementation of the LBM: Architectural Requirements and Performance Result,
    Parallel CFD Conference 2011, Barcelona, Spain, May 2011.

  • C. Feichtinger, J. Habich, H. Köstler, U. Rüde G. Wellein, WaLBerla: Heterogeneous Simulation of Particulate Flows on GPU Clusters,
    Parallel CFD Conference 2011, Barcelona, Spain, May 2011.

  • J.Habich, Ch. Feichtinger and G. Wellein, GPU optimizations at RRZE,
    invited Talk, ZISC GPU Workshop, Erlangen, Germany, April, 2011.

  • G. Wellein, G. Hager and J.Habich, The Lattice Boltzmann Method: Basic Performance Characteristics and Performance Modeling,
    invited Minisymposia talk, SIAM CSE 2011, Reno, Nevada, USA, March, 2011.

  • J.Habich and Ch. Feichtinger, Performance Optimizations for Heterogeneous and Hybrid 3D Lattice Boltzmann Simulations on Highly Parallel On-Chip Architectures,
    invited Minisymposia talk, SIAM CSE 2011, Reno, Nevada, USA, March, 2011.

  • J.Habich, Ch. Feichtinger, T. Zeiser, G. Wellein, Optimizations on Highly Parallel On-Chip Architectures: GPUs vs. Multi-Core CPUs (for stencil codes),
    iRMB TU-Braunschweig, invited Seminar talk, Braunschweig, Germany, July 2010.

  • J.Habich, Ch. Feichtinger, T. Zeiser, G. Wellein, Optimizations on Highly Parallel On-Chip Architectures: GPUs vs. Multi-Core CPUs (for stencil codes),
    iRMB TU-Braunschweig, invited Seminar talk, Braunschweig, Germany, July 2010.


  • J.Habich, Ch. Feichtinger, T. Zeiser, G. Hager, G. Wellein, Performance Modeling and Optimization for 3D Lattice Boltzmann Simulations on Highly Parallel On-Chip Architectures: GPUs Vs. Multi-Core CPUs,
    ECCOMAS CFD Lisboa, Lisbon, Portugal, June 2010.


  • J.Habich, T. Zeiser, G. Hager, G. Wellein, Performance Modeling and Multicore-aware Optimization for 3D Parallel Lattice Boltzmann Simulations,
    Facing the Multicore-Challenge, Heidelberger Akademie der Wissenschaften, Heidelberg, Germany, March 2010.


  • J. Habich, T. Zeiser, G. Hager, G. Wellein: Performance Evaluation of Numerical Compute Kernels on GPUs,
    First International Workshop on Computational Engineering - Special Topic Fluid-Structure Interaction, Herrsching am Ammersee, Germany, October, 2009.


  • J.Habich, T. Zeiser, G. Hager, G. Wellein: Towards multicore-aware wavefront parallelization of a lattice Boltzmann flow solver,
    5th Erlangen High-End-Computing Symposium, Erlangen, Germany, June 2009.


  • J. Habich, T. Zeiser, G. Hager, G. Wellein: Enabling temporal blocking for a lattice Boltzmann flow solver through multicore-aware wavefront parallelization, submitted to Parallel CFD Conference,
    Moffett Field, California, USA, May 18-22, 2009.


  • J. Habich, T. Zeiser, G. Hager, G. Wellein: Speeding up a Lattice Boltzmann Kernel on nVIDIA GPUs,
    First International Conference on Parallel, Distributed and Grid Computing for Engineering (PARENG09-S01), Pecs, Hungary, April 2009.


  • J. Habich, G. Hager: Erfahrungsbericht Windows HPC in Erlangen,
    WindowsHPC User Group 2nd Meeting, Dresden, Germany, March 2009.


  • J. Habich, G. Hager: Windows CCS im Produktionsbetrieb und erste Erfahrungen mit HPC Server 2008,
    WindowsHPC User Group 1st Meeting, Aachen, Germany, April 2008.

  • T. Zeiser, J. Habich, G. Hager, G. Wellein: Vector computers in a world of commodity clusters, massively parallel systems and many-core many-threaded CPUs: recent experience based on advanced lattice Boltzmann flow solvers,
    HLRS Results and Review Workshop, Stuttgart, Germany, September 2008.

  • S. Donath, T. Zeiser, G. Hager, J. Habich, G. Wellein: On cache-optimized implementations of the lattice Boltzmann method on complex geometries,
    ASIM, Erlangen, Germany, September 2005.

Conference, workshop and tutorial participation without own presentation


  • WindowsHPC User Group 3rd Meeting, St. Augustin, March 2010.

  • WindowsHPC User Group 2nd Meeting, Dresden, March 2009.

  • Introduction to Unified Parallel C (UPC) and Co-array Fortran (CAF) HLRS, October 2008

  • Course on Microfluidics University of Erlangen-Nuremberg Computer Science 10, System Simulation, October 2008

  • IBM Power6 Programming Workshop at RZG, September, 2008

  • PRACE Petascale Summer School (P2S2), Stockholm, Sweden, August, 2008.