8 October 2026 to 18 February 2027
Europe/Brussels timezone

Session

Day 5: Organise your work and storage on HPC

19 Nov 2026, 09:30
CYCL09a

CYCL09a

chemin du cyclotron 2, 1348 Louvain-La-Neuve

Presentation materials

There are no materials yet.

  1. Dr Olivier Mattelaer (UCLouvain/CISM)
    19/11/2026, 09:30

    Checkpointing and restarting, or the art of stopping computations and resuming them later, is a powerful technique for overcoming job time limits, surviving hardware or software failures, and improving the robustness of long-running computations on HPC clusters. This session introduces the main checkpointing approaches and teaches how to use them effectively on CÉCI clusters.

    | Contents |...

    Go to contribution page
  2. Damien François (UCLouvain/CISM)
    19/11/2026, 11:15

    Whenever one has to deal with multiple jobs on an HPC system, the idea of automating parts or all of the job management process naturally leads to the concept of workflows. Workflow management solutions range from basic scheduler features such as job arrays and dependencies to sophisticated workflow engines capable of coordinating large-scale computational campaigns. This session helps...

    Go to contribution page
  3. Ariel Lozano (ULB)
    19/11/2026, 13:30

    Data storage on CÉCI clusters differs from that of a standalone workstation. Thanks to the shared filesystem infrastructure, users can seamlessly access the same files from multiple clusters, making it easier to move computations, collaborate with colleagues, and manage data across the entire CÉCI ecosystem. This session explains how storage is organized, how to use it efficiently, and how to...

    Go to contribution page
Building timetable...