The CERN Node is not in production yet. Access to services is not guaranteed, and content and features may still change.
CERN | Accelerating scienceDirectory
EOSC Node CERN
Sign in

Training materials

Browse HEP training materials from the HEP Training catalogue.
Scoped to the CERN EOSC Node Space.

9 materials found

Hepdata Lib

Intermediate

See page: [https://hepdata-lib.readthedocs.io/en/latest/](https://hepdata-lib.readthedocs.io/en/latest/) Library for getting your data into HEPData

Github Repositoryhacktoberfesthepdata
Overview of Rucio – ATLAS Software Tutorial

Intermediate

Soon you will find yourself in the situation that the input/output sandboxes for regular grid jobs are not big enough anymore. You will need to start using distributed mass storage systems on the grid. Rucio coordinates the ATLAS distributed data management systems and provides a convenient way to manage your files. More details about Rucio can be found in the Rucio users guide. The basic unit in Rucio is a data identifier (DID). A DID is nothing but a registered file, dataset (set of files), or container (set of datasets). These DIDs are stored at certain grid sites (like CERN, or BNL) and are registered in a central location (the DDM central catalogues). The logical mapping of files in a dataset to the physical location of the files on these grid sites is done by distributed file catalogues local to a certain site. (Multiple datasets can be aggregated into containers, but we will not cover this right now.)

Tutorialrucioatlasdata management system
How-to articles Rucio

Intermediate

A list of How-to to get started with Rucio: - Deploying OpenSearch monitoring for Rucio - Deploying ElasticSearch monitoring for Rucio - Getting Started - Rucio Command-Line Interface

Tutorialruciodeployopensearchmonitoringelasticsearchcli
Hands-on reinterpretability: HEPData

Intermediate

This is a practical hands-on for analysis reinterpretability, following the first LHC BSM WG general meeting on Nov 10-13, 2025. The aim is to discuss show-case examples of reinterpretability as well of practical questions for specific analyses participants are working on. This specific session on Nov 13 will focus on preparing the HEPData record for your analyses. Part of the workshops of the LHC REI WG.

Eventshepdatareinterpretabilityhands-onexamples
sPHENIX-Collaboration tutorials HEPData

Intermediate

This short tutorial give a example macro on how to generate HEPData submission using hepdata_lib (https://github.com/HEPData/hepdata_lib/)

Tutorialhepdatagithubsubmission
The SWAN Examples

Intermediate

A set of examples for CERN SWAN a Service for Web based ANalysis. With SWAN you can: - Analyse data without the need to install any software - Access via Jupyter notebook interface or a shell from the browser - Use CERNBOX as your home directory and synchronise your local user space with the cloud storage - Access LCG and experiments' software on CVMFS - Share your work with your colleagues thanks to CERNBox - Document and preserve science - create catalogues of analyses: encourage reproducible studies and learning by example

Examplesswananalysesjupyter notebookcernboxcloud
Hadoop Tutorials 2016

Intermediate

Hadoop Tutorials The Hadoop ecosystem is the leading opensource platform for distributed storage and processing of "big data". The Hadoop platform is available at CERN as a central service provided by the IT department. These tutorials organized by the IT Hadoop service, aims to introduce the main concepts about Hadoop technology in a practical way and is targeted to those who would like to start using the service for distributed parallel data processing. Attendees will have the possibility to access a test Hadoop system where they will be able to perform hands-on exercises. Instructions will be provided by the speakers. To facilitate the preparation of the test environment, please register if you plan to attend. [alt text] (https://cern.ch/swanserver/cgi-bin/go?projurl=https://github.com/prasanthkothuri/hadoop-tutorials-2016.git)

Github Repositoryhadoopopen sourcebig dataswan
Training

Intermediate

SWAN training material Physics Analysis Use Case The notebooks in this folder guide through an examplary analysis producing a dimuon spectrum from CMS Open Data, both in C++ and Python. The tools used for the analysis include ROOT's RDataFrame and graphics, numpy, pandas and matplotlib, all of them available in SWAN via the LCG releases on CVMFS. This material is based on the analysis done by Stefan Wunsch, available here in CERN's Open Data portal

Github Repositoryswantraining material
CERN Open Data – CMS Guide to education use of CMS Open Data, Make histograms with collision data

Intermediate

You can get an overview of our education resources through this search query with your keyword of interest (you can change the pre-defined keywords here).

Guidecern open datacmsguidecms open datacollisionshistograms