Installing, Running and Maintaining Large Linux Clusters at CERN

dc.creatorBahyl, Vladimir
dc.creatorChardi, Benjamin
dc.creatorvan Eldik, Jan
dc.creatorFuchs, Ulrich
dc.creatorKleinwort, Thorsten
dc.creatorMurth, Martin
dc.creatorSmith, Tim
dc.date2003-06-12
dc.date.accessioned2026-07-07T03:19:50Z
dc.date.available2026-07-07T03:19:50Z
dc.descriptionHaving built up Linux clusters to more than 1000 nodes over the past five years, we already have practical experience confronting some of the LHC scale computing challenges: scalability, automation, hardware diversity, security, and rolling OS upgrades. This paper describes the tools and processes we have implemented, working in close collaboration with the EDG project [1], especially with the WP4 subtask, to improve the manageability of our clusters, in particular in the areas of system installation, configuration, and monitoring. In addition to the purely technical issues, providing shared interactive and batch services which can adapt to meet the diverse and changing requirements of our users is a significant challenge. We describe the developments and tuning that we have introduced on our LSF based systems to maximise both responsiveness to users and overall system utilisation. Finally, this paper will describe the problems we are facing in enlarging our heterogeneous Linux clusters, the progress we have made in dealing with the current issues and the steps we are taking to gridify the clusters
dc.description5 pages, Proceedings for the CHEP 2003 conference, La Jolla, California, March 24 - 28, 2003
dc.identifierhttps://arxiv.org/abs/cs/0306058
dc.identifierhttp://arxiv.org/abs/cs/0306058
dc.identifier.urihttp://salesiana.dossiersoluciones.com/handle/123456789/31622
dc.subjectDistributed, Parallel, and Cluster Computing
dc.subjectC.5.3; K.6.3
dc.titleInstalling, Running and Maintaining Large Linux Clusters at CERN
dc.typetext

Files

Collections