Linux - HPC Administrator
ATOS International · Pune, IN · India · On-site
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
About Bull
Bull is a story. One with a century of European innovation and a working environment where experts design powerful, sustainable, and sovereign digital solutions, enabling states and industries to retain full control over their data and their AI.
Bull is also thousands of engineers, researchers and passionate tech people shaping the future of high‑performance computing, AI, and quantum technologies.
Every day, our teams push the boundaries of what is technologically possible – from next‑generation HPC architectures to exascale supercomputers – supported by world‑class R&D, more than 1,600 patents, and unique end‑to‑end capabilities spanning hardware design, software engineering, data science and quantum research.
We are a people‑centric, innovation‑driven company, where collaboration spans Europe, the Americas and India. We share a common vision of a responsible and sustainable innovation that delivers concrete impact for our customers.
We're Hiring: System Engineer (HPC Administration)
Location: Pune / Noida (Onsite) Experience: 3 to 7 Years - Immediate Joiners Work Environment: 24x7 Support Operations
About the Role
We are seeking a System Engineer with hands-on experience in High-Performance Computing (HPC) environments to support, maintain, and optimize critical infrastructure. The ideal candidate will possess strong Linux administration skills, HPC cluster management expertise, and experience managing enterprise-grade servers, storage, and networking technologies.
Main Responsibilities
Administer and support High-Performance Computing (HPC) environments.
Manage Linux-based infrastructure and system administration activities.
Perform hardware troubleshooting and maintenance of high-end servers and storage systems.
Administer and support parallel file systems such as Lustre and/or GPFS .
Manage and troubleshoot InfiniBand (IB) networking technologies.
Support HPC cluster solutions including Torque, Moab, OpenHPC, Bright Cluster , and similar platforms.
Administer workload management and job scheduling tools such as PBS and Slurm .
Develop and maintain shell scripts for automation and operational efficiency.
Support configuration management and automation tools such as Ansible (preferred).
Participate in incident management, root cause analysis, and problem resolution.
Provide support in a 24x7 operational environment as required.
Skills & Qualifications
Required
3 - 7 years of relevant experience in Linux System Administration and/or HPC Administration.
Strong expertise in Linux operating systems and server administration.
Experience with HPC cluster administration (Torque, Moab, OpenHPC, Bright Cluster, etc.).
Hands-on experience troubleshooting high-end servers and storage infrastructure.
Knowledge of Lustre and/or GPFS administration.
Good understanding of InfiniBand (IB) technology.
Experience with workload…