Linux & HPC Systems Engineer
Enregistrez cette offre et organisez votre recherche
Créez un compte gratuit pour enregistrer des offres d'emploi, créer des alertes et revenir à cette liste depuis votre tableau de bord.
Are you passionate about Linux, automation, and high-performance computing? Do you enjoy solving complex technical challenges while supporting cutting-edge research? Join imec as a Linux & HPC Systems Engineer and play a key role in operating, improving, and future‑proofing our Linux and High-Performance Computing environments.
In this role, you will help ensure the reliability, security, and performance of critical research computing platforms used by leading scientists and engineers. You will collaborate closely with researchers, infrastructure specialists, and technology partners while contributing to the automation, modernization, and evolution of our HPC ecosystem.
What you will do
As part of the operations team, you will help ensure the smooth day-to-day functioning of our Linux and HPC environments.
You Will:
- Provide second-line support for Linux and HPC platforms.
- Diagnose and resolve system, performance, scheduling, and connectivity issues.
- Administer Linux servers and HPC infrastructure in production environments.
- Perform system patching, security hardening, and compliance-related activities.
- Monitor platform health, capacity, availability, and performance.
- Contribute to incident, problem, and change management processes.
- Create and maintain high‑quality operational documentation, procedures, and runbooks
HPC Platform Management
You will help operate and further enhance our HPC services, enabling researchers to maximize the value of computational resources.
You Will:
- Manage and develop HPC platforms based on Slurm and Open OnDemand.
- Support workload scheduling, project allocations, priorities, and reservations.
- Assist users with batch processing, interactive workloads, and scientific applications.
- Support HPC software environments, containerized workloads, and FlexLM‑based license services.
- Contribute to capacity planning, utilization reporting, and software license management.
Automation & Continuous Improvement
We believe automation is key to operational excellence.
You Will:
- Automate repetitive operational tasks using Ansible, Python, and shell scripting.
- Support infrastructure‑as‑code practices and version‑controlled platform configuration.
- Participate in platform upgrades, migrations, and transformative HPC projects.
- Contribute to the development of hybrid HPC capabilities.
- Share knowledge within the team and actively improve platform reliability, scalability, and user experience.
What we do for you
We offer you the opportunity to join one of the world’s premier research centers in nanotechnology at its headquarters in Leuven, Belgium. With your talent, passion and expertise, you’ll become part of a team that makes the impossible possible. Together, we shape the technology that will determine the society of tomorrow.
We are committed to being an inclusive employer and proud of our open, multicultural, and informal working environment with ample possibilities to take initiative and show responsibility. We commit to supporting and guiding you in this process; not only with words but also with tangible actions. Through imec.academy, 'our corporate university', we actively invest in your development to further your technical and personal growth.
We are aware that your valuable contribution makes imec a top player in its field. Your energy and commitment are therefore appreciated by means of a market appropriate salary with many fringe benefits.
Who you are
Required Experience & Skills
- Hands‑on experience administering Linux systems in production environments.
- Strong knowledge of Red Hat Enterprise Linux or comparable Linux distributions.
- Experience with Python and shell scripting.
- Experience with Ansible or similar automation technologies.
- Familiarity with monitoring and observability tools such as Prometheus, Grafana, and Signoz.
- Familiarity with container technologies including Apptainer/Singularity, Docker, or OpenShift.
- Strong analytical and troubleshooting abilities.
- Excellent communication and documentation skills.
- A collaborative, customer‑focused, and service‑oriented mindset.
Nice to Have
- Experience with Slurm or another HPC workload scheduler.
- Familiarity with Open OnDemand or NoMachine.
- Experience supporting scientific, engineering, or research computing environments.
- Knowledge of FlexLM, EDA applications, or HPC software lifecycle management.
- Experience with Azure or hybrid HPC environments.
- Familiarity with ITIL processes and best practices.
Language
- Professional proficiency