Jobiglo

No results.

Staff Platform Engineer - High Performance Computing Platform Management

csit · Singapore

New
🇬🇧 English
Slurm Torque LSF SAN NAS Object storage InfiniBand AWS Azure Google Cloud Python Perl Docker Kubernetes Knative Run:AI Grafana Prometheus Kyverno ArgoCD Rancher NVIDIA BCM NVIDIA Superpod architecture

Job description

About the role

You will be part of the dynamic team responsible for building resilient network infrastructure using cutting‑edge technologies such as cloud‑based and software‑defined networking (e.g., SD‑WAN, ACI, NSX). You must have a good understanding of IT infrastructure systems and the latest networking platforms. You will act as a technical specialist, taking on new challenges and staying current with rapidly evolving technology.

Key responsibilities

  • Lead the design, implementation and management of a resilient, scalable and secure HPC platform, including compute nodes, storage, networks and job‑scheduling systems.
  • Plan and manage HPC resource capacity, forecast needs, procure and deploy new hardware and software.
  • Design and implement high‑performance networking solutions such as InfiniBand and Ethernet.
  • Optimize, monitor and troubleshoot cluster performance, and manage job scheduling and resource allocation.
  • Ensure security and compliance through access controls, patches and regular checks.
  • Collaborate with data scientists and developers to optimise application performance on the HPC platform.

Required profile

  • Degree in Computer Science, Computer Engineering or related field.
  • 8+ years experience managing HPC systems on Linux/Unix.
  • Strong knowledge of HPC architectures, clusters, grids and cloud‑based HPC.
  • Experience with job‑scheduling systems such as Slurm, Torque or LSF.
  • Proficiency in storage technologies (SAN, NAS, object storage) and high‑performance networking (InfiniBand, Ethernet).
  • Hands‑on experience with cloud platforms (AWS, Azure, Google Cloud) and scripting languages (Python, Perl, Bash).
  • Experience leading engineering teams.

Required skills

  • Linux/Unix administration
  • HPC cluster management
  • Slurm, Torque, LSF
  • SAN, NAS, object storage
  • InfiniBand, Ethernet networking
  • AWS, Azure, Google Cloud
  • Python, Perl, Bash scripting
  • Docker, Kubernetes
  • Knative, Run:AI, Grafana, Prometheus, Kyverno, ArgoCD, Rancher
  • NVIDIA BCM, NVIDIA Superpod architecture

What we offer

  • Purposeful and meaningful work
  • Collaboration with top engineers
  • Modern technology stack
  • Excellent engineering culture and work‑life balance
  • Focus on engineering and operational excellence
  • Opportunities to innovate and grow together as a family

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec csit.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.
Source : ats:lever

Why are you reporting this job?

Thank you for your report. We will review this job.

Explore further

Salaries, guides and searches in Singapore.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

By continuing, you accept our terms of use.

Already have an account? Login

💬 Chat with us on Telegram Chat on WhatsApp

Published 4 hours ago

Expires 1 month from now

2 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

csit

Singapore