Jobiglo

No results.

Senior Software Engineer, Infrastructure

clera · Singapore

New
Onsite Senior 150,000 - 250,000 USD/year 🇬🇧 English
AWS Terraform Kubernetes EKS Helm Docker EC2 CodeBuild ECR S3 IAM Networking Secrets management CI/CD Observability Alerting Incident response

Job description

About the role

This is a backend‑architecture‑heavy platform engineering role sitting within a tight‑knit engineering team of roughly 15 people. You will own the reliability, scale, performance, and developer experience of core infrastructure and systems for an AI/ML evaluation and reinforcement learning platform. The work has direct, measurable impact on how fast, reliable, and cost‑effective the platform is to build on and operate.

Key responsibilities

  • Own production uptime, latency, provisioning speed, infrastructure cost, and incident response for core platform services.
  • Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Helm, Docker, EC2, CodeBuild, ECR, S3, IAM, networking, and secrets management.
  • Design and improve backend and platform systems for scale, covering capacity planning, autoscaling, queueing, backpressure, cleanup jobs, retries, and rollback paths.
  • Define and improve dashboards, alerts, logs, traces, SLOs, runbooks, and on‑call workflows so failures are detected, debugged, and resolved quickly.
  • Build reliable CI/CD pipelines, release automation, environment management, and deployment workflows that improve developer productivity and reduce production risk.
  • Write clean, maintainable code to automate systems, improve backend services, and create internal developer tooling.

Required profile

  • 2‑4 years of experience owning production cloud infrastructure for a high‑availability, user‑facing platform.
  • Deep hands‑on experience with AWS and containerized systems; Terraform, Kubernetes/EKS, Docker, EC2, networking, load balancers, and secrets management.
  • Track record of building or operating CI/CD, release automation, observability, alerting, and incident response systems.
  • Strong backend engineering judgment across service architecture, APIs, databases, async systems, queues, scaling limits, and production failure modes.
  • Experience designing systems for bursty workloads, long‑running jobs, sandboxed execution, distributed workers, or high‑concurrency services.
  • Demonstrated focus on reducing cloud spend through better architecture, autoscaling, workload placement, caching, cleanup systems, or observability.

Required skills

  • AWS
  • Terraform
  • Kubernetes
  • EKS
  • Helm
  • Docker
  • EC2
  • CodeBuild
  • ECR
  • S3
  • IAM
  • Networking
  • Secrets management
  • CI/CD
  • Observability
  • Alerting
  • Incident response

What we offer

  • Salary range $150,000–$250,000 USD annually.
  • Visa sponsorship is available.

Questions fréquentes

Le salaire proposé pour ce poste est de 150-250k USD par an. Le détail figure dans l'annonce.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.
Source : ats:ashby

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

By continuing, you accept our terms of use.

Already have an account? Login

💬 Chat with us on Telegram Chat on WhatsApp

Published 3 hours ago

Expires 1 month from now

4 views · 0 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

clera

Singapore