Senior Software Engineer, Infrastructure
clera · Singapore · $150,000 to $250,000
Open. First seen 17 September 2026.
Description
About the Role
This is a backend-architecture-heavy platform engineering role sitting within a tight-knit engineering team of roughly 15 people. You will own the reliability, scale, performance, and developer experience of core infrastructure and systems for an AI/ML evaluation and reinforcement learning platform. The work has direct, measurable impact on how fast, reliable, and cost-effective the platform is to build on and operate.
What You'll Do
Own production uptime, latency, provisioning speed, infrastructure cost, and incident response for core platform services. Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Helm, Docker, EC2, CodeBuild, ECR, S3, IAM, networking, and secrets management. Design and improve backend and platform systems for scale, covering capacity planning, autoscaling, queueing, backpressure, cleanup jobs, retries, and rollback paths.
Define and improve dashboards, alerts, logs, traces, SLOs, runbooks, and on-call workflows so failures are detected, debugged, and resolved quickly. Build reliable CI/CD pipelines, release automation, environment management, and deployment workflows that improve developer productivity and reduce production risk. Write clean, maintainable code to automate systems, improve backend services, and create internal developer tooling.
What We're Looking For
2 to 4 years of experience owning production cloud infrastructure for a high-availability, user-facing platform, with responsibility for uptime, performance, deployment safety, and cost. Deep hands-on experience with AWS and containerized systems; Terraform, Kubernetes/EKS, Docker, EC2, networking, load balancers, and secrets management strongly preferred. A track record of building or operating CI/CD, release automation, observability, alerting, and incident response systems.
Strong backend engineering judgment across service architecture, APIs, databases, async systems, queues, scaling limits, and production failure modes. Experience designing systems for bursty workloads, long-running jobs, sandboxed execution, distributed workers, or high-concurrency services. Background operating infrastructure for AI/ML, data-heavy, marketplace, workflow, developer-tools, or enterprise platforms.
Demonstrated focus on reducing cloud spend through better architecture, autoscaling, workload placement, caching, cleanup systems, or observability. Ability to write clean, production-quality code; generic DevOps or infrastructure-only backgrounds without backend software engineering depth are not a strong fit. Compensation & Benefits Salary range: $150,000 to $250,000 USD annually.
Visa sponsorship is available. Location On-site in Singapore . Candidates based in San Francisco are also considered for on-site work there.
Candidates outside these locations, particularly in Europe, may be considered as fully remote independent contractors.
ABOUT THE ROLE
This is a backend-architecture-heavy platform engineering role sitting within a tight-knit engineering team of roughly 15 people. You will own the reliability, scale, performance, and developer experience of core infrastructure and systems for an AI/ML evaluation and reinforcement learning platform. The work has direct, measurable impact on how fast, reliable, and cost-effective the platform is to build on and operate.
WHAT YOU'LL DO
- Own production uptime, latency, provisioning speed, infrastructure cost, and incident response for core platform services.
- Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Helm, Docker, EC2, CodeBuild, ECR, S3, IAM, networking, and secrets management.
- Design and improve backend and platform systems for scale, covering capacity planning, autoscaling, queueing, backpressure, cleanup jobs, retries, and rollback paths.
- Define and improve dashboards, alerts, logs, traces, SLOs, runbooks, and on-call workflows so failures are detected, debugged, and resolved quickly.
- Build reliable CI/CD pipelines, release automation, environment management, and deployment workflows that improve developer productivity and reduce production risk.
- Write clean, maintainable code to automate systems, improve backend services, and create internal developer tooling.
WHAT WE'RE LOOKING FOR
- 2 to 4 years of experience owning production cloud infrastructure for a high-availability, user-facing platform, with responsibility for uptime, performance, deployment safety, and cost.
- Deep hands-on experience with AWS and containerized systems; Terraform, Kubernetes/EKS, Docker, EC2, networking, load balancers, and secrets management strongly preferred.
- A track record of building or operating CI/CD, release automation, observability, alerting, and incident response systems.
- Strong backend engineering judgment across service architecture, APIs, databases, async systems, queues, scaling limits, and production failure modes.
- Experience designing systems for bursty workloads, long-running jobs, sandboxed execution, distributed workers, or high-concurrency services.
- Background operating infrastructure for AI/ML, data-heavy, marketplace, workflow, developer-tools, or enterprise platforms.
- Demonstrated focus on reducing cloud spend through better architecture, autoscaling, workload placement, caching, cleanup systems, or observability.
- Ability to write clean, production-quality code; generic DevOps or infrastructure-only backgrounds without backend software engineering depth are not a strong fit. COMPENSATION & BENEFITS Salary range: $150,000 to $250,000 USD annually. Visa sponsorship is available. LOCATION On-site in Singapore. Candidates based in San Francisco are also considered for on-site work there. Candidates outside these locations, particularly in Europe, may be considered as fully remote independent contractors.
Other openings for this role
2 open postings for this role in Singapore, first posted 17 September 2026. The company is likely hiring for more than one position. None has been taken down and posted again.
- 25 September 2026 – open
- 17 September 2026 – open (this posting)
Salary context
9 other open postings titled Senior Software Engineer, Infrastructure state a salary: median USD 175,000 to USD 230,000 a year.
Similar jobs
- Senior Software Engineer, Infrastructure at CLEAR - Corporate
- Senior Software Engineer, Infrastructure at hud
- Senior Software Engineer, Infrastructure at skydio
- Senior Software Engineer, Infrastructure at siftstack
- Senior Software Engineer, Infrastructure at siftstack
- Senior Software Engineer, Infrastructure at artemis
Is this your posting and you want it taken down? Email info@careerholo.com.