Match your resume skills with our AI powered skill match!
You will own production uptime, infrastructure cost, and incident response for core platform services. Additionally, you will build and maintain scalable AWS infrastructure while developing CI/CD pipelines and internal developer tooling.
This is a backend-architecture-heavy platform engineering role responsible for the reliability, scale, performance, and developer experience of core infrastructure and systems. You'll join a tight-knit engineering team of ~15 and own everything from production uptime and incident response to CI/CD and cost efficiency — work that directly shapes how fast and dependably the platform can be built on and operated.
Own production uptime, latency, provisioning speed, infrastructure cost, and incident response for core platform services.
Build and maintain AWS infrastructure using Terraform, Kubernetes/EKS, Helm, Docker, EC2, CodeBuild, ECR, S3, IAM, networking, and secrets management.
Design and improve backend and platform systems for scale, including capacity planning, autoscaling, queueing, backpressure, cleanup jobs, retries, and rollback paths.
Define and improve dashboards, alerts, logs, traces, SLOs, runbooks, and on-call workflows so failures are detected, debugged, and resolved quickly.
Build reliable CI/CD pipelines, release automation, environment management, and deployment workflows that improve developer productivity and reduce production risk.
Write clean, maintainable code to automate systems, improve backend services, and create internal developer tooling.
2–4 years of experience owning production cloud infrastructure for a high-availability, user-facing platform, with responsibility for uptime, performance, deployment safety, and cost.
Deep hands-on experience with AWS and containerized systems; strong familiarity with Terraform, Kubernetes/EKS, Docker, EC2, networking, load balancers, and secrets management.
Track record of building or operating CI/CD, release automation, observability, alerting, and incident response systems.
Strong backend engineering judgment — able to reason about service architecture, APIs, databases, async systems, queues, scaling limits, and production failure modes.
Experience designing for bursty workloads, long-running jobs, sandboxed execution, distributed workers, or high-concurrency services.
Background operating infrastructure for data-heavy, ML/AI, workflow, marketplace, developer-tools, or enterprise platforms is a strong plus.
Demonstrated focus on reducing cloud spend through better architecture, autoscaling, workload placement, caching, or cleanup systems.
Comfortable writing production-quality code — this is not a pure ops role.
Salary range: $150,000 – $250,000 USD annually. Visa sponsorship is available.
On-site in San Francisco, CA. Singapore-based candidates are also welcome (on-site). Candidates based elsewhere — particularly in Europe — may be considered as fully remote independent contractors.
Stop the endless job search. Our AI finds and applies to the best jobs for you.
Discover remote opportunities in Platform Engineer
Answer easy questions
200,000+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”