For Employers

Range

Senior Infrastructure & Operations Engineer (Kubernetes / Platform Reliability)

Posted 21 days ago
$75000 - $130K per year
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

You will design, deploy, and maintain production Kubernetes clusters while ensuring system reliability, observability, and security. Additionally, you will build automated CI/CD pipelines and optimize infrastructure performance, including CDN and edge configurations.

We’re looking for a senior infrastructure and operations engineer to own and evolve our platform reliability. You’ll design, operate, and maintain our Kubernetes-based infrastructure, build reliable monitoring and alerting pipelines, and ensure our systems remain stable under real-world load and failure conditions. This is a hands-on role for someone with deep experience running production systems at scale and who focuses on making infrastructure predictable and stable. You’ll work across Kubernetes, networking, CI/CD, Cloudflare, and observability to create a platform engineers can trust.
What You’ll Do

  • Design, deploy, and maintain production Kubernetes clusters.

  • Own cluster reliability, upgrades, security, and performance.

  • Build and operate monitoring, logging, and alerting pipelines.

  • Ensure full-stack observability across infrastructure and services.

  • Design and maintain CI/CD pipelines that are fast, reproducible, and safe.

  • Improve deployment strategies (rollouts, canaries, rollbacks).

  • Automate infrastructure provisioning and configuration.

  • Investigate and resolve production incidents.

  • Improve system resilience, redundancy, and recovery strategies.

  • Define SLOs/SLIs and track reliability targets.

  • Optimize and maintain our Cloudflare setup (caching, routing, security, edge behavior).

  • Work closely with engineering teams to improve operational practices.

  • Identify and remove single points of failure.

What We’re Looking For
Must-have

  • Senior-level experience operating production infrastructure.

  • Deep, hands-on expertise with Kubernetes (cluster internals, networking, storage, security).

  • Strong networking fundamentals (TCP/IP, routing, DNS, TLS, load balancing).

  • Experience debugging distributed systems and network-related issues.

  • Experience optimizing CDN and edge setups, including Cloudflare.

  • Strong experience building monitoring and observability systems.

  • Experience with metrics, logs, traces, and alerting pipelines.

  • Experience designing reliable CI/CD pipelines.

  • Strong Linux fundamentals.

  • Experience with infrastructure as code and automation.

  • Ability to debug issues across the entire stack.

  • Experience handling incidents and conducting postmortems.

Nice-to-have

  • Experience with multi-cluster or multi-region setups.

  • Experience with high-throughput or data-heavy systems.

  • Experience with Elasticsearch or large-scale data infrastructure.

  • Experience with service meshes.

  • Experience with cost optimization and capacity planning.

  • Experience in regulated or reliability-focused environments.

How You Work

  • You assume infrastructure will fail and design accordingly.

  • You prioritize reliability, visibility, and recoverability.

  • You build systems that engineers trust in production.

  • You automate carefully and deliberately.

  • You are calm and methodical during incidents.

  • You focus on long-term stability over short-term fixes.

  • You document and standardize important processes.

Example Problems You Might Work On

  • Hardening Kubernetes clusters for high availability and safe upgrades.

  • Debugging network latency or connectivity issues across services.

  • Optimizing Cloudflare caching, routing, and edge security rules.

  • Building monitoring pipelines that provide reliable signals.

  • Designing alerting that is actionable and low-noise.

  • Improving deployment reliability and rollback safety.

  • Removing single points of failure in production systems.

  • Ensuring observability across all services and data pipelines.

Why Join Range

  • Join one of the fastest-growing sectors in Web3 as stablecoins reach mass adoption.

  • Competitive compensation with meaningful equity upside.

  • Strong potential for growth and leadership opportunities.

  • Remote-first culture with bi-yearly international off-sites.

  • Opportunities for global conference travel and ecosystem engagement.

  • Health and wellness benefits.

How to Apply
Send us:

  • A short introduction and your background.

  • Examples of infrastructure or platform work you’ve led.

  • Any public write-ups, repositories, or talks (if available).

  • We’re particularly interested in engineers who have built and operated Kubernetes platforms, improved network reliability, and optimized CDN/edge setups such as Cloudflare in production.

Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

Senior Salesforce Developer (Remote)

Full Time India Software Development

Senior Project Manager - Cloud & Compliance Advisory

Full Time United States $85000 - $141K per year Software Development

Senior Spectrum Engineer

Full Time Canada, United States 151K - 176K per year Software Development

Senior Software Application Developer

Full Time Canada 110K - 140K per year Software Development

Application Developer - Senior

Full Time United States $120K - $175K per year Software Development

Quality Assurance Manager - Senior

Full Time United States $120K - $165K per year Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 219,042+ Jobs in Software Development

Answer easy questions

Answer easy questions

219,042+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified