Match your resume skills with our AI powered skill match!
The Senior Site Reliability Engineer will build, transition, and operate a multi-cloud production platform while ensuring system reliability and scalability. They will also develop automation, manage GitOps workflows, and troubleshoot complex issues across cloud and Kubernetes environments.
Job Title:
Senior Site Reliability EngineerJob Description
We're Concentrix. The intelligent transformation partner. Solution-focused. Tech-powered. Intelligence-fueled.We’re looking for a hands-on Senior Site Reliability Engineer to help build, transition, and operate a multi-cloud production platform supporting GPO products. This role is ideal for someone who enjoys working across cloud infrastructure, Kubernetes, GitOps, reliability, and production improvement—bringing strong engineering fundamentals, sound operational judgment, and a passion for building resilient, scalable platforms.
Responsibilities
Partner with GPO architects and planners to ensure target platform designs are deployable, secure, observable, supportable, and recoverable in production.
Assess and document the current multi-cloud and Kubernetes estate across Azure, GCP, AWS, and GitOps environments, identifying dependencies, migration needs, risks, and operational gaps.
Build, maintain, and enhance cloud infrastructure and shared platform services using infrastructure as code, peer-reviewed workflows, and safe production change practices.
Operate and improve AKS, GKE, and EKS environments, including networking, identity, upgrades, scalability, availability, recovery, and shared platform capabilities.
Develop and support GitOps and CI/CD workflows that make platform changes repeatable, reviewable, observable, and easy to validate and roll back.
Troubleshoot production issues across cloud, network, Kubernetes, GitOps, database, and shared platform layers, while collaborating effectively across team boundaries.
Reduce operational toil through automation and continuously improve SLOs, alerts, dashboards, runbooks, disaster recovery procedures, and cost controls.
Complete all assigned, mandatory training within the timeframe provided.
Conduct and/or participate in regularly scheduled 1:1 meetings with your direct manager and/or direct reports.
Qualifications
Strong Linux and networking fundamentals, with practical understanding of how distributed production systems behave and fail.
Hands-on experience in at least one key area such as Azure, GCP, AWS, Kubernetes, infrastructure as code, GitOps, CI/CD, observability, or reliability engineering.
Ability to build, test, review, and maintain infrastructure automation using tools or languages such as Terraform/OpenTofu, Terragrunt, Ansible, Python, Go, shell scripting, or Kubernetes configuration.
Experience with Git-based workflows, peer review, CI validation, rollback planning, and safe production change management.
Demonstrated troubleshooting and incident response skills, with a structured, risk-aware approach to solving unfamiliar technical problems.
Familiarity with managed Kubernetes platforms such as AKS, GKE, or EKS, and exposure to shared platform services including secrets, certificates, private connectivity, messaging, caching, or observability tooling.
Experience supporting brownfield platform environments, staged migrations, and database technologies such as MongoDB, PostgreSQL, or MySQL is preferred.
Strong written and verbal English communication skills, with the ability to collaborate effectively across global teams and meet requirements for privileged production access.
#li-remote
Location:
BGR Work-at-HomeLanguage Requirements:
Time Type:
Full timeStop the endless job search. Our AI finds and applies to the best jobs for you.
Discover remote opportunities in Site Reliability Engineer
Answer easy questions
200,000+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”