For Employers

Publicis Groupe Holdings B.V

Senior Site Reliability Engineer

Posted 24 days ago
5-10 years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review
AI Summary

Design and implement reliability strategies for distributed systems while leading incident response and root cause analysis. Collaborate with engineering teams to improve system performance, scalability, and operational readiness through automation.

Overview

 Publicis Sapient is seeking a Senior Site Reliability Engineer to help build, operate, and evolve highly scalable, resilient, and secure cloud platforms supporting critical enterprise applications. As part of a large-scale cloud transformation initiative, you will partner closely with Engineering, DevOps, Platform, and Security teams to establish reliability practices, improve operational excellence, and ensure systems meet performance, availability, and scalability objectives. This is a hands-on technical leadership role requiring deep expertise in cloud infrastructure, Kubernetes, observability, incident management, and reliability engineering. You will drive technical decisions, influence engineering practices, and help teams design systems that are resilient by design.

Responsibilities

 

Your Impact

  • Design and implement reliability strategies for distributed systems running across AWS and GCP.
  • Define and measure Service Level Indicators (SLIs), Service Level Objectives (SLOs), and reliability metrics.
  • Build and enhance observability solutions using monitoring, logging, tracing, and alerting platforms.
  • Lead incident response, root cause analysis, and postmortem processes to improve system reliability.
  • Collaborate with engineering teams to improve system performance, resiliency, scalability, and operational readiness.
  • Automate operational processes and reduce toil through engineering solutions.
  • Guide teams on reliability-focused architecture decisions, capacity planning, and non-functional requirements.

Qualifications

Skills & Experience

  • 7+ years of experience in Site Reliability Engineering, Cloud Engineering, DevOps, or Platform Engineering.
  • Strong experience supporting production systems in AWS and/or GCP environments.
  • Deep understanding of SRE principles, including SLIs, SLOs, error budgets, and operational excellence.
  • Experience operating and troubleshooting Kubernetes platforms such as EKS and/or GKE.
  • Strong knowledge of observability tools such as Prometheus, Grafana, CloudWatch, Cloud Monitoring, Datadog, Splunk, or similar.
  • Experience with Infrastructure as Code tools such as Terraform.
  • Strong scripting and automation skills using Python, Bash, or comparable languages.
  • Solid understanding of networking, distributed systems, cloud security, and performance optimization.

Qualifications

Set Yourself Apart With

  • Experience supporting large-scale cloud migration or modernization programs.
  • Expertise in incident management and production operations for high-availability systems.
  • Experience implementing chaos engineering or resilience testing practices.
  • Knowledge of service mesh technologies such as Istio.
  • AWS and/or GCP certifications.
  • Experience working in Agile, DevOps, or DevSecOps environments.

**This position is only available for candidates based in LATAM**

Additional Information

  •  

 

Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

Senior Software Engineer II, Developer Experience / Operational Excellence

Full Time United Kingdom Software Development

Senior Software Engineer(java)

Other, Full Time United States $105K - $130K per year Software Development

Sr. PCBA Aerospace Manufacturing Test Engineer, Amazon Leo

Full Time United States $137K - $185K per year Software Development

ServiceNow Administrator / ITSM Analyst

Freelance United States $48.74 - $71.39 per hour Software Development

Sr Data Scientist II

Full Time Canada 108K - 158K per year Software Development

Sr Data Scientist II

Full Time Canada 108K - 158K per year Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Discover remote opportunities in Site Reliability Engineer

Answer easy questions

Answer easy questions

200,000+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified