Vice President, Global Production Operations & Reliability

 Posted 14 hours ago
     
 $195K - $270K per year
  
10+ years experience
Apply Now

Please mention DailyRemote when applying

AI Summary

The Vice President will lead global production operations and reliability strategy to ensure the scalability, security, and performance of the company's cloud-native SaaS platform. This role involves managing global SRE and DRE teams while driving operational excellence across incident management, disaster recovery, and platform engineering.

Everbridge is seeking a Vice President, Global Production Operations & Reliability to lead the strategy, execution, and continuous improvement of our global production operations. Reporting to the Chief Technology Officer, this executive will be responsible for ensuring the reliability, scalability, security, and operational excellence of our cloud-native SaaS platform.

 

Everbridge empowers organizations to keep people safe and operations running during critical events. Our Critical Event Management platform is trusted by enterprises, governments, healthcare providers, financial institutions, and public safety organizations around the world to deliver resilient, mission-critical services when they matter most.

 

This role will lead global Site Reliability Engineering (SRE) and Development & Reliability Engineering (DRE) teams while driving operational excellence across production operations, incident management, change governance, disaster recovery, observability, and platform engineering. The successful candidate will play a key role in strengthening customer trust by delivering highly available, resilient services on a global scale.

 

\n


What you'll do:
  • Define and execute Everbridge's global production operations and reliability strategy.
  • Lead and develop high-performing global SRE and DRE teams responsible for platform reliability and operational engineering.
  • Own the operational excellence of Everbridge's AWS and Kubernetes-based cloud platform, ensuring scalability, resilience, security, and performance.
  • Establish best practices for service reliability, including Service Level Objectives (SLOs), error budgets, production readiness, capacity planning, observability, and operational automation.
  • Drive disciplined incident management, change governance, release management, and post-incident reviews to continuously improve platform stability and reduce operational risk.
  • Lead disaster recovery planning, resilience testing, business continuity initiatives, and operational readiness across global production environments.
  • Partner closely with Engineering, Product, Security, Customer Support, and Customer Success to embed reliability into the software development lifecycle and improve customer outcomes.
  • Champion automation, cloud-native engineering practices, and continuous improvement to enhance operational efficiency and platform performance.
  • Provide executive leadership and reporting on operational health, reliability metrics, customer-impacting incidents, and strategic initiatives.


What you'll bring:
  • 15+ years of experience in production operations, cloud infrastructure, platform engineering, SRE, or related technology leadership roles.
  • Proven experience leading global production operations for large-scale, mission-critical SaaS or cloud platforms.
  • Deep expertise in AWS, Kubernetes, cloud-native architectures, distributed systems, and high-availability environments.
  • Strong understanding of SRE principles, incident management, observability, disaster recovery, change management, CI/CD, and Infrastructure as Code.
  • Demonstrated success building and leading high-performing global engineering and operations teams.
  • Excellent executive communication, stakeholder management, and cross-functional leadership skills.


Preferred Qualifications
  • Experience in mission-critical industries such as public safety, critical communications, healthcare, financial services, security, or enterprise SaaS.
  • Familiarity with modern SRE practices, progressive delivery, platform engineering, service mesh, and cloud cost optimization.
  • Knowledge of regulatory and operational frameworks including ISO 27001, SOC 2, NIST, or FedRAMP.


\n
$195,300 - $270,000 a year
#LI-HG1
 
\n

The reasonably estimated salary for this role at Everbridge ranges from $195,300 - $270,000 and may also include variable compensation. Actual compensation is based on factors such as the candidate's skills, qualifications, and experience. In addition, Everbridge offers a wide range of best in class, comprehensive and inclusive employee benefits for this role including healthcare, dental, parental planning, and mental health benefits, disability income benefits, life and AD&D insurance, a 401(k) plan and match, paid time off, and fitness reimbursements.

 

Fair Chance Statement US & Canada

We are committed to providing equal employment opportunities in compliance with all applicable Federal, Provincial/State and Local laws, including the California Fair Chance Act and any local County Fair Chance Ordinance (or local equivalent). Pursuant to these and other relevant regulations, we consider qualified applicants with criminal histories in a manner consistent with the law.

 

For roles subject to background checks, the following material job duties may be affected by an applicant’s criminal history:

- Access to sensitive or confidential information, such as financial records, proprietary data, or client information.

- Management of cash, company funds, or other valuable assets.

- Work in environments requiring heightened security measures.

- Compliance with contractual or regulatory requirements specific to the position.

 

We evaluate each applicant's criminal history individually, considering its nature, timing, and relevance to the specific job duties, while maintaining our commitment to fair hiring practices and promoting workplace equity.

 

 


About Everbridge

Everbridge empowers enterprises and government organizations to anticipate, mitigate, respond to, and recover from critical events. In an unpredictable world, resilient organizations protect their people and operations, adapt under pressure, and return to productivity faster. Our Critical Event Management (CEM) technology combines intelligent automation with comprehensive risk intelligence to help organizations strengthen resilience, keep people safe, and maintain operations.

 

Learn more at everbridge.com, explore the Everbridge blog, and connect with us on social media.

 

Equal Employment Opportunity

Everbridge is an Equal Opportunity Employer. We consider all qualified applicants for employment without regard to race, color, creed, religion, national origin, ancestry, age, sex, pregnancy, sexual orientation, gender identity or expression, disability, protected veteran status, genetic information, marital status, or any other characteristic protected by applicable federal, state, or local law.

 

Employment Practices

Everbridge does not require or administer lie detector (polygraph) tests as a condition of employment or continued employment. In Massachusetts, it is unlawful to require or administer a lie detector test as a condition of employment or continued employment. Employers who violate this law may be subject to criminal penalties and civil liability.

Similar Jobs

See all Remote Others jobs →

Personalize your Remote Job Search in 3 Easy Steps!

Discover remote opportunities in Others

Answer easy questions

Answer easy questions

200,000+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified