About Tapestry Health
Tapestry Health is transforming post-acute and long-term care through innovative technology-enabled healthcare solutions. Our mission is to improve patient outcomes, empower clinicians, and help healthcare organizations deliver exceptional care at scale.
As we continue to grow, we are investing in a modern technology organization built on reliability, operational excellence, automation, security, and data-driven decision making.
Position Summary
The Technology Operations Lead is responsible for ensuring the reliability, scalability, security, and operational effectiveness of Tapestry Health's technology platforms and services.
This leader will build and mature the Technology Operations function, including Site Reliability Engineering (SRE), Infrastructure & Cloud Operations, Application Support, Release Management, Incident Management, Help Desk, and Disaster Recovery.
The ideal candidate is a hands-on technology leader who combines strong operational discipline with a modern DevOps, AI and cloud-first mindset. They will partner closely with Software Engineering, Data Engineering, Security, Product, and Business Operations to deliver a highly available, secure, and scalable healthcare technology platform.
Key Responsibilities
Production Reliability & Operations
- Own availability, stability, and operational health of all production platforms.
- Establish and lead a Site Reliability Engineering (SRE) capability.
- Define and manage SLAs, SLOs, and operational KPIs.
- Create operational dashboards and executive reporting.
- Drive observability, monitoring, alerting, and capacity planning initiatives.
- Reduce service interruptions through proactive operational practices.
- Establish, improve, maintain, and monitor resiliency, recovery and availability metrics.
Incident Management
- Build and operationalize an enterprise Incident Management framework.
- Lead major incident response and post-incident reviews.
- Establish root-cause analysis processes and corrective action tracking.
- Develop communication protocols for technology incidents and customer-impacting events.
Infrastructure & Cloud Operations
- Lead management of on-prem and cloud infrastructure, networking, storage, backup, and recovery capabilities.
- Drive automation through Infrastructure as Code (IaC) and cloud-native operational practices.
- Ensure systems are scalable, secure, and compliant with healthcare regulatory requirements.
- Optimize operational cost and cloud resource utilization
Release & Environment Management
- Establish formal release management processes.
- Partner with Engineering to improve CI/CD maturity and deployment reliability.
- Establish Environment strategy and maintain parity.
- Coordinate release readiness reviews and deployment planning.
- Implement change governance and production safeguards.
Product Support & Service Delivery
- Build and lead Product Support and Application Support functions.
- Define support operating models, escalation paths, and service ownership.
- Improve customer issue resolution processes.
- Partner with Customer Success and Clinical Operations teams on customer-impacting issues.
- Establish support metrics and continuous improvement programs.
Disaster Recovery & Business Continuity
- Own disaster recovery planning, testing, and operational execution.
- Establish recovery objectives and testing schedules.
- Coordinate business continuity planning across Technology.
- Lead resilience and platform hardening initiatives.
Technology Governance & Metrics
- Develop and maintain operational scorecards.
- Define technology performance metrics and executive dashboards.
- Establish governance processes for operational reviews.
- Drive a culture of accountability, measurement, and continuous improvement.
Leadership & Team Development
- Recruit, mentor, and develop a high-performing operations organization.
- Foster a culture of reliability, ownership, collaboration, and customer focus.
- Build organizational capability across:
- Infrastructure
- SRE
- Application Support
- Help Desk
- Release Management
- Technology Project Management
- Network Operations
- Cloud and On-prem Operations
Required Qualifications
- 10+ years of experience in Technology Operations, Infrastructure, SRE, DevOps, or IT Operations leadership roles.
- 5+ years leading operational teams in a SaaS, healthcare technology, or regulated environment.
- Experience building or scaling Technology Operations organizations.
- Strong knowledge of:
- Cloud platforms (Azure preferred)
- Site Reliability Engineering
- Incident Management
- Infrastructure Operations
- IT Service Management (ITSM)
- Release Management
- Disaster Recovery
- Observability and Monitoring
- Experience implementing operational metrics and executive dashboards.
- Strong cross-functional leadership and communication skills.
- Ability to operate in a fast-growing, transformation-focused organization.
Preferred Qualifications
- Experience in healthcare technology, post-acute care, RPM, CCM, value-based care, or clinical operations.
- Experience supporting HIPAA, SOC 2, HITRUST, or other regulated environments.
- Experience implementing SRE practices at scale.
- Familiarity with Azure DevOps, GitLab, CI/CD pipelines, and DevSecOps practices.
- Experience leading cloud modernization and operational transformation initiatives.
- AI-first mindset with experience using automation and AI-assisted operational workflows.
This remote position follows a location-based compensation structure. The posted salary range is from 200-225k and represents the potential pay range across various U.S. geographic markets. Actual compensation will be determined based on the candidate’s primary work location, experience, qualifications, and internal equity considerations, in accordance with applicable pay transparency laws.