The SRE Delivery Manager will lead a team of six across two pods to oversee infrastructure reliability, deployment, and operational support. This role involves a mix of management and hands-on technical work, including defining best practices and managing on-call rotations.
Why this role, why now
Ninety is at an inflection point: the product is evolving quickly and the business is ready to level up how we take new capabilities to market. We’re opening a new SRE/Delivery Manager seat for someone who can build repeatable systems and raise the bar for what great and reliable systems look like in a lean, high-impact environment.
What you’ll own
Manage and lead a team of 6, organized into two pods, covering both SRE and Delivery
Provide career development and mentorship through continuous checkins and quarterly reviews
Drive independent workstreams for each pod while providing reliable infrastructure, deployment, and operational support to the broader engineering team and business
Define and implement best practices for delivery, infrastructure reliability, and observability in a lean, high-impact environment
Coordinate with application development team (software engineers, designers, and product owners) to ensure deployment, infrastructure, and monitoring needs are met
Manage on-call rotation and scheduling
Ensure infrastructure and systems operate securely, following industry best practices
Fulfill a player-coach role, anticipating 80% management and 20% hands on
What Success Looks Like:
Systems and infrastructure operate efficiently with the bounds of our SLA’s
Production incidents are properly managed and routed to responsible parties to be mitigated and resolved within expected timeframes
Engineering pods are equipped with tools needed to develop and iterate quickly
We routinely scan and identify security vulnerabilities and address appropriately
We are keeping up to date with industry standards, best practices, and technology as it pertains to our infrastructure and delivery processes
What we’re looking for
Relevant work experience:
8+ years of experience in software development, with focus on SRE/devops
Demonstrated success leading teams through architectural transitions (e.g., serverless to EKS).
Backend service architecture experience is a plus (monolith, microservices, etc).
Skills:
Proven experience managing engineering teams
Deep understanding of Infrastructure as Code (IaC)
Observability best practices
Ability to drive outcomes in a lean, fast-paced environment.
Technologies:
Expertise in the AWS stack, specifically Lambda/serverless and Kubernetes/EKS
Strong proficiency in Datadog (APM/RUM/Synthetics)
GitHub; experience with Harness is a plus.
Rootly (or other major on-call, paging toolsets)
Feature-flagging & experimentation platforms like Launch Darkly
This won’t be a fit if you prefer
Pure strategy with a separate team to “do the doing”
Highly resourced environments where systems already exist
Narrow, order-taking roles vs. owning outcomes
Defining the how instead of what and why
Slow-paced, siloed teams
Micromanaging vs providing coaching and direction to your direct reports
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”