Who We Are
Magnit is the future of work. Serving hundreds of the world’s most recognizable brands for the past 30+ years, Magnit offers the industry’s first holistic platform for the modern workforce. Magnit's integrated workforce management (IWM) platform supported by data, software, intelligence, and best-in-class services team is key to our clients’ success. It can adapt quickly to regional or industry economic shifts, and provides the speed, scale, flexibility, transparency, and expertise required to meet an organization’s contingent workforce management, talent strategy and broader organization goals. At Magnit, you’ll work with passionate colleagues who collaborate and deliver meaningful results that positively transform the largest companies around the globe.
About the Role
The Senior Director, Cloud Platforms will own Magnit's migration from on-premise infrastructure into the public cloud and build the cloud platform that our engineering organization runs on. This is a hands-on leadership role: you will set the strategy, then work shoulder-to-shoulder with your teams in the code, the pipelines, and the Kubernetes manifests. You will own the data center exit roadmap, the AWS-first landing zones underneath it, and the paved-road developer platform (GitLab CI/CD, Argo CD, EKS, Terraform) that makes doing the right thing the easy thing for product teams. Reporting to the Chief Information Officer, you will partner across Engineering, Security, Data, and Operations to move Magnit to an API-first, cloud-native architecture while improving reliability, security posture, and unit economics along the way.
What You Will Do
Cloud Migration & Data Center Exit
- Own the end-to-end on-premise to public cloud migration roadmap — wave planning, sequencing, readiness validation, cutover, and stabilization — with clear milestones, exit criteria, and success metrics.
- Lead application portfolio assessment and disposition (rehost, replatform, refactor, retire, retain), building the business case and target architecture for each wave.
- Design and operate secure, multi-account cloud foundations: landing zones, identity, networking, compute, storage, and Kubernetes.
- Act as the integration layer across infrastructure, application engineering, security, networking, finance, and vendor partners to keep migration waves on schedule and de-risked.
- Retire on-premise footprint on a committed timeline, and report progress, risk, and dependencies to executive leadership in plain business terms.
Platform Engineering & Developer Experience
- Run the cloud platform as an internal product with product teams as its customers — not as a ticket queue.
- Deliver golden paths for service creation, CI/CD, secrets, observability, and security through reusable templates, modules, and self-service tooling, with sensible escape hatches for teams with genuine edge cases.
- Own GitLab-based source control and CI/CD standards, Kubernetes manifests, and GitOps delivery via Argo CD.
- Drive infrastructure as code and policy as code as the only way infrastructure changes, eliminating manual configuration and drift.
- Advance containerized microservices and serverless patterns on AWS (EKS, Lambda, DynamoDB, Secrets Manager, ALB/NLB) as the default for new standalone services.
- Measure and improve developer experience using DORA metrics, time-to-first-deploy, platform adoption, and direct developer feedback.
- Partner with the enterprise workflow effort to shift manual, human-in-the-loop processes to software-defined workflows (Temporal), removing people from steps that do not need them.
Reliability, Observability & Operations
- Establish the SRE charter: SLIs, SLOs, error budgets, incident response, blameless postmortems, and disaster recovery with tested RTO/RPO targets.
- Consolidate fragmented monitoring into a unified observability platform across metrics, logs, traces, and alerting, standardizing instrumentation on OpenTelemetry and structured logging.
- Lead observability platform direction and rationalization (New Relic as the strategic direction alongside existing Datadog usage), including agent fleet management, telemetry governance, and cost control.
- Move the organization from reactive monitoring to proactive, data-driven observability across business uptime, engineering excellence, and digital experience, with dashboards that serve both engineers and executives.
- Drive down MTTD and MTTR, and tie reliability goals directly to customer experience and business KPIs.
Security, Networking & Resilience
- Embed security and compliance by design: least-privilege IAM, secrets management, encryption in transit and at rest, and continuous control monitoring.
- Own cloud network architecture end to end — VPC design, segmentation, transit connectivity, private link patterns, DNS, and native load balancing (ALB/NLB).
- Operate edge protection with Cloudflare for WAF, DDoS mitigation, bot management, and DNS, in partnership with the security organization.
- Integrate enterprise identity into platform access patterns (Okta for external user authentication, Microsoft Entra for internal).
- Partner with the CISO organization on cloud security posture, audit readiness, and operational risk management.
FinOps & Vendor Management
- Establish cloud cost transparency, tagging discipline, showback/chargeback, and forecasting the business can trust.
- Build unit economics (cost per workload, per environment, per client) and run a standing optimization backlog that lowers total cost of ownership as consumption grows.
- Manage cloud, observability, and tooling vendor relationships, contracts, commitments, and renewals.
Leadership & Influence
- Build, lead, and develop a high-performing global organization spanning cloud engineering, platform engineering, SRE, and network engineering.
- Set the engineering bar through architecture review, code review, and hands-on participation — credibility at this level comes from doing the work, not describing it.
- Define clear ownership boundaries between platform and product teams, and the engagement model for prioritization, funding, and decision-making.
- Translate complex technical trade-offs for technical, non-technical, and executive audiences, including CIO and CDTO staff.
- Champion adoption of AI-assisted engineering and AI-driven operations, and ensure the platform supports enterprise AI, GenAI, and analytics workloads.
What You Will Need
- 12+ years of progressive experience in cloud, infrastructure, or platform engineering, with at least 5 years leading engineers and engineering managers.
- Bachelor’s degree in Computer Science, Engineering, or a related field required; advanced degree preferred, equivalent professional experience considered.
- Demonstrated success leading a large-scale on-premise to public cloud migration or data center exit in a complex, matrixed organization — owned it, not adjacent to it.
- Proven experience building or operating an internal developer platform at scale, with real adoption to show for it.
- Genuinely hands-on: you can still read and write the Terraform, the pipeline config, and the Kubernetes manifests — and in this role you will.
- Deep expertise in Kubernetes (EKS strongly preferred) and GitOps delivery patterns.
- Strong cloud networking depth — VPC architecture, segmentation, native load balancing, DNS, and hybrid connectivity.
- Experience with cloud security and edge protection, including WAF and DDoS prevention.
- Track record of driving adoption and non-trivial organizational change, with comfort operating in ambiguity.
- Ability to convey complex technical information to technical, non-technical, and executive audiences.
- Preferred tech stack experience includes:
- Cloud: AWS-first (not AWS-only); Azure for internal IT and Power BI workloads
- Containers & Orchestration: Kubernetes/EKS, Argo CD, Helm, serverless patterns (Lambda)
- CI/CD & IaC: GitLab source control and CI/CD, Terraform, policy as code
- AWS Services: DynamoDB, Secrets Manager, ALB/NLB, IAM, S3, CloudWatch
- Networking & Edge: Cloudflare (WAF, DDoS, DNS), VPC design, transit and private connectivity
- Data: PostgreSQL (relational standard), Snowflake (enterprise data warehouse), Power BI
- Observability: New Relic, Datadog, OpenTelemetry
- Workflow & Integration: Temporal, API-first service design, ServiceNow
- Identity: Okta (external), Microsoft Entra (internal)
- Languages: Python primary; polyglot where the use case warrants
Compensation: Salary range is $220,000-$230,000 USD annually.
Salary rates are based on experience, skills, and geographic location.
What Magnit will Offer You
At Magnit,you’ll be joining an innovative, high-growth environment and can quickly make an impact to help transform the largest companies in the world. You will work with passionate colleagues who collaborate and deliver. Magnit offers all employees the opportunity for growth and development, and we want individuals to fulfill their potential and blaze their own trails!
Magnit will offer you a competitive benefits package, including unlimited PTO, medical, dental, and vision coverage, retirement planning, as well as discounts and perks for tickets, travel, merchandise and more! Magnit encourages employees to participate in giving back, and we will match employee contributions to favorite charities and support corporate volunteering hours to make a difference in your community!
If this role isn’t for you
Stay in touch, we will let you know when we have new positions on the team. To see a complete list of our open career opportunities please visit
https://magnitglobal.com/us/en/company/careers.html
To do our best work we need different viewpoints. Therefore, we celebrate diversity and embrace inclusion.
As an equal opportunity employer, we are dedicated to building a team that represents a variety of backgrounds, perspectives, and skills. We strive to ensure that we maintain a positive and enriching work environment for all.
By applying to this role, you consent to Magnit safely storing and managing your personal data. Please read this link to learn more.
https://magnitglobal.com/us/en/privacy-notice.html