For Employers

CDIT

Principal Dev-Ops Architect

Posted 8 days ago
Worldwide
10+ years experience
Apply Now

Please mention DailyRemote when applying

?/100
Resume Match Score

Match your resume skills with our AI powered skill match!

Get professional review

Questions interviewers often ask for this role, with sample answers.

Create a cover letter for this job

Upload your resume and we draft a letter for this exact role, tailored to what it asks for.

  • Tailored to this role
  • Based on your resume
  • Fully editable
AI Summary

The architect owns the platform architecture, technical roadmap, and engineering standards for infrastructure, deployment, and AI/ML platforms. They are responsible for hands-on implementation of IaC, CI/CD pipelines, and ensuring the environment remains compliant with HIPAA and SOC 2 standards.

This is a remote position.

Senior technical authority for a cloud platform running global, multi-tenant SaaS services. This is a hands-on individual-contributor architect role — not people management. The architect designs the platform, sets standards and reference implementations other teams build on, and still writes Terraform, builds CI/CD pipelines, and stands up the AI/ML platform personally. Defines how reliability is measured against SLOs, how releases ship, and how the AI/ML platform is built and governed while keeping the environment HIPAA-compliant and SOC 2 Type 2 audit-ready. Influences products, software, and QA through architecture and example.

Core Responsibilities

     Own platform architecture and technical roadmap for infrastructure, deployment, observability, and the AI/ML platform.

     Set engineering standards, patterns, and golden paths for Infrastructure as Code (IaC), CI/CD, and AI tooling; drive adoption through reference implementations and architecture reviews.

     Manage all cloud infrastructure as code in Terraform — reusable modules, remote state, peer-reviewed PRs, drift detection, and automated plan/apply in CI/CD.

     Enforce policy-as-code (OPA, Sentinel, or equivalent) so infrastructure changes meet security and cost guardrails before merge.

     Design and deploy AWS infrastructure across dev, UAT, staging, and production for performance, availability, recoverability, and security (CIS Critical Security Controls).

     Build and operate CI/CD pipelines for large-scale applications on AWS; own release management, rollback, blue/green, canary, and release gates.

     Package and run containerized workloads on Docker and Kubernetes (EKS).

     Lead the SLI/SLO/SLA program and modern observability using OpenTelemetry; drive down MTTD and MTTR; lead blameless post-incident reviews and participate in on-call.

     Provision and operate the AI/ML platform — Anthropic Claude via AWS Bedrock and internal MCP services — all managed as IaC.

     Build LLMOps practices: prompt versioning, evaluation pipelines, token cost attribution, guardrails, and audit logging of agent actions; enforce the PHI data boundary to BAA-covered providers only.

     Operate and evidence the platform controls required for SOC 2 Type 2 and HIPAA; own secrets management, supply-chain security (SBOM, image and dependency scanning), and FinOps.

Required Qualifications

     Bachelor's degree in Software Engineering or equivalent combination of technical education and work experience.

     10+ years in SRE / DevOps / Platform Engineering delivering CI/CD, REST API deployment, containerization, IaaS/PaaS, data pipelines, and application observability — including time at a senior IC or architect level (Staff, Principal, or Architect).

     Proven technical authority across teams: sets architecture and standards and influences delivery through expertise and example rather than direct management.

     Demonstrated experience driving adoption of a new practice or platform (IaC, CI/CD overhaul, or an AI/ML platform) across multiple teams.

     Hands-on Terraform, including reusable modules other teams consume via self-service, remote state, and change management in a CI/CD pipeline.

     Building and operating CI/CD pipelines for large-scale applications on AWS (GitHub Actions, Jenkins, GitLab, or AWS-native).

     Running containerized workloads on Docker and Kubernetes.

     Monitoring and troubleshooting using cloud-native tooling and OpenTelemetry.

     Linux system administration, Unix scripting, and automation.

     Experience working in a HIPAA / HITECH / HITRUST / PHI / PII or PCI DSS environment.



Automatically Apply to the Best Remote Jobs

Stop the endless job search. Our AI finds and applies to the best jobs for you.

Try it Now
Keep looking

Similar Jobs

See all Remote Software Development jobs →

AI-LLM Systems Engineer

Freelance Worldwide $2000 - $3000 per month Software Development

Senior Site Reliability Engineer

Full Time United States $160K - $208K per year Software Development

Senior Product Analytics Engineer

Full Time United Kingdom Software Development

Senior Software Engineer

Full Time France Software Development

Senior Full-Stack Product Engineer

Full Time France Software Development

Staff Software Engineer, Frontend

Full Time India Software Development
Apply Now

Personalize your Remote Job Search in 3 Easy Steps!

Featuring 214,939+ Jobs in Architect

Answer easy questions

Answer easy questions

214,939+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified