Please mention DailyRemote when applying
See how much of this job your resume covers, and what’s missing.
Want a recruiter to go through it line by line?
Get professional reviewQuestions interviewers often ask for this role, with sample answers.
Upload your resume and we draft a letter for this exact role, tailored to what it asks for.
You will design, deploy, and manage cloud infrastructure while building internal DevOps tools to improve efficiency and reduce operational costs. The role involves managing web applications, implementing security best practices, and participating in on-call rotations for platform reliability.
Managing, and scaling web applications and data platforms
Building tools to implement devops and security best practices
Creating reusable and immutable infrastructure with Terraform
Continuously improve our infrastructure with monitoring, logging, and alerting
Developing authentication and gateway solutions for our infrastructure and applications
Investigating, identifying application issues and advising development teams on design, deployment and infrastructure choices.
Participating in on-call rotations, post-mortems and root cause analysis (RCA)
5+ years of professional experience
Take full ownership of significant system components with responsibility for their reliability and performance
Manage Lifecycle of core project infrastructure, from design through to deployment, maintenance, and performance optimization
Familiar with industry standards and devops best practices
Experience supporting the overall platform in on-call rotations.
Able to operate with minimal supervision
Managing deployment infrastructure and automations including Github, CI/CD pipelines,and other deployment tooling
Experience managing workflows and pipelines using tools like CircleCI, Argo Workflows etc
Cloud computing providers such as AWS
Hands-on experience with Kubernetes (EKS, GKE) or similar container orchestration platforms
Infrastructure as code tools such as Terraform
Experience coding with at least one language such as Bash, Python required
Hands-on experience with authentication and authorization technologies required
Automation experience of cloud environments
Containerization technologies and tools such as Docker
Monitoring tools such as Datadog, Opensearch, Sentry
Good at using AI agents and writing project specific standard guidelines
Experience operating data pipelines with Databricks, Kafka
Experience with Java, Javascript
Experience with managing infrastructure costs and budgets
Experience with databases like AWS RDS/Postgres
Stop the endless job search. Our AI finds and applies to the best jobs for you.
Featuring 212,130+ Jobs in Site Reliability Engineer
Answer easy questions
212,130+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”