The DevOps Engineer will design, implement, and maintain secure, scalable cloud infrastructure while building robust CI/CD pipelines. They will also partner with engineering teams to improve observability, automate deployment processes, and ensure high availability of production systems.
We are seeking a hands-on DevOps Engineer to improve the reliability, scalability, security, and delivery performance of our SaaS platform. This role will own and improve the systems that allow engineering teams to build, test, deploy, observe, and operate software safely. The DevOps Engineer will partner closely with application engineers and technical leadership to automate infrastructure, strengthen CI/CD, improve environment consistency, increase observability, reduce operational risk, and support production systems. This is not a ticket-based infrastructure support role. We are looking for an engineer who treats the developer platform and production environment as products, proactively identifies sources of friction and risk, and implements sustainable improvements.
Location: LATAM 100% Remote. Working hours are based on the US Central-Pacific Time Zone, with at least 6 working hours of overlap.
About Us:
Abstra is a fast-growing, Nearshore Tech Talent services company, providing top Latin American tech talent to U.S. companies and beyond. Founded by U.S.-bred engineers with over 15 years of experience, Abstra specializes in sourcing skilled professionals across a wide range of technologies to meet our clients’ needs, driving innovation and efficiency.
Key Responsibilities
- Design, implement, and maintain secure, reliable, and scalable cloud infrastructure.
- Build and improve CI/CD pipelines that support frequent, repeatable, and low-risk deployments.
- Automate infrastructure provisioning, configuration, deployment, and environment management.
- Maintain infrastructure as code using tools such as Terraform, CloudFormation, Pulumi, or comparable technologies.
- Improve consistency across development, testing, staging, and production environments.
- Build and maintain containerized application environments and orchestration capabilities.
- Implement and improve monitoring, logging, tracing, alerting, dashboards, and production-health reporting.
- Partner with engineers to establish service-level indicators, service-level objectives, and actionable operational metrics.
- Improve application and infrastructure resiliency, scalability, availability, and disaster-recovery readiness.
- Manage secrets, certificates, identity, permissions, network controls, and cloud security configurations.
- Support vulnerability management, dependency security, patching, and infrastructure hardening.
- Improve deployment strategies, including automated rollback, blue-green deployment, canary releases, or comparable approaches where appropriate.
- Diagnose and resolve production incidents involving infrastructure, networking, application deployment, performance, capacity, or cloud services.
- Participate in incident response, post-incident reviews, root-cause analysis, and corrective-action planning.
- Reduce cloud waste and improve infrastructure cost visibility.
- Build self-service tools and reusable deployment patterns for application engineering teams.
- Document infrastructure architecture, operational procedures, recovery processes, and troubleshooting guidance.
- Participate in production support and an appropriate on-call rotation.
- Use AI-assisted engineering and operational tools to accelerate scripting, troubleshooting, documentation, and infrastructure analysis while applying appropriate validation and security controls.
Required Technical Qualifications
- Approximately 5 or more years of experience in DevOps, Site Reliability Engineering, Cloud Engineering, Platform Engineering, or a closely related role.
- Strong experience operating production workloads in AWS.
- Strong hands-on experience with infrastructure as code.
- Experience building and maintaining CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, CircleCI, AWS CodePipeline, or comparable tools.
- Strong knowledge of Docker and containerized application delivery.
- Experience with Kubernetes, Amazon ECS, or another container orchestration platform.
- Strong understanding of cloud networking, including virtual networks, subnets, routing, load balancers, DNS, firewalls, gateways, and private connectivity.
- Experience with cloud identity and access management, role-based access, least-privilege design, and secrets management.
- Experience implementing logging, metrics, dashboards, alerts, and distributed tracing.
- Familiarity with observability platforms such as Datadog, New Relic, Grafana, Prometheus, CloudWatch, OpenTelemetry, or comparable technologies.
- Strong Linux and command-line skills.
- Scripting experience with Python, Bash, JavaScript, or a comparable language.
- Understanding of application deployment patterns for modern web applications, APIs, background workers, and event-driven systems.
- Experience supporting relational databases, backups, restore processes, high availability, and database operational health.
- Knowledge of incident-management practices, root-cause analysis, capacity planning, and production-readiness reviews.
- Working knowledge of cloud security, vulnerability remediation, encryption, certificates, key management, and compliance-oriented controls.
Preferred Technical Experience
- Experience supporting a multi-tenant SaaS platform.
- Experience with serverless infrastructure and event-driven cloud services.
- Experience with PostgreSQL and managed database services.
- Experience with content delivery networks, media delivery, large-file processing, video workloads, or storage-intensive applications.
- Experience with security and compliance frameworks relevant to enterprise SaaS.
- Experience building internal developer platforms or self-service engineering capabilities.
- Experience with automated performance, resilience, chaos, or disaster-recovery testing.
- Experience improving cloud cost allocation, forecasting, and optimization.
- Experience supporting globally distributed engineering teams.
DevOps-Level Expectations
A successful DevOps Engineer should be able to:
- Independently own infrastructure and delivery improvements from design through implementation.
- Diagnose production issues across cloud infrastructure, networking, deployment pipelines, containers, databases, and application runtime behavior.
- Replace repetitive manual processes with reliable automation.
- Identify operational risks before they result in incidents.
- Build guardrails that improve engineering speed without sacrificing security or stability.
- Explain infrastructure and production concerns clearly to application engineers.
- Produce actionable alerts rather than excessive operational noise.
- Balance availability, security, developer productivity, and cloud cost.
- Treat incident follow-up and permanent corrective action as part of the role, not optional work.
What We Offer
- Competitive compensation paid in USD.
- 20 days of paid time off (PTO) per year.
- Opportunities for professional growth and career development.
- Company-provided equipment.
- A collaborative, inclusive, and multicultural work environment.
- The opportunity to contribute to meaningful projects alongside a talented and supportive team.
Pre-Employment Verification
As part of our standard onboarding process, candidates who successfully complete the interview process and accept an employment offer will be required to complete an employment verification check, and background check. This process will confirm job titles and dates of employment with two previous employers and is a standard requirement for all new employees joining the company.