DevOps & Infrastructure Engineer (Part Time)

 Posted 8 hours ago
     
5-10 years experience
Apply Now

Please mention DailyRemote when applying

AI Summary

The engineer will own the AWS infrastructure, ensuring reliability, security, and scalability for webinar products. They are responsible for managing infrastructure as code, CI/CD pipelines, incident response, and observability to support a growing engineering team.

DevOps & Infrastructure Engineer (Part-Time)

About the Role

The DevOps & Infrastructure Engineer owns the infrastructure that WebinarJam and EverWebinar run on. This role is responsible for the reliability, security, and scalability of the AWS environment serving every live and automated webinar our customers run, and for building the automation, monitoring, and deployment practices that allow a small engineering team to ship quickly and safely.Working closely with Engineering, Product, Leadership, and Customer Support, this individual will take full ownership of the platform layer for both products, and will build the operational foundation the company grows on. Success in this role means infrastructure that is predictable and observable, where failures are rare, recovery is fast, and engineers can deploy with confidence.


Key Responsibilities

Infrastructure Ownership

  • Own the AWS environment end to end, including compute, networking, storage, and cost.
  • Lead the separation of WebinarJam's infrastructure into its own accounts, with independent billing, access, and governance.
  • Maintain and evolve infrastructure as code across CloudFormation and CDK.
  • Manage third-party infrastructure services, including CDN, DNS, identity, and certificate management.
  • Plan capacity and scaling for live events, including peak and burst demand.

Reliability & Incident Response

  • Serve as the primary responder for infrastructure incidents.
  • Build and maintain alerting and paging so issues reach the right person quickly.
  • Establish incident response practices, including escalation paths, severity definitions, and post-incident reviews.
  • Reduce the frequency and duration of customer-facing outages.
  • Build and regularly test disaster recovery and backup restoration procedures.

Deployment & Automation

  • Complete and deploy the platform's containerization initiative.
  • Own CI/CD pipelines and improve build reliability and speed.
  • Replace manual deployment steps with automated, repeatable, reversible releases.
  • Enable safe rollback and progressive rollout of changes.

Observability

  • Own monitoring, logging, and tracing across the platform.
  • Build dashboards and alerts that give the team visibility into system and customer-facing health.
  • Introduce distributed tracing to shorten diagnosis time for complex failures.
  • Make performance and reliability measurable rather than anecdotal.

Security & Access

  • Manage credentials, secrets, and access control across all infrastructure and vendor accounts.
  • Maintain least-privilege access and audit it regularly.
  • Keep systems patched and machine images current.
  • Support compliance and data protection requirements as they arise.

Collaboration & Documentation

  • Partner with Engineering on architecture decisions and technical feasibility.
  • Document infrastructure, runbooks, and operational procedures so knowledge does not sit with one person.
  • Advise Leadership on infrastructure cost, risk, and resourcing.
  • Support Product and Customer Support in diagnosing platform-related customer issues.

Technical Environment

  • AWS: EC2, autoscaling groups and warm pools, ALB, CloudFormation, CDK, EC2 Image Builder, SSM, Lambda, API Gateway, DynamoDB, S3, CloudFront, SQS, ElastiCache, Aurora
  • Containers and orchestration: Docker, Kubernetes
  • CI/CD: GitHub Actions
  • Observability: CloudWatch, Loki, Grafana, OpenTelemetry
  • Supporting services: Cloudflare, Auth0, Firebase, Vault
  • Application stack: Laravel (PHP), Vue, Node

What You Bring

  • Deep, hands-on AWS experience running production systems at scale.
  • Experience separating, migrating, or consolidating AWS accounts between organizations.
  • Strong infrastructure as code discipline.
  • Container and orchestration experience in production, not only in development.
  • Comfort auditing unfamiliar systems and inherited work, and reporting clearly on what you find.
  • Sound judgment operating independently, with a careful and communicative approach to production change.
  • Experience supporting real-time or media-heavy workloads is an advantage.

Success Looks Like

  • Infrastructure incidents are rare, and those that occur are detected and resolved quickly.
  • Deployments are automated, routine, and reversible.
  • The platform scales predictably through peak webinar demand.
  • WebinarJam operates fully independent infrastructure with clean ownership of every account and credential.
  • Engineers ship without waiting on manual infrastructure steps.
  • Infrastructure knowledge is documented rather than held by one person.
  • Infrastructure cost is understood, tracked, and optimized.

Similar Jobs

See all Remote Software Development jobs →

Personalize your Remote Job Search in 3 Easy Steps!

Discover remote opportunities in DevOps Engineer

Answer easy questions

Answer easy questions

200,000+ jobs across 15+ categories

Get your best job matches

Get your best job matches

Only hand-screened, legit jobs

Find a remote job faster

Find a remote job faster

No ads, scams, or junk

I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!

Sarah J. — Sarah J. · Marketing Manager ★★★★★ Verified